Test-Time Augmentation for Pose-invariant Face Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jung, Jaemin, Jang, Youngjoon, Chung, Joon Son |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AVCD: Mitigating Hallucinations in Audio-Visual Large Language Models through Contrastive Decoding
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025)
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025)
Fork-Merge Decoding: Enhancing Multimodal Understanding in Audio-Visual Large Language Models
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025)
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025)
On the Nature of Attention Sink that Shapes Decoding Strategy in Omni-LLMs
von: Yoo, Suho, et al.
Veröffentlicht: (2026)
von: Yoo, Suho, et al.
Veröffentlicht: (2026)
Deep Understanding of Sign Language for Sign to Subtitle Alignment
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
Let Me Finish My Sentence: Video Temporal Grounding with Holistic Text Understanding
von: Woo, Jongbhin, et al.
Veröffentlicht: (2024)
von: Woo, Jongbhin, et al.
Veröffentlicht: (2024)
Lost in Translation, Found in Embeddings: Sign Language Translation and Alignment
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
Keep What Audio Cannot Say: Context-Preserving Token Pruning for Omni-LLMs
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2026)
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2026)
Towards Large-Scale Pose-Invariant Face Recognition Using Face Defrontalization
von: Mesec, Patrik, et al.
Veröffentlicht: (2025)
von: Mesec, Patrik, et al.
Veröffentlicht: (2025)
Unified Negative Pair Generation toward Well-discriminative Feature Space for Face Recognition
von: Jung, Junuk, et al.
Veröffentlicht: (2022)
von: Jung, Junuk, et al.
Veröffentlicht: (2022)
Domain-generalizable Face Anti-Spoofing with Patch-based Multi-tasking and Artifact Pattern Conversion
von: Jung, Seungjin, et al.
Veröffentlicht: (2026)
von: Jung, Seungjin, et al.
Veröffentlicht: (2026)
Faces that Speak: Jointly Synthesising Talking Face and Speech from Text
von: Jang, Youngjoon, et al.
Veröffentlicht: (2024)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2024)
From Faces to Voices: Learning Hierarchical Representations for High-quality Video-to-Speech
von: Kim, Ji-Hoon, et al.
Veröffentlicht: (2025)
von: Kim, Ji-Hoon, et al.
Veröffentlicht: (2025)
KRAST: Knowledge-Augmented Robotic Action Recognition with Structured Text for Vision-Language Models
von: Nguyen, Son Hai, et al.
Veröffentlicht: (2025)
von: Nguyen, Son Hai, et al.
Veröffentlicht: (2025)
Testing the Performance of Face Recognition for People with Down Syndrome
von: Rathgeb, Christian, et al.
Veröffentlicht: (2024)
von: Rathgeb, Christian, et al.
Veröffentlicht: (2024)
Pose-dIVE: Pose-Diversified Augmentation with Diffusion Model for Person Re-Identification
von: Kim, Inès Hyeonsu, et al.
Veröffentlicht: (2024)
von: Kim, Inès Hyeonsu, et al.
Veröffentlicht: (2024)
Ranked Entropy Minimization for Continual Test-Time Adaptation
von: Han, Jisu, et al.
Veröffentlicht: (2025)
von: Han, Jisu, et al.
Veröffentlicht: (2025)
SqueezeFacePoseNet: Lightweight Face Verification Across Different Poses for Mobile Platforms
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2020)
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2020)
Advancing Cross-Domain Generalizability in Face Anti-Spoofing: Insights, Design, and Metrics
von: Kim, Hyojin, et al.
Veröffentlicht: (2024)
von: Kim, Hyojin, et al.
Veröffentlicht: (2024)
Full-range Head Pose Geometric Data Augmentations
von: Hu, Huei-Chung, et al.
Veröffentlicht: (2024)
von: Hu, Huei-Chung, et al.
Veröffentlicht: (2024)
Improving the Transferability of Adversarial Attacks on Face Recognition with Diverse Parameters Augmentation
von: Zhou, Fengfan, et al.
Veröffentlicht: (2024)
von: Zhou, Fengfan, et al.
Veröffentlicht: (2024)
Protego: User-Centric Pose-Invariant Privacy Protection Against Face Recognition-Induced Digital Footprint Exposure
von: Wang, Ziling, et al.
Veröffentlicht: (2025)
von: Wang, Ziling, et al.
Veröffentlicht: (2025)
Revisiting Misalignment in Multispectral Pedestrian Detection: A Language-Driven Approach for Cross-modal Alignment Fusion
von: Kim, Taeheon, et al.
Veröffentlicht: (2024)
von: Kim, Taeheon, et al.
Veröffentlicht: (2024)
Label Distribution Shift-Aware Prediction Refinement for Test-Time Adaptation
von: Jang, Minguk, et al.
Veröffentlicht: (2024)
von: Jang, Minguk, et al.
Veröffentlicht: (2024)
Test-Time Domain Generalization for Face Anti-Spoofing
von: Zhou, Qianyu, et al.
Veröffentlicht: (2024)
von: Zhou, Qianyu, et al.
Veröffentlicht: (2024)
Seeing Through Touch: Tactile-Driven Visual Localization of Material Regions
von: Kim, Seongyu, et al.
Veröffentlicht: (2026)
von: Kim, Seongyu, et al.
Veröffentlicht: (2026)
Lost in Translation, Found in Context: Sign Language Translation with Contextual Cues
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
Index-Preserving Lightweight Token Pruning for Efficient Document Understanding in Vision-Language Models
von: Son, Jaemin, et al.
Veröffentlicht: (2025)
von: Son, Jaemin, et al.
Veröffentlicht: (2025)
AVHBench: A Cross-Modal Hallucination Benchmark for Audio-Visual Large Language Models
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
von: Sung-Bin, Kim, et al.
Veröffentlicht: (2024)
One-Shot Pose-Driving Face Animation Platform
von: Feng, He, et al.
Veröffentlicht: (2024)
von: Feng, He, et al.
Veröffentlicht: (2024)
SPARK: Multi-Vision Sensor Perception and Reasoning Benchmark for Large-scale Vision-Language Models
von: Yu, Youngjoon, et al.
Veröffentlicht: (2024)
von: Yu, Youngjoon, et al.
Veröffentlicht: (2024)
ControlFace: Harnessing Facial Parametric Control for Face Rigging
von: Jang, Wooseok, et al.
Veröffentlicht: (2024)
von: Jang, Wooseok, et al.
Veröffentlicht: (2024)
Hearing and Seeing Through CLIP: A Framework for Self-Supervised Sound Source Localization
von: Park, Sooyoung, et al.
Veröffentlicht: (2025)
von: Park, Sooyoung, et al.
Veröffentlicht: (2025)
FunFace: Feature Utility and Norm Estimation for Face Recognition
von: Babnik, Žiga, et al.
Veröffentlicht: (2026)
von: Babnik, Žiga, et al.
Veröffentlicht: (2026)
Feature Augmentation based Test-Time Adaptation
von: Cho, Younggeol, et al.
Veröffentlicht: (2024)
von: Cho, Younggeol, et al.
Veröffentlicht: (2024)
EFHQ: Multi-purpose ExtremePose-Face-HQ dataset
von: Dao, Trung Tuan, et al.
Veröffentlicht: (2023)
von: Dao, Trung Tuan, et al.
Veröffentlicht: (2023)
A Rapid Test for Accuracy and Bias of Face Recognition Technology
von: Knott, Manuel, et al.
Veröffentlicht: (2025)
von: Knott, Manuel, et al.
Veröffentlicht: (2025)
HyperFace: Generating Synthetic Face Recognition Datasets by Exploring Face Embedding Hypersphere
von: Shahreza, Hatef Otroshi, et al.
Veröffentlicht: (2024)
von: Shahreza, Hatef Otroshi, et al.
Veröffentlicht: (2024)
CemiFace: Center-based Semi-hard Synthetic Face Generation for Face Recognition
von: Sun, Zhonglin, et al.
Veröffentlicht: (2024)
von: Sun, Zhonglin, et al.
Veröffentlicht: (2024)
Collaborative Face Experts Fusion in Video Generation: Boosting Identity Consistency Across Large Face Poses
von: Wang, Yuji, et al.
Veröffentlicht: (2025)
von: Wang, Yuji, et al.
Veröffentlicht: (2025)
MambaVideo for Discrete Video Tokenization with Channel-Split Quantization
von: Argaw, Dawit Mureja, et al.
Veröffentlicht: (2025)
von: Argaw, Dawit Mureja, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AVCD: Mitigating Hallucinations in Audio-Visual Large Language Models through Contrastive Decoding
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025) -
Fork-Merge Decoding: Enhancing Multimodal Understanding in Audio-Visual Large Language Models
von: Jung, Chaeyoung, et al.
Veröffentlicht: (2025) -
On the Nature of Attention Sink that Shapes Decoding Strategy in Omni-LLMs
von: Yoo, Suho, et al.
Veröffentlicht: (2026) -
Deep Understanding of Sign Language for Sign to Subtitle Alignment
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025) -
Let Me Finish My Sentence: Video Temporal Grounding with Holistic Text Understanding
von: Woo, Jongbhin, et al.
Veröffentlicht: (2024)