ASDnB: Merging Face with Body Cues For Robust Active Speaker Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Roxo, Tiago, Costa, Joana C., Inácio, Pedro, Proença, Hugo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BIAS: A Body-based Interpretable Active Speaker Approach
by: Roxo, Tiago, et al.
Published: (2024)
by: Roxo, Tiago, et al.
Published: (2024)
ZQBA: Zero Query Black-box Adversarial Attack
by: Costa, Joana C., et al.
Published: (2025)
by: Costa, Joana C., et al.
Published: (2025)
How Deep Learning Sees the World: A Survey on Adversarial Attacks & Defenses
by: Costa, Joana C., et al.
Published: (2023)
by: Costa, Joana C., et al.
Published: (2023)
How to Squeeze An Explanation Out of Your Model
by: Roxo, Tiago, et al.
Published: (2024)
by: Roxo, Tiago, et al.
Published: (2024)
LISArD: Learning Image Similarity to Defend Against Gray-box Adversarial Attacks
by: Costa, Joana C., et al.
Published: (2025)
by: Costa, Joana C., et al.
Published: (2025)
Are DeepFakes Realistic Enough? Exploring Semantic Mismatch as a Novel Challenge
by: Deshmukh, Sharayu Nilesh, et al.
Published: (2026)
by: Deshmukh, Sharayu Nilesh, et al.
Published: (2026)
SortWaste: A Densely Annotated Dataset for Object Detection in Industrial Waste Sorting
by: Inácio, Sara, et al.
Published: (2026)
by: Inácio, Sara, et al.
Published: (2026)
FD-MAD: Frequency-Domain Residual Analysis for Face Morphing Attack Detection
by: Paulo, Diogo J., et al.
Published: (2026)
by: Paulo, Diogo J., et al.
Published: (2026)
Bias Analysis for Synthetic Face Detection: A Case Study of the Impact of Facial Attributes
by: Lamsaf, Asmae, et al.
Published: (2025)
by: Lamsaf, Asmae, et al.
Published: (2025)
The Good, the Better, and the Best: Improving the Discriminability of Face Embeddings through Attribute-aware Learning
by: Dias, Ana, et al.
Published: (2026)
by: Dias, Ana, et al.
Published: (2026)
Learning to Discover Forgery Cues for Face Forgery Detection
by: Tian, Jiahe, et al.
Published: (2024)
by: Tian, Jiahe, et al.
Published: (2024)
Benchmarking Joint Face Spoofing and Forgery Detection with Visual and Physiological Cues
by: Yu, Zitong, et al.
Published: (2022)
by: Yu, Zitong, et al.
Published: (2022)
SRL-MAD: Structured Residual Latents for One-Class Morphing Attack Detection
by: Paulo, Diogo J., et al.
Published: (2026)
by: Paulo, Diogo J., et al.
Published: (2026)
Towards Zero-Shot Interpretable Human Recognition: A 2D-3D Registration Framework
by: Jesus, Henrique, et al.
Published: (2024)
by: Jesus, Henrique, et al.
Published: (2024)
Robust Active Speaker Detection in Noisy Environments
by: Vasireddy, Siva Sai Nagender, et al.
Published: (2024)
by: Vasireddy, Siva Sai Nagender, et al.
Published: (2024)
Emotion Detection through Body Gesture and Face
by: Liu, Haoyang
Published: (2024)
by: Liu, Haoyang
Published: (2024)
Exploiting Polarized Material Cues for Robust Car Detection
by: Dong, Wen, et al.
Published: (2024)
by: Dong, Wen, et al.
Published: (2024)
Generalization Under Scrutiny: Cross-Domain Detection Progresses, Pitfalls, and Persistent Challenges
by: Deshmukh, Saniya M., et al.
Published: (2026)
by: Deshmukh, Saniya M., et al.
Published: (2026)
Rectifying Geometry-Induced Similarity Distortions for Real-World Aerial-Ground Person Re-Identification
by: Hambarde, Kailash A., et al.
Published: (2026)
by: Hambarde, Kailash A., et al.
Published: (2026)
Seeing Across Time and Views: Multi-Temporal Cross-View Learning for Robust Video Person Re-Identification
by: Rashidunnabi, Md, et al.
Published: (2025)
by: Rashidunnabi, Md, et al.
Published: (2025)
Human Re-ID Meets LVLMs: What can we expect?
by: Hambarde, Kailash, et al.
Published: (2025)
by: Hambarde, Kailash, et al.
Published: (2025)
Causality and "In-the-Wild" Video-Based Person Re-ID: A Survey
by: Rashidunnabi, Md, et al.
Published: (2025)
by: Rashidunnabi, Md, et al.
Published: (2025)
RobustMerge: Parameter-Efficient Model Merging for MLLMs with Direction Robustness
by: Zeng, Fanhu, et al.
Published: (2025)
by: Zeng, Fanhu, et al.
Published: (2025)
GateFusion: Hierarchical Gated Cross-Modal Fusion for Active Speaker Detection
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
LoCoNet: Long-Short Context Network for Active Speaker Detection
by: Wang, Xizi, et al.
Published: (2023)
by: Wang, Xizi, et al.
Published: (2023)
FUSE: Unifying Spectral and Semantic Cues for Robust AI-Generated Image Detection
by: Hossain, Md. Zahid, et al.
Published: (2025)
by: Hossain, Md. Zahid, et al.
Published: (2025)
FA^{3}-CLIP: Frequency-Aware Cues Fusion and Attack-Agnostic Prompt Learning for Unified Face Attack Detection
by: Li, Yongze, et al.
Published: (2025)
by: Li, Yongze, et al.
Published: (2025)
Massively Annotated Datasets for Assessment of Synthetic and Real Data in Face Recognition
by: Neto, Pedro C., et al.
Published: (2024)
by: Neto, Pedro C., et al.
Published: (2024)
Detection of Synthetic Face Images: Accuracy, Robustness, Generalization
by: Petrzelkova, Nela, et al.
Published: (2024)
by: Petrzelkova, Nela, et al.
Published: (2024)
Robust AI-Generated Face Detection with Imbalanced Data
by: Krubha, Yamini Sri, et al.
Published: (2025)
by: Krubha, Yamini Sri, et al.
Published: (2025)
StreetView-Waste: A Multi-Task Dataset for Urban Waste Management
by: Paulo, Diogo J., et al.
Published: (2025)
by: Paulo, Diogo J., et al.
Published: (2025)
LASER: Lip Landmark Assisted Speaker Detection for Robustness
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
Lightweight Complementary-Cue Fusion for Robust Video Face Forgery Detection
by: Baek, Sunghwan, et al.
Published: (2026)
by: Baek, Sunghwan, et al.
Published: (2026)
Beyond Appearance: Geometric Cues for Robust Video Instance Segmentation
by: Niu, Quanzhu, et al.
Published: (2025)
by: Niu, Quanzhu, et al.
Published: (2025)
Robust Face Liveness Detection for Biometric Authentication using Single Image
by: Raha, Poulami, et al.
Published: (2025)
by: Raha, Poulami, et al.
Published: (2025)
RCDN: Real-Centered Detection Network for Robust Face Forgery Identification
by: McCurdy, Wyatt, et al.
Published: (2026)
by: McCurdy, Wyatt, et al.
Published: (2026)
EmoSpeaker: One-shot Fine-grained Emotion-Controlled Talking Face Generation
by: Feng, Guanwen, et al.
Published: (2024)
by: Feng, Guanwen, et al.
Published: (2024)
UniTalk: Towards Universal Active Speaker Detection in Real World Scenarios
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
An Efficient and Streaming Audio Visual Active Speaker Detection System
by: Kundu, Arnav, et al.
Published: (2024)
by: Kundu, Arnav, et al.
Published: (2024)
Depth-Copy-Paste: Multimodal and Depth-Aware Compositing for Robust Face Detection
by: Guo, Qiushi
Published: (2025)
by: Guo, Qiushi
Published: (2025)
Similar Items
-
BIAS: A Body-based Interpretable Active Speaker Approach
by: Roxo, Tiago, et al.
Published: (2024) -
ZQBA: Zero Query Black-box Adversarial Attack
by: Costa, Joana C., et al.
Published: (2025) -
How Deep Learning Sees the World: A Survey on Adversarial Attacks & Defenses
by: Costa, Joana C., et al.
Published: (2023) -
How to Squeeze An Explanation Out of Your Model
by: Roxo, Tiago, et al.
Published: (2024) -
LISArD: Learning Image Similarity to Defend Against Gray-box Adversarial Attacks
by: Costa, Joana C., et al.
Published: (2025)