BIAS: A Body-based Interpretable Active Speaker Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Roxo, Tiago, Costa, Joana C., Inácio, Pedro R. M., Proença, Hugo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ASDnB: Merging Face with Body Cues For Robust Active Speaker Detection
by: Roxo, Tiago, et al.
Published: (2024)
by: Roxo, Tiago, et al.
Published: (2024)
How Deep Learning Sees the World: A Survey on Adversarial Attacks & Defenses
by: Costa, Joana C., et al.
Published: (2023)
by: Costa, Joana C., et al.
Published: (2023)
ZQBA: Zero Query Black-box Adversarial Attack
by: Costa, Joana C., et al.
Published: (2025)
by: Costa, Joana C., et al.
Published: (2025)
How to Squeeze An Explanation Out of Your Model
by: Roxo, Tiago, et al.
Published: (2024)
by: Roxo, Tiago, et al.
Published: (2024)
LISArD: Learning Image Similarity to Defend Against Gray-box Adversarial Attacks
by: Costa, Joana C., et al.
Published: (2025)
by: Costa, Joana C., et al.
Published: (2025)
Are DeepFakes Realistic Enough? Exploring Semantic Mismatch as a Novel Challenge
by: Deshmukh, Sharayu Nilesh, et al.
Published: (2026)
by: Deshmukh, Sharayu Nilesh, et al.
Published: (2026)
SortWaste: A Densely Annotated Dataset for Object Detection in Industrial Waste Sorting
by: Inácio, Sara, et al.
Published: (2026)
by: Inácio, Sara, et al.
Published: (2026)
Towards Zero-Shot Interpretable Human Recognition: A 2D-3D Registration Framework
by: Jesus, Henrique, et al.
Published: (2024)
by: Jesus, Henrique, et al.
Published: (2024)
BIAS: A Biologically Inspired Algorithm for Video Saliency Detection
by: Zhang, Zhao-ji, et al.
Published: (2026)
by: Zhang, Zhao-ji, et al.
Published: (2026)
Rectifying Geometry-Induced Similarity Distortions for Real-World Aerial-Ground Person Re-Identification
by: Hambarde, Kailash A., et al.
Published: (2026)
by: Hambarde, Kailash A., et al.
Published: (2026)
Causality and "In-the-Wild" Video-Based Person Re-ID: A Survey
by: Rashidunnabi, Md, et al.
Published: (2025)
by: Rashidunnabi, Md, et al.
Published: (2025)
Human Re-ID Meets LVLMs: What can we expect?
by: Hambarde, Kailash, et al.
Published: (2025)
by: Hambarde, Kailash, et al.
Published: (2025)
SRL-MAD: Structured Residual Latents for One-Class Morphing Attack Detection
by: Paulo, Diogo J., et al.
Published: (2026)
by: Paulo, Diogo J., et al.
Published: (2026)
FD-MAD: Frequency-Domain Residual Analysis for Face Morphing Attack Detection
by: Paulo, Diogo J., et al.
Published: (2026)
by: Paulo, Diogo J., et al.
Published: (2026)
Generalization Under Scrutiny: Cross-Domain Detection Progresses, Pitfalls, and Persistent Challenges
by: Deshmukh, Saniya M., et al.
Published: (2026)
by: Deshmukh, Saniya M., et al.
Published: (2026)
StreetView-Waste: A Multi-Task Dataset for Urban Waste Management
by: Paulo, Diogo J., et al.
Published: (2025)
by: Paulo, Diogo J., et al.
Published: (2025)
Bias Analysis for Synthetic Face Detection: A Case Study of the Impact of Facial Attributes
by: Lamsaf, Asmae, et al.
Published: (2025)
by: Lamsaf, Asmae, et al.
Published: (2025)
The Good, the Better, and the Best: Improving the Discriminability of Face Embeddings through Attribute-aware Learning
by: Dias, Ana, et al.
Published: (2026)
by: Dias, Ana, et al.
Published: (2026)
GateFusion: Hierarchical Gated Cross-Modal Fusion for Active Speaker Detection
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
Seeing Across Time and Views: Multi-Temporal Cross-View Learning for Robust Video Person Re-Identification
by: Rashidunnabi, Md, et al.
Published: (2025)
by: Rashidunnabi, Md, et al.
Published: (2025)
LoCoNet: Long-Short Context Network for Active Speaker Detection
by: Wang, Xizi, et al.
Published: (2023)
by: Wang, Xizi, et al.
Published: (2023)
A Super-pixel-based Approach to the Stable Interpretation of Neural Networks
by: Gong, Shizhan, et al.
Published: (2024)
by: Gong, Shizhan, et al.
Published: (2024)
A Deep Learning-based Global and Segmentation-based Semantic Feature Fusion Approach for Indoor Scene Classification
by: Pereira, Ricardo, et al.
Published: (2023)
by: Pereira, Ricardo, et al.
Published: (2023)
Approaching Test Time Augmentation in the Context of Uncertainty Calibration for Deep Neural Networks
by: Conde, Pedro, et al.
Published: (2023)
by: Conde, Pedro, et al.
Published: (2023)
UniTalk: Towards Universal Active Speaker Detection in Real World Scenarios
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
An Efficient and Streaming Audio Visual Active Speaker Detection System
by: Kundu, Arnav, et al.
Published: (2024)
by: Kundu, Arnav, et al.
Published: (2024)
Robust Active Speaker Detection in Noisy Environments
by: Vasireddy, Siva Sai Nagender, et al.
Published: (2024)
by: Vasireddy, Siva Sai Nagender, et al.
Published: (2024)
Deep Model Interpretation with Limited Data : A Coreset-based Approach
by: Behzadi-Khormouji, Hamed, et al.
Published: (2024)
by: Behzadi-Khormouji, Hamed, et al.
Published: (2024)
Similarity-as-Evidence: Calibrating Overconfident VLMs for Interpretable and Label-Efficient Medical Active Learning
by: Xie, Zhuofan, et al.
Published: (2026)
by: Xie, Zhuofan, et al.
Published: (2026)
Causal Interpretability for Adversarial Robustness: A Hybrid Generative Classification Approach
by: Zhao, Chunheng, et al.
Published: (2024)
by: Zhao, Chunheng, et al.
Published: (2024)
Massively Annotated Datasets for Assessment of Synthetic and Real Data in Face Recognition
by: Neto, Pedro C., et al.
Published: (2024)
by: Neto, Pedro C., et al.
Published: (2024)
A Unified Approach Towards Active Learning and Out-of-Distribution Detection
by: Schmidt, Sebastian, et al.
Published: (2024)
by: Schmidt, Sebastian, et al.
Published: (2024)
ReTracing: An Archaeological Approach Through Body, Machine, and Generative Systems
by: Wang, Yitong, et al.
Published: (2026)
by: Wang, Yitong, et al.
Published: (2026)
Rethinking Whole-Body CT Image Interpretation: An Abnormality-Centric Approach
by: Zhao, Ziheng, et al.
Published: (2025)
by: Zhao, Ziheng, et al.
Published: (2025)
A Fully Interpretable Statistical Approach for Roadside LiDAR Background Subtraction
by: Iglesias, Aitor, et al.
Published: (2025)
by: Iglesias, Aitor, et al.
Published: (2025)
An Interpretable Deep Learning Approach for Morphological Script Type Analysis
by: Vlachou-Efstathiou, Malamatenia, et al.
Published: (2024)
by: Vlachou-Efstathiou, Malamatenia, et al.
Published: (2024)
ControlCol: Controllability in Automatic Speaker Video Colorization
by: Ward, Rory, et al.
Published: (2024)
by: Ward, Rory, et al.
Published: (2024)
Divide to Conquer: A Field Decomposition Approach for Multi-Organ Whole-Body CT Image Registration
by: Pham, Xuan Loc, et al.
Published: (2025)
by: Pham, Xuan Loc, et al.
Published: (2025)
Towards Concept-based Interpretability of Skin Lesion Diagnosis using Vision-Language Models
by: Patrício, Cristiano, et al.
Published: (2023)
by: Patrício, Cristiano, et al.
Published: (2023)
Active Learning for GCN-based Action Recognition
by: Sahbi, Hichem
Published: (2025)
by: Sahbi, Hichem
Published: (2025)
Similar Items
-
ASDnB: Merging Face with Body Cues For Robust Active Speaker Detection
by: Roxo, Tiago, et al.
Published: (2024) -
How Deep Learning Sees the World: A Survey on Adversarial Attacks & Defenses
by: Costa, Joana C., et al.
Published: (2023) -
ZQBA: Zero Query Black-box Adversarial Attack
by: Costa, Joana C., et al.
Published: (2025) -
How to Squeeze An Explanation Out of Your Model
by: Roxo, Tiago, et al.
Published: (2024) -
LISArD: Learning Image Similarity to Defend Against Gray-box Adversarial Attacks
by: Costa, Joana C., et al.
Published: (2025)