PAEFF: Precise Alignment and Enhanced Gated Feature Fusion for Face-Voice Association
Fuente:
arXiv
Saved in:
| Main Authors: | Hannan, Abdul, Manzoor, Muhammad Arslan, Nawaz, Shah, Liaqat, Muhammad Irzam, Schedl, Markus, Noman, Mubashir |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Face-Voice Association with Inductive Bias for Maximum Class Separation
by: Moscati, Marta, et al.
Published: (2026)
by: Moscati, Marta, et al.
Published: (2026)
RFOP: Rethinking Fusion and Orthogonal Projection for Face-Voice Association
by: Hannan, Abdul, et al.
Published: (2025)
by: Hannan, Abdul, et al.
Published: (2025)
Distillation-based Layer Dropping (DLD): Effective End-to-end Framework for Dynamic Speech Networks
by: Hannan, Abdul, et al.
Published: (2026)
by: Hannan, Abdul, et al.
Published: (2026)
An Effective Training Framework for Light-Weight Automatic Speech Recognition Models
by: Hannan, Abdul, et al.
Published: (2025)
by: Hannan, Abdul, et al.
Published: (2025)
SB-BEVFusion: Enhancing the Robustness against Sensor Malfunction and Corruptions
by: Essl, Markus, et al.
Published: (2026)
by: Essl, Markus, et al.
Published: (2026)
Chameleon: Images Are What You Need For Multimodal Learning Robust To Missing Modalities
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
by: Liaqat, Muhammad Irzam, et al.
Published: (2024)
EoCD: Encoder only Remote Sensing Change Detection
by: Noman, Mubashir, et al.
Published: (2026)
by: Noman, Mubashir, et al.
Published: (2026)
Linking Faces and Voices Across Languages: Insights from the FAME 2026 Challenge
by: Moscati, Marta, et al.
Published: (2025)
by: Moscati, Marta, et al.
Published: (2025)
Face-voice Association in Multilingual Environments (FAME) 2026 Challenge Evaluation Plan
by: Moscati, Marta, et al.
Published: (2025)
by: Moscati, Marta, et al.
Published: (2025)
InceptionMamba: Efficient Multi-Stage Feature Enhancement with Selective State Space Model for Microscopic Medical Image Segmentation
by: Kareem, Daniya Najiha Abdul, et al.
Published: (2025)
by: Kareem, Daniya Najiha Abdul, et al.
Published: (2025)
Face-voice Association in Multilingual Environments (FAME) Challenge 2024 Evaluation Plan
by: Saeed, Muhammad Saad, et al.
Published: (2024)
by: Saeed, Muhammad Saad, et al.
Published: (2024)
RobustA: Robust Anomaly Detection in Multimodal Data
by: AlMarri, Salem, et al.
Published: (2025)
by: AlMarri, Salem, et al.
Published: (2025)
POLY-SIM: Polyglot Speaker Identification with Missing Modality Grand Challenge 2026 Evaluation Plan
by: Moscati, Marta, et al.
Published: (2026)
by: Moscati, Marta, et al.
Published: (2026)
FANet: Feature Amplification Network for Semantic Segmentation in Cluttered Background
by: Ali, Muhammad, et al.
Published: (2024)
by: Ali, Muhammad, et al.
Published: (2024)
COSNet: A Novel Semantic Segmentation Network using Enhanced Boundaries in Cluttered Scenes
by: Ali, Muhammad, et al.
Published: (2024)
by: Ali, Muhammad, et al.
Published: (2024)
Modality Invariant Multimodal Learning to Handle Missing Modalities: A Single-Branch Approach
by: Saeed, Muhammad Saad, et al.
Published: (2024)
by: Saeed, Muhammad Saad, et al.
Published: (2024)
Robust Harmful Meme Detection under Missing Modalities via Shared Representation Learning
by: Breiteneder, Felix, et al.
Published: (2026)
by: Breiteneder, Felix, et al.
Published: (2026)
ChangeBind: A Hybrid Change Encoder for Remote Sensing Change Detection
by: Noman, Mubashir, et al.
Published: (2024)
by: Noman, Mubashir, et al.
Published: (2024)
TinyNeRV: Compact Neural Video Representations via Capacity Scaling, Distillation, and Low-Precision Inference
by: Akhtar, Muhammad Hannan, et al.
Published: (2026)
by: Akhtar, Muhammad Hannan, et al.
Published: (2026)
Gated-Attention Feature-Fusion Based Framework for Poverty Prediction
by: Ramzan, Muhammad Umer, et al.
Published: (2024)
by: Ramzan, Muhammad Umer, et al.
Published: (2024)
EMF: Event Meta Formers for Event-based Real-time Traffic Object Detection
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2025)
A Deep Features-Based Approach Using Modified ResNet50 and Gradient Boosting for Visual Sentiments Classification
by: Arslan, Muhammad, et al.
Published: (2024)
by: Arslan, Muhammad, et al.
Published: (2024)
Rethinking Transformers Pre-training for Multi-Spectral Satellite Imagery
by: Noman, Mubashir, et al.
Published: (2024)
by: Noman, Mubashir, et al.
Published: (2024)
FaceScore: Benchmarking and Enhancing Face Quality in Human Generation
by: Liao, Zhenyi, et al.
Published: (2024)
by: Liao, Zhenyi, et al.
Published: (2024)
Locally-Focused Face Representation for Sketch-to-Image Generation Using Noise-Induced Refinement
by: Ramzan, Muhammad Umer, et al.
Published: (2024)
by: Ramzan, Muhammad Umer, et al.
Published: (2024)
XM-ALIGN: Unified Cross-Modal Embedding Alignment for Face-Voice Association
by: Fang, Zhihua, et al.
Published: (2025)
by: Fang, Zhihua, et al.
Published: (2025)
CDChat: A Large Multimodal Model for Remote Sensing Change Description
by: Noman, Mubashir, et al.
Published: (2024)
by: Noman, Mubashir, et al.
Published: (2024)
AgriCLIP: Adapting CLIP for Agriculture and Livestock via Domain-Specialized Cross-Model Alignment
by: Nawaz, Umair, et al.
Published: (2024)
by: Nawaz, Umair, et al.
Published: (2024)
HyRet-Change: A hybrid retentive network for remote sensing change detection
by: Fiaz, Mustansar, et al.
Published: (2025)
by: Fiaz, Mustansar, et al.
Published: (2025)
ELGC-Net: Efficient Local-Global Context Aggregation for Remote Sensing Change Detection
by: Noman, Mubashir, et al.
Published: (2024)
by: Noman, Mubashir, et al.
Published: (2024)
Split-Fuse-Transport: Annotation-Free Saliency via Dual Clustering and Optimal Transport Alignment
by: Ramzan, Muhammad Umer, et al.
Published: (2025)
by: Ramzan, Muhammad Umer, et al.
Published: (2025)
Lightweight Deepfake Detection Based on Multi-Feature Fusion
by: Yasir, Siddiqui Muhammad, et al.
Published: (2025)
by: Yasir, Siddiqui Muhammad, et al.
Published: (2025)
WoundFormer: Multi-Scale Spatial Feature Fusion for Multi-Class Wound Tissue Segmentation
by: Kabir, Muhammad Ashad, et al.
Published: (2026)
by: Kabir, Muhammad Ashad, et al.
Published: (2026)
Shared Multi-modal Embedding Space for Face-Voice Association
by: Simic, Christopher, et al.
Published: (2025)
by: Simic, Christopher, et al.
Published: (2025)
MHAFF: Multi-Head Attention Feature Fusion of CNN and Transformer for Cattle Identification
by: Dulal, Rabin, et al.
Published: (2025)
by: Dulal, Rabin, et al.
Published: (2025)
Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos
by: Ding, Bonan, et al.
Published: (2026)
by: Ding, Bonan, et al.
Published: (2026)
VerLM: Explaining Face Verification Using Natural Language
by: Hannan, Syed Abdul, et al.
Published: (2026)
by: Hannan, Syed Abdul, et al.
Published: (2026)
A Tumor Aware DenseNet Swin Hybrid Learning with Boosted and Hierarchical Feature Spaces for Large-Scale Brain MRI Classification
by: Shah, Muhammad Ali, et al.
Published: (2026)
by: Shah, Muhammad Ali, et al.
Published: (2026)
Alignment-Aware and Reliability-Gated Multimodal Fusion for Unmanned Aerial Vehicle Detection Across Heterogeneous Thermal-Visual Sensors
by: Jahan, Ishrat, et al.
Published: (2026)
by: Jahan, Ishrat, et al.
Published: (2026)
FineFACE: Fair Facial Attribute Classification Leveraging Fine-grained Features
by: Manzoor, Ayesha, et al.
Published: (2024)
by: Manzoor, Ayesha, et al.
Published: (2024)
Similar Items
-
Face-Voice Association with Inductive Bias for Maximum Class Separation
by: Moscati, Marta, et al.
Published: (2026) -
RFOP: Rethinking Fusion and Orthogonal Projection for Face-Voice Association
by: Hannan, Abdul, et al.
Published: (2025) -
Distillation-based Layer Dropping (DLD): Effective End-to-end Framework for Dynamic Speech Networks
by: Hannan, Abdul, et al.
Published: (2026) -
An Effective Training Framework for Light-Weight Automatic Speech Recognition Models
by: Hannan, Abdul, et al.
Published: (2025) -
SB-BEVFusion: Enhancing the Robustness against Sensor Malfunction and Corruptions
by: Essl, Markus, et al.
Published: (2026)