CAF-Mamba: Mamba-Based Cross-Modal Adaptive Attention Fusion for Multimodal Depression Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Bowen, Fiedler, Marc-André, Al-Hamadi, Ayoub |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DepMamba: Progressive Fusion Mamba for Multimodal Depression Detection
by: Ye, Jiaxin, et al.
Published: (2024)
by: Ye, Jiaxin, et al.
Published: (2024)
MambaTalk: Efficient Holistic Gesture Synthesis with Selective State Space Models
by: Xu, Zunnan, et al.
Published: (2024)
by: Xu, Zunnan, et al.
Published: (2024)
Multimodal Attention-Aware Fusion for Diagnosing Distal Myopathy: Evaluating Model Interpretability and Clinician Trust
by: Onari, Mohsen Abbaspour, et al.
Published: (2025)
by: Onari, Mohsen Abbaspour, et al.
Published: (2025)
Beyond Questionnaires: Video Analysis for Social Anxiety Detection
by: Sahu, Nilesh Kumar, et al.
Published: (2024)
by: Sahu, Nilesh Kumar, et al.
Published: (2024)
AI-based Multimodal Biometrics for Detecting Smartphone Distractions: Application to Online Learning
by: Becerra, Alvaro, et al.
Published: (2025)
by: Becerra, Alvaro, et al.
Published: (2025)
Adaptive Modality Balanced Online Knowledge Distillation for Brain-Eye-Computer based Dim Object Detection
by: Li, Zixing, et al.
Published: (2024)
by: Li, Zixing, et al.
Published: (2024)
BabyMamba-HAR: Lightweight Selective State Space Models for Efficient Human Activity Recognition on Resource Constrained Devices
by: Mandal, Mridankan
Published: (2026)
by: Mandal, Mridankan
Published: (2026)
MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion
by: Fu, Chencan, et al.
Published: (2024)
by: Fu, Chencan, et al.
Published: (2024)
Ocular Authentication: Fusion of Gaze and Periocular Modalities
by: Lohr, Dillon, et al.
Published: (2025)
by: Lohr, Dillon, et al.
Published: (2025)
Learning Multimodal Cues of Children's Uncertainty
by: Cheng, Qi, et al.
Published: (2024)
by: Cheng, Qi, et al.
Published: (2024)
mEBAL: A Multimodal Database for Eye Blink Detection and Attention Level Estimation
by: Daza, Roberto, et al.
Published: (2020)
by: Daza, Roberto, et al.
Published: (2020)
Cross-View Cross-Modal Unsupervised Domain Adaptation for Driver Monitoring System
by: Bhalla, Aditi, et al.
Published: (2025)
by: Bhalla, Aditi, et al.
Published: (2025)
DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimation
by: Li, Jia, et al.
Published: (2024)
by: Li, Jia, et al.
Published: (2024)
DeepFace-Attention: Multimodal Face Biometrics for Attention Estimation with Application to e-Learning
by: Daza, Roberto, et al.
Published: (2024)
by: Daza, Roberto, et al.
Published: (2024)
Enhancing Apparent Personality Trait Analysis with Cross-Modal Embeddings
by: Fodor, Ádám, et al.
Published: (2024)
by: Fodor, Ádám, et al.
Published: (2024)
Semantic and Expressive Variation in Image Captions Across Languages
by: Ye, Andre, et al.
Published: (2023)
by: Ye, Andre, et al.
Published: (2023)
A Call to Arms: AI Should be Critical for Social Media Analysis of Conflict Zones
by: Abedin, Afia, et al.
Published: (2023)
by: Abedin, Afia, et al.
Published: (2023)
Improved Digital Therapy for Developmental Pediatrics Using Domain-Specific Artificial Intelligence: Machine Learning Study
by: Washington, Peter, et al.
Published: (2020)
by: Washington, Peter, et al.
Published: (2020)
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
by: Hall, Melissa, et al.
Published: (2024)
by: Hall, Melissa, et al.
Published: (2024)
A Comparison of Human and Machine Learning Errors in Face Recognition
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
Classification of the lunar surface pattern by AI architectures: Does AI see a rabbit in the Moon?
by: Shoji, Daigo
Published: (2023)
by: Shoji, Daigo
Published: (2023)
AIDEN: Design and Pilot Study of an AI Assistant for the Visually Impaired
by: Marquez-Carpintero, Luis, et al.
Published: (2025)
by: Marquez-Carpintero, Luis, et al.
Published: (2025)
The Cadaver in the Machine: The Social Practices of Measurement and Validation in Motion Capture Technology
by: Harvey, Emma, et al.
Published: (2024)
by: Harvey, Emma, et al.
Published: (2024)
Intuitions of Machine Learning Researchers about Transfer Learning for Medical Image Classification
by: Lu, Yucheng, et al.
Published: (2025)
by: Lu, Yucheng, et al.
Published: (2025)
Listen to Rhythm, Choose Movements: Autoregressive Multimodal Dance Generation via Diffusion and Mamba with Decoupled Dance Dataset
by: Duan, Oran, et al.
Published: (2026)
by: Duan, Oran, et al.
Published: (2026)
Negative Shanshui: Real-time Interactive Ink Painting Synthesis
by: Zhou, Aven-Le
Published: (2025)
by: Zhou, Aven-Le
Published: (2025)
Using Salient Object Detection to Identify Manipulative Cookie Banners that Circumvent GDPR
by: Grossman, Riley, et al.
Published: (2025)
by: Grossman, Riley, et al.
Published: (2025)
SasMamba: A Lightweight Structure-Aware Stride State Space Model for 3D Human Pose Estimation
by: Cui, Hu, et al.
Published: (2025)
by: Cui, Hu, et al.
Published: (2025)
Voting-based Multimodal Automatic Deception Detection
by: Touma, Lana, et al.
Published: (2023)
by: Touma, Lana, et al.
Published: (2023)
AttentionBender: Manipulating Cross-Attention in Video Diffusion Transformers as a Creative Probe
by: Cole, Adam, et al.
Published: (2026)
by: Cole, Adam, et al.
Published: (2026)
AI-Based Facial Emotion Recognition Solutions for Education: A Study of Teacher-User and Other Categories
by: Ravenor, R. Yamamoto
Published: (2023)
by: Ravenor, R. Yamamoto
Published: (2023)
Harmonious Color Pairings: Insights from Human Preference and Natural Hue Statistics
by: Forni, Ortensia, et al.
Published: (2025)
by: Forni, Ortensia, et al.
Published: (2025)
Multi-scale structural complexity as a quantitative measure of visual complexity
by: Kravchenko, Anna, et al.
Published: (2024)
by: Kravchenko, Anna, et al.
Published: (2024)
Deciphering Emotions in Children Storybooks: A Comparative Analysis of Multimodal LLMs in Educational Applications
by: Asseri, Bushra, et al.
Published: (2025)
by: Asseri, Bushra, et al.
Published: (2025)
Towards Fairness in AI for Melanoma Detection: Systemic Review and Recommendations
by: Montoya, Laura N, et al.
Published: (2024)
by: Montoya, Laura N, et al.
Published: (2024)
Joining Forces for Pathology Diagnostics with AI Assistance: The EMPAIA Initiative
by: Zerbe, Norman, et al.
Published: (2023)
by: Zerbe, Norman, et al.
Published: (2023)
Investigating Disability Representations in Text-to-Image Models
by: Tian, Yang, et al.
Published: (2026)
by: Tian, Yang, et al.
Published: (2026)
Real-Time Hand Gesture Recognition: Integrating Skeleton-Based Data Fusion and Multi-Stream CNN
by: Yusuf, Oluwaleke, et al.
Published: (2024)
by: Yusuf, Oluwaleke, et al.
Published: (2024)
AKRMap: Adaptive Kernel Regression for Trustworthy Visualization of Cross-Modal Embeddings
by: Ye, Yilin, et al.
Published: (2025)
by: Ye, Yilin, et al.
Published: (2025)
Tell Me Without Telling Me: Two-Way Prediction of Visualization Literacy and Visual Attention
by: Chang, Minsuk, et al.
Published: (2025)
by: Chang, Minsuk, et al.
Published: (2025)
Similar Items
-
DepMamba: Progressive Fusion Mamba for Multimodal Depression Detection
by: Ye, Jiaxin, et al.
Published: (2024) -
MambaTalk: Efficient Holistic Gesture Synthesis with Selective State Space Models
by: Xu, Zunnan, et al.
Published: (2024) -
Multimodal Attention-Aware Fusion for Diagnosing Distal Myopathy: Evaluating Model Interpretability and Clinician Trust
by: Onari, Mohsen Abbaspour, et al.
Published: (2025) -
Beyond Questionnaires: Video Analysis for Social Anxiety Detection
by: Sahu, Nilesh Kumar, et al.
Published: (2024) -
AI-based Multimodal Biometrics for Detecting Smartphone Distractions: Application to Online Learning
by: Becerra, Alvaro, et al.
Published: (2025)