Foundation Model Embeddings Meet Blended Emotions: A Multimodal Fusion Approach for the BLEMORE Challenge
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chapariniya, Masoumeh, Farhadipour, Aref, Ebling, Sarah, Dellwo, Volker, Vukovic, Teodora |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Comparative Analysis of Modality Fusion Approaches for Audio-Visual Person Identification and Verification
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
Multimodal Emotion Recognition and Sentiment Analysis in Multi-Party Conversation Contexts
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
Beyond Appearance: Transformer-based Person Identification from Conversational Dynamics
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2025)
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2025)
Two-Stream Spatial-Temporal Transformer Framework for Person Identification via Natural Conversational Keypoints
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2025)
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2025)
Micro-Expression-Aware Avatar Fingerprinting via Inter-Frame Feature Differencing
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2026)
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2026)
Investigating Identity Signals in Conversational Facial Dynamics via Disentangled Expression Features
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2025)
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2025)
Adaptive Multimodal Person Recognition: A Robust Framework for Handling Missing Modalities
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
Towards Language-Independent Face-Voice Association with Multimodal Foundation Models
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
CL-UZH submission to the NIST SRE 2024 Speaker Recognition Evaluation
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
Not all Blends are Equal: The BLEMORE Dataset of Blended Emotion Expressions with Relative Salience Annotations
von: Lachmann, Tim, et al.
Veröffentlicht: (2026)
von: Lachmann, Tim, et al.
Veröffentlicht: (2026)
Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
EmoLLM: Multimodal Emotional Understanding Meets Large Language Models
von: Yang, Qu, et al.
Veröffentlicht: (2024)
von: Yang, Qu, et al.
Veröffentlicht: (2024)
Deep Neural Networks for Automatic Speaker Recognition Do Not Learn Supra-Segmental Temporal Features
von: Neururer, Daniel, et al.
Veröffentlicht: (2023)
von: Neururer, Daniel, et al.
Veröffentlicht: (2023)
Ordering Matters: Rank-Aware Selective Fusion for Blended Emotion Recognition
von: Lee, Junghyun, et al.
Veröffentlicht: (2026)
von: Lee, Junghyun, et al.
Veröffentlicht: (2026)
Survey of Multimodal Geospatial Foundation Models: Techniques, Applications, and Challenges
von: Yang, Liling, et al.
Veröffentlicht: (2025)
von: Yang, Liling, et al.
Veröffentlicht: (2025)
ChefFusion: Multimodal Foundation Model Integrating Recipe and Food Image Generation
von: Li, Peiyu, et al.
Veröffentlicht: (2024)
von: Li, Peiyu, et al.
Veröffentlicht: (2024)
ECMF: Enhanced Cross-Modal Fusion for Multimodal Emotion Recognition in MER-SEMI Challenge
von: Hu, Juewen, et al.
Veröffentlicht: (2025)
von: Hu, Juewen, et al.
Veröffentlicht: (2025)
BlendFusion -- Scalable Synthetic Data Generation for Diffusion Model Training
von: Venkatesh, Thejas, et al.
Veröffentlicht: (2026)
von: Venkatesh, Thejas, et al.
Veröffentlicht: (2026)
Audio Description Generation in the Era of LLMs and VLMs: A Review of Transferable Generative AI Technologies
von: Gao, Yingqiang, et al.
Veröffentlicht: (2024)
von: Gao, Yingqiang, et al.
Veröffentlicht: (2024)
A Multimodal Fusion Network For Student Emotion Recognition Based on Transformer and Tensor Product
von: Xiang, Ao, et al.
Veröffentlicht: (2024)
von: Xiang, Ao, et al.
Veröffentlicht: (2024)
Sleep Stage Classification using Multimodal Embedding Fusion from EOG and PSM
von: Papillon, Olivier, et al.
Veröffentlicht: (2025)
von: Papillon, Olivier, et al.
Veröffentlicht: (2025)
Anchoring Emotions in Text: Robust Multimodal Fusion for Mimicry Intensity Estimation
von: Zhu, Lingsi, et al.
Veröffentlicht: (2026)
von: Zhu, Lingsi, et al.
Veröffentlicht: (2026)
Expanding the Content-Style Frontier: a Balanced Subspace Blending Approach for Content-Style LoRA Fusion
von: Huang, Linhao
Veröffentlicht: (2025)
von: Huang, Linhao
Veröffentlicht: (2025)
ERIT Lightweight Multimodal Dataset for Elderly Emotion Recognition and Multimodal Fusion Evaluation
von: Frieske, Rita, et al.
Veröffentlicht: (2024)
von: Frieske, Rita, et al.
Veröffentlicht: (2024)
State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
Beyond Imperfections: A Conditional Inpainting Approach for End-to-End Artifact Removal in VTON and Pose Transfer
von: Tabatabaei, Aref, et al.
Veröffentlicht: (2024)
von: Tabatabaei, Aref, et al.
Veröffentlicht: (2024)
Interactive Multimodal Fusion with Temporal Modeling
von: Yu, Jun, et al.
Veröffentlicht: (2025)
von: Yu, Jun, et al.
Veröffentlicht: (2025)
AdaFusion: Prompt-Guided Inference with Adaptive Fusion of Pathology Foundation Models
von: Xiao, Yuxiang, et al.
Veröffentlicht: (2025)
von: Xiao, Yuxiang, et al.
Veröffentlicht: (2025)
HFMF: Hierarchical Fusion Meets Multi-Stream Models for Deepfake Detection
von: Mehta, Anant, et al.
Veröffentlicht: (2025)
von: Mehta, Anant, et al.
Veröffentlicht: (2025)
TeEFusion: Blending Text Embeddings to Distill Classifier-Free Guidance
von: Fu, Minghao, et al.
Veröffentlicht: (2025)
von: Fu, Minghao, et al.
Veröffentlicht: (2025)
TidyVoice 2026 Challenge Evaluation Plan
von: Farhadipour, Aref, et al.
Veröffentlicht: (2026)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2026)
Investigating Disability Representations in Text-to-Image Models
von: Tian, Yang, et al.
Veröffentlicht: (2026)
von: Tian, Yang, et al.
Veröffentlicht: (2026)
Foundational Question Generation for Video Question Answering via an Embedding-Integrated Approach
von: Oh, Ju-Young
Veröffentlicht: (2025)
von: Oh, Ju-Young
Veröffentlicht: (2025)
MANGO: Multimodal Attention-based Normalizing Flow Approach to Fusion Learning
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2025)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2025)
Multimodal Models Meet Presentation Attack Detection on ID Documents
von: Villanueva, Marina, et al.
Veröffentlicht: (2026)
von: Villanueva, Marina, et al.
Veröffentlicht: (2026)
MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled Embedding
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
von: Ma, Yifeng, et al.
Veröffentlicht: (2023)
Low-Resource Vision Challenges for Foundation Models
von: Zhang, Yunhua, et al.
Veröffentlicht: (2024)
von: Zhang, Yunhua, et al.
Veröffentlicht: (2024)
EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models
von: Xing, Bohao, et al.
Veröffentlicht: (2025)
von: Xing, Bohao, et al.
Veröffentlicht: (2025)
Plug-and-Play Logit Fusion for Heterogeneous Pathology Foundation Models
von: Huang, Gexin, et al.
Veröffentlicht: (2026)
von: Huang, Gexin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Comparative Analysis of Modality Fusion Approaches for Audio-Visual Person Identification and Verification
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024) -
Multimodal Emotion Recognition and Sentiment Analysis in Multi-Party Conversation Contexts
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025) -
Beyond Appearance: Transformer-based Person Identification from Conversational Dynamics
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2025) -
Two-Stream Spatial-Temporal Transformer Framework for Person Identification via Natural Conversational Keypoints
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2025) -
Micro-Expression-Aware Avatar Fingerprinting via Inter-Frame Feature Differencing
von: Chapariniya, Masoumeh, et al.
Veröffentlicht: (2026)