Evaluating Vision-Language Models for Emotion Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhattacharyya, Sree, Wang, James Z. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Heterogeneous Multimodal Graph Learning Framework for Recognizing User Emotions in Social Networks
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025)
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025)
Unsupervised Memorability Modeling from Tip-of-the-Tongue Retrieval Queries
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025)
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025)
Artwork Interpretation with Vision Language Models: A Case Study on Emotions and Emotion Symbols
von: Padó, Sebastian, et al.
Veröffentlicht: (2025)
von: Padó, Sebastian, et al.
Veröffentlicht: (2025)
EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models
von: Xing, Bohao, et al.
Veröffentlicht: (2025)
von: Xing, Bohao, et al.
Veröffentlicht: (2025)
Texture or Semantics? Vision-Language Models Get Lost in Font Recognition
von: Li, Zhecheng, et al.
Veröffentlicht: (2025)
von: Li, Zhecheng, et al.
Veröffentlicht: (2025)
Evaluating Vision-Language Models as Evaluators in Path Planning
von: Aghzal, Mohamed, et al.
Veröffentlicht: (2024)
von: Aghzal, Mohamed, et al.
Veröffentlicht: (2024)
GlyphPattern: An Abstract Pattern Recognition Benchmark for Vision-Language Models
von: Wu, Zixuan, et al.
Veröffentlicht: (2024)
von: Wu, Zixuan, et al.
Veröffentlicht: (2024)
Evaluation of Cultural Competence of Vision-Language Models
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
Anatomy of a Feeling: Narrating Embodied Emotions via Large Vision-Language Models
von: Saim, Mohammad, et al.
Veröffentlicht: (2025)
von: Saim, Mohammad, et al.
Veröffentlicht: (2025)
Can Vision-Language Models Evaluate Handwritten Math?
von: Nath, Oikantik, et al.
Veröffentlicht: (2025)
von: Nath, Oikantik, et al.
Veröffentlicht: (2025)
Instruction-Following Evaluation of Large Vision-Language Models
von: Shiono, Daiki, et al.
Veröffentlicht: (2025)
von: Shiono, Daiki, et al.
Veröffentlicht: (2025)
NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Fine-Grained Evaluation of Large Vision-Language Models in Autonomous Driving
von: Li, Yue, et al.
Veröffentlicht: (2025)
von: Li, Yue, et al.
Veröffentlicht: (2025)
Chartographer: Counterfactual Chart Generation for Evaluating Vision-Language Models
von: Jiang, Yifan, et al.
Veröffentlicht: (2026)
von: Jiang, Yifan, et al.
Veröffentlicht: (2026)
Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models
von: Lu, Jiaying, et al.
Veröffentlicht: (2023)
von: Lu, Jiaying, et al.
Veröffentlicht: (2023)
VLMInferSlow: Evaluating the Efficiency Robustness of Large Vision-Language Models as a Service
von: Wang, Xiasi, et al.
Veröffentlicht: (2025)
von: Wang, Xiasi, et al.
Veröffentlicht: (2025)
AlignMMBench: Evaluating Chinese Multimodal Alignment in Large Vision-Language Models
von: Wu, Yuhang, et al.
Veröffentlicht: (2024)
von: Wu, Yuhang, et al.
Veröffentlicht: (2024)
Finer: Investigating and Enhancing Fine-Grained Visual Concept Recognition in Large Vision Language Models
von: Kim, Jeonghwan, et al.
Veröffentlicht: (2024)
von: Kim, Jeonghwan, et al.
Veröffentlicht: (2024)
Evaluating Fairness in Large Vision-Language Models Across Diverse Demographic Attributes and Prompts
von: Wu, Xuyang, et al.
Veröffentlicht: (2024)
von: Wu, Xuyang, et al.
Veröffentlicht: (2024)
Evaluating Vision Language Model Adaptations for Radiology Report Generation in Low-Resource Languages
von: Salmè, Marco, et al.
Veröffentlicht: (2025)
von: Salmè, Marco, et al.
Veröffentlicht: (2025)
NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples
von: Li, Baiqi, et al.
Veröffentlicht: (2024)
von: Li, Baiqi, et al.
Veröffentlicht: (2024)
Heron-Bench: A Benchmark for Evaluating Vision Language Models in Japanese
von: Inoue, Yuichi, et al.
Veröffentlicht: (2024)
von: Inoue, Yuichi, et al.
Veröffentlicht: (2024)
ERIT Lightweight Multimodal Dataset for Elderly Emotion Recognition and Multimodal Fusion Evaluation
von: Frieske, Rita, et al.
Veröffentlicht: (2024)
von: Frieske, Rita, et al.
Veröffentlicht: (2024)
"See the World, Discover Knowledge": A Chinese Factuality Evaluation for Large Vision Language Models
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation
von: Wang, Xintong, et al.
Veröffentlicht: (2025)
von: Wang, Xintong, et al.
Veröffentlicht: (2025)
Granular Privacy Control for Geolocation with Vision Language Models
von: Mendes, Ethan, et al.
Veröffentlicht: (2024)
von: Mendes, Ethan, et al.
Veröffentlicht: (2024)
EMOVA: Empowering Language Models to See, Hear and Speak with Vivid Emotions
von: Chen, Kai, et al.
Veröffentlicht: (2024)
von: Chen, Kai, et al.
Veröffentlicht: (2024)
The Visual Iconicity Challenge: Evaluating Vision-Language Models on Sign Language Form-Meaning Mapping
von: Keleş, Onur, et al.
Veröffentlicht: (2025)
von: Keleş, Onur, et al.
Veröffentlicht: (2025)
RoundTripOCR: A Data Generation Technique for Enhancing Post-OCR Error Correction in Low-Resource Devanagari Languages
von: Kashid, Harshvivek, et al.
Veröffentlicht: (2024)
von: Kashid, Harshvivek, et al.
Veröffentlicht: (2024)
Evaluating Reasoning Faithfulness in Medical Vision-Language Models using Multimodal Perturbations
von: Moll, Johannes, et al.
Veröffentlicht: (2025)
von: Moll, Johannes, et al.
Veröffentlicht: (2025)
Leveraging Vision-Language Pre-training for Human Activity Recognition in Still Images
von: Mahanta, Cristina, et al.
Veröffentlicht: (2025)
von: Mahanta, Cristina, et al.
Veröffentlicht: (2025)
CROPE: Evaluating In-Context Adaptation of Vision and Language Models to Culture-Specific Concepts
von: Nikandrou, Malvina, et al.
Veröffentlicht: (2024)
von: Nikandrou, Malvina, et al.
Veröffentlicht: (2024)
VALOR-EVAL: Holistic Coverage and Faithfulness Evaluation of Large Vision-Language Models
von: Qiu, Haoyi, et al.
Veröffentlicht: (2024)
von: Qiu, Haoyi, et al.
Veröffentlicht: (2024)
Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models
von: Khan, Mohammed Safi Ur Rahman, et al.
Veröffentlicht: (2026)
von: Khan, Mohammed Safi Ur Rahman, et al.
Veröffentlicht: (2026)
WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences
von: Lu, Yujie, et al.
Veröffentlicht: (2024)
von: Lu, Yujie, et al.
Veröffentlicht: (2024)
Beyond Words: Enhancing Desire, Emotion, and Sentiment Recognition with Non-Verbal Cues
von: Chen, Wei, et al.
Veröffentlicht: (2025)
von: Chen, Wei, et al.
Veröffentlicht: (2025)
Emotion Recognition in Signers
von: Funakoshi, Kotaro, et al.
Veröffentlicht: (2025)
von: Funakoshi, Kotaro, et al.
Veröffentlicht: (2025)
Enhancing Large Vision Language Models with Self-Training on Image Comprehension
von: Deng, Yihe, et al.
Veröffentlicht: (2024)
von: Deng, Yihe, et al.
Veröffentlicht: (2024)
TemMed-Bench: Evaluating Temporal Medical Image Reasoning in Vision-Language Models
von: Zhang, Junyi, et al.
Veröffentlicht: (2025)
von: Zhang, Junyi, et al.
Veröffentlicht: (2025)
MMLongCite: A Benchmark for Evaluating Fidelity of Long-Context Vision-Language Models
von: Zhou, Keyan, et al.
Veröffentlicht: (2025)
von: Zhou, Keyan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Heterogeneous Multimodal Graph Learning Framework for Recognizing User Emotions in Social Networks
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025) -
Unsupervised Memorability Modeling from Tip-of-the-Tongue Retrieval Queries
von: Bhattacharyya, Sree, et al.
Veröffentlicht: (2025) -
Artwork Interpretation with Vision Language Models: A Case Study on Emotions and Emotion Symbols
von: Padó, Sebastian, et al.
Veröffentlicht: (2025) -
EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models
von: Xing, Bohao, et al.
Veröffentlicht: (2025) -
Texture or Semantics? Vision-Language Models Get Lost in Font Recognition
von: Li, Zhecheng, et al.
Veröffentlicht: (2025)