AICA-Bench: Holistically Examining the Capabilities of VLMs in Affective Image Content Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | She, Dong, Yao, Xianrong, Chen, Liqun, Yu, Jinghe, Gao, Yang, Jin, Zhanpeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EmoMM: Benchmarking and Steering MLLM for Multimodal Emotion Recognition under Conflict and Missingness
by: Sun, Yueru, et al.
Published: (2026)
by: Sun, Yueru, et al.
Published: (2026)
VisRes Bench: On Evaluating the Visual Reasoning Capabilities of VLMs
by: Törtei, Brigitta Malagurski, et al.
Published: (2025)
by: Törtei, Brigitta Malagurski, et al.
Published: (2025)
MyoSem: Aligning Electromyography to Natural-Language Action Semantics for Hand Action Understanding
by: Wang, Chiyue, et al.
Published: (2026)
by: Wang, Chiyue, et al.
Published: (2026)
Knowledge Distillation for Underwater Feature Extraction and Matching via GAN-synthesized Images
by: Yang, Jinghe, et al.
Published: (2025)
by: Yang, Jinghe, et al.
Published: (2025)
LookBench: A Live and Holistic Open Benchmark for Fashion Image Retrieval
by: ai, Gensmo., et al.
Published: (2026)
by: ai, Gensmo., et al.
Published: (2026)
AIM-Bench: Benchmarking and Improving Affective Image Manipulation via Fine-Grained Hierarchical Control
by: Chen, Shi, et al.
Published: (2026)
by: Chen, Shi, et al.
Published: (2026)
Prism: A Framework for Decoupling and Assessing the Capabilities of VLMs
by: Qiao, Yuxuan, et al.
Published: (2024)
by: Qiao, Yuxuan, et al.
Published: (2024)
VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents
by: Yi, Hongzhu, et al.
Published: (2026)
by: Yi, Hongzhu, et al.
Published: (2026)
Omni IIE Bench: Benchmarking the Practical Capabilities of Image Editing Models
by: Yang, Yujia, et al.
Published: (2026)
by: Yang, Yujia, et al.
Published: (2026)
MissBench: Benchmarking Multimodal Affective Analysis under Imbalanced Missing Modalities
by: Pham, Tien Anh, et al.
Published: (2026)
by: Pham, Tien Anh, et al.
Published: (2026)
CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs
by: Fang, Zhengru, et al.
Published: (2026)
by: Fang, Zhengru, et al.
Published: (2026)
Rethinking the Paradigm of Content Constraints in Unpaired Image-to-Image Translation
by: Cai, Xiuding, et al.
Published: (2022)
by: Cai, Xiuding, et al.
Published: (2022)
Affective Behaviour Analysis via Progressive Learning
by: Liu, Chen, et al.
Published: (2024)
by: Liu, Chen, et al.
Published: (2024)
Affective Video Content Analysis: Decade Review and New Perspectives
by: Xue, Junxiao, et al.
Published: (2023)
by: Xue, Junxiao, et al.
Published: (2023)
Facial Affective Behavior Analysis with Instruction Tuning
by: Li, Yifan, et al.
Published: (2024)
by: Li, Yifan, et al.
Published: (2024)
Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality
by: Oh, Youngtaek, et al.
Published: (2024)
by: Oh, Youngtaek, et al.
Published: (2024)
OCRGenBench: A Comprehensive Benchmark for Evaluating OCR Generative Capabilities
by: Zhang, Peirong, et al.
Published: (2025)
by: Zhang, Peirong, et al.
Published: (2025)
VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects
by: Gao, Xiangbo, et al.
Published: (2026)
by: Gao, Xiangbo, et al.
Published: (2026)
Deep Pre-Alignment for VLMs
by: Yu, Tianyu, et al.
Published: (2026)
by: Yu, Tianyu, et al.
Published: (2026)
VisualActBench: Can VLMs See and Act like a Human?
by: Zhang, Daoan, et al.
Published: (2025)
by: Zhang, Daoan, et al.
Published: (2025)
SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks
by: Song, Zijian, et al.
Published: (2025)
by: Song, Zijian, et al.
Published: (2025)
Solution for 8th Competition on Affective & Behavior Analysis in-the-wild
by: Yu, Jun, et al.
Published: (2025)
by: Yu, Jun, et al.
Published: (2025)
Are VLMs Lost Between Sky and Space? LinkS$^2$Bench for UAV-Satellite Dynamic Cross-View Spatial Intelligence
by: Liu, Dian, et al.
Published: (2026)
by: Liu, Dian, et al.
Published: (2026)
One-shot Optimized Steering Vector for Hallucination Mitigation for VLMs
by: Shi, Youxu, et al.
Published: (2026)
by: Shi, Youxu, et al.
Published: (2026)
ViC-Bench: Benchmarking Visual-Interleaved Chain-of-Thought Capability in MLLMs with Free-Style Intermediate State Representations
by: Wu, Xuecheng, et al.
Published: (2025)
by: Wu, Xuecheng, et al.
Published: (2025)
EmoSpace: Fine-Grained Emotion Prototype Learning for Immersive Affective Content Generation
by: Wang, Bingyuan, et al.
Published: (2026)
by: Wang, Bingyuan, et al.
Published: (2026)
Analyzing Image Beyond Visual Aspect: Image Emotion Classification via Multiple-Affective Captioning
by: Zhou, Zibo, et al.
Published: (2025)
by: Zhou, Zibo, et al.
Published: (2025)
VideoASMR-Bench: Can AI-Generated ASMR Videos Fool VLMs and Humans?
by: Wang, Jiaqi, et al.
Published: (2025)
by: Wang, Jiaqi, et al.
Published: (2025)
Affective Behaviour Analysis via Integrating Multi-Modal Knowledge
by: Zhang, Wei, et al.
Published: (2024)
by: Zhang, Wei, et al.
Published: (2024)
EmoAgent: A Multi-Agent Framework for Diverse Affective Image Manipulation
by: Mao, Qi, et al.
Published: (2025)
by: Mao, Qi, et al.
Published: (2025)
Exposing Hallucinations To Suppress Them: VLMs Representation Editing With Generative Anchors
by: Shi, Youxu, et al.
Published: (2025)
by: Shi, Youxu, et al.
Published: (2025)
WeatherBench: A Real-World Benchmark Dataset for All-in-One Adverse Weather Image Restoration
by: Guan, Qiyuan, et al.
Published: (2025)
by: Guan, Qiyuan, et al.
Published: (2025)
Image Recognition with Vision and Language Embeddings of VLMs
by: Volkov, Illia, et al.
Published: (2025)
by: Volkov, Illia, et al.
Published: (2025)
Should VLMs be Pre-trained with Image Data?
by: Keh, Sedrick, et al.
Published: (2025)
by: Keh, Sedrick, et al.
Published: (2025)
GIR-Bench: Versatile Benchmark for Generating Images with Reasoning
by: Li, Hongxiang, et al.
Published: (2025)
by: Li, Hongxiang, et al.
Published: (2025)
Are VLMs Ready for Lane Topology Awareness in Autonomous Driving?
by: Chen, Xin, et al.
Published: (2025)
by: Chen, Xin, et al.
Published: (2025)
Bridging Visual Affective Gap: Borrowing Textual Knowledge by Learning from Noisy Image-Text Pairs
by: Wu, Daiqing, et al.
Published: (2025)
by: Wu, Daiqing, et al.
Published: (2025)
SpinBench: Perspective and Rotation as a Lens on Spatial Reasoning in VLMs
by: Zhang, Yuyou, et al.
Published: (2025)
by: Zhang, Yuyou, et al.
Published: (2025)
Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports
by: Yang, Yuchen, et al.
Published: (2026)
by: Yang, Yuchen, et al.
Published: (2026)
QwenSafe: Multimodal Content Rating Description Identification via Preference-Aligned VLMs
by: Denipitiyage, Dishanika, et al.
Published: (2026)
by: Denipitiyage, Dishanika, et al.
Published: (2026)
Similar Items
-
EmoMM: Benchmarking and Steering MLLM for Multimodal Emotion Recognition under Conflict and Missingness
by: Sun, Yueru, et al.
Published: (2026) -
VisRes Bench: On Evaluating the Visual Reasoning Capabilities of VLMs
by: Törtei, Brigitta Malagurski, et al.
Published: (2025) -
MyoSem: Aligning Electromyography to Natural-Language Action Semantics for Hand Action Understanding
by: Wang, Chiyue, et al.
Published: (2026) -
Knowledge Distillation for Underwater Feature Extraction and Matching via GAN-synthesized Images
by: Yang, Jinghe, et al.
Published: (2025) -
LookBench: A Live and Holistic Open Benchmark for Fashion Image Retrieval
by: ai, Gensmo., et al.
Published: (2026)