VLLMs Provide Better Context for Emotion Understanding Through Common Sense Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Xenos, Alexandros, Foteinopoulou, Niki Maria, Ntinou, Ioanna, Patras, Ioannis, Tzimiropoulos, Georgios |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
EmoCLIP: A Vision-Language Method for Zero-Shot Video Facial Expression Recognition
por: Foteinopoulou, Niki Maria, et al.
Publicado: (2023)
por: Foteinopoulou, Niki Maria, et al.
Publicado: (2023)
Vision-Free Retrieval: Rethinking Multimodal Search with Textual Scene Descriptions
por: Ntinou, Ioanna, et al.
Publicado: (2025)
por: Ntinou, Ioanna, et al.
Publicado: (2025)
Multiscale Vision Transformers meet Bipartite Matching for efficient single-stage Action Localization
por: Ntinou, Ioanna, et al.
Publicado: (2023)
por: Ntinou, Ioanna, et al.
Publicado: (2023)
MeMSVD: Long-Range Temporal Structure Capturing Using Incremental SVD
por: Ntinou, Ioanna, et al.
Publicado: (2024)
por: Ntinou, Ioanna, et al.
Publicado: (2024)
CLIPCleaner: Cleaning Noisy Labels with CLIP
por: Feng, Chen, et al.
Publicado: (2024)
por: Feng, Chen, et al.
Publicado: (2024)
Efficient Unsupervised Visual Representation Learning with Explicit Cluster Balancing
por: Metaxas, Ioannis Maniadis, et al.
Publicado: (2024)
por: Metaxas, Ioannis Maniadis, et al.
Publicado: (2024)
A Hitchhikers Guide to Fine-Grained Face Forgery Detection Using Common Sense Reasoning
por: Foteinopoulou, Niki Maria, et al.
Publicado: (2024)
por: Foteinopoulou, Niki Maria, et al.
Publicado: (2024)
RWKV-UI: UI Understanding with Enhanced Perception and Reasoning
por: Yang, Jiaxi, et al.
Publicado: (2025)
por: Yang, Jiaxi, et al.
Publicado: (2025)
SSR: An Efficient and Robust Framework for Learning with Unknown Label Noise
por: Feng, Chen, et al.
Publicado: (2021)
por: Feng, Chen, et al.
Publicado: (2021)
Emotion Based Prediction in the Context of Optimized Trajectory Planning for Immersive Learning
por: Sungheetha, Akey, et al.
Publicado: (2023)
por: Sungheetha, Akey, et al.
Publicado: (2023)
Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision
por: Li, Chentao, et al.
Publicado: (2026)
por: Li, Chentao, et al.
Publicado: (2026)
AVERE: Improving Audiovisual Emotion Reasoning with Preference Optimization
por: Chaubey, Ashutosh, et al.
Publicado: (2026)
por: Chaubey, Ashutosh, et al.
Publicado: (2026)
CemiFace: Center-based Semi-hard Synthetic Face Generation for Face Recognition
por: Sun, Zhonglin, et al.
Publicado: (2024)
por: Sun, Zhonglin, et al.
Publicado: (2024)
MM2Latent: Text-to-facial image generation and editing in GANs with multimodal assistance
por: Meng, Debin, et al.
Publicado: (2024)
por: Meng, Debin, et al.
Publicado: (2024)
Learning Annotation Consensus for Continuous Emotion Recognition
por: Shoer, Ibrahim, et al.
Publicado: (2025)
por: Shoer, Ibrahim, et al.
Publicado: (2025)
Seeing Eye to AI: Human Alignment via Gaze-Based Response Rewards for Large Language Models
por: Lopez-Cardona, Angela, et al.
Publicado: (2024)
por: Lopez-Cardona, Angela, et al.
Publicado: (2024)
Visual Affect Analysis: Predicting Emotions of Image Viewers with Vision-Language Models
por: Nowicki, Filip, et al.
Publicado: (2026)
por: Nowicki, Filip, et al.
Publicado: (2026)
Not all Blends are Equal: The BLEMORE Dataset of Blended Emotion Expressions with Relative Salience Annotations
por: Lachmann, Tim, et al.
Publicado: (2026)
por: Lachmann, Tim, et al.
Publicado: (2026)
The Latency Wall: Benchmarking Off-the-Shelf Emotion Recognition for Real-Time Virtual Avatars
por: Benyamin, Yarin
Publicado: (2026)
por: Benyamin, Yarin
Publicado: (2026)
Exploring Thermography Technology: A Comprehensive Facial Dataset for Face Detection, Recognition, and Emotion
por: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Publicado: (2024)
por: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Publicado: (2024)
egoEMOTION: Egocentric Vision and Physiological Signals for Emotion and Personality Recognition in Real-World Tasks
por: Jammot, Matthias, et al.
Publicado: (2025)
por: Jammot, Matthias, et al.
Publicado: (2025)
Low Latency Gaze Tracking via Latent Optical Sensing
por: Zheng, Yidan, et al.
Publicado: (2026)
por: Zheng, Yidan, et al.
Publicado: (2026)
Panda or not Panda? Understanding Adversarial Attacks with Interactive Visualization
por: You, Yuzhe, et al.
Publicado: (2023)
por: You, Yuzhe, et al.
Publicado: (2023)
LAFS: Landmark-based Facial Self-supervised Learning for Face Recognition
por: Sun, Zhonglin, et al.
Publicado: (2024)
por: Sun, Zhonglin, et al.
Publicado: (2024)
Accelerating Physical Property Reasoning for Augmented Visual Cognition
por: Lan, Hongbo, et al.
Publicado: (2025)
por: Lan, Hongbo, et al.
Publicado: (2025)
Across-Game Engagement Modelling via Few-Shot Learning
por: Pinitas, Kosmas, et al.
Publicado: (2024)
por: Pinitas, Kosmas, et al.
Publicado: (2024)
A Comparative Study of Scanpath Models in Graph-Based Visualization
por: Lopez-Cardona, Angela, et al.
Publicado: (2025)
por: Lopez-Cardona, Angela, et al.
Publicado: (2025)
GLIMPSE : Real-Time Text Recognition and Contextual Understanding for VQA in Wearables
por: Ramachandran, Akhil, et al.
Publicado: (2026)
por: Ramachandran, Akhil, et al.
Publicado: (2026)
Beyond Object Categories: Multi-Attribute Reference Understanding for Visual Grounding
por: Guo, Hao, et al.
Publicado: (2025)
por: Guo, Hao, et al.
Publicado: (2025)
SimVecVis: A Dataset for Enhancing MLLMs in Visualization Understanding
por: Liu, Can, et al.
Publicado: (2025)
por: Liu, Can, et al.
Publicado: (2025)
AgentSense: Virtual Sensor Data Generation Using LLM Agents in Simulated Home Environments
por: Leng, Zikang, et al.
Publicado: (2025)
por: Leng, Zikang, et al.
Publicado: (2025)
Towards Context-aware Support for Color Vision Deficiency: An Approach Integrating LLM and AR
por: Morita, Shogo, et al.
Publicado: (2024)
por: Morita, Shogo, et al.
Publicado: (2024)
Shape vs. Context: Examining Human--AI Gaps in Ambiguous Japanese Character Recognition
por: Haraguchi, Daichi
Publicado: (2026)
por: Haraguchi, Daichi
Publicado: (2026)
Enhanced Automated Quality Assessment Network for Interactive Building Segmentation in High-Resolution Remote Sensing Imagery
por: Zhang, Zhili, et al.
Publicado: (2024)
por: Zhang, Zhili, et al.
Publicado: (2024)
CoEditor++: Instruction-based Visual Editing via Cognitive Reasoning
por: Ni, Minheng, et al.
Publicado: (2026)
por: Ni, Minheng, et al.
Publicado: (2026)
Breaking Coordinate Overfitting: Geometry-Aware WiFi Sensing for Cross-Layout 3D Pose Estimation
por: Jia, Songming, et al.
Publicado: (2026)
por: Jia, Songming, et al.
Publicado: (2026)
LLM4Brain: Training a Large Language Model for Brain Video Understanding
por: Zheng, Ruizhe, et al.
Publicado: (2024)
por: Zheng, Ruizhe, et al.
Publicado: (2024)
Vitron: A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing
por: Fei, Hao, et al.
Publicado: (2024)
por: Fei, Hao, et al.
Publicado: (2024)
AltChart: Enhancing VLM-based Chart Summarization Through Multi-Pretext Tasks
por: Moured, Omar, et al.
Publicado: (2024)
por: Moured, Omar, et al.
Publicado: (2024)
Interactivity x Explainability: Toward Understanding How Interactivity Can Improve Computer Vision Explanations
por: Panigrahi, Indu, et al.
Publicado: (2025)
por: Panigrahi, Indu, et al.
Publicado: (2025)
Ejemplares similares
-
EmoCLIP: A Vision-Language Method for Zero-Shot Video Facial Expression Recognition
por: Foteinopoulou, Niki Maria, et al.
Publicado: (2023) -
Vision-Free Retrieval: Rethinking Multimodal Search with Textual Scene Descriptions
por: Ntinou, Ioanna, et al.
Publicado: (2025) -
Multiscale Vision Transformers meet Bipartite Matching for efficient single-stage Action Localization
por: Ntinou, Ioanna, et al.
Publicado: (2023) -
MeMSVD: Long-Range Temporal Structure Capturing Using Incremental SVD
por: Ntinou, Ioanna, et al.
Publicado: (2024) -
CLIPCleaner: Cleaning Noisy Labels with CLIP
por: Feng, Chen, et al.
Publicado: (2024)