GPT-4V with Emotion: A Zero-shot Benchmark for Generalized Emotion Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lian, Zheng, Sun, Licai, Sun, Haiyang, Chen, Kang, Wen, Zhuofan, Gu, Hao, Liu, Bin, Tao, Jianhua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HiCMAE: Hierarchical Contrastive Masked Autoencoder for Self-Supervised Audio-Visual Emotion Recognition
von: Sun, Licai, et al.
Veröffentlicht: (2024)
von: Sun, Licai, et al.
Veröffentlicht: (2024)
ITEACH-Net: Inverted Teacher-studEnt seArCH Network for Emotion Recognition in Conversation
von: Sun, Haiyang, et al.
Veröffentlicht: (2023)
von: Sun, Haiyang, et al.
Veröffentlicht: (2023)
Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
SVFAP: Self-supervised Video Facial Affect Perceiver
von: Sun, Licai, et al.
Veröffentlicht: (2023)
von: Sun, Licai, et al.
Veröffentlicht: (2023)
Multi-modal Speech Emotion Recognition via Feature Distribution Adaptation Network
von: Li, Shaokai, et al.
Veröffentlicht: (2024)
von: Li, Shaokai, et al.
Veröffentlicht: (2024)
Emotion-Qwen: A Unified Framework for Emotion and Vision Understanding
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
von: Huang, Dawei, et al.
Veröffentlicht: (2025)
GAIA: Zero-shot Talking Avatar Generation
von: He, Tianyu, et al.
Veröffentlicht: (2023)
von: He, Tianyu, et al.
Veröffentlicht: (2023)
XEmoGPT: An Explainable Multimodal Emotion Recognition Framework with Cue-Level Perception and Reasoning
von: Zhang, Hanwen, et al.
Veröffentlicht: (2026)
von: Zhang, Hanwen, et al.
Veröffentlicht: (2026)
AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
MCIHN: A Hybrid Network Model Based on Multi-path Cross-modal Interaction for Multimodal Emotion Recognition
von: Zhang, Haoyang, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyang, et al.
Veröffentlicht: (2025)
Enhancing Modal Fusion by Alignment and Label Matching for Multimodal Emotion Recognition
von: Li, Qifei, et al.
Veröffentlicht: (2024)
von: Li, Qifei, et al.
Veröffentlicht: (2024)
MicroEmo: Time-Sensitive Multimodal Emotion Recognition with Micro-Expression Dynamics in Video Dialogues
von: Zhang, Liyun
Veröffentlicht: (2024)
von: Zhang, Liyun
Veröffentlicht: (2024)
MERBench: A Unified Evaluation Benchmark for Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
von: Lian, Zheng, et al.
Veröffentlicht: (2024)
Zero-shot Video Moment Retrieval via Off-the-shelf Multimodal Large Language Models
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
OT-DETECTOR: Delving into Optimal Transport for Zero-shot Out-of-Distribution Detection
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
Towards Emotion Analysis in Short-form Videos: A Large-Scale Dataset and Baseline
von: Wu, Xuecheng, et al.
Veröffentlicht: (2023)
von: Wu, Xuecheng, et al.
Veröffentlicht: (2023)
Anchoring Emotions in Text: Robust Multimodal Fusion for Mimicry Intensity Estimation
von: Zhu, Lingsi, et al.
Veröffentlicht: (2026)
von: Zhu, Lingsi, et al.
Veröffentlicht: (2026)
Calibrating Multimodal Consensus for Emotion Recognition
von: Zhong, Guowei, et al.
Veröffentlicht: (2025)
von: Zhong, Guowei, et al.
Veröffentlicht: (2025)
EmotionGesture: Audio-Driven Diverse Emotional Co-Speech 3D Gesture Generation
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
Modularized Zero-shot VQA with Pre-trained Models
von: Cao, Rui, et al.
Veröffentlicht: (2023)
von: Cao, Rui, et al.
Veröffentlicht: (2023)
Multimodal Fusion with Pre-Trained Model Features in Affective Behaviour Analysis In-the-wild
von: Wen, Zhuofan, et al.
Veröffentlicht: (2024)
von: Wen, Zhuofan, et al.
Veröffentlicht: (2024)
SZTU-CMU at MER2024: Improving Emotion-LLaMA with Conv-Attention for Multimodal Emotion Recognition
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
EALD-MLLM: Emotion Analysis in Long-sequential and De-identity videos with Multi-modal Large Language Model
von: Li, Deng, et al.
Veröffentlicht: (2024)
von: Li, Deng, et al.
Veröffentlicht: (2024)
Cross-domain Multi-step Thinking: Zero-shot Fine-grained Traffic Sign Recognition in the Wild
von: Gan, Yaozong, et al.
Veröffentlicht: (2024)
von: Gan, Yaozong, et al.
Veröffentlicht: (2024)
Interpretable Zero-shot Referring Expression Comprehension with Query-driven Scene Graphs
von: Wu, Yike, et al.
Veröffentlicht: (2026)
von: Wu, Yike, et al.
Veröffentlicht: (2026)
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
von: Cheng, Zebang, et al.
Veröffentlicht: (2024)
Enhancing Emotion Recognition in Incomplete Data: A Novel Cross-Modal Alignment, Reconstruction, and Refinement Framework
von: Sun, Haoqin, et al.
Veröffentlicht: (2024)
von: Sun, Haoqin, et al.
Veröffentlicht: (2024)
GANonymization: A GAN-based Face Anonymization Framework for Preserving Emotional Expressions
von: Hellmann, Fabio, et al.
Veröffentlicht: (2023)
von: Hellmann, Fabio, et al.
Veröffentlicht: (2023)
Do LLMs Understand Visual Anomalies? Uncovering LLM's Capabilities in Zero-shot Anomaly Detection
von: Zhu, Jiaqi, et al.
Veröffentlicht: (2024)
von: Zhu, Jiaqi, et al.
Veröffentlicht: (2024)
MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled Embedding
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
OralGPT-Omni: A Versatile Dental Multimodal Large Language Model
von: Hao, Jing, et al.
Veröffentlicht: (2025)
von: Hao, Jing, et al.
Veröffentlicht: (2025)
From Data Deluge to Data Curation: A Filtering-WoRA Paradigm for Efficient Text-based Person Search
von: Sun, Jintao, et al.
Veröffentlicht: (2024)
von: Sun, Jintao, et al.
Veröffentlicht: (2024)
Towards Robust Multimodal Emotion Recognition under Missing Modalities and Distribution Shifts
von: Zhong, Guowei, et al.
Veröffentlicht: (2025)
von: Zhong, Guowei, et al.
Veröffentlicht: (2025)
GalleryGPT: Analyzing Paintings with Large Multimodal Models
von: Bin, Yi, et al.
Veröffentlicht: (2024)
von: Bin, Yi, et al.
Veröffentlicht: (2024)
GAOT: Generating Articulated Objects Through Text-Guided Diffusion Models
von: Sun, Hao, et al.
Veröffentlicht: (2025)
von: Sun, Hao, et al.
Veröffentlicht: (2025)
Attributes-aware Visual Emotion Representation Learning
von: Maharjan, Rahul Singh, et al.
Veröffentlicht: (2025)
von: Maharjan, Rahul Singh, et al.
Veröffentlicht: (2025)
Pseudo-triplet Guided Few-shot Composed Image Retrieval
von: Hou, Bohan, et al.
Veröffentlicht: (2024)
von: Hou, Bohan, et al.
Veröffentlicht: (2024)
Zero-shot image privacy classification with Vision-Language Models
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
Diverse Sign Language Translation
von: Shen, Xin, et al.
Veröffentlicht: (2024)
von: Shen, Xin, et al.
Veröffentlicht: (2024)
EntroAD: Structural Entropy-Guided Prompt Adaptation for Zero-Shot Anomaly Detection
von: Zhao, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhao, Xinyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
HiCMAE: Hierarchical Contrastive Masked Autoencoder for Self-Supervised Audio-Visual Emotion Recognition
von: Sun, Licai, et al.
Veröffentlicht: (2024) -
ITEACH-Net: Inverted Teacher-studEnt seArCH Network for Emotion Recognition in Conversation
von: Sun, Haiyang, et al.
Veröffentlicht: (2023) -
Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023) -
SVFAP: Self-supervised Video Facial Affect Perceiver
von: Sun, Licai, et al.
Veröffentlicht: (2023) -
Multi-modal Speech Emotion Recognition via Feature Distribution Adaptation Network
von: Li, Shaokai, et al.
Veröffentlicht: (2024)