Visual Prompting in LLMs for Enhancing Emotion Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Qixuan, Wang, Zhifeng, Zhang, Dylan, Niu, Wenjia, Caldwell, Sabrina, Gedeon, Tom, Liu, Yang, Qin, Zhenyue |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
di: Wang, Zhifeng, et al.
Pubblicazione: (2025)
di: Wang, Zhifeng, et al.
Pubblicazione: (2025)
Authentic Emotion Mapping: Benchmarking Facial Expressions in Real News
di: Zhang, Qixuan, et al.
Pubblicazione: (2024)
di: Zhang, Qixuan, et al.
Pubblicazione: (2024)
HandCraft: Anatomically Correct Restoration of Malformed Hands in Diffusion Generated Images
di: Qin, Zhenyue, et al.
Pubblicazione: (2024)
di: Qin, Zhenyue, et al.
Pubblicazione: (2024)
Representation-Centric Survey of Supervised Skeletal Action Recognition and the New Benchmark
di: Liu, Yang, et al.
Pubblicazione: (2022)
di: Liu, Yang, et al.
Pubblicazione: (2022)
LLDif: Diffusion Models for Low-light Emotion Recognition
di: Wang, Zhifeng, et al.
Pubblicazione: (2024)
di: Wang, Zhifeng, et al.
Pubblicazione: (2024)
LRDif: Diffusion Models for Under-Display Camera Emotion Recognition
di: Wang, Zhifeng, et al.
Pubblicazione: (2024)
di: Wang, Zhifeng, et al.
Pubblicazione: (2024)
Taylor Videos for Action Recognition
di: Wang, Lei, et al.
Pubblicazione: (2024)
di: Wang, Lei, et al.
Pubblicazione: (2024)
Bipartite Mode Matching for Vision Training Set Search from a Hierarchical Data Server
di: Yao, Yue, et al.
Pubblicazione: (2026)
di: Yao, Yue, et al.
Pubblicazione: (2026)
Improving Visual Prompt Tuning by Gaussian Neighborhood Minimization for Long-Tailed Visual Recognition
di: Li, Mengke, et al.
Pubblicazione: (2024)
di: Li, Mengke, et al.
Pubblicazione: (2024)
When Spatial meets Temporal in Action Recognition
di: Chen, Huilin, et al.
Pubblicazione: (2024)
di: Chen, Huilin, et al.
Pubblicazione: (2024)
Motion meets Attention: Video Motion Prompts
di: Chen, Qixiang, et al.
Pubblicazione: (2024)
di: Chen, Qixiang, et al.
Pubblicazione: (2024)
Cross-modal Prompting for Balanced Incomplete Multi-modal Emotion Recognition
di: He, Wen-Jue, et al.
Pubblicazione: (2025)
di: He, Wen-Jue, et al.
Pubblicazione: (2025)
Feature-Based Dual Visual Feature Extraction Model for Compound Multimodal Emotion Recognition
di: Liu, Ran, et al.
Pubblicazione: (2025)
di: Liu, Ran, et al.
Pubblicazione: (2025)
Multimodal Emotion Recognition with Vision-language Prompting and Modality Dropout
di: QI, Anbin, et al.
Pubblicazione: (2024)
di: QI, Anbin, et al.
Pubblicazione: (2024)
Synergistic Prompting for Robust Visual Recognition with Missing Modalities
di: Zhang, Zhihui, et al.
Pubblicazione: (2025)
di: Zhang, Zhihui, et al.
Pubblicazione: (2025)
Gems: Group Emotion Profiling Through Multimodal Situational Understanding
di: Kataria, Anubhav, et al.
Pubblicazione: (2025)
di: Kataria, Anubhav, et al.
Pubblicazione: (2025)
Ranked from Within: Ranking Large Multimodal Models Without Labels
di: Tu, Weijie, et al.
Pubblicazione: (2024)
di: Tu, Weijie, et al.
Pubblicazione: (2024)
Toward a Holistic Evaluation of Robustness in CLIP Models
di: Tu, Weijie, et al.
Pubblicazione: (2024)
di: Tu, Weijie, et al.
Pubblicazione: (2024)
An Empirical Study Into What Matters for Calibrating Vision-Language Models
di: Tu, Weijie, et al.
Pubblicazione: (2024)
di: Tu, Weijie, et al.
Pubblicazione: (2024)
TrackNetV4: Enhancing Fast Sports Object Tracking with Motion Attention Maps
di: Raj, Arjun, et al.
Pubblicazione: (2024)
di: Raj, Arjun, et al.
Pubblicazione: (2024)
SonoSelect: Efficient Ultrasound Perception via Active Probe Exploration
di: Zhang, Yixin, et al.
Pubblicazione: (2026)
di: Zhang, Yixin, et al.
Pubblicazione: (2026)
IF-Bench: Benchmarking and Enhancing MLLMs for Infrared Images with Generative Visual Prompting
di: Zhang, Tao, et al.
Pubblicazione: (2025)
di: Zhang, Tao, et al.
Pubblicazione: (2025)
Detail-Enhanced Intra- and Inter-modal Interaction for Audio-Visual Emotion Recognition
di: Shi, Tong, et al.
Pubblicazione: (2024)
di: Shi, Tong, et al.
Pubblicazione: (2024)
Enhancing Micro Gesture Recognition for Emotion Understanding via Context-aware Visual-Text Contrastive Learning
di: Li, Deng, et al.
Pubblicazione: (2024)
di: Li, Deng, et al.
Pubblicazione: (2024)
E-InMeMo: Enhanced Prompting for Visual In-Context Learning
di: Zhang, Jiahao, et al.
Pubblicazione: (2025)
di: Zhang, Jiahao, et al.
Pubblicazione: (2025)
Hints of Prompt: Enhancing Visual Representation for Multimodal LLMs in Autonomous Driving
di: Zhou, Hao, et al.
Pubblicazione: (2024)
di: Zhou, Hao, et al.
Pubblicazione: (2024)
PromptHub: Enhancing Multi-Prompt Visual In-Context Learning with Locality-Aware Fusion, Concentration and Alignment
di: Luo, Tianci, et al.
Pubblicazione: (2026)
di: Luo, Tianci, et al.
Pubblicazione: (2026)
Grounding Emotion Recognition with Visual Prototypes: VEGA -- Revisiting CLIP in MERC
di: Hu, Guanyu, et al.
Pubblicazione: (2025)
di: Hu, Guanyu, et al.
Pubblicazione: (2025)
A Survey on Facial Expression Recognition of Static and Dynamic Emotions
di: Wang, Yan, et al.
Pubblicazione: (2024)
di: Wang, Yan, et al.
Pubblicazione: (2024)
LMOD: A Large Multimodal Ophthalmology Dataset and Benchmark for Large Vision-Language Models
di: Qin, Zhenyue, et al.
Pubblicazione: (2024)
di: Qin, Zhenyue, et al.
Pubblicazione: (2024)
Visual Prompt-Agnostic Evolution
di: Wang, Junze, et al.
Pubblicazione: (2026)
di: Wang, Junze, et al.
Pubblicazione: (2026)
SEP: Self-Enhanced Prompt Tuning for Visual-Language Model
di: Yao, Hantao, et al.
Pubblicazione: (2024)
di: Yao, Hantao, et al.
Pubblicazione: (2024)
Multiverse Through Deepfakes: The MultiFakeVerse Dataset of Person-Centric Visual and Conceptual Manipulations
di: Gupta, Parul, et al.
Pubblicazione: (2025)
di: Gupta, Parul, et al.
Pubblicazione: (2025)
Robust Emotion Recognition in Context Debiasing
di: Yang, Dingkang, et al.
Pubblicazione: (2024)
di: Yang, Dingkang, et al.
Pubblicazione: (2024)
Knowledge-Aligned Counterfactual-Enhancement Diffusion Perception for Unsupervised Cross-Domain Visual Emotion Recognition
di: Yin, Wen, et al.
Pubblicazione: (2025)
di: Yin, Wen, et al.
Pubblicazione: (2025)
Attend and Enrich: Enhanced Visual Prompt for Zero-Shot Learning
di: Liu, Man, et al.
Pubblicazione: (2024)
di: Liu, Man, et al.
Pubblicazione: (2024)
A Multimodal Fusion Network For Student Emotion Recognition Based on Transformer and Tensor Product
di: Xiang, Ao, et al.
Pubblicazione: (2024)
di: Xiang, Ao, et al.
Pubblicazione: (2024)
XR-VLM: Cross-Relationship Modeling with Multi-part Prompts and Visual Features for Fine-Grained Recognition
di: Wang, Chuanming, et al.
Pubblicazione: (2025)
di: Wang, Chuanming, et al.
Pubblicazione: (2025)
Panther: Illuminate the Sight of Multimodal LLMs with Instruction-Guided Visual Prompts
di: Li, Honglin, et al.
Pubblicazione: (2024)
di: Li, Honglin, et al.
Pubblicazione: (2024)
A Closer Look at the Robustness of Contrastive Language-Image Pre-Training (CLIP)
di: Tu, Weijie, et al.
Pubblicazione: (2024)
di: Tu, Weijie, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
di: Wang, Zhifeng, et al.
Pubblicazione: (2025) -
Authentic Emotion Mapping: Benchmarking Facial Expressions in Real News
di: Zhang, Qixuan, et al.
Pubblicazione: (2024) -
HandCraft: Anatomically Correct Restoration of Malformed Hands in Diffusion Generated Images
di: Qin, Zhenyue, et al.
Pubblicazione: (2024) -
Representation-Centric Survey of Supervised Skeletal Action Recognition and the New Benchmark
di: Liu, Yang, et al.
Pubblicazione: (2022) -
LLDif: Diffusion Models for Low-light Emotion Recognition
di: Wang, Zhifeng, et al.
Pubblicazione: (2024)