EmoVIT: Revolutionizing Emotion Insights with Visual Instruction Tuning
Fuente:
arXiv
Guardado en:
| Autores principales: | Xie, Hongxia, Peng, Chu-Jun, Tseng, Yu-Wen, Chen, Hung-Jen, Hsu, Chan-Feng, Shuai, Hong-Han, Cheng, Wen-Huang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Future Sight and Tough Fights: Revolutionizing Sequential Recommendation with FENRec
por: Huang, Yu-Hsuan, et al.
Publicado: (2024)
por: Huang, Yu-Hsuan, et al.
Publicado: (2024)
EmoArt: A Multidimensional Dataset for Emotion-Aware Artistic Generation
por: Zhang, Cheng, et al.
Publicado: (2025)
por: Zhang, Cheng, et al.
Publicado: (2025)
Perspective-Aware Teaching: Adapting Knowledge for Heterogeneous Distillation
por: Lin, Jhe-Hao, et al.
Publicado: (2025)
por: Lin, Jhe-Hao, et al.
Publicado: (2025)
Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
por: Chen, Pei-Chi, et al.
Publicado: (2025)
por: Chen, Pei-Chi, et al.
Publicado: (2025)
NUS-Emo at SemEval-2024 Task 3: Instruction-Tuning LLM for Multimodal Emotion-Cause Analysis in Conversations
por: Luo, Meng, et al.
Publicado: (2024)
por: Luo, Meng, et al.
Publicado: (2024)
EL-VIT: Probing Vision Transformer with Interactive Visualization
por: Zhou, Hong, et al.
Publicado: (2024)
por: Zhou, Hong, et al.
Publicado: (2024)
EmoVerse: A MLLMs-Driven Emotion Representation Dataset for Interpretable Visual Emotion Analysis
por: Guo, Yijie, et al.
Publicado: (2025)
por: Guo, Yijie, et al.
Publicado: (2025)
The Fabrication of Reality and Fantasy: Scene Generation with LLM-Assisted Prompt Interpretation
por: Yao, Yi, et al.
Publicado: (2024)
por: Yao, Yi, et al.
Publicado: (2024)
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
por: Cheng, Zebang, et al.
Publicado: (2024)
por: Cheng, Zebang, et al.
Publicado: (2024)
Visual Instruction Tuning with Chain of Region-of-Interest
por: Chen, Yixin, et al.
Publicado: (2025)
por: Chen, Yixin, et al.
Publicado: (2025)
RecipeGen: A Step-Aligned Multimodal Benchmark for Real-World Recipe Generation
por: Zhang, Ruoxuan, et al.
Publicado: (2025)
por: Zhang, Ruoxuan, et al.
Publicado: (2025)
Attribute-Grounded Selective Reasoning for Artwork Emotion Understanding with Multimodal Large Language Models
por: Zhang, Cheng, et al.
Publicado: (2026)
por: Zhang, Cheng, et al.
Publicado: (2026)
Emo-Tune: Harnessing Emotion-based Music for Patient Wellness
por: N Jothy, et al.
Publicado: (2026)
por: N Jothy, et al.
Publicado: (2026)
EmoAssist: Emotional Assistant for Visual Impairment Community
por: Qi, Xingyu, et al.
Publicado: (2025)
por: Qi, Xingyu, et al.
Publicado: (2025)
Learning to Instruct for Visual Instruction Tuning
por: Zhou, Zhihan, et al.
Publicado: (2025)
por: Zhou, Zhihan, et al.
Publicado: (2025)
Emo-bias: A Large Scale Evaluation of Social Bias on Speech Emotion Recognition
por: Lin, Yi-Cheng, et al.
Publicado: (2024)
por: Lin, Yi-Cheng, et al.
Publicado: (2024)
MemEmo: Evaluating Emotion in Memory Systems of Agents
por: Liu, Peng, et al.
Publicado: (2026)
por: Liu, Peng, et al.
Publicado: (2026)
A Dataset and Baselines for Measuring and Predicting the Music Piece Memorability
por: Tseng, Li-Yang, et al.
Publicado: (2024)
por: Tseng, Li-Yang, et al.
Publicado: (2024)
EmoGist: Efficient In-Context Learning for Visual Emotion Understanding
por: Seoh, Ronald, et al.
Publicado: (2025)
por: Seoh, Ronald, et al.
Publicado: (2025)
EmoSEM: Segment and Explain Emotion Stimuli in Visual Art
por: Zhang, Jing, et al.
Publicado: (2025)
por: Zhang, Jing, et al.
Publicado: (2025)
Single Document Image Highlight Removal via A Large-Scale Real-World Dataset and A Location-Aware Network
por: Pan, Lu, et al.
Publicado: (2025)
por: Pan, Lu, et al.
Publicado: (2025)
What Makes for Good Visual Instructions? Synthesizing Complex Visual Reasoning Instructions for Visual Instruction Tuning
por: Du, Yifan, et al.
Publicado: (2023)
por: Du, Yifan, et al.
Publicado: (2023)
Synergistic Effects of Interfacial Electric Fields and Cryogenic Temperatures on the Performance of Li‐Ion Electronic Synapses
por: Chao‐Hung Wang, et al.
Publicado: (2025)
por: Chao‐Hung Wang, et al.
Publicado: (2025)
EmoTalker: Emotionally Editable Talking Face Generation via Diffusion Model
por: Zhang, Bingyuan, et al.
Publicado: (2024)
por: Zhang, Bingyuan, et al.
Publicado: (2024)
LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning
por: You, Zebin, et al.
Publicado: (2025)
por: You, Zebin, et al.
Publicado: (2025)
EmoEdit: Evoking Emotions through Image Manipulation
por: Yang, Jingyuan, et al.
Publicado: (2024)
por: Yang, Jingyuan, et al.
Publicado: (2024)
Seismocardiography for Emotion Recognition: A Study on EmoWear with Insights from DEAP
por: Rahmani, Mohammad Hasan, et al.
Publicado: (2024)
por: Rahmani, Mohammad Hasan, et al.
Publicado: (2024)
CrossVIT-augmented Geospatial-Intelligence Visualization System for Tracking Economic Development Dynamics
por: Bai, Yanbing, et al.
Publicado: (2024)
por: Bai, Yanbing, et al.
Publicado: (2024)
CookAnything: A Framework for Flexible and Consistent Multi-Step Recipe Image Generation
por: Zhang, Ruoxuan, et al.
Publicado: (2025)
por: Zhang, Ruoxuan, et al.
Publicado: (2025)
VoxEmo: Benchmarking Speech Emotion Recognition with Speech LLMs
por: Zhang, Hezhao, et al.
Publicado: (2026)
por: Zhang, Hezhao, et al.
Publicado: (2026)
AV-EmoDialog: Chat with Audio-Visual Users Leveraging Emotional Cues
por: Park, Se Jin, et al.
Publicado: (2024)
por: Park, Se Jin, et al.
Publicado: (2024)
Transforming Redaction: How AI is Revolutionizing Data Protection
por: Peng, Sida, et al.
Publicado: (2024)
por: Peng, Sida, et al.
Publicado: (2024)
A DeNoising FPN With Transformer R-CNN for Tiny Object Detection
por: Liu, Hou-I, et al.
Publicado: (2024)
por: Liu, Hou-I, et al.
Publicado: (2024)
Personalized Visual Instruction Tuning
por: Pi, Renjie, et al.
Publicado: (2024)
por: Pi, Renjie, et al.
Publicado: (2024)
EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting
por: Yang, Guanrou, et al.
Publicado: (2025)
por: Yang, Guanrou, et al.
Publicado: (2025)
Multi-modal Instruction Tuned LLMs with Fine-grained Visual Perception
por: He, Junwen, et al.
Publicado: (2024)
por: He, Junwen, et al.
Publicado: (2024)
EmoAttack: Utilizing Emotional Voice Conversion for Speech Backdoor Attacks on Deep Speech Classification Models
por: Yao, Wenhan, et al.
Publicado: (2024)
por: Yao, Wenhan, et al.
Publicado: (2024)
Video-LaVIT: Unified Video-Language Pre-training with Decoupled Visual-Motional Tokenization
por: Jin, Yang, et al.
Publicado: (2024)
por: Jin, Yang, et al.
Publicado: (2024)
Real‐World Evidence That Non‐Smokers With High PD‐L1 Non‐Squamous NSCLC Have Poorer Outcomes With Immune Checkpoint Inhibitors
por: Yu‐Chu Kuo, et al.
Publicado: (2025)
por: Yu‐Chu Kuo, et al.
Publicado: (2025)
EmoPrefer: Can Large Language Models Understand Human Emotion Preferences?
por: Lian, Zheng, et al.
Publicado: (2025)
por: Lian, Zheng, et al.
Publicado: (2025)
Ejemplares similares
-
Future Sight and Tough Fights: Revolutionizing Sequential Recommendation with FENRec
por: Huang, Yu-Hsuan, et al.
Publicado: (2024) -
EmoArt: A Multidimensional Dataset for Emotion-Aware Artistic Generation
por: Zhang, Cheng, et al.
Publicado: (2025) -
Perspective-Aware Teaching: Adapting Knowledge for Heterogeneous Distillation
por: Lin, Jhe-Hao, et al.
Publicado: (2025) -
Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
por: Chen, Pei-Chi, et al.
Publicado: (2025) -
NUS-Emo at SemEval-2024 Task 3: Instruction-Tuning LLM for Multimodal Emotion-Cause Analysis in Conversations
por: Luo, Meng, et al.
Publicado: (2024)