UniEmo: Unifying Emotional Understanding and Generation with Learnable Expert Queries
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Yijie, Zhang, Lingsen, Yu, Zitong, Shao, Rui, Tan, Tao, Nie, Liqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey
von: Shao, Rui, et al.
Veröffentlicht: (2025)
von: Shao, Rui, et al.
Veröffentlicht: (2025)
PUMA: Layer-Pruned Language Model for Efficient Unified Multimodal Retrieval with Modality-Adaptive Learning
von: Lyu, Yibo, et al.
Veröffentlicht: (2025)
von: Lyu, Yibo, et al.
Veröffentlicht: (2025)
CoEmoGen: Towards Semantically-Coherent and Scalable Emotional Image Content Generation
von: Yuan, Kaishen, et al.
Veröffentlicht: (2025)
von: Yuan, Kaishen, et al.
Veröffentlicht: (2025)
UniFork: Exploring Modality Alignment for Unified Multimodal Understanding and Generation
von: Li, Teng, et al.
Veröffentlicht: (2025)
von: Li, Teng, et al.
Veröffentlicht: (2025)
$Δ$VLA: Prior-Guided Vision-Language-Action Models via World Knowledge Variation
von: Zhu, Yijie, et al.
Veröffentlicht: (2026)
von: Zhu, Yijie, et al.
Veröffentlicht: (2026)
MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models
von: Shen, Leyang, et al.
Veröffentlicht: (2024)
von: Shen, Leyang, et al.
Veröffentlicht: (2024)
EmoVid: A Multimodal Emotion Video Dataset for Emotion-Centric Video Understanding and Generation
von: Qiu, Zongyang, et al.
Veröffentlicht: (2025)
von: Qiu, Zongyang, et al.
Veröffentlicht: (2025)
UniCVR: From Alignment to Reranking for Unified Zero-Shot Composed Visual Retrieval
von: Wen, Haokun, et al.
Veröffentlicht: (2026)
von: Wen, Haokun, et al.
Veröffentlicht: (2026)
EmoStory: Emotion-Aware Story Generation
von: Yang, Jingyuan, et al.
Veröffentlicht: (2026)
von: Yang, Jingyuan, et al.
Veröffentlicht: (2026)
Token-level Correlation-guided Compression for Efficient Multimodal Document Understanding
von: Zhang, Renshan, et al.
Veröffentlicht: (2024)
von: Zhang, Renshan, et al.
Veröffentlicht: (2024)
OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation
von: Wu, Size, et al.
Veröffentlicht: (2025)
von: Wu, Size, et al.
Veröffentlicht: (2025)
UniMesh: Unifying 3D Mesh Understanding and Generation
von: Huang, Peng, et al.
Veröffentlicht: (2026)
von: Huang, Peng, et al.
Veröffentlicht: (2026)
UniX: Unifying Autoregression and Diffusion for Chest X-Ray Understanding and Generation
von: Zhang, Ruiheng, et al.
Veröffentlicht: (2026)
von: Zhang, Ruiheng, et al.
Veröffentlicht: (2026)
UniVerse-1: Unified Audio-Video Generation via Stitching of Experts
von: Wang, Duomin, et al.
Veröffentlicht: (2025)
von: Wang, Duomin, et al.
Veröffentlicht: (2025)
EmoLLM: Multimodal Emotional Understanding Meets Large Language Models
von: Yang, Qu, et al.
Veröffentlicht: (2024)
von: Yang, Qu, et al.
Veröffentlicht: (2024)
EmoAttack: Emotion-to-Image Diffusion Models for Emotional Backdoor Generation
von: Wei, Tianyu, et al.
Veröffentlicht: (2024)
von: Wei, Tianyu, et al.
Veröffentlicht: (2024)
UniVideo: Unified Understanding, Generation, and Editing for Videos
von: Wei, Cong, et al.
Veröffentlicht: (2025)
von: Wei, Cong, et al.
Veröffentlicht: (2025)
EmoCtrl: Controllable Emotional Image Content Generation
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
EmoArt: A Multidimensional Dataset for Emotion-Aware Artistic Generation
von: Zhang, Cheng, et al.
Veröffentlicht: (2025)
von: Zhang, Cheng, et al.
Veröffentlicht: (2025)
EmoNet-Face: An Expert-Annotated Benchmark for Synthetic Emotion Recognition
von: Schuhmann, Christoph, et al.
Veröffentlicht: (2025)
von: Schuhmann, Christoph, et al.
Veröffentlicht: (2025)
Nano-EmoX: Unifying Multimodal Emotional Intelligence from Perception to Empathy
von: Huang, Jiahao, et al.
Veröffentlicht: (2026)
von: Huang, Jiahao, et al.
Veröffentlicht: (2026)
UniPaint: Unified Space-time Video Inpainting via Mixture-of-Experts
von: Wan, Zhen, et al.
Veröffentlicht: (2024)
von: Wan, Zhen, et al.
Veröffentlicht: (2024)
Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
von: Wang, Peiyu, et al.
Veröffentlicht: (2025)
UniGRPO: Unified Policy Optimization for Reasoning-Driven Visual Generation
von: Liu, Jie, et al.
Veröffentlicht: (2026)
von: Liu, Jie, et al.
Veröffentlicht: (2026)
EmoSign: A Multimodal Dataset for Understanding Emotions in American Sign Language
von: Chua, Phoebe, et al.
Veröffentlicht: (2025)
von: Chua, Phoebe, et al.
Veröffentlicht: (2025)
UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation
von: Tian, Rui, et al.
Veröffentlicht: (2025)
von: Tian, Rui, et al.
Veröffentlicht: (2025)
UniEval: Unified Holistic Evaluation for Unified Multimodal Understanding and Generation
von: Li, Yi, et al.
Veröffentlicht: (2025)
von: Li, Yi, et al.
Veröffentlicht: (2025)
Image Aesthetics Assessment via Learnable Queries
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2023)
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2023)
DeepFake-Adapter: Dual-Level Adapter for DeepFake Detection
von: Shao, Rui, et al.
Veröffentlicht: (2023)
von: Shao, Rui, et al.
Veröffentlicht: (2023)
UniUGG: Unified 3D Understanding and Generation via Geometric-Semantic Encoding
von: Xu, Yueming, et al.
Veröffentlicht: (2025)
von: Xu, Yueming, et al.
Veröffentlicht: (2025)
UniCTokens: Boosting Personalized Understanding and Generation via Unified Concept Tokens
von: An, Ruichuan, et al.
Veröffentlicht: (2025)
von: An, Ruichuan, et al.
Veröffentlicht: (2025)
UniVS: Unified and Universal Video Segmentation with Prompts as Queries
von: Li, Minghan, et al.
Veröffentlicht: (2024)
von: Li, Minghan, et al.
Veröffentlicht: (2024)
UniTok: A Unified Tokenizer for Visual Generation and Understanding
von: Ma, Chuofan, et al.
Veröffentlicht: (2025)
von: Ma, Chuofan, et al.
Veröffentlicht: (2025)
UniUGP: Unifying Understanding, Generation, and Planing For End-to-end Autonomous Driving
von: Lu, Hao, et al.
Veröffentlicht: (2025)
von: Lu, Hao, et al.
Veröffentlicht: (2025)
UniModel: A Visual-Only Framework for Unified Multimodal Understanding and Generation
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
VidEmo: Affective-Tree Reasoning for Emotion-Centric Video Foundation Models
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
UniShield: Unified Face Attack Detection via KG-Informed Multimodal Reasoning
von: Li, Hongrui, et al.
Veröffentlicht: (2026)
von: Li, Hongrui, et al.
Veröffentlicht: (2026)
MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models
von: Xie, Wulin, et al.
Veröffentlicht: (2025)
von: Xie, Wulin, et al.
Veröffentlicht: (2025)
UniFlow: A Unified Pixel Flow Tokenizer for Visual Understanding and Generation
von: Yue, Zhengrong, et al.
Veröffentlicht: (2025)
von: Yue, Zhengrong, et al.
Veröffentlicht: (2025)
UniCompress: Token Compression for Unified Vision-Language Understanding and Generation
von: Wang, Ziyao, et al.
Veröffentlicht: (2026)
von: Wang, Ziyao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey
von: Shao, Rui, et al.
Veröffentlicht: (2025) -
PUMA: Layer-Pruned Language Model for Efficient Unified Multimodal Retrieval with Modality-Adaptive Learning
von: Lyu, Yibo, et al.
Veröffentlicht: (2025) -
CoEmoGen: Towards Semantically-Coherent and Scalable Emotional Image Content Generation
von: Yuan, Kaishen, et al.
Veröffentlicht: (2025) -
UniFork: Exploring Modality Alignment for Unified Multimodal Understanding and Generation
von: Li, Teng, et al.
Veröffentlicht: (2025) -
$Δ$VLA: Prior-Guided Vision-Language-Action Models via World Knowledge Variation
von: Zhu, Yijie, et al.
Veröffentlicht: (2026)