EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tian, Linrui, Hu, Siqi, Wang, Qi, Zhang, Bang, Bo, Liefeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
von: Tian, Linrui, et al.
Veröffentlicht: (2024)
von: Tian, Linrui, et al.
Veröffentlicht: (2024)
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025)
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025)
Wan-S2V: Audio-Driven Cinematic Video Generation
von: Gao, Xin, et al.
Veröffentlicht: (2025)
von: Gao, Xin, et al.
Veröffentlicht: (2025)
MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025)
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025)
Controllable and Expressive One-Shot Video Head Swapping
von: Ji, Chaonan, et al.
Veröffentlicht: (2025)
von: Ji, Chaonan, et al.
Veröffentlicht: (2025)
ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
von: Qi, Jinwei, et al.
Veröffentlicht: (2025)
von: Qi, Jinwei, et al.
Veröffentlicht: (2025)
UniLS: End-to-End Audio-Driven Avatars for Unified Listening and Speaking
von: Chu, Xuangeng, et al.
Veröffentlicht: (2025)
von: Chu, Xuangeng, et al.
Veröffentlicht: (2025)
A Unit Enhancement and Guidance Framework for Audio-Driven Avatar Video Generation
von: Zhou, S. Z., et al.
Veröffentlicht: (2025)
von: Zhou, S. Z., et al.
Veröffentlicht: (2025)
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
von: Hu, Li, et al.
Veröffentlicht: (2023)
von: Hu, Li, et al.
Veröffentlicht: (2023)
ViSAudio: End-to-End Video-Driven Binaural Spatial Audio Generation
von: Zhang, Mengchen, et al.
Veröffentlicht: (2025)
von: Zhang, Mengchen, et al.
Veröffentlicht: (2025)
OutfitAnyone: Ultra-high Quality Virtual Try-On for Any Clothing and Any Person
von: Sun, Ke, et al.
Veröffentlicht: (2024)
von: Sun, Ke, et al.
Veröffentlicht: (2024)
Baton: Explicit Semantic Blueprints for Joint Video-Audio Generation
von: Tu, Shuyuan, et al.
Veröffentlicht: (2026)
von: Tu, Shuyuan, et al.
Veröffentlicht: (2026)
Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length
von: Huang, Yubo, et al.
Veröffentlicht: (2025)
von: Huang, Yubo, et al.
Veröffentlicht: (2025)
Exploring Timeline Control for Facial Motion Generation
von: Ma, Yifeng, et al.
Veröffentlicht: (2025)
von: Ma, Yifeng, et al.
Veröffentlicht: (2025)
Synchronized Video-to-Audio Generation via Mel Quantization-Continuum Decomposition
von: Wang, Juncheng, et al.
Veröffentlicht: (2025)
von: Wang, Juncheng, et al.
Veröffentlicht: (2025)
OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models
von: Yu, Haojie, et al.
Veröffentlicht: (2025)
von: Yu, Haojie, et al.
Veröffentlicht: (2025)
AnyText2: Visual Text Generation and Editing With Customizable Attributes
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2024)
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2024)
Audio-Driven Universal Gaussian Head Avatars
von: Teotia, Kartik, et al.
Veröffentlicht: (2025)
von: Teotia, Kartik, et al.
Veröffentlicht: (2025)
HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
von: Chen, Yi, et al.
Veröffentlicht: (2025)
von: Chen, Yi, et al.
Veröffentlicht: (2025)
Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation
von: Chen, Yuheng, et al.
Veröffentlicht: (2026)
von: Chen, Yuheng, et al.
Veröffentlicht: (2026)
InstructAV2AV: Instruction-Guided Audio-Video Joint Editing
von: Zheng, Haojie, et al.
Veröffentlicht: (2026)
von: Zheng, Haojie, et al.
Veröffentlicht: (2026)
JoyStreamer-Flash: Real-time and Infinite Audio-Driven Avatar Generation with Autoregressive Diffusion
von: Li, Chaochao, et al.
Veröffentlicht: (2025)
von: Li, Chaochao, et al.
Veröffentlicht: (2025)
I4VGen: Image as Free Stepping Stone for Text-to-Video Generation
von: Guo, Xiefan, et al.
Veröffentlicht: (2024)
von: Guo, Xiefan, et al.
Veröffentlicht: (2024)
DiffuEraser: A Diffusion Model for Video Inpainting
von: Li, Xiaowen, et al.
Veröffentlicht: (2025)
von: Li, Xiaowen, et al.
Veröffentlicht: (2025)
Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance
von: Hu, Li, et al.
Veröffentlicht: (2025)
von: Hu, Li, et al.
Veröffentlicht: (2025)
MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling
von: Men, Yifang, et al.
Veröffentlicht: (2024)
von: Men, Yifang, et al.
Veröffentlicht: (2024)
AMG: Avatar Motion Guided Video Generation
von: Yang, Zhangsihao, et al.
Veröffentlicht: (2024)
von: Yang, Zhangsihao, et al.
Veröffentlicht: (2024)
UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control
von: Sun, Wenzhang, et al.
Veröffentlicht: (2024)
von: Sun, Wenzhang, et al.
Veröffentlicht: (2024)
Video-Driven Animation of Neural Head Avatars
von: Paier, Wolfgang, et al.
Veröffentlicht: (2024)
von: Paier, Wolfgang, et al.
Veröffentlicht: (2024)
TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
AdaptiveDrag: Semantic-Driven Dragging on Diffusion-Based Image Editing
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
LAM: Large Avatar Model for One-shot Animatable Gaussian Head
von: He, Yisheng, et al.
Veröffentlicht: (2025)
von: He, Yisheng, et al.
Veröffentlicht: (2025)
CoGenAV: Versatile Audio-Visual Representation Learning via Contrastive-Generative Synchronization
von: Bai, Detao, et al.
Veröffentlicht: (2025)
von: Bai, Detao, et al.
Veröffentlicht: (2025)
OmniWeaving: Towards Unified Video Generation with Free-form Composition and Reasoning
von: Pan, Kaihang, et al.
Veröffentlicht: (2026)
von: Pan, Kaihang, et al.
Veröffentlicht: (2026)
GoalFlow: Goal-Driven Flow Matching for Multimodal Trajectories Generation in End-to-End Autonomous Driving
von: Xing, Zebin, et al.
Veröffentlicht: (2025)
von: Xing, Zebin, et al.
Veröffentlicht: (2025)
Foundation Feature-Driven Online End-Effector Pose Estimation: A Marker-Free and Learning-Free Approach
von: Wu, Tianshu, et al.
Veröffentlicht: (2025)
von: Wu, Tianshu, et al.
Veröffentlicht: (2025)
SmartAvatar: Text- and Image-Guided Human Avatar Generation with VLM AI Agents
von: Huang-Menders, Alexander, et al.
Veröffentlicht: (2025)
von: Huang-Menders, Alexander, et al.
Veröffentlicht: (2025)
Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
von: Jiang, Jianwen, et al.
Veröffentlicht: (2024)
von: Jiang, Jianwen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
von: Tian, Linrui, et al.
Veröffentlicht: (2024) -
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025) -
Wan-S2V: Audio-Driven Cinematic Video Generation
von: Gao, Xin, et al.
Veröffentlicht: (2025) -
MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation
von: Meng, Dechao, et al.
Veröffentlicht: (2025) -
StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025)