X-Actor: Emotional and Expressive Long-Range Portrait Acting from Audio
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Chenxu, Li, Zenan, Xu, Hongyi, Xie, You, Zhao, Xiaochen, Gu, Tianpei, Song, Guoxian, Chen, Xin, Liang, Chao, Jiang, Jianwen, Luo, Linjie |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
X-UniMotion: Animating Human Images with Expressive, Unified and Identity-Agnostic Motion Latents
por: Song, Guoxian, et al.
Publicado: (2025)
por: Song, Guoxian, et al.
Publicado: (2025)
X-Streamer: Unified Human World Modeling with Audiovisual Interaction
por: Xie, You, et al.
Publicado: (2025)
por: Xie, You, et al.
Publicado: (2025)
X-Portrait: Expressive Portrait Animation with Hierarchical Motion Attention
por: Xie, You, et al.
Publicado: (2024)
por: Xie, You, et al.
Publicado: (2024)
Plan-X: Instruct Video Generation via Semantic Planning
por: Huang, Lun, et al.
Publicado: (2025)
por: Huang, Lun, et al.
Publicado: (2025)
X-Dancer: Expressive Music to Human Dance Video Generation
por: Chen, Zeyuan, et al.
Publicado: (2025)
por: Chen, Zeyuan, et al.
Publicado: (2025)
X-NeMo: Expressive Neural Motion Reenactment via Disentangled Latent Attention
por: Zhao, Xiaochen, et al.
Publicado: (2025)
por: Zhao, Xiaochen, et al.
Publicado: (2025)
DiffPortrait3D: Controllable Diffusion for Zero-Shot Portrait View Synthesis
por: Gu, Yuming, et al.
Publicado: (2023)
por: Gu, Yuming, et al.
Publicado: (2023)
X-Dyna: Expressive Dynamic Human Image Animation
por: Chang, Di, et al.
Publicado: (2025)
por: Chang, Di, et al.
Publicado: (2025)
Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
por: Jiang, Jianwen, et al.
Publicado: (2024)
por: Jiang, Jianwen, et al.
Publicado: (2024)
Lynx: Towards High-Fidelity Personalized Video Generation
por: Sang, Shen, et al.
Publicado: (2025)
por: Sang, Shen, et al.
Publicado: (2025)
Bridging Your Imagination with Audio-Video Generation via a Unified Director
por: Zhang, Jiaxu, et al.
Publicado: (2025)
por: Zhang, Jiaxu, et al.
Publicado: (2025)
High Quality Human Image Animation using Regional Supervision and Motion Blur Condition
por: Xu, Zhongcong, et al.
Publicado: (2024)
por: Xu, Zhongcong, et al.
Publicado: (2024)
DeX-Portrait: Disentangled and Expressive Portrait Animation via Explicit and Latent Motion Representations
por: Shi, Yuxiang, et al.
Publicado: (2025)
por: Shi, Yuxiang, et al.
Publicado: (2025)
GMTalker: Gaussian Mixture-based Audio-Driven Emotional Talking Video Portraits
por: Xia, Yibo, et al.
Publicado: (2023)
por: Xia, Yibo, et al.
Publicado: (2023)
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
por: Tian, Linrui, et al.
Publicado: (2024)
por: Tian, Linrui, et al.
Publicado: (2024)
ExpPortrait: Expressive Portrait Generation via Personalized Representation
por: Wang, Junyi, et al.
Publicado: (2026)
por: Wang, Junyi, et al.
Publicado: (2026)
Acting Emotions
por: Konijn, Elly
Publicado: (2010)
por: Konijn, Elly
Publicado: (2010)
Expressive Range Characterization of Open Text-to-Audio Models
por: Morse, Jonathan, et al.
Publicado: (2025)
por: Morse, Jonathan, et al.
Publicado: (2025)
DiffPortrait360: Consistent Portrait Diffusion for 360 View Synthesis
por: Gu, Yuming, et al.
Publicado: (2025)
por: Gu, Yuming, et al.
Publicado: (2025)
MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation
por: Liu, Huaize, et al.
Publicado: (2025)
por: Liu, Huaize, et al.
Publicado: (2025)
CartoonAlive: Towards Expressive Live2D Modeling from Single Portraits
por: He, Chao, et al.
Publicado: (2025)
por: He, Chao, et al.
Publicado: (2025)
MobilePortrait: Real-Time One-Shot Neural Head Avatars on Mobile Devices
por: Jiang, Jianwen, et al.
Publicado: (2024)
por: Jiang, Jianwen, et al.
Publicado: (2024)
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation
por: Yang, Shurong, et al.
Publicado: (2024)
por: Yang, Shurong, et al.
Publicado: (2024)
Learning Feature-Preserving Portrait Editing from Generated Pairs
por: Chen, Bowei, et al.
Publicado: (2024)
por: Chen, Bowei, et al.
Publicado: (2024)
AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
por: Wei, Huawei, et al.
Publicado: (2024)
por: Wei, Huawei, et al.
Publicado: (2024)
CU-Mamba: Selective State Space Models with Channel Learning for Image Restoration
por: Deng, Rui, et al.
Publicado: (2024)
por: Deng, Rui, et al.
Publicado: (2024)
How Does Mandated Non‐Financial Disclosure Affect Corporate Cash Holdings? Evidence From the CSR Disclosure Mandate in China
por: Adrian Cheung, et al.
Publicado: (2026)
por: Adrian Cheung, et al.
Publicado: (2026)
PersonaLive! Expressive Portrait Image Animation for Live Streaming
por: Li, Zhiyuan, et al.
Publicado: (2025)
por: Li, Zhiyuan, et al.
Publicado: (2025)
FlowPortrait: Reinforcement Learning for Audio-Driven Portrait Video Generation
por: Tan, Weiting, et al.
Publicado: (2026)
por: Tan, Weiting, et al.
Publicado: (2026)
SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers
por: Fei, Zhengcong, et al.
Publicado: (2025)
por: Fei, Zhengcong, et al.
Publicado: (2025)
Seeing is Believing: Emotion-Aware Audio-Visual Language Modeling for Expressive Speech Generation
por: Tan, Weiting, et al.
Publicado: (2025)
por: Tan, Weiting, et al.
Publicado: (2025)
Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation
por: Cui, Jiahao, et al.
Publicado: (2024)
por: Cui, Jiahao, et al.
Publicado: (2024)
Video-As-Prompt: Unified Semantic Control for Video Generation
por: Bian, Yuxuan, et al.
Publicado: (2025)
por: Bian, Yuxuan, et al.
Publicado: (2025)
Follow-Your-Emoji: Fine-Controllable and Expressive Freestyle Portrait Animation
por: Ma, Yue, et al.
Publicado: (2024)
por: Ma, Yue, et al.
Publicado: (2024)
MegActor-$Σ$: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
por: Yang, Shurong, et al.
Publicado: (2024)
por: Yang, Shurong, et al.
Publicado: (2024)
Breaking the Trade-Off Between Faithfulness and Expressiveness for Large Language Models
por: Yang, Chenxu, et al.
Publicado: (2025)
por: Yang, Chenxu, et al.
Publicado: (2025)
Revealing the Long‐Range Coupling for Multi‐Dimensional Metasurface Multiplexer
por: Ouling Wu, et al.
Publicado: (2026)
por: Ouling Wu, et al.
Publicado: (2026)
Sonic: Shifting Focus to Global Audio Perception in Portrait Animation
por: Ji, Xiaozhong, et al.
Publicado: (2024)
por: Ji, Xiaozhong, et al.
Publicado: (2024)
Planarized Sidewall‐Integrated Metafibers for Enhanced Evanescent Field Interaction via Long‐Range Resonant Near‐Field Coupling
por: Chao Zeng, et al.
Publicado: (2026)
por: Chao Zeng, et al.
Publicado: (2026)
CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention
por: Lin, Gaojie, et al.
Publicado: (2024)
por: Lin, Gaojie, et al.
Publicado: (2024)
Ejemplares similares
-
X-UniMotion: Animating Human Images with Expressive, Unified and Identity-Agnostic Motion Latents
por: Song, Guoxian, et al.
Publicado: (2025) -
X-Streamer: Unified Human World Modeling with Audiovisual Interaction
por: Xie, You, et al.
Publicado: (2025) -
X-Portrait: Expressive Portrait Animation with Hierarchical Motion Attention
por: Xie, You, et al.
Publicado: (2024) -
Plan-X: Instruct Video Generation via Semantic Planning
por: Huang, Lun, et al.
Publicado: (2025) -
X-Dancer: Expressive Music to Human Dance Video Generation
por: Chen, Zeyuan, et al.
Publicado: (2025)