FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, MengChao, Wang, Qiang, Jiang, Fan, Xu, Mu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
von: Wang, Mengchao, et al.
Veröffentlicht: (2025)
von: Wang, Mengchao, et al.
Veröffentlicht: (2025)
FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers
von: Wang, Qiang, et al.
Veröffentlicht: (2025)
von: Wang, Qiang, et al.
Veröffentlicht: (2025)
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation
von: Liang, Chao, et al.
Veröffentlicht: (2025)
von: Liang, Chao, et al.
Veröffentlicht: (2025)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
von: Wang, Baiqin, et al.
Veröffentlicht: (2025)
AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
von: Wei, Huawei, et al.
Veröffentlicht: (2024)
von: Wei, Huawei, et al.
Veröffentlicht: (2024)
Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation
von: Xu, Mingwang, et al.
Veröffentlicht: (2024)
von: Xu, Mingwang, et al.
Veröffentlicht: (2024)
GMTalker: Gaussian Mixture-based Audio-Driven Emotional Talking Video Portraits
von: Xia, Yibo, et al.
Veröffentlicht: (2023)
von: Xia, Yibo, et al.
Veröffentlicht: (2023)
Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation
von: Cui, Jiahao, et al.
Veröffentlicht: (2024)
von: Cui, Jiahao, et al.
Veröffentlicht: (2024)
SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation
von: Liu, Yujian, et al.
Veröffentlicht: (2025)
von: Liu, Yujian, et al.
Veröffentlicht: (2025)
FantasyID: Face Knowledge Enhanced ID-Preserving Video Generation
von: Zhang, Yunpeng, et al.
Veröffentlicht: (2025)
von: Zhang, Yunpeng, et al.
Veröffentlicht: (2025)
Exploiting Temporal Audio-Visual Correlation Embedding for Audio-Driven One-Shot Talking Head Animation
von: Xu, Zhihua, et al.
Veröffentlicht: (2025)
von: Xu, Zhihua, et al.
Veröffentlicht: (2025)
SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers
von: Fei, Zhengcong, et al.
Veröffentlicht: (2025)
von: Fei, Zhengcong, et al.
Veröffentlicht: (2025)
MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
von: Liu, Huaize, et al.
Veröffentlicht: (2025)
Hallo4: High-Fidelity Dynamic Portrait Animation via Direct Preference Optimization
von: Cui, Jiahao, et al.
Veröffentlicht: (2025)
von: Cui, Jiahao, et al.
Veröffentlicht: (2025)
FantasyHSI: Video-Generation-Centric 4D Human Synthesis In Any Scene through A Graph-based Multi-Agent Framework
von: Mu, Lingzhou, et al.
Veröffentlicht: (2025)
von: Mu, Lingzhou, et al.
Veröffentlicht: (2025)
LinguaLinker: Audio-Driven Portraits Animation with Implicit Facial Control Enhancement
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
Controllable Longer Image Animation with Diffusion Models
von: Wang, Qiang, et al.
Veröffentlicht: (2024)
von: Wang, Qiang, et al.
Veröffentlicht: (2024)
X-Portrait: Expressive Portrait Animation with Hierarchical Motion Attention
von: Xie, You, et al.
Veröffentlicht: (2024)
von: Xie, You, et al.
Veröffentlicht: (2024)
EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Chen, Zhiyuan, et al.
Veröffentlicht: (2024)
Teller: Real-Time Streaming Audio-Driven Portrait Animation with Autoregressive Motion Generation
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
UniTalking: A Unified Audio-Video Framework for Talking Portrait Generation
von: Li, Hebeizi, et al.
Veröffentlicht: (2026)
von: Li, Hebeizi, et al.
Veröffentlicht: (2026)
FantasyWorld: Geometry-Consistent World Modeling via Unified Video and 3D Prediction
von: Dai, Yixiang, et al.
Veröffentlicht: (2025)
von: Dai, Yixiang, et al.
Veröffentlicht: (2025)
JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation
von: Cao, Xuyang, et al.
Veröffentlicht: (2024)
von: Cao, Xuyang, et al.
Veröffentlicht: (2024)
LayerAnimate: Layer-level Control for Animation
von: Yang, Yuxue, et al.
Veröffentlicht: (2025)
von: Yang, Yuxue, et al.
Veröffentlicht: (2025)
FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent Prediction
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025)
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025)
Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
von: Jiang, Jianwen, et al.
Veröffentlicht: (2024)
von: Jiang, Jianwen, et al.
Veröffentlicht: (2024)
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models
von: Yu, Haojie, et al.
Veröffentlicht: (2025)
von: Yu, Haojie, et al.
Veröffentlicht: (2025)
FantasyVLN: Unified Multimodal Chain-of-Thought Reasoning for Vision-Language Navigation
von: Zuo, Jing, et al.
Veröffentlicht: (2026)
von: Zuo, Jing, et al.
Veröffentlicht: (2026)
FactorPortrait: Controllable Portrait Animation via Disentangled Expression, Pose, and Viewpoint
von: Tang, Jiapeng, et al.
Veröffentlicht: (2025)
von: Tang, Jiapeng, et al.
Veröffentlicht: (2025)
FG-Portrait: 3D Flow Guided Editable Portrait Animation
von: Xu, Yating, et al.
Veröffentlicht: (2026)
von: Xu, Yating, et al.
Veröffentlicht: (2026)
HunyuanPortrait: Implicit Condition Control for Enhanced Portrait Animation
von: Xu, Zunnan, et al.
Veröffentlicht: (2025)
von: Xu, Zunnan, et al.
Veröffentlicht: (2025)
MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
LIA-X: Interpretable Latent Portrait Animator
von: Wang, Yaohui, et al.
Veröffentlicht: (2025)
von: Wang, Yaohui, et al.
Veröffentlicht: (2025)
Sonic: Shifting Focus to Global Audio Perception in Portrait Animation
von: Ji, Xiaozhong, et al.
Veröffentlicht: (2024)
von: Ji, Xiaozhong, et al.
Veröffentlicht: (2024)
Disentangle Identity, Cooperate Emotion: Correlation-Aware Emotional Talking Portrait Generation
von: Tan, Weipeng, et al.
Veröffentlicht: (2025)
von: Tan, Weipeng, et al.
Veröffentlicht: (2025)
LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control
von: Guo, Jianzhu, et al.
Veröffentlicht: (2024)
von: Guo, Jianzhu, et al.
Veröffentlicht: (2024)
Tuning Timestep-Distilled Diffusion Model Using Pairwise Sample Optimization
von: Miao, Zichen, et al.
Veröffentlicht: (2024)
von: Miao, Zichen, et al.
Veröffentlicht: (2024)
EmoCAST: Emotional Talking Portrait via Emotive Text Description
von: Jiang, Yiguo, et al.
Veröffentlicht: (2025)
von: Jiang, Yiguo, et al.
Veröffentlicht: (2025)
CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning
von: Feng, He, et al.
Veröffentlicht: (2026)
von: Feng, He, et al.
Veröffentlicht: (2026)
PortraitTalk: Towards Customizable One-Shot Audio-to-Talking Face Generation
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2024)
von: Nazarieh, Fatemeh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
von: Wang, Mengchao, et al.
Veröffentlicht: (2025) -
FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers
von: Wang, Qiang, et al.
Veröffentlicht: (2025) -
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation
von: Liang, Chao, et al.
Veröffentlicht: (2025) -
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
von: Wang, Baiqin, et al.
Veröffentlicht: (2025) -
AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
von: Wei, Huawei, et al.
Veröffentlicht: (2024)