ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qi, Jinwei, Ji, Chaonan, Xu, Sheng, Zhang, Peng, Zhang, Bang, Bo, Liefeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment
von: Ji, Chaonan, et al.
Veröffentlicht: (2026)
von: Ji, Chaonan, et al.
Veröffentlicht: (2026)
Controllable and Expressive One-Shot Video Head Swapping
von: Ji, Chaonan, et al.
Veröffentlicht: (2025)
von: Ji, Chaonan, et al.
Veröffentlicht: (2025)
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025)
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025)
Exploring Timeline Control for Facial Motion Generation
von: Ma, Yifeng, et al.
Veröffentlicht: (2025)
von: Ma, Yifeng, et al.
Veröffentlicht: (2025)
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
von: Tian, Linrui, et al.
Veröffentlicht: (2024)
von: Tian, Linrui, et al.
Veröffentlicht: (2024)
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
von: Hu, Li, et al.
Veröffentlicht: (2023)
von: Hu, Li, et al.
Veröffentlicht: (2023)
Knot Forcing: Taming Autoregressive Video Diffusion Models for Real-time Infinite Interactive Portrait Animation
von: Xiao, Steven, et al.
Veröffentlicht: (2025)
von: Xiao, Steven, et al.
Veröffentlicht: (2025)
EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
von: Tian, Linrui, et al.
Veröffentlicht: (2025)
von: Tian, Linrui, et al.
Veröffentlicht: (2025)
Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance
von: Hu, Li, et al.
Veröffentlicht: (2025)
von: Hu, Li, et al.
Veröffentlicht: (2025)
OutfitAnyone: Ultra-high Quality Virtual Try-On for Any Clothing and Any Person
von: Sun, Ke, et al.
Veröffentlicht: (2024)
von: Sun, Ke, et al.
Veröffentlicht: (2024)
Real-time One-Step Diffusion-based Expressive Portrait Videos Generation
von: Guo, Hanzhong, et al.
Veröffentlicht: (2024)
von: Guo, Hanzhong, et al.
Veröffentlicht: (2024)
UniPortrait: A Unified Framework for Identity-Preserving Single- and Multi-Human Image Personalization
von: He, Junjie, et al.
Veröffentlicht: (2024)
von: He, Junjie, et al.
Veröffentlicht: (2024)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
von: Song, Yafei, et al.
Veröffentlicht: (2025)
von: Song, Yafei, et al.
Veröffentlicht: (2025)
SyncAnyone: Implicit Disentanglement via Progressive Self-Correction for Lip-Syncing in the wild
von: Zhang, Xindi, et al.
Veröffentlicht: (2025)
von: Zhang, Xindi, et al.
Veröffentlicht: (2025)
MaTe3D: Mask-guided Text-based 3D-aware Portrait Editing
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
X-Portrait: Expressive Portrait Animation with Hierarchical Motion Attention
von: Xie, You, et al.
Veröffentlicht: (2024)
von: Xie, You, et al.
Veröffentlicht: (2024)
DiT as Real-Time Rerenderer: Streaming Video Stylization with Autoregressive Diffusion Transformer
von: Lyu, Hengye, et al.
Veröffentlicht: (2026)
von: Lyu, Hengye, et al.
Veröffentlicht: (2026)
A Self-supervised Motion Representation for Portrait Video Generation
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2025)
Teller: Real-Time Streaming Audio-Driven Portrait Animation with Autoregressive Motion Generation
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
von: Zhen, Dingcheng, et al.
Veröffentlicht: (2025)
DiffuEraser: A Diffusion Model for Video Inpainting
von: Li, Xiaowen, et al.
Veröffentlicht: (2025)
von: Li, Xiaowen, et al.
Veröffentlicht: (2025)
RASA: Replace Anyone, Say Anything -- A Training-Free Framework for Audio-Driven and Universal Portrait Video Editing
von: Pan, Tianrui, et al.
Veröffentlicht: (2025)
von: Pan, Tianrui, et al.
Veröffentlicht: (2025)
RAP: Real-time Audio-driven Portrait Animation with Video Diffusion Transformer
von: Du, Fangyu, et al.
Veröffentlicht: (2025)
von: Du, Fangyu, et al.
Veröffentlicht: (2025)
Generative Human Motion Stylization in Latent Space
von: Guo, Chuan, et al.
Veröffentlicht: (2024)
von: Guo, Chuan, et al.
Veröffentlicht: (2024)
Wan-S2V: Audio-Driven Cinematic Video Generation
von: Gao, Xin, et al.
Veröffentlicht: (2025)
von: Gao, Xin, et al.
Veröffentlicht: (2025)
CPA: Camera-pose-awareness Diffusion Transformer for Video Generation
von: Wang, Yuelei, et al.
Veröffentlicht: (2024)
von: Wang, Yuelei, et al.
Veröffentlicht: (2024)
HDRFlow: Real-Time HDR Video Reconstruction with Large Motions
von: Xu, Gangwei, et al.
Veröffentlicht: (2024)
von: Xu, Gangwei, et al.
Veröffentlicht: (2024)
Replace Anyone in Videos
von: Wang, Xiang, et al.
Veröffentlicht: (2024)
von: Wang, Xiang, et al.
Veröffentlicht: (2024)
MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
MagicStyle: Portrait Stylization Based on Reference Image
von: Deng, Zhaoli, et al.
Veröffentlicht: (2024)
von: Deng, Zhaoli, et al.
Veröffentlicht: (2024)
SMooDi: Stylized Motion Diffusion Model
von: Zhong, Lei, et al.
Veröffentlicht: (2024)
von: Zhong, Lei, et al.
Veröffentlicht: (2024)
FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
von: Wang, Mengchao, et al.
Veröffentlicht: (2025)
von: Wang, Mengchao, et al.
Veröffentlicht: (2025)
EchoShot: Multi-Shot Portrait Video Generation
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
AdaptiveDrag: Semantic-Driven Dragging on Diffusion-Based Image Editing
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
Generative Motion Stylization of Cross-structure Characters within Canonical Motion Space
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
Realtime Data-Efficient Portrait Stylization Based On Geometric Alignment
von: Wang, Xinrui, et al.
Veröffentlicht: (2022)
von: Wang, Xinrui, et al.
Veröffentlicht: (2022)
ZePo: Zero-Shot Portrait Stylization with Faster Sampling
von: Liu, Jin, et al.
Veröffentlicht: (2024)
von: Liu, Jin, et al.
Veröffentlicht: (2024)
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2024)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2024)
MotionRAG-Diff: A Retrieval-Augmented Diffusion Framework for Long-Term Music-to-Dance Generation
von: Huang, Mingyang, et al.
Veröffentlicht: (2025)
von: Huang, Mingyang, et al.
Veröffentlicht: (2025)
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs
von: Xu, Sicheng, et al.
Veröffentlicht: (2026)
von: Xu, Sicheng, et al.
Veröffentlicht: (2026)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2025)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment
von: Ji, Chaonan, et al.
Veröffentlicht: (2026) -
Controllable and Expressive One-Shot Video Head Swapping
von: Ji, Chaonan, et al.
Veröffentlicht: (2025) -
OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking
von: Wang, Zhongjian, et al.
Veröffentlicht: (2025) -
Exploring Timeline Control for Facial Motion Generation
von: Ma, Yifeng, et al.
Veröffentlicht: (2025) -
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
von: Tian, Linrui, et al.
Veröffentlicht: (2024)