SkyReels-A1: Expressive Portrait Animation in Video Diffusion Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Qiu, Di, Fei, Zhengcong, Wang, Rui, Bai, Jialin, Yu, Changqian, Fan, Mingyuan, Chen, Guibin, Wen, Xiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers
by: Fei, Zhengcong, et al.
Published: (2025)
by: Fei, Zhengcong, et al.
Published: (2025)
SkyReels-A2: Compose Anything in Video Diffusion Transformers
by: Fei, Zhengcong, et al.
Published: (2025)
by: Fei, Zhengcong, et al.
Published: (2025)
Video Diffusion Transformers are In-Context Learners
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
SkyReels-V3 Technique Report
by: Li, Debang, et al.
Published: (2026)
by: Li, Debang, et al.
Published: (2026)
Ingredients: Blending Custom Photos with Video Diffusion Transformers
by: Fei, Zhengcong, et al.
Published: (2025)
by: Fei, Zhengcong, et al.
Published: (2025)
SkyReels-V2: Infinite-length Film Generative Model
by: Chen, Guibin, et al.
Published: (2025)
by: Chen, Guibin, et al.
Published: (2025)
Scaling Diffusion Transformers to 16 Billion Parameters
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
SkyReels-Text: Fine-Grained Font-Controllable Text Editing for Poster Design
by: Yu, Yunjie, et al.
Published: (2025)
by: Yu, Yunjie, et al.
Published: (2025)
Scalable Diffusion Models with State Space Backbone
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
SkyReels-V4: Multi-modal Video-Audio Generation, Inpainting and Editing model
by: Chen, Guibin, et al.
Published: (2026)
by: Chen, Guibin, et al.
Published: (2026)
Dimba: Transformer-Mamba Diffusion Models
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
Diffusion-RWKV: Scaling RWKV-Like Architectures for Diffusion Models
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
FLUX that Plays Music
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
MovieCharacter: A Tuning-Free Framework for Controllable Character Video Synthesis
by: Qiu, Di, et al.
Published: (2024)
by: Qiu, Di, et al.
Published: (2024)
FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers
by: Wang, Qiang, et al.
Published: (2025)
by: Wang, Qiang, et al.
Published: (2025)
A-JEPA: Joint-Embedding Predictive Architecture Can Listen
by: Fei, Zhengcong, et al.
Published: (2023)
by: Fei, Zhengcong, et al.
Published: (2023)
X-Portrait: Expressive Portrait Animation with Hierarchical Motion Attention
by: Xie, You, et al.
Published: (2024)
by: Xie, You, et al.
Published: (2024)
PersonaLive! Expressive Portrait Image Animation for Live Streaming
by: Li, Zhiyuan, et al.
Published: (2025)
by: Li, Zhiyuan, et al.
Published: (2025)
DeX-Portrait: Disentangled and Expressive Portrait Animation via Explicit and Latent Motion Representations
by: Shi, Yuxiang, et al.
Published: (2025)
by: Shi, Yuxiang, et al.
Published: (2025)
RAP: Real-time Audio-driven Portrait Animation with Video Diffusion Transformer
by: Du, Fangyu, et al.
Published: (2025)
by: Du, Fangyu, et al.
Published: (2025)
Follow-Your-Emoji: Fine-Controllable and Expressive Freestyle Portrait Animation
by: Ma, Yue, et al.
Published: (2024)
by: Ma, Yue, et al.
Published: (2024)
Real-time One-Step Diffusion-based Expressive Portrait Videos Generation
by: Guo, Hanzhong, et al.
Published: (2024)
by: Guo, Hanzhong, et al.
Published: (2024)
Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
by: Cui, Jiahao, et al.
Published: (2024)
by: Cui, Jiahao, et al.
Published: (2024)
MegActor-$Σ$: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
by: Yang, Shurong, et al.
Published: (2024)
by: Yang, Shurong, et al.
Published: (2024)
EMO: Emote Portrait Alive -- Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
by: Tian, Linrui, et al.
Published: (2024)
by: Tian, Linrui, et al.
Published: (2024)
LIA-X: Interpretable Latent Portrait Animator
by: Wang, Yaohui, et al.
Published: (2025)
by: Wang, Yaohui, et al.
Published: (2025)
CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning
by: Feng, He, et al.
Published: (2026)
by: Feng, He, et al.
Published: (2026)
LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control
by: Guo, Jianzhu, et al.
Published: (2024)
by: Guo, Jianzhu, et al.
Published: (2024)
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation
by: Yang, Shurong, et al.
Published: (2024)
by: Yang, Shurong, et al.
Published: (2024)
High-Fidelity Relightable Monocular Portrait Animation with Lighting-Controllable Video Diffusion Model
by: Guo, Mingtao, et al.
Published: (2025)
by: Guo, Mingtao, et al.
Published: (2025)
Follow-Your-Emoji-Faster: Towards Efficient, Fine-Controllable, and Expressive Freestyle Portrait Animation
by: Ma, Yue, et al.
Published: (2025)
by: Ma, Yue, et al.
Published: (2025)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
by: Wang, Xiang, et al.
Published: (2025)
by: Wang, Xiang, et al.
Published: (2025)
HunyuanPortrait: Implicit Condition Control for Enhanced Portrait Animation
by: Xu, Zunnan, et al.
Published: (2025)
by: Xu, Zunnan, et al.
Published: (2025)
FactorPortrait: Controllable Portrait Animation via Disentangled Expression, Pose, and Viewpoint
by: Tang, Jiapeng, et al.
Published: (2025)
by: Tang, Jiapeng, et al.
Published: (2025)
Loki: Representation over Architecture for Diffusion-Based Portrait Animation
by: Navard, Pouyan, et al.
Published: (2026)
by: Navard, Pouyan, et al.
Published: (2026)
FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent Prediction
by: Tu, Shuyuan, et al.
Published: (2025)
by: Tu, Shuyuan, et al.
Published: (2025)
Knot Forcing: Taming Autoregressive Video Diffusion Models for Real-time Infinite Interactive Portrait Animation
by: Xiao, Steven, et al.
Published: (2025)
by: Xiao, Steven, et al.
Published: (2025)
AnimateAnyMesh++: A Flexible 4D Foundation Model for High-Fidelity Text-Driven Mesh Animation
by: Wu, Zijie, et al.
Published: (2026)
by: Wu, Zijie, et al.
Published: (2026)
AnimateAnyMesh: A Feed-Forward 4D Foundation Model for Text-Driven Universal Mesh Animation
by: Wu, Zijie, et al.
Published: (2025)
by: Wu, Zijie, et al.
Published: (2025)
Embedded Representation Learning Network for Animating Styled Video Portrait
by: Wang, Tianyong, et al.
Published: (2024)
by: Wang, Tianyong, et al.
Published: (2024)
Similar Items
-
SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers
by: Fei, Zhengcong, et al.
Published: (2025) -
SkyReels-A2: Compose Anything in Video Diffusion Transformers
by: Fei, Zhengcong, et al.
Published: (2025) -
Video Diffusion Transformers are In-Context Learners
by: Fei, Zhengcong, et al.
Published: (2024) -
SkyReels-V3 Technique Report
by: Li, Debang, et al.
Published: (2026) -
Ingredients: Blending Custom Photos with Video Diffusion Transformers
by: Fei, Zhengcong, et al.
Published: (2025)