AnimaX: Animating the Inanimate in 3D with Joint Video-Pose Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Zehuan, Feng, Haoran, Sun, Yangtian, Guo, Yuanchen, Cao, Yanpei, Sheng, Lu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Personalize Anything for Free with Diffusion Transformer
by: Feng, Haoran, et al.
Published: (2025)
by: Feng, Haoran, et al.
Published: (2025)
Repurposing 3D Generative Model for Autoregressive Layout Generation
by: Feng, Haoran, et al.
Published: (2026)
by: Feng, Haoran, et al.
Published: (2026)
VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space
by: Li, Lin, et al.
Published: (2025)
by: Li, Lin, et al.
Published: (2025)
Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion
by: Wen, Hao, et al.
Published: (2024)
by: Wen, Hao, et al.
Published: (2024)
Animate3D: Animating Any 3D Model with Multi-view Video Diffusion
by: Jiang, Yanqin, et al.
Published: (2024)
by: Jiang, Yanqin, et al.
Published: (2024)
DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models
by: Sun, Zhiyao, et al.
Published: (2023)
by: Sun, Zhiyao, et al.
Published: (2023)
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
by: Wang, Qilin, et al.
Published: (2024)
by: Wang, Qilin, et al.
Published: (2024)
HYPNOS : Highly Precise Foreground-focused Diffusion Finetuning for Inanimate Objects
by: Nathanael, Oliverio Theophilus, et al.
Published: (2024)
by: Nathanael, Oliverio Theophilus, et al.
Published: (2024)
MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
by: Huang, Zehuan, et al.
Published: (2024)
by: Huang, Zehuan, et al.
Published: (2024)
FastAnimate: Towards Learnable Template Construction and Pose Deformation for Fast 3D Human Avatar Animation
by: Shu, Jian, et al.
Published: (2025)
by: Shu, Jian, et al.
Published: (2025)
SegviGen: Repurposing 3D Generative Model for Part Segmentation
by: Li, Lin, et al.
Published: (2026)
by: Li, Lin, et al.
Published: (2026)
MV-Adapter: Multi-view Consistent Image Generation Made Easy
by: Huang, Zehuan, et al.
Published: (2024)
by: Huang, Zehuan, et al.
Published: (2024)
UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation
by: Sun, Yang-Tian, et al.
Published: (2025)
by: Sun, Yang-Tian, et al.
Published: (2025)
KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes
by: Wu, Jingchao, et al.
Published: (2025)
by: Wu, Jingchao, et al.
Published: (2025)
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
by: Gan, Qijun, et al.
Published: (2025)
by: Gan, Qijun, et al.
Published: (2025)
ALIVE: Animate Your World with Lifelike Audio-Video Generation
by: Guo, Ying, et al.
Published: (2026)
by: Guo, Ying, et al.
Published: (2026)
DisPose: Disentangling Pose Guidance for Controllable Human Image Animation
by: Li, Hongxiang, et al.
Published: (2024)
by: Li, Hongxiang, et al.
Published: (2024)
From Parts to Whole: A Unified Reference Framework for Controllable Human Image Generation
by: Huang, Zehuan, et al.
Published: (2024)
by: Huang, Zehuan, et al.
Published: (2024)
One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer
by: Shi, Shijun, et al.
Published: (2025)
by: Shi, Shijun, et al.
Published: (2025)
InterMoE: Individual-Specific 3D Human Interaction Generation via Dynamic Temporal-Selective MoE
by: Wang, Lipeng, et al.
Published: (2025)
by: Wang, Lipeng, et al.
Published: (2025)
SCIPaD: Incorporating Spatial Clues into Unsupervised Pose-Depth Joint Learning
by: Feng, Yi, et al.
Published: (2024)
by: Feng, Yi, et al.
Published: (2024)
Joint-Motion Mutual Learning for Pose Estimation in Videos
by: Wu, Sifan, et al.
Published: (2024)
by: Wu, Sifan, et al.
Published: (2024)
Animating the Uncaptured: Humanoid Mesh Animation with Video Diffusion Models
by: Millán, Marc Benedí San, et al.
Published: (2025)
by: Millán, Marc Benedí San, et al.
Published: (2025)
RAIN: Real-time Animation of Infinite Video Stream
by: Shu, Zhilei, et al.
Published: (2024)
by: Shu, Zhilei, et al.
Published: (2024)
JGHand: Joint-Driven Animatable Hand Avater via 3D Gaussian Splatting
by: Sun, Zhoutao, et al.
Published: (2025)
by: Sun, Zhoutao, et al.
Published: (2025)
VividAnimator: An End-to-End Audio and Pose-driven Half-Body Human Animation Framework
by: Huang, Donglin, et al.
Published: (2025)
by: Huang, Donglin, et al.
Published: (2025)
One-Shot Pose-Driving Face Animation Platform
by: Feng, He, et al.
Published: (2024)
by: Feng, He, et al.
Published: (2024)
SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representations
by: Yan, Wenhao, et al.
Published: (2025)
by: Yan, Wenhao, et al.
Published: (2025)
MultiAnimate: Pose-Guided Image Animation Made Extensible
by: Hu, Yingcheng, et al.
Published: (2026)
by: Hu, Yingcheng, et al.
Published: (2026)
StarPose: 3D Human Pose Estimation via Spatial-Temporal Autoregressive Diffusion
by: Yang, Haoxin, et al.
Published: (2025)
by: Yang, Haoxin, et al.
Published: (2025)
CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving
by: Zhang, Tianrui, et al.
Published: (2025)
by: Zhang, Tianrui, et al.
Published: (2025)
TalkingPose: Efficient Face and Gesture Animation with Feedback-guided Diffusion Model
by: Javanmardi, Alireza, et al.
Published: (2025)
by: Javanmardi, Alireza, et al.
Published: (2025)
AnimationBench: Are Video Models Good at Character-Centric Animation?
by: Wu, Leyi, et al.
Published: (2026)
by: Wu, Leyi, et al.
Published: (2026)
Controllable Expressive 3D Facial Animation via Diffusion in a Unified Multimodal Space
by: Liu, Kangwei, et al.
Published: (2025)
by: Liu, Kangwei, et al.
Published: (2025)
AniGen: Unified $S^3$ Fields for Animatable 3D Asset Generation
by: Huang, Yi-Hua, et al.
Published: (2026)
by: Huang, Yi-Hua, et al.
Published: (2026)
OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation
by: Wang, Junke, et al.
Published: (2024)
by: Wang, Junke, et al.
Published: (2024)
DreamCraft3D++: Efficient Hierarchical 3D Generation with Multi-Plane Reconstruction Model
by: Sun, Jingxiang, et al.
Published: (2024)
by: Sun, Jingxiang, et al.
Published: (2024)
DPoser-X: Diffusion Model as Robust 3D Whole-body Human Pose Prior
by: Lu, Junzhe, et al.
Published: (2025)
by: Lu, Junzhe, et al.
Published: (2025)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
by: Wang, Xiang, et al.
Published: (2024)
by: Wang, Xiang, et al.
Published: (2024)
SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models
by: Yang, Ruolin, et al.
Published: (2025)
by: Yang, Ruolin, et al.
Published: (2025)
Similar Items
-
Personalize Anything for Free with Diffusion Transformer
by: Feng, Haoran, et al.
Published: (2025) -
Repurposing 3D Generative Model for Autoregressive Layout Generation
by: Feng, Haoran, et al.
Published: (2026) -
VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space
by: Li, Lin, et al.
Published: (2025) -
Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion
by: Wen, Hao, et al.
Published: (2024) -
Animate3D: Animating Any 3D Model with Multi-view Video Diffusion
by: Jiang, Yanqin, et al.
Published: (2024)