RealisDance-DiT: Simple yet Strong Baseline towards Controllable Character Animation in the Wild
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhou, Jingkai, Wu, Yifan, Li, Shikai, Wei, Min, Fan, Chao, Chen, Weihua, Jiang, Wei, Wang, Fan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RealisDance: Equip controllable character animation with realistic hands
di: Zhou, Jingkai, et al.
Pubblicazione: (2024)
di: Zhou, Jingkai, et al.
Pubblicazione: (2024)
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
di: Liang, Jingyun, et al.
Pubblicazione: (2025)
di: Liang, Jingyun, et al.
Pubblicazione: (2025)
DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation
di: Feng, He, et al.
Pubblicazione: (2025)
di: Feng, He, et al.
Pubblicazione: (2025)
RealisHuman: A Two-Stage Approach for Refining Malformed Human Parts in Generated Images
di: Wang, Benzhi, et al.
Pubblicazione: (2024)
di: Wang, Benzhi, et al.
Pubblicazione: (2024)
RealisVSR: Detail-enhanced Diffusion for Real-World 4K Video Super-Resolution
di: Zhao, Weisong, et al.
Pubblicazione: (2025)
di: Zhao, Weisong, et al.
Pubblicazione: (2025)
DiVE: DiT-based Video Generation with Enhanced Control
di: Jiang, Junpeng, et al.
Pubblicazione: (2024)
di: Jiang, Junpeng, et al.
Pubblicazione: (2024)
VITON-DiT: Learning In-the-Wild Video Try-On from Human Dance Videos via Diffusion Transformers
di: Zheng, Jun, et al.
Pubblicazione: (2024)
di: Zheng, Jun, et al.
Pubblicazione: (2024)
DiT4DiT: Jointly Modeling Video Dynamics and Actions for Generalizable Robot Control
di: Ma, Teli, et al.
Pubblicazione: (2026)
di: Ma, Teli, et al.
Pubblicazione: (2026)
MikuDance: Animating Character Art with Mixed Motion Dynamics
di: Zhang, Jiaxu, et al.
Pubblicazione: (2024)
di: Zhang, Jiaxu, et al.
Pubblicazione: (2024)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
di: Wang, Xiang, et al.
Pubblicazione: (2025)
di: Wang, Xiang, et al.
Pubblicazione: (2025)
RealisID: Scale-Robust and Fine-Controllable Identity Customization via Local and Global Complementation
di: Sun, Zhaoyang, et al.
Pubblicazione: (2024)
di: Sun, Zhaoyang, et al.
Pubblicazione: (2024)
DreamDance: Animating Character Art via Inpainting Stable Gaussian Worlds
di: Zhang, Jiaxu, et al.
Pubblicazione: (2025)
di: Zhang, Jiaxu, et al.
Pubblicazione: (2025)
EverybodyDance: Bipartite Graph-Based Identity Correspondence for Multi-Character Animation
di: Ling, Haotian, et al.
Pubblicazione: (2025)
di: Ling, Haotian, et al.
Pubblicazione: (2025)
Rethinking Multi-Condition DiTs: Eliminating Redundant Attention via Position-Alignment and Keyword-Scoping
di: Zhou, Chao, et al.
Pubblicazione: (2026)
di: Zhou, Chao, et al.
Pubblicazione: (2026)
Pruned Adaptation Modules: A Simple yet Strong Baseline for Continual Foundation Models
di: Yildirim, Elif Ceren Gok, et al.
Pubblicazione: (2026)
di: Yildirim, Elif Ceren Gok, et al.
Pubblicazione: (2026)
FIS-DiT: Breaking the Few-Step Video Inference Barrier via Training-Free Frame Interleaved Sparsity
di: Tang, Jian, et al.
Pubblicazione: (2026)
di: Tang, Jian, et al.
Pubblicazione: (2026)
HyperMotionX: The Dataset and Benchmark with DiT-Based Pose-Guided Human Image Animation of Complex Motions
di: Xu, Shuolin, et al.
Pubblicazione: (2025)
di: Xu, Shuolin, et al.
Pubblicazione: (2025)
DiT-HC: Enabling Efficient Training of Visual Generation Model DiT on HPC-oriented CPU Cluster
di: Zhang, Jinxiao, et al.
Pubblicazione: (2026)
di: Zhang, Jinxiao, et al.
Pubblicazione: (2026)
GACA-DiT: Diffusion-based Dance-to-Music Generation with Genre-Adaptive Rhythm and Context-Aware Alignment
di: Wang, Jinting, et al.
Pubblicazione: (2025)
di: Wang, Jinting, et al.
Pubblicazione: (2025)
OUSAC: Optimized Guidance Scheduling with Adaptive Caching for DiT Acceleration
di: Sun, Ruitong, et al.
Pubblicazione: (2025)
di: Sun, Ruitong, et al.
Pubblicazione: (2025)
DiTReducio: A Training-Free Acceleration for DiT-Based TTS via Progressive Calibration
di: Huo, Yanru, et al.
Pubblicazione: (2025)
di: Huo, Yanru, et al.
Pubblicazione: (2025)
SketchColour: Channel Concat Guided DiT-based Sketch-to-Colour Pipeline for 2D Animation
di: Sadihin, Bryan Constantine, et al.
Pubblicazione: (2025)
di: Sadihin, Bryan Constantine, et al.
Pubblicazione: (2025)
3DV-TON: Textured 3D-Guided Consistent Video Try-on via Diffusion Models
di: Wei, Min, et al.
Pubblicazione: (2025)
di: Wei, Min, et al.
Pubblicazione: (2025)
FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers
di: Wang, Qiang, et al.
Pubblicazione: (2025)
di: Wang, Qiang, et al.
Pubblicazione: (2025)
Towards 3D-Aware Video Diffusion Models: Render-Free Human Motion Control with Mesh Tokenization
di: Liang, Jingyun, et al.
Pubblicazione: (2026)
di: Liang, Jingyun, et al.
Pubblicazione: (2026)
DanceCamAnimator: Keyframe-Based Controllable 3D Dance Camera Synthesis
di: Wang, Zixuan, et al.
Pubblicazione: (2024)
di: Wang, Zixuan, et al.
Pubblicazione: (2024)
Rethinking Irregular Time Series Forecasting: A Simple yet Effective Baseline
di: Liu, Xvyuan, et al.
Pubblicazione: (2025)
di: Liu, Xvyuan, et al.
Pubblicazione: (2025)
MaterialPicker: Multi-Modal DiT-Based Material Generation
di: Ma, Xiaohe, et al.
Pubblicazione: (2024)
di: Ma, Xiaohe, et al.
Pubblicazione: (2024)
Attend to Not Attended: Structure-then-Detail Token Merging for Post-training DiT Acceleration
di: Fang, Haipeng, et al.
Pubblicazione: (2025)
di: Fang, Haipeng, et al.
Pubblicazione: (2025)
FD-DiT: Frequency Domain-Directed Diffusion Transformer for Low-Dose CT Reconstruction
di: Liu, Qiqing, et al.
Pubblicazione: (2025)
di: Liu, Qiqing, et al.
Pubblicazione: (2025)
Revisiting Simple Baselines for In-The-Wild Deepfake Detection
di: Castaneda, Orlando, et al.
Pubblicazione: (2025)
di: Castaneda, Orlando, et al.
Pubblicazione: (2025)
Animate-X++: Universal Character Image Animation with Dynamic Backgrounds
di: Tan, Shuai, et al.
Pubblicazione: (2025)
di: Tan, Shuai, et al.
Pubblicazione: (2025)
Insert Anything: Image Insertion via In-Context Editing in DiT
di: Song, Wensong, et al.
Pubblicazione: (2025)
di: Song, Wensong, et al.
Pubblicazione: (2025)
GPG: A Simple and Strong Reinforcement Learning Baseline for Model Reasoning
di: Chu, Xiangxiang, et al.
Pubblicazione: (2025)
di: Chu, Xiangxiang, et al.
Pubblicazione: (2025)
On Denoising Walking Videos for Gait Recognition
di: Jin, Dongyang, et al.
Pubblicazione: (2025)
di: Jin, Dongyang, et al.
Pubblicazione: (2025)
Animate Any Character in Any World
di: Wang, Yitong, et al.
Pubblicazione: (2025)
di: Wang, Yitong, et al.
Pubblicazione: (2025)
Untwisting RoPE: Frequency Control for Shared Attention in DiTs
di: Mikaeili, Aryan, et al.
Pubblicazione: (2026)
di: Mikaeili, Aryan, et al.
Pubblicazione: (2026)
U-DiTs: Downsample Tokens in U-Shaped Diffusion Transformers
di: Tian, Yuchuan, et al.
Pubblicazione: (2024)
di: Tian, Yuchuan, et al.
Pubblicazione: (2024)
DancingBox: A Lightweight MoCap System for Character Animation from Physical Proxies
di: Yuan, Haocheng, et al.
Pubblicazione: (2026)
di: Yuan, Haocheng, et al.
Pubblicazione: (2026)
XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation
di: Chen, Bowen, et al.
Pubblicazione: (2025)
di: Chen, Bowen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
RealisDance: Equip controllable character animation with realistic hands
di: Zhou, Jingkai, et al.
Pubblicazione: (2024) -
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
di: Liang, Jingyun, et al.
Pubblicazione: (2025) -
DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation
di: Feng, He, et al.
Pubblicazione: (2025) -
RealisHuman: A Two-Stage Approach for Refining Malformed Human Parts in Generated Images
di: Wang, Benzhi, et al.
Pubblicazione: (2024) -
RealisVSR: Detail-enhanced Diffusion for Real-World 4K Video Super-Resolution
di: Zhao, Weisong, et al.
Pubblicazione: (2025)