CoDance: An Unbind-Rebind Paradigm for Robust Multi-Subject Animation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tan, Shuai, Gong, Biao, Ma, Ke, Feng, Yutong, Zhang, Qiyuan, Wang, Yan, Shen, Yujun, Zhao, Hengshuang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Animate-X++: Universal Character Image Animation with Dynamic Backgrounds
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
UKnow: A Unified Knowledge Protocol with Multimodal Knowledge Graph Datasets for Reasoning and Vision-Language Pre-Training
von: Gong, Biao, et al.
Veröffentlicht: (2023)
von: Gong, Biao, et al.
Veröffentlicht: (2023)
Mimir: Improving Video Diffusion Models for Precise Text Understanding
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
Ranni: Taming Text-to-Image Diffusion for Accurate Instruction Following
von: Feng, Yutong, et al.
Veröffentlicht: (2023)
von: Feng, Yutong, et al.
Veröffentlicht: (2023)
PhysRVG: Physics-Aware Unified Reinforcement Learning for Video Generative Models
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Qiyuan, et al.
Veröffentlicht: (2026)
Animate-X: Universal Character Image Animation with Enhanced Motion Representation
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
Zero-shot Image Editing with Reference Imitation
von: Chen, Xi, et al.
Veröffentlicht: (2024)
von: Chen, Xi, et al.
Veröffentlicht: (2024)
EverybodyDance: Bipartite Graph-Based Identity Correspondence for Multi-Character Animation
von: Ling, Haotian, et al.
Veröffentlicht: (2025)
von: Ling, Haotian, et al.
Veröffentlicht: (2025)
AnyDoor: Zero-shot Object-level Image Customization
von: Chen, Xi, et al.
Veröffentlicht: (2023)
von: Chen, Xi, et al.
Veröffentlicht: (2023)
DanceCamAnimator: Keyframe-Based Controllable 3D Dance Camera Synthesis
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
MultiMotion: Multi Subject Video Motion Transfer via Video Diffusion Transformer
von: Liu, Penghui, et al.
Veröffentlicht: (2025)
von: Liu, Penghui, et al.
Veröffentlicht: (2025)
MikuDance: Animating Character Art with Mixed Motion Dynamics
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
CoReS: Orchestrating the Dance of Reasoning and Segmentation
von: Bao, Xiaoyi, et al.
Veröffentlicht: (2024)
von: Bao, Xiaoyi, et al.
Veröffentlicht: (2024)
Expressive Speech-driven Facial Animation with controllable emotions
von: Chen, Yutong, et al.
Veröffentlicht: (2023)
von: Chen, Yutong, et al.
Veröffentlicht: (2023)
MultiAnimate: Pose-Guided Image Animation Made Extensible
von: Hu, Yingcheng, et al.
Veröffentlicht: (2026)
von: Hu, Yingcheng, et al.
Veröffentlicht: (2026)
DiffDoctor: Diagnosing Image Diffusion Models Before Treating
von: Wang, Yiyang, et al.
Veröffentlicht: (2025)
von: Wang, Yiyang, et al.
Veröffentlicht: (2025)
Troika: Multi-Path Cross-Modal Traction for Compositional Zero-Shot Learning
von: Huang, Siteng, et al.
Veröffentlicht: (2023)
von: Huang, Siteng, et al.
Veröffentlicht: (2023)
Scanning Only Once: An End-to-end Framework for Fast Temporal Grounding in Long Videos
von: Pan, Yulin, et al.
Veröffentlicht: (2023)
von: Pan, Yulin, et al.
Veröffentlicht: (2023)
DreamDance: Animating Character Art via Inpainting Stable Gaussian Worlds
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2025)
ESGaussianFace: Emotional and Stylized Audio-Driven Facial Animation via 3D Gaussian Splatting
von: Ma, Chuhang, et al.
Veröffentlicht: (2026)
von: Ma, Chuhang, et al.
Veröffentlicht: (2026)
AnimateScene: Camera-controllable Animation in Any Scene
von: Liu, Qingyang, et al.
Veröffentlicht: (2025)
von: Liu, Qingyang, et al.
Veröffentlicht: (2025)
Every Image Listens, Every Image Dances: Music-Driven Image Animation
von: Dong, Zhikang, et al.
Veröffentlicht: (2025)
von: Dong, Zhikang, et al.
Veröffentlicht: (2025)
DisCo: Towards Distinct and Coherent Visual Encapsulation in Video MLLMs
von: Zhao, Jiahe, et al.
Veröffentlicht: (2025)
von: Zhao, Jiahe, et al.
Veröffentlicht: (2025)
FocalClick-XL: Towards Unified and High-quality Interactive Segmentation
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
AniDoc: Animation Creation Made Easier
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
von: Meng, Yihao, et al.
Veröffentlicht: (2024)
AnimateZoo: Zero-shot Video Generation of Cross-Species Animation via Subject Alignment
von: Xu, Yuanfeng, et al.
Veröffentlicht: (2024)
von: Xu, Yuanfeng, et al.
Veröffentlicht: (2024)
MiCo: Multi-image Contrast for Reinforcement Visual Reasoning
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
RealisDance-DiT: Simple yet Strong Baseline towards Controllable Character Animation in the Wild
von: Zhou, Jingkai, et al.
Veröffentlicht: (2025)
von: Zhou, Jingkai, et al.
Veröffentlicht: (2025)
Human Geometry Distribution for 3D Animation Generation
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
Check, Locate, Rectify: A Training-Free Layout Calibration System for Text-to-Image Generation
von: Gong, Biao, et al.
Veröffentlicht: (2023)
von: Gong, Biao, et al.
Veröffentlicht: (2023)
InstanceAnimator: Multi-Instance Sketch Video Colorization
von: Zhang, Yinhan, et al.
Veröffentlicht: (2026)
von: Zhang, Yinhan, et al.
Veröffentlicht: (2026)
DisCo: Disentangled Control for Realistic Human Dance Generation
von: Wang, Tan, et al.
Veröffentlicht: (2023)
von: Wang, Tan, et al.
Veröffentlicht: (2023)
OmniLottie: Generating Vector Animations via Parameterized Lottie Tokens
von: Yang, Yiying, et al.
Veröffentlicht: (2026)
von: Yang, Yiying, et al.
Veröffentlicht: (2026)
Interacted Planes Reveal 3D Line Mapping
von: Ke, Zeran, et al.
Veröffentlicht: (2026)
von: Ke, Zeran, et al.
Veröffentlicht: (2026)
Scaling Up Audio-Synchronized Visual Animation: An Efficient Training Paradigm
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
Efficient 3D Perception on Multi-Sweep Point Cloud with Gumbel Spatial Pruning
von: Sun, Tianyu, et al.
Veröffentlicht: (2024)
von: Sun, Tianyu, et al.
Veröffentlicht: (2024)
One for All: Multi-Domain Joint Training for Point Cloud Based 3D Object Detection
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
Media2Face: Co-speech Facial Animation Generation With Multi-Modality Guidance
von: Zhao, Qingcheng, et al.
Veröffentlicht: (2024)
von: Zhao, Qingcheng, et al.
Veröffentlicht: (2024)
UniMatch V2: Pushing the Limit of Semi-Supervised Semantic Segmentation
von: Yang, Lihe, et al.
Veröffentlicht: (2024)
von: Yang, Lihe, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Animate-X++: Universal Character Image Animation with Dynamic Backgrounds
von: Tan, Shuai, et al.
Veröffentlicht: (2025) -
SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation
von: Tan, Shuai, et al.
Veröffentlicht: (2025) -
UKnow: A Unified Knowledge Protocol with Multimodal Knowledge Graph Datasets for Reasoning and Vision-Language Pre-Training
von: Gong, Biao, et al.
Veröffentlicht: (2023) -
Mimir: Improving Video Diffusion Models for Precise Text Understanding
von: Tan, Shuai, et al.
Veröffentlicht: (2024) -
Ranni: Taming Text-to-Image Diffusion for Accurate Instruction Following
von: Feng, Yutong, et al.
Veröffentlicht: (2023)