Controllable Longer Image Animation with Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Qiang, Liu, Minghua, Hu, Junjun, Jiang, Fan, Xu, Mu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers
von: Wang, Qiang, et al.
Veröffentlicht: (2025)
von: Wang, Qiang, et al.
Veröffentlicht: (2025)
FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation
von: Wang, MengChao, et al.
Veröffentlicht: (2025)
von: Wang, MengChao, et al.
Veröffentlicht: (2025)
Animate3D: Animating Any 3D Model with Multi-view Video Diffusion
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
Cinemo: Consistent and Controllable Image Animation with Motion Diffusion Models
von: Ma, Xin, et al.
Veröffentlicht: (2024)
von: Ma, Xin, et al.
Veröffentlicht: (2024)
NavForesee: A Unified Vision-Language World Model for Hierarchical Planning and Dual-Horizon Navigation Prediction
von: Liu, Fei, et al.
Veröffentlicht: (2025)
von: Liu, Fei, et al.
Veröffentlicht: (2025)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
von: Wang, Xiang, et al.
Veröffentlicht: (2024)
von: Wang, Xiang, et al.
Veröffentlicht: (2024)
Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers
von: Ma, Xin, et al.
Veröffentlicht: (2025)
von: Ma, Xin, et al.
Veröffentlicht: (2025)
MultiAnimate: Pose-Guided Image Animation Made Extensible
von: Hu, Yingcheng, et al.
Veröffentlicht: (2026)
von: Hu, Yingcheng, et al.
Veröffentlicht: (2026)
LayerAnimate: Layer-level Control for Animation
von: Yang, Yuxue, et al.
Veröffentlicht: (2025)
von: Yang, Yuxue, et al.
Veröffentlicht: (2025)
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
von: Hu, Li, et al.
Veröffentlicht: (2023)
von: Hu, Li, et al.
Veröffentlicht: (2023)
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
von: Wang, Qilin, et al.
Veröffentlicht: (2024)
von: Wang, Qilin, et al.
Veröffentlicht: (2024)
AnimateAnything: Consistent and Controllable Animation for Video Generation
von: Lei, Guojun, et al.
Veröffentlicht: (2024)
von: Lei, Guojun, et al.
Veröffentlicht: (2024)
AstraNav-World: World Model for Foresight Control and Consistency
von: Chen, Jintao, et al.
Veröffentlicht: (2025)
von: Chen, Jintao, et al.
Veröffentlicht: (2025)
Learning from History: Task-agnostic Model Contrastive Learning for Image Restoration
von: Wu, Gang, et al.
Veröffentlicht: (2023)
von: Wu, Gang, et al.
Veröffentlicht: (2023)
High-Fidelity and Long-Duration Human Image Animation with Diffusion Transformer
von: Zheng, Shen, et al.
Veröffentlicht: (2025)
von: Zheng, Shen, et al.
Veröffentlicht: (2025)
ControlFusion: A Controllable Image Fusion Framework with Language-Vision Degradation Prompts
von: Tang, Linfeng, et al.
Veröffentlicht: (2025)
von: Tang, Linfeng, et al.
Veröffentlicht: (2025)
High-Fidelity Relightable Monocular Portrait Animation with Lighting-Controllable Video Diffusion Model
von: Guo, Mingtao, et al.
Veröffentlicht: (2025)
von: Guo, Mingtao, et al.
Veröffentlicht: (2025)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion
von: Zhou, Zhenglin, et al.
Veröffentlicht: (2025)
von: Zhou, Zhenglin, et al.
Veröffentlicht: (2025)
One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer
von: Shi, Shijun, et al.
Veröffentlicht: (2025)
von: Shi, Shijun, et al.
Veröffentlicht: (2025)
Playmate: Flexible Control of Portrait Animation via 3D-Implicit Space Guided Diffusion
von: Ma, Xingpei, et al.
Veröffentlicht: (2025)
von: Ma, Xingpei, et al.
Veröffentlicht: (2025)
FantasyHSI: Video-Generation-Centric 4D Human Synthesis In Any Scene through A Graph-based Multi-Agent Framework
von: Mu, Lingzhou, et al.
Veröffentlicht: (2025)
von: Mu, Lingzhou, et al.
Veröffentlicht: (2025)
Balancing Task-invariant Interaction and Task-specific Adaptation for Unified Image Fusion
von: Hu, Xingyu, et al.
Veröffentlicht: (2025)
von: Hu, Xingyu, et al.
Veröffentlicht: (2025)
Towards Unified Semantic and Controllable Image Fusion: A Diffusion Transformer Approach
von: Li, Jiayang, et al.
Veröffentlicht: (2025)
von: Li, Jiayang, et al.
Veröffentlicht: (2025)
VFX Creator: Animated Visual Effect Generation with Controllable Diffusion Transformer
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
AnimateAnywhere: Rouse the Background in Human Image Animation
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2025)
Animate-X++: Universal Character Image Animation with Dynamic Backgrounds
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
von: Tan, Shuai, et al.
Veröffentlicht: (2025)
FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling
von: Qiu, Haonan, et al.
Veröffentlicht: (2023)
von: Qiu, Haonan, et al.
Veröffentlicht: (2023)
EvAnimate: Event-conditioned Image-to-Video Generation for Human Animation
von: Qu, Qiang, et al.
Veröffentlicht: (2025)
von: Qu, Qiang, et al.
Veröffentlicht: (2025)
Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance
von: Hu, Li, et al.
Veröffentlicht: (2025)
von: Hu, Li, et al.
Veröffentlicht: (2025)
MegActor-$Σ$: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
von: Yang, Shurong, et al.
Veröffentlicht: (2024)
von: Yang, Shurong, et al.
Veröffentlicht: (2024)
Animating the Uncaptured: Humanoid Mesh Animation with Video Diffusion Models
von: Millán, Marc Benedí San, et al.
Veröffentlicht: (2025)
von: Millán, Marc Benedí San, et al.
Veröffentlicht: (2025)
InterAnimate: Taming Region-aware Diffusion Model for Realistic Human Interaction Animation
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
RealisDance-DiT: Simple yet Strong Baseline towards Controllable Character Animation in the Wild
von: Zhou, Jingkai, et al.
Veröffentlicht: (2025)
von: Zhou, Jingkai, et al.
Veröffentlicht: (2025)
Boosting All-in-One Image Restoration via Self-Improved Privilege Learning
von: Wu, Gang, et al.
Veröffentlicht: (2025)
von: Wu, Gang, et al.
Veröffentlicht: (2025)
Grounded Forcing: Bridging Time-Independent Semantics and Proximal Dynamics in Autoregressive Video Synthesis
von: Chen, Jintao, et al.
Veröffentlicht: (2026)
von: Chen, Jintao, et al.
Veröffentlicht: (2026)
Continuous Piecewise-Affine Based Motion Model for Image Animation
von: Wang, Hexiang, et al.
Veröffentlicht: (2024)
von: Wang, Hexiang, et al.
Veröffentlicht: (2024)
MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model
von: Niu, Muyao, et al.
Veröffentlicht: (2024)
von: Niu, Muyao, et al.
Veröffentlicht: (2024)
SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models
von: Yang, Ruolin, et al.
Veröffentlicht: (2025)
von: Yang, Ruolin, et al.
Veröffentlicht: (2025)
KeyVID: Keyframe-Aware Video Diffusion for Audio-Synchronized Visual Animation
von: Wang, Xingrui, et al.
Veröffentlicht: (2025)
von: Wang, Xingrui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers
von: Wang, Qiang, et al.
Veröffentlicht: (2025) -
FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation
von: Wang, MengChao, et al.
Veröffentlicht: (2025) -
Animate3D: Animating Any 3D Model with Multi-view Video Diffusion
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024) -
Cinemo: Consistent and Controllable Image Animation with Motion Diffusion Models
von: Ma, Xin, et al.
Veröffentlicht: (2024) -
NavForesee: A Unified Vision-Language World Model for Hierarchical Planning and Dual-Horizon Navigation Prediction
von: Liu, Fei, et al.
Veröffentlicht: (2025)