High-Fidelity and Long-Duration Human Image Animation with Diffusion Transformer
Fuente:
arXiv
Guardado en:
| Autores principales: | Zheng, Shen, Cai, Jiaran, Guan, Yuansheng, Huang, Shenneng, Ma, Xingpei, Cao, Junjie, Zhao, Hanfeng, Zhang, Qiang, Zhang, Shunsi, Zhang, Xiao-Ping |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Playmate2: Training-Free Multi-Character Audio-Driven Animation via Diffusion Transformer with Reward Feedback
por: Ma, Xingpei, et al.
Publicado: (2025)
por: Ma, Xingpei, et al.
Publicado: (2025)
Playmate: Flexible Control of Portrait Animation via 3D-Implicit Space Guided Diffusion
por: Ma, Xingpei, et al.
Publicado: (2025)
por: Ma, Xingpei, et al.
Publicado: (2025)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
por: Wang, Xiang, et al.
Publicado: (2025)
por: Wang, Xiang, et al.
Publicado: (2025)
MultiGO: Towards Multi-level Geometry Learning for Monocular 3D Textured Human Reconstruction
por: Zhang, Gangjian, et al.
Publicado: (2024)
por: Zhang, Gangjian, et al.
Publicado: (2024)
Human Multi-View Synthesis from a Single-View Model:Transferred Body and Face Representations
por: Feng, Yu, et al.
Publicado: (2024)
por: Feng, Yu, et al.
Publicado: (2024)
Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance
por: Hu, Li, et al.
Publicado: (2025)
por: Hu, Li, et al.
Publicado: (2025)
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
por: Wang, Qilin, et al.
Publicado: (2024)
por: Wang, Qilin, et al.
Publicado: (2024)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
por: Wang, Xiang, et al.
Publicado: (2024)
por: Wang, Xiang, et al.
Publicado: (2024)
CDKFormer: Contextual Deviation Knowledge-Based Transformer for Long-Tail Trajectory Prediction
por: Lian, Yuansheng, et al.
Publicado: (2025)
por: Lian, Yuansheng, et al.
Publicado: (2025)
Kling-Avatar: Grounding Multimodal Instructions for Cascaded Long-Duration Avatar Animation Synthesis
por: Ding, Yikang, et al.
Publicado: (2025)
por: Ding, Yikang, et al.
Publicado: (2025)
Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation
por: Cui, Jiahao, et al.
Publicado: (2024)
por: Cui, Jiahao, et al.
Publicado: (2024)
Adaptive Duration Model for Text Speech Alignment
por: Cao, Junjie
Publicado: (2025)
por: Cao, Junjie
Publicado: (2025)
EvAnimate: Event-conditioned Image-to-Video Generation for Human Animation
por: Qu, Qiang, et al.
Publicado: (2025)
por: Qu, Qiang, et al.
Publicado: (2025)
Hydrophilic Polyethers Derived from Functional Epoxides: Beyond Poly(ethylene glycol)
por: Xingpei Hong, et al.
Publicado: (2025)
por: Xingpei Hong, et al.
Publicado: (2025)
AnimateAnywhere: Rouse the Background in Human Image Animation
por: Liu, Xiaoyu, et al.
Publicado: (2025)
por: Liu, Xiaoyu, et al.
Publicado: (2025)
Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers
por: Ma, Xin, et al.
Publicado: (2025)
por: Ma, Xin, et al.
Publicado: (2025)
MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation
por: Meng, Dechao, et al.
Publicado: (2025)
por: Meng, Dechao, et al.
Publicado: (2025)
Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation
por: Zhang, Haojie, et al.
Publicado: (2024)
por: Zhang, Haojie, et al.
Publicado: (2024)
DisPose: Disentangling Pose Guidance for Controllable Human Image Animation
por: Li, Hongxiang, et al.
Publicado: (2024)
por: Li, Hongxiang, et al.
Publicado: (2024)
MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer
por: Wang, Yuancheng, et al.
Publicado: (2024)
por: Wang, Yuancheng, et al.
Publicado: (2024)
Controllable Longer Image Animation with Diffusion Models
por: Wang, Qiang, et al.
Publicado: (2024)
por: Wang, Qiang, et al.
Publicado: (2024)
FideDiff: Efficient Diffusion Model for High-Fidelity Image Motion Deblurring
por: Liu, Xiaoyang, et al.
Publicado: (2025)
por: Liu, Xiaoyang, et al.
Publicado: (2025)
Implicit Preference Alignment for Human Image Animation
por: Wang, Yuanzhi, et al.
Publicado: (2026)
por: Wang, Yuanzhi, et al.
Publicado: (2026)
DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer
por: Ma, Zhiyuan, et al.
Publicado: (2024)
por: Ma, Zhiyuan, et al.
Publicado: (2024)
Taming Consistency Distillation for Accelerated Human Image Animation
por: Wang, Xiang, et al.
Publicado: (2025)
por: Wang, Xiang, et al.
Publicado: (2025)
PGAHum: Prior-Guided Geometry and Appearance Learning for High-Fidelity Animatable Human Reconstruction
por: Wang, Hao, et al.
Publicado: (2024)
por: Wang, Hao, et al.
Publicado: (2024)
Animate-X: Universal Character Image Animation with Enhanced Motion Representation
por: Tan, Shuai, et al.
Publicado: (2024)
por: Tan, Shuai, et al.
Publicado: (2024)
SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation
por: Zhang, Jiaming, et al.
Publicado: (2025)
por: Zhang, Jiaming, et al.
Publicado: (2025)
Adaptive-avg-pooling based Attention Vision Transformer for Face Anti-spoofing
por: Yang, Jichen, et al.
Publicado: (2024)
por: Yang, Jichen, et al.
Publicado: (2024)
AniCrafter: Customizing Realistic Human-Centric Animation via Avatar-Background Conditioning in Video Diffusion Models
por: Niu, Muyao, et al.
Publicado: (2025)
por: Niu, Muyao, et al.
Publicado: (2025)
X-Dyna: Expressive Dynamic Human Image Animation
por: Chang, Di, et al.
Publicado: (2025)
por: Chang, Di, et al.
Publicado: (2025)
Multi-identity Human Image Animation with Structural Video Diffusion
por: Wang, Zhenzhi, et al.
Publicado: (2025)
por: Wang, Zhenzhi, et al.
Publicado: (2025)
InterAnimate: Taming Region-aware Diffusion Model for Realistic Human Interaction Animation
por: Lin, Yukang, et al.
Publicado: (2025)
por: Lin, Yukang, et al.
Publicado: (2025)
AnimeColor: Reference-based Animation Colorization with Diffusion Transformers
por: Zhang, Yuhong, et al.
Publicado: (2025)
por: Zhang, Yuhong, et al.
Publicado: (2025)
TIMI: Training-Free Image-to-3D Multi-Instance Generation with Spatial Fidelity
por: Cai, Xiao, et al.
Publicado: (2026)
por: Cai, Xiao, et al.
Publicado: (2026)
How Do Human Activities Impact Soil Erosion? A Long‐Term (1985–2020) Study in the Three Gorges Reservoir Area
por: Yuansheng Huang, et al.
Publicado: (2026)
por: Yuansheng Huang, et al.
Publicado: (2026)
HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
por: Chen, Yi, et al.
Publicado: (2025)
por: Chen, Yi, et al.
Publicado: (2025)
Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
por: Cui, Jiahao, et al.
Publicado: (2024)
por: Cui, Jiahao, et al.
Publicado: (2024)
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
por: Hu, Li, et al.
Publicado: (2023)
por: Hu, Li, et al.
Publicado: (2023)
Disco4D: Disentangled 4D Human Generation and Animation from a Single Image
por: Pang, Hui En, et al.
Publicado: (2024)
por: Pang, Hui En, et al.
Publicado: (2024)
Ejemplares similares
-
Playmate2: Training-Free Multi-Character Audio-Driven Animation via Diffusion Transformer with Reward Feedback
por: Ma, Xingpei, et al.
Publicado: (2025) -
Playmate: Flexible Control of Portrait Animation via 3D-Implicit Space Guided Diffusion
por: Ma, Xingpei, et al.
Publicado: (2025) -
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
por: Wang, Xiang, et al.
Publicado: (2025) -
MultiGO: Towards Multi-level Geometry Learning for Monocular 3D Textured Human Reconstruction
por: Zhang, Gangjian, et al.
Publicado: (2024) -
Human Multi-View Synthesis from a Single-View Model:Transferred Body and Face Representations
por: Feng, Yu, et al.
Publicado: (2024)