High-Fidelity and Long-Duration Human Image Animation with Diffusion Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Shen, Cai, Jiaran, Guan, Yuansheng, Huang, Shenneng, Ma, Xingpei, Cao, Junjie, Zhao, Hanfeng, Zhang, Qiang, Zhang, Shunsi, Zhang, Xiao-Ping |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Playmate2: Training-Free Multi-Character Audio-Driven Animation via Diffusion Transformer with Reward Feedback
von: Ma, Xingpei, et al.
Veröffentlicht: (2025)
von: Ma, Xingpei, et al.
Veröffentlicht: (2025)
Playmate: Flexible Control of Portrait Animation via 3D-Implicit Space Guided Diffusion
von: Ma, Xingpei, et al.
Veröffentlicht: (2025)
von: Ma, Xingpei, et al.
Veröffentlicht: (2025)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
MultiGO: Towards Multi-level Geometry Learning for Monocular 3D Textured Human Reconstruction
von: Zhang, Gangjian, et al.
Veröffentlicht: (2024)
von: Zhang, Gangjian, et al.
Veröffentlicht: (2024)
Human Multi-View Synthesis from a Single-View Model:Transferred Body and Face Representations
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance
von: Hu, Li, et al.
Veröffentlicht: (2025)
von: Hu, Li, et al.
Veröffentlicht: (2025)
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
von: Wang, Qilin, et al.
Veröffentlicht: (2024)
von: Wang, Qilin, et al.
Veröffentlicht: (2024)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
von: Wang, Xiang, et al.
Veröffentlicht: (2024)
von: Wang, Xiang, et al.
Veröffentlicht: (2024)
CDKFormer: Contextual Deviation Knowledge-Based Transformer for Long-Tail Trajectory Prediction
von: Lian, Yuansheng, et al.
Veröffentlicht: (2025)
von: Lian, Yuansheng, et al.
Veröffentlicht: (2025)
Kling-Avatar: Grounding Multimodal Instructions for Cascaded Long-Duration Avatar Animation Synthesis
von: Ding, Yikang, et al.
Veröffentlicht: (2025)
von: Ding, Yikang, et al.
Veröffentlicht: (2025)
Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation
von: Cui, Jiahao, et al.
Veröffentlicht: (2024)
von: Cui, Jiahao, et al.
Veröffentlicht: (2024)
Adaptive Duration Model for Text Speech Alignment
von: Cao, Junjie
Veröffentlicht: (2025)
von: Cao, Junjie
Veröffentlicht: (2025)
EvAnimate: Event-conditioned Image-to-Video Generation for Human Animation
von: Qu, Qiang, et al.
Veröffentlicht: (2025)
von: Qu, Qiang, et al.
Veröffentlicht: (2025)
Hydrophilic Polyethers Derived from Functional Epoxides: Beyond Poly(ethylene glycol)
von: Xingpei Hong, et al.
Veröffentlicht: (2025)
von: Xingpei Hong, et al.
Veröffentlicht: (2025)
AnimateAnywhere: Rouse the Background in Human Image Animation
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2025)
Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers
von: Ma, Xin, et al.
Veröffentlicht: (2025)
von: Ma, Xin, et al.
Veröffentlicht: (2025)
MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
von: Meng, Dechao, et al.
Veröffentlicht: (2025)
Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation
von: Zhang, Haojie, et al.
Veröffentlicht: (2024)
von: Zhang, Haojie, et al.
Veröffentlicht: (2024)
DisPose: Disentangling Pose Guidance for Controllable Human Image Animation
von: Li, Hongxiang, et al.
Veröffentlicht: (2024)
von: Li, Hongxiang, et al.
Veröffentlicht: (2024)
MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec Transformer
von: Wang, Yuancheng, et al.
Veröffentlicht: (2024)
von: Wang, Yuancheng, et al.
Veröffentlicht: (2024)
Controllable Longer Image Animation with Diffusion Models
von: Wang, Qiang, et al.
Veröffentlicht: (2024)
von: Wang, Qiang, et al.
Veröffentlicht: (2024)
FideDiff: Efficient Diffusion Model for High-Fidelity Image Motion Deblurring
von: Liu, Xiaoyang, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoyang, et al.
Veröffentlicht: (2025)
Implicit Preference Alignment for Human Image Animation
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2026)
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2026)
DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
Taming Consistency Distillation for Accelerated Human Image Animation
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
PGAHum: Prior-Guided Geometry and Appearance Learning for High-Fidelity Animatable Human Reconstruction
von: Wang, Hao, et al.
Veröffentlicht: (2024)
von: Wang, Hao, et al.
Veröffentlicht: (2024)
Animate-X: Universal Character Image Animation with Enhanced Motion Representation
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
von: Tan, Shuai, et al.
Veröffentlicht: (2024)
SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
Adaptive-avg-pooling based Attention Vision Transformer for Face Anti-spoofing
von: Yang, Jichen, et al.
Veröffentlicht: (2024)
von: Yang, Jichen, et al.
Veröffentlicht: (2024)
AniCrafter: Customizing Realistic Human-Centric Animation via Avatar-Background Conditioning in Video Diffusion Models
von: Niu, Muyao, et al.
Veröffentlicht: (2025)
von: Niu, Muyao, et al.
Veröffentlicht: (2025)
X-Dyna: Expressive Dynamic Human Image Animation
von: Chang, Di, et al.
Veröffentlicht: (2025)
von: Chang, Di, et al.
Veröffentlicht: (2025)
Multi-identity Human Image Animation with Structural Video Diffusion
von: Wang, Zhenzhi, et al.
Veröffentlicht: (2025)
von: Wang, Zhenzhi, et al.
Veröffentlicht: (2025)
InterAnimate: Taming Region-aware Diffusion Model for Realistic Human Interaction Animation
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
AnimeColor: Reference-based Animation Colorization with Diffusion Transformers
von: Zhang, Yuhong, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhong, et al.
Veröffentlicht: (2025)
TIMI: Training-Free Image-to-3D Multi-Instance Generation with Spatial Fidelity
von: Cai, Xiao, et al.
Veröffentlicht: (2026)
von: Cai, Xiao, et al.
Veröffentlicht: (2026)
How Do Human Activities Impact Soil Erosion? A Long‐Term (1985–2020) Study in the Three Gorges Reservoir Area
von: Yuansheng Huang, et al.
Veröffentlicht: (2026)
von: Yuansheng Huang, et al.
Veröffentlicht: (2026)
HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
von: Chen, Yi, et al.
Veröffentlicht: (2025)
von: Chen, Yi, et al.
Veröffentlicht: (2025)
Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
von: Cui, Jiahao, et al.
Veröffentlicht: (2024)
von: Cui, Jiahao, et al.
Veröffentlicht: (2024)
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
von: Hu, Li, et al.
Veröffentlicht: (2023)
von: Hu, Li, et al.
Veröffentlicht: (2023)
Disco4D: Disentangled 4D Human Generation and Animation from a Single Image
von: Pang, Hui En, et al.
Veröffentlicht: (2024)
von: Pang, Hui En, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Playmate2: Training-Free Multi-Character Audio-Driven Animation via Diffusion Transformer with Reward Feedback
von: Ma, Xingpei, et al.
Veröffentlicht: (2025) -
Playmate: Flexible Control of Portrait Animation via 3D-Implicit Space Guided Diffusion
von: Ma, Xingpei, et al.
Veröffentlicht: (2025) -
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
von: Wang, Xiang, et al.
Veröffentlicht: (2025) -
MultiGO: Towards Multi-level Geometry Learning for Monocular 3D Textured Human Reconstruction
von: Zhang, Gangjian, et al.
Veröffentlicht: (2024) -
Human Multi-View Synthesis from a Single-View Model:Transferred Body and Face Representations
von: Feng, Yu, et al.
Veröffentlicht: (2024)