HumanVid: Demystifying Training Data for Camera-controllable Human Image Animation
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Zhenzhi, Li, Yixuan, Zeng, Yanhong, Fang, Youqing, Guo, Yuwei, Liu, Wenran, Tan, Jing, Chen, Kai, Xue, Tianfan, Dai, Bo, Lin, Dahua |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi-identity Human Image Animation with Structural Video Diffusion
por: Wang, Zhenzhi, et al.
Publicado: (2025)
por: Wang, Zhenzhi, et al.
Publicado: (2025)
PIA: Your Personalized Image Animator via Plug-and-Play Modules in Text-to-Image Models
por: Zhang, Yiming, et al.
Publicado: (2023)
por: Zhang, Yiming, et al.
Publicado: (2023)
InterControl: Zero-shot Human Interaction Generation by Controlling Every Joint
por: Wang, Zhenzhi, et al.
Publicado: (2023)
por: Wang, Zhenzhi, et al.
Publicado: (2023)
InterActHuman: Multi-Concept Human Animation with Layout-Aligned Audio Conditions
por: Wang, Zhenzhi, et al.
Publicado: (2025)
por: Wang, Zhenzhi, et al.
Publicado: (2025)
CharacterShot: Controllable and Consistent 4D Character Animation
por: Gao, Junyao, et al.
Publicado: (2025)
por: Gao, Junyao, et al.
Publicado: (2025)
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
por: Guo, Yuwei, et al.
Publicado: (2023)
por: Guo, Yuwei, et al.
Publicado: (2023)
A Task is Worth One Word: Learning with Task Prompts for High-Quality Versatile Image Inpainting
por: Zhuang, Junhao, et al.
Publicado: (2023)
por: Zhuang, Junhao, et al.
Publicado: (2023)
VidAnimator: User-Guided Stylized 3D Character Animation from Human Videos
por: Ye, Xinwu, et al.
Publicado: (2025)
por: Ye, Xinwu, et al.
Publicado: (2025)
LEO: Generative Latent Image Animator for Human Video Synthesis
por: Wang, Yaohui, et al.
Publicado: (2023)
por: Wang, Yaohui, et al.
Publicado: (2023)
Make-It-Vivid: Dressing Your Animatable Biped Cartoon Characters from Text
por: Tang, Junshu, et al.
Publicado: (2024)
por: Tang, Junshu, et al.
Publicado: (2024)
Sagiri: Low Dynamic Range Image Enhancement with Generative Diffusion Prior
por: Li, Baiang, et al.
Publicado: (2024)
por: Li, Baiang, et al.
Publicado: (2024)
StableAnimator: High-Quality Identity-Preserving Human Image Animation
por: Tu, Shuyuan, et al.
Publicado: (2024)
por: Tu, Shuyuan, et al.
Publicado: (2024)
Improving Human Image Animation via Semantic Representation Alignment
por: Liu, Chang, et al.
Publicado: (2026)
por: Liu, Chang, et al.
Publicado: (2026)
DynamicTree: Interactive Real Tree Animation via Sparse Voxel Spectrum
por: Li, Yaokun, et al.
Publicado: (2025)
por: Li, Yaokun, et al.
Publicado: (2025)
Implicit Preference Alignment for Human Image Animation
por: Wang, Yuanzhi, et al.
Publicado: (2026)
por: Wang, Yuanzhi, et al.
Publicado: (2026)
StableAnimator++: Overcoming Pose Misalignment and Face Distortion for Human Image Animation
por: Tu, Shuyuan, et al.
Publicado: (2025)
por: Tu, Shuyuan, et al.
Publicado: (2025)
AnimateAnywhere: Rouse the Background in Human Image Animation
por: Liu, Xiaoyu, et al.
Publicado: (2025)
por: Liu, Xiaoyu, et al.
Publicado: (2025)
AnimateScene: Camera-controllable Animation in Any Scene
por: Liu, Qingyang, et al.
Publicado: (2025)
por: Liu, Qingyang, et al.
Publicado: (2025)
Proc-GS: Procedural Building Generation for City Assembly with 3D Gaussians
por: Li, Yixuan, et al.
Publicado: (2024)
por: Li, Yixuan, et al.
Publicado: (2024)
Dormant: Defending against Pose-driven Human Image Animation
por: Zhou, Jiachen, et al.
Publicado: (2024)
por: Zhou, Jiachen, et al.
Publicado: (2024)
FuseGrasp: Radar-Camera Fusion for Robotic Grasping of Transparent Objects
por: Deng, Hongyu, et al.
Publicado: (2025)
por: Deng, Hongyu, et al.
Publicado: (2025)
EvAnimate: Event-conditioned Image-to-Video Generation for Human Animation
por: Qu, Qiang, et al.
Publicado: (2025)
por: Qu, Qiang, et al.
Publicado: (2025)
PACER+: On-Demand Pedestrian Animation Controller in Driving Scenarios
por: Wang, Jingbo, et al.
Publicado: (2024)
por: Wang, Jingbo, et al.
Publicado: (2024)
SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation
por: Liu, Yujian, et al.
Publicado: (2025)
por: Liu, Yujian, et al.
Publicado: (2025)
VidCRAFT3: Camera, Object, and Lighting Control for Image-to-Video Generation
por: Zheng, Sixiao, et al.
Publicado: (2025)
por: Zheng, Sixiao, et al.
Publicado: (2025)
A Brain-inspired Computational Model for Human-like Concept Learning
por: Wang, Yuwei, et al.
Publicado: (2024)
por: Wang, Yuwei, et al.
Publicado: (2024)
Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
por: Zhu, Shenhao, et al.
Publicado: (2024)
por: Zhu, Shenhao, et al.
Publicado: (2024)
Taming Consistency Distillation for Accelerated Human Image Animation
por: Wang, Xiang, et al.
Publicado: (2025)
por: Wang, Xiang, et al.
Publicado: (2025)
X-Dyna: Expressive Dynamic Human Image Animation
por: Chang, Di, et al.
Publicado: (2025)
por: Chang, Di, et al.
Publicado: (2025)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
por: Wang, Xiang, et al.
Publicado: (2024)
por: Wang, Xiang, et al.
Publicado: (2024)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
por: Wang, Xiang, et al.
Publicado: (2025)
por: Wang, Xiang, et al.
Publicado: (2025)
OpenHumanVid: A Large-Scale High-Quality Dataset for Enhancing Human-Centric Video Generation
por: Li, Hui, et al.
Publicado: (2024)
por: Li, Hui, et al.
Publicado: (2024)
AniCrafter: Customizing Realistic Human-Centric Animation via Avatar-Background Conditioning in Video Diffusion Models
por: Niu, Muyao, et al.
Publicado: (2025)
por: Niu, Muyao, et al.
Publicado: (2025)
RelightVid: Temporal-Consistent Diffusion Model for Video Relighting
por: Fang, Ye, et al.
Publicado: (2025)
por: Fang, Ye, et al.
Publicado: (2025)
SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation
por: Zhang, Jiaming, et al.
Publicado: (2025)
por: Zhang, Jiaming, et al.
Publicado: (2025)
TalkVerse: Democratizing Minute-Long Audio-Driven Video Generation
por: Wang, Zhenzhi, et al.
Publicado: (2025)
por: Wang, Zhenzhi, et al.
Publicado: (2025)
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
por: He, Hao, et al.
Publicado: (2024)
por: He, Hao, et al.
Publicado: (2024)
Human iPSC‐based models of Alzheimer’s disease
por: Yanhong Shi
Publicado: (2024)
por: Yanhong Shi
Publicado: (2024)
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
por: Qiu, Lingteng, et al.
Publicado: (2025)
por: Qiu, Lingteng, et al.
Publicado: (2025)
Demystifying Video Reasoning
por: Wang, Ruisi, et al.
Publicado: (2026)
por: Wang, Ruisi, et al.
Publicado: (2026)
Ejemplares similares
-
Multi-identity Human Image Animation with Structural Video Diffusion
por: Wang, Zhenzhi, et al.
Publicado: (2025) -
PIA: Your Personalized Image Animator via Plug-and-Play Modules in Text-to-Image Models
por: Zhang, Yiming, et al.
Publicado: (2023) -
InterControl: Zero-shot Human Interaction Generation by Controlling Every Joint
por: Wang, Zhenzhi, et al.
Publicado: (2023) -
InterActHuman: Multi-Concept Human Animation with Layout-Aligned Audio Conditions
por: Wang, Zhenzhi, et al.
Publicado: (2025) -
CharacterShot: Controllable and Consistent 4D Character Animation
por: Gao, Junyao, et al.
Publicado: (2025)