HumanVid: Demystifying Training Data for Camera-controllable Human Image Animation
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Zhenzhi, Li, Yixuan, Zeng, Yanhong, Fang, Youqing, Guo, Yuwei, Liu, Wenran, Tan, Jing, Chen, Kai, Xue, Tianfan, Dai, Bo, Lin, Dahua |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-identity Human Image Animation with Structural Video Diffusion
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
PIA: Your Personalized Image Animator via Plug-and-Play Modules in Text-to-Image Models
di: Zhang, Yiming, et al.
Pubblicazione: (2023)
di: Zhang, Yiming, et al.
Pubblicazione: (2023)
InterControl: Zero-shot Human Interaction Generation by Controlling Every Joint
di: Wang, Zhenzhi, et al.
Pubblicazione: (2023)
di: Wang, Zhenzhi, et al.
Pubblicazione: (2023)
InterActHuman: Multi-Concept Human Animation with Layout-Aligned Audio Conditions
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
CharacterShot: Controllable and Consistent 4D Character Animation
di: Gao, Junyao, et al.
Pubblicazione: (2025)
di: Gao, Junyao, et al.
Pubblicazione: (2025)
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
di: Guo, Yuwei, et al.
Pubblicazione: (2023)
di: Guo, Yuwei, et al.
Pubblicazione: (2023)
A Task is Worth One Word: Learning with Task Prompts for High-Quality Versatile Image Inpainting
di: Zhuang, Junhao, et al.
Pubblicazione: (2023)
di: Zhuang, Junhao, et al.
Pubblicazione: (2023)
VidAnimator: User-Guided Stylized 3D Character Animation from Human Videos
di: Ye, Xinwu, et al.
Pubblicazione: (2025)
di: Ye, Xinwu, et al.
Pubblicazione: (2025)
LEO: Generative Latent Image Animator for Human Video Synthesis
di: Wang, Yaohui, et al.
Pubblicazione: (2023)
di: Wang, Yaohui, et al.
Pubblicazione: (2023)
Make-It-Vivid: Dressing Your Animatable Biped Cartoon Characters from Text
di: Tang, Junshu, et al.
Pubblicazione: (2024)
di: Tang, Junshu, et al.
Pubblicazione: (2024)
Sagiri: Low Dynamic Range Image Enhancement with Generative Diffusion Prior
di: Li, Baiang, et al.
Pubblicazione: (2024)
di: Li, Baiang, et al.
Pubblicazione: (2024)
StableAnimator: High-Quality Identity-Preserving Human Image Animation
di: Tu, Shuyuan, et al.
Pubblicazione: (2024)
di: Tu, Shuyuan, et al.
Pubblicazione: (2024)
Improving Human Image Animation via Semantic Representation Alignment
di: Liu, Chang, et al.
Pubblicazione: (2026)
di: Liu, Chang, et al.
Pubblicazione: (2026)
DynamicTree: Interactive Real Tree Animation via Sparse Voxel Spectrum
di: Li, Yaokun, et al.
Pubblicazione: (2025)
di: Li, Yaokun, et al.
Pubblicazione: (2025)
Implicit Preference Alignment for Human Image Animation
di: Wang, Yuanzhi, et al.
Pubblicazione: (2026)
di: Wang, Yuanzhi, et al.
Pubblicazione: (2026)
StableAnimator++: Overcoming Pose Misalignment and Face Distortion for Human Image Animation
di: Tu, Shuyuan, et al.
Pubblicazione: (2025)
di: Tu, Shuyuan, et al.
Pubblicazione: (2025)
AnimateAnywhere: Rouse the Background in Human Image Animation
di: Liu, Xiaoyu, et al.
Pubblicazione: (2025)
di: Liu, Xiaoyu, et al.
Pubblicazione: (2025)
AnimateScene: Camera-controllable Animation in Any Scene
di: Liu, Qingyang, et al.
Pubblicazione: (2025)
di: Liu, Qingyang, et al.
Pubblicazione: (2025)
Proc-GS: Procedural Building Generation for City Assembly with 3D Gaussians
di: Li, Yixuan, et al.
Pubblicazione: (2024)
di: Li, Yixuan, et al.
Pubblicazione: (2024)
Dormant: Defending against Pose-driven Human Image Animation
di: Zhou, Jiachen, et al.
Pubblicazione: (2024)
di: Zhou, Jiachen, et al.
Pubblicazione: (2024)
FuseGrasp: Radar-Camera Fusion for Robotic Grasping of Transparent Objects
di: Deng, Hongyu, et al.
Pubblicazione: (2025)
di: Deng, Hongyu, et al.
Pubblicazione: (2025)
EvAnimate: Event-conditioned Image-to-Video Generation for Human Animation
di: Qu, Qiang, et al.
Pubblicazione: (2025)
di: Qu, Qiang, et al.
Pubblicazione: (2025)
PACER+: On-Demand Pedestrian Animation Controller in Driving Scenarios
di: Wang, Jingbo, et al.
Pubblicazione: (2024)
di: Wang, Jingbo, et al.
Pubblicazione: (2024)
SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation
di: Liu, Yujian, et al.
Pubblicazione: (2025)
di: Liu, Yujian, et al.
Pubblicazione: (2025)
VidCRAFT3: Camera, Object, and Lighting Control for Image-to-Video Generation
di: Zheng, Sixiao, et al.
Pubblicazione: (2025)
di: Zheng, Sixiao, et al.
Pubblicazione: (2025)
A Brain-inspired Computational Model for Human-like Concept Learning
di: Wang, Yuwei, et al.
Pubblicazione: (2024)
di: Wang, Yuwei, et al.
Pubblicazione: (2024)
Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
di: Zhu, Shenhao, et al.
Pubblicazione: (2024)
di: Zhu, Shenhao, et al.
Pubblicazione: (2024)
Taming Consistency Distillation for Accelerated Human Image Animation
di: Wang, Xiang, et al.
Pubblicazione: (2025)
di: Wang, Xiang, et al.
Pubblicazione: (2025)
X-Dyna: Expressive Dynamic Human Image Animation
di: Chang, Di, et al.
Pubblicazione: (2025)
di: Chang, Di, et al.
Pubblicazione: (2025)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
di: Wang, Xiang, et al.
Pubblicazione: (2024)
di: Wang, Xiang, et al.
Pubblicazione: (2024)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
di: Wang, Xiang, et al.
Pubblicazione: (2025)
di: Wang, Xiang, et al.
Pubblicazione: (2025)
OpenHumanVid: A Large-Scale High-Quality Dataset for Enhancing Human-Centric Video Generation
di: Li, Hui, et al.
Pubblicazione: (2024)
di: Li, Hui, et al.
Pubblicazione: (2024)
AniCrafter: Customizing Realistic Human-Centric Animation via Avatar-Background Conditioning in Video Diffusion Models
di: Niu, Muyao, et al.
Pubblicazione: (2025)
di: Niu, Muyao, et al.
Pubblicazione: (2025)
RelightVid: Temporal-Consistent Diffusion Model for Video Relighting
di: Fang, Ye, et al.
Pubblicazione: (2025)
di: Fang, Ye, et al.
Pubblicazione: (2025)
SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation
di: Zhang, Jiaming, et al.
Pubblicazione: (2025)
di: Zhang, Jiaming, et al.
Pubblicazione: (2025)
TalkVerse: Democratizing Minute-Long Audio-Driven Video Generation
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
di: He, Hao, et al.
Pubblicazione: (2024)
di: He, Hao, et al.
Pubblicazione: (2024)
Human iPSC‐based models of Alzheimer’s disease
di: Yanhong Shi
Pubblicazione: (2024)
di: Yanhong Shi
Pubblicazione: (2024)
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
di: Qiu, Lingteng, et al.
Pubblicazione: (2025)
di: Qiu, Lingteng, et al.
Pubblicazione: (2025)
Demystifying Video Reasoning
di: Wang, Ruisi, et al.
Pubblicazione: (2026)
di: Wang, Ruisi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Multi-identity Human Image Animation with Structural Video Diffusion
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025) -
PIA: Your Personalized Image Animator via Plug-and-Play Modules in Text-to-Image Models
di: Zhang, Yiming, et al.
Pubblicazione: (2023) -
InterControl: Zero-shot Human Interaction Generation by Controlling Every Joint
di: Wang, Zhenzhi, et al.
Pubblicazione: (2023) -
InterActHuman: Multi-Concept Human Animation with Layout-Aligned Audio Conditions
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025) -
CharacterShot: Controllable and Consistent 4D Character Animation
di: Gao, Junyao, et al.
Pubblicazione: (2025)