Learning to Animate Images from A Few Videos to Portray Delicate Human Actions
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Haoxin, Yu, Yingchen, Wu, Qilong, Zhang, Hanwang, Bai, Song, Li, Boyang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Versatile Transition Generation with Image-to-Video Diffusion
di: Yang, Zuhao, et al.
Pubblicazione: (2025)
di: Yang, Zuhao, et al.
Pubblicazione: (2025)
Enhancing Vision-Language Compositional Understanding with Multimodal Synthetic Data
di: Li, Haoxin, et al.
Pubblicazione: (2025)
di: Li, Haoxin, et al.
Pubblicazione: (2025)
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding
di: Yang, Zuhao, et al.
Pubblicazione: (2025)
di: Yang, Zuhao, et al.
Pubblicazione: (2025)
An Animation-based Augmentation Approach for Action Recognition from Discontinuous Video
di: Song, Xingyu, et al.
Pubblicazione: (2024)
di: Song, Xingyu, et al.
Pubblicazione: (2024)
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
di: Wang, Xiang, et al.
Pubblicazione: (2024)
di: Wang, Xiang, et al.
Pubblicazione: (2024)
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
di: Hu, Li, et al.
Pubblicazione: (2023)
di: Hu, Li, et al.
Pubblicazione: (2023)
AnimateAnywhere: Rouse the Background in Human Image Animation
di: Liu, Xiaoyu, et al.
Pubblicazione: (2025)
di: Liu, Xiaoyu, et al.
Pubblicazione: (2025)
LEO: Generative Latent Image Animator for Human Video Synthesis
di: Wang, Yaohui, et al.
Pubblicazione: (2023)
di: Wang, Yaohui, et al.
Pubblicazione: (2023)
EvAnimate: Event-conditioned Image-to-Video Generation for Human Animation
di: Qu, Qiang, et al.
Pubblicazione: (2025)
di: Qu, Qiang, et al.
Pubblicazione: (2025)
HumanEdit: A High-Quality Human-Rewarded Dataset for Instruction-based Image Editing
di: Bai, Jinbin, et al.
Pubblicazione: (2024)
di: Bai, Jinbin, et al.
Pubblicazione: (2024)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
di: Wang, Xiang, et al.
Pubblicazione: (2025)
di: Wang, Xiang, et al.
Pubblicazione: (2025)
SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models
di: Yang, Ruolin, et al.
Pubblicazione: (2025)
di: Yang, Ruolin, et al.
Pubblicazione: (2025)
TAMT: Temporal-Aware Model Tuning for Cross-Domain Few-Shot Action Recognition
di: Wang, Yilong, et al.
Pubblicazione: (2024)
di: Wang, Yilong, et al.
Pubblicazione: (2024)
Debiasing Text-to-Image Diffusion Models
di: He, Ruifei, et al.
Pubblicazione: (2024)
di: He, Ruifei, et al.
Pubblicazione: (2024)
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
di: Wang, Qilin, et al.
Pubblicazione: (2024)
di: Wang, Qilin, et al.
Pubblicazione: (2024)
Multi-identity Human Image Animation with Structural Video Diffusion
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
AnimateAnything: Consistent and Controllable Animation for Video Generation
di: Lei, Guojun, et al.
Pubblicazione: (2024)
di: Lei, Guojun, et al.
Pubblicazione: (2024)
SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation
di: Zhang, Jiaming, et al.
Pubblicazione: (2025)
di: Zhang, Jiaming, et al.
Pubblicazione: (2025)
A$^2$M$^2$-Net: Adaptively Aligned Multi-Scale Moment for Few-Shot Action Recognition
di: Gao, Zilin, et al.
Pubblicazione: (2025)
di: Gao, Zilin, et al.
Pubblicazione: (2025)
VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation
di: Chen, Hanzhi, et al.
Pubblicazione: (2025)
di: Chen, Hanzhi, et al.
Pubblicazione: (2025)
Controllable Relation Disentanglement for Few-Shot Class-Incremental Learning
di: Zhou, Yuan, et al.
Pubblicazione: (2024)
di: Zhou, Yuan, et al.
Pubblicazione: (2024)
Paint Outside the Box: Synthesizing and Selecting Training Data for Visual Grounding
di: Du, Zilin, et al.
Pubblicazione: (2024)
di: Du, Zilin, et al.
Pubblicazione: (2024)
Animate Your Motion: Turning Still Images into Dynamic Videos
di: Li, Mingxiao, et al.
Pubblicazione: (2024)
di: Li, Mingxiao, et al.
Pubblicazione: (2024)
Thinking with Images as Continuous Actions: Numerical Visual Chain-of-Thought
di: Zhao, Kesen, et al.
Pubblicazione: (2026)
di: Zhao, Kesen, et al.
Pubblicazione: (2026)
X-Dyna: Expressive Dynamic Human Image Animation
di: Chang, Di, et al.
Pubblicazione: (2025)
di: Chang, Di, et al.
Pubblicazione: (2025)
DisPose: Disentangling Pose Guidance for Controllable Human Image Animation
di: Li, Hongxiang, et al.
Pubblicazione: (2024)
di: Li, Hongxiang, et al.
Pubblicazione: (2024)
InstrAct: Towards Action-Centric Understanding in Instructional Videos
di: Yang, Zhuoyi, et al.
Pubblicazione: (2026)
di: Yang, Zhuoyi, et al.
Pubblicazione: (2026)
Taming Consistency Distillation for Accelerated Human Image Animation
di: Wang, Xiang, et al.
Pubblicazione: (2025)
di: Wang, Xiang, et al.
Pubblicazione: (2025)
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling
di: Lan, Mengcheng, et al.
Pubblicazione: (2025)
di: Lan, Mengcheng, et al.
Pubblicazione: (2025)
InsTaG: Learning Personalized 3D Talking Head from Few-Second Video
di: Li, Jiahe, et al.
Pubblicazione: (2025)
di: Li, Jiahe, et al.
Pubblicazione: (2025)
Few-shot Learner Parameterization by Diffusion Time-steps
di: Yue, Zhongqi, et al.
Pubblicazione: (2024)
di: Yue, Zhongqi, et al.
Pubblicazione: (2024)
StableAnimator++: Overcoming Pose Misalignment and Face Distortion for Human Image Animation
di: Tu, Shuyuan, et al.
Pubblicazione: (2025)
di: Tu, Shuyuan, et al.
Pubblicazione: (2025)
Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition
di: Guo, Hanyu, et al.
Pubblicazione: (2024)
di: Guo, Hanyu, et al.
Pubblicazione: (2024)
AnimateAnyMesh++: A Flexible 4D Foundation Model for High-Fidelity Text-Driven Mesh Animation
di: Wu, Zijie, et al.
Pubblicazione: (2026)
di: Wu, Zijie, et al.
Pubblicazione: (2026)
AnimateAnyMesh: A Feed-Forward 4D Foundation Model for Text-Driven Universal Mesh Animation
di: Wu, Zijie, et al.
Pubblicazione: (2025)
di: Wu, Zijie, et al.
Pubblicazione: (2025)
AnimationBench: Are Video Models Good at Character-Centric Animation?
di: Wu, Leyi, et al.
Pubblicazione: (2026)
di: Wu, Leyi, et al.
Pubblicazione: (2026)
Hierarchical Compositional Representations for Few-shot Action Recognition
di: Li, Changzhen, et al.
Pubblicazione: (2022)
di: Li, Changzhen, et al.
Pubblicazione: (2022)
LoRA of Change: Learning to Generate LoRA for the Editing Instruction from A Single Before-After Image Pair
di: Song, Xue, et al.
Pubblicazione: (2024)
di: Song, Xue, et al.
Pubblicazione: (2024)
Bayesian Evidential Learning for Few-Shot Classification
di: Linghu, Xiongkun, et al.
Pubblicazione: (2022)
di: Linghu, Xiongkun, et al.
Pubblicazione: (2022)
FAGhead: Fully Animate Gaussian Head from Monocular Videos
di: Xuan, Yixin, et al.
Pubblicazione: (2024)
di: Xuan, Yixin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Versatile Transition Generation with Image-to-Video Diffusion
di: Yang, Zuhao, et al.
Pubblicazione: (2025) -
Enhancing Vision-Language Compositional Understanding with Multimodal Synthetic Data
di: Li, Haoxin, et al.
Pubblicazione: (2025) -
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding
di: Yang, Zuhao, et al.
Pubblicazione: (2025) -
An Animation-based Augmentation Approach for Action Recognition from Discontinuous Video
di: Song, Xingyu, et al.
Pubblicazione: (2024) -
UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
di: Wang, Xiang, et al.
Pubblicazione: (2024)