3D-Aware Implicit Motion Control for View-Adaptive Human Video Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Fang, Zhixue, He, Xu, Tang, Songlin, Zhang, Haoxian, Li, Qingfeng, Liu, Xiaoqiang, Wan, Pengfei, Gai, Kun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Semantic-Aware Prefix Learning for Token-Efficient Image Generation
di: Li, Qingfeng, et al.
Pubblicazione: (2026)
di: Li, Qingfeng, et al.
Pubblicazione: (2026)
IM-Animation: An Implicit Motion Representation for Identity-decoupled Character Animation
di: Xu, Zhufeng, et al.
Pubblicazione: (2026)
di: Xu, Zhufeng, et al.
Pubblicazione: (2026)
Kling-MotionControl Technical Report
di: Kling Team, et al.
Pubblicazione: (2026)
di: Kling Team, et al.
Pubblicazione: (2026)
CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation
di: Wang, Qinghe, et al.
Pubblicazione: (2025)
di: Wang, Qinghe, et al.
Pubblicazione: (2025)
From Inpainting to Editing: Unlocking Robust Mask-Free Visual Dubbing via Generative Bootstrapping
di: He, Xu, et al.
Pubblicazione: (2025)
di: He, Xu, et al.
Pubblicazione: (2025)
MotionCharacter: Fine-Grained Motion Controllable Human Video Generation
di: Fang, Haopeng, et al.
Pubblicazione: (2024)
di: Fang, Haopeng, et al.
Pubblicazione: (2024)
MoViD: View-Invariant 3D Human Pose Estimation via Motion-View Disentanglement
di: Liu, Yejia, et al.
Pubblicazione: (2026)
di: Liu, Yejia, et al.
Pubblicazione: (2026)
OmniSync: Towards Universal Lip Synchronization via Diffusion Transformers
di: Peng, Ziqiao, et al.
Pubblicazione: (2025)
di: Peng, Ziqiao, et al.
Pubblicazione: (2025)
Cafe-Talk: Generating 3D Talking Face Animation with Multimodal Coarse- and Fine-grained Control
di: Chen, Hejia, et al.
Pubblicazione: (2025)
di: Chen, Hejia, et al.
Pubblicazione: (2025)
MIDAS: Multimodal Interactive Digital-humAn Synthesis via Real-time Autoregressive Video Generation
di: Chen, Ming, et al.
Pubblicazione: (2025)
di: Chen, Ming, et al.
Pubblicazione: (2025)
VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization
di: Cheng, Junhao, et al.
Pubblicazione: (2026)
di: Cheng, Junhao, et al.
Pubblicazione: (2026)
Scaling Image and Video Generation via Test-Time Evolutionary Search
di: He, Haoran, et al.
Pubblicazione: (2025)
di: He, Haoran, et al.
Pubblicazione: (2025)
MultiShotMaster: A Controllable Multi-Shot Video Generation Framework
di: Wang, Qinghe, et al.
Pubblicazione: (2025)
di: Wang, Qinghe, et al.
Pubblicazione: (2025)
UniVideo: Unified Understanding, Generation, and Editing for Videos
di: Wei, Cong, et al.
Pubblicazione: (2025)
di: Wei, Cong, et al.
Pubblicazione: (2025)
CamCloneMaster: Enabling Reference-based Camera Control for Video Generation
di: Luo, Yawen, et al.
Pubblicazione: (2025)
di: Luo, Yawen, et al.
Pubblicazione: (2025)
A Reason-then-Describe Instruction Interpreter for Controllable Video Generation
di: Wu, Shengqiong, et al.
Pubblicazione: (2025)
di: Wu, Shengqiong, et al.
Pubblicazione: (2025)
MotionFlow:Learning Implicit Motion Flow for Complex Camera Trajectory Control in Video Generation
di: Lei, Guojun, et al.
Pubblicazione: (2025)
di: Lei, Guojun, et al.
Pubblicazione: (2025)
VINO: A Unified Visual Generator with Interleaved OmniModal Context
di: Chen, Junyi, et al.
Pubblicazione: (2026)
di: Chen, Junyi, et al.
Pubblicazione: (2026)
Geometry-Aware Implicit Memory for Video World Models
di: Wei, Zhengxuan, et al.
Pubblicazione: (2026)
di: Wei, Zhengxuan, et al.
Pubblicazione: (2026)
Implicit Priors Editing in Stable Diffusion via Targeted Token Adjustment
di: He, Feng, et al.
Pubblicazione: (2024)
di: He, Feng, et al.
Pubblicazione: (2024)
Rethinking Generative Human Video Coding with Implicit Motion Transformation
di: Chen, Bolin, et al.
Pubblicazione: (2025)
di: Chen, Bolin, et al.
Pubblicazione: (2025)
CineScene: Implicit 3D as Effective Scene Representation for Cinematic Video Generation
di: Huang, Kaiyi, et al.
Pubblicazione: (2026)
di: Huang, Kaiyi, et al.
Pubblicazione: (2026)
FullDiT: Multi-Task Video Generative Foundation Model with Full Attention
di: Ju, Xuan, et al.
Pubblicazione: (2025)
di: Ju, Xuan, et al.
Pubblicazione: (2025)
Motion-Aware Video Frame Interpolation
di: Han, Pengfei, et al.
Pubblicazione: (2024)
di: Han, Pengfei, et al.
Pubblicazione: (2024)
AdaptiveFusion: Adaptive Multi-Modal Multi-View Fusion for 3D Human Body Reconstruction
di: Chen, Anjun, et al.
Pubblicazione: (2024)
di: Chen, Anjun, et al.
Pubblicazione: (2024)
A Survey of Interactive Generative Video
di: Yu, Jiwen, et al.
Pubblicazione: (2025)
di: Yu, Jiwen, et al.
Pubblicazione: (2025)
3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation
di: Fu, Xiao, et al.
Pubblicazione: (2024)
di: Fu, Xiao, et al.
Pubblicazione: (2024)
Generating Attribute-Aware Human Motions from Textual Prompt
di: Wang, Xinghan, et al.
Pubblicazione: (2025)
di: Wang, Xinghan, et al.
Pubblicazione: (2025)
Human Video Generation from a Single Image with 3D Pose and View Control
di: Wang, Tiantian, et al.
Pubblicazione: (2026)
di: Wang, Tiantian, et al.
Pubblicazione: (2026)
Any2Caption:Interpreting Any Condition to Caption for Controllable Video Generation
di: Wu, Shengqiong, et al.
Pubblicazione: (2025)
di: Wu, Shengqiong, et al.
Pubblicazione: (2025)
HumanScore: Benchmarking Human Motions in Generated Videos
di: Fang, Yusu, et al.
Pubblicazione: (2026)
di: Fang, Yusu, et al.
Pubblicazione: (2026)
ViewFormer: Exploring Spatiotemporal Modeling for Multi-View 3D Occupancy Perception via View-Guided Transformers
di: Li, Jinke, et al.
Pubblicazione: (2024)
di: Li, Jinke, et al.
Pubblicazione: (2024)
SceneExpander: Expanding 3D Scenes with Free-Form Inserted Views
di: He, Zijian, et al.
Pubblicazione: (2026)
di: He, Zijian, et al.
Pubblicazione: (2026)
HumanDreamer: Generating Controllable Human-Motion Videos via Decoupled Generation
di: Wang, Boyuan, et al.
Pubblicazione: (2025)
di: Wang, Boyuan, et al.
Pubblicazione: (2025)
Uni3C: Unifying Precisely 3D-Enhanced Camera and Human Motion Controls for Video Generation
di: Cao, Chenjie, et al.
Pubblicazione: (2025)
di: Cao, Chenjie, et al.
Pubblicazione: (2025)
KlingAvatar 2.0 Technical Report
di: Kling Team, et al.
Pubblicazione: (2025)
di: Kling Team, et al.
Pubblicazione: (2025)
PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion
di: Gao, Heyuan, et al.
Pubblicazione: (2026)
di: Gao, Heyuan, et al.
Pubblicazione: (2026)
Improving Video Generation with Human Feedback
di: Liu, Jie, et al.
Pubblicazione: (2025)
di: Liu, Jie, et al.
Pubblicazione: (2025)
DreamForge: Motion-Aware Autoregressive Video Generation for Multi-View Driving Scenes
di: Mei, Jianbiao, et al.
Pubblicazione: (2024)
di: Mei, Jianbiao, et al.
Pubblicazione: (2024)
UNIC: Unified In-Context Video Editing
di: Ye, Zixuan, et al.
Pubblicazione: (2025)
di: Ye, Zixuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Semantic-Aware Prefix Learning for Token-Efficient Image Generation
di: Li, Qingfeng, et al.
Pubblicazione: (2026) -
IM-Animation: An Implicit Motion Representation for Identity-decoupled Character Animation
di: Xu, Zhufeng, et al.
Pubblicazione: (2026) -
Kling-MotionControl Technical Report
di: Kling Team, et al.
Pubblicazione: (2026) -
CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation
di: Wang, Qinghe, et al.
Pubblicazione: (2025) -
From Inpainting to Editing: Unlocking Robust Mask-Free Visual Dubbing via Generative Bootstrapping
di: He, Xu, et al.
Pubblicazione: (2025)