3D-Aware Implicit Motion Control for View-Adaptive Human Video Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Fang, Zhixue, He, Xu, Tang, Songlin, Zhang, Haoxian, Li, Qingfeng, Liu, Xiaoqiang, Wan, Pengfei, Gai, Kun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Semantic-Aware Prefix Learning for Token-Efficient Image Generation
por: Li, Qingfeng, et al.
Publicado: (2026)
por: Li, Qingfeng, et al.
Publicado: (2026)
IM-Animation: An Implicit Motion Representation for Identity-decoupled Character Animation
por: Xu, Zhufeng, et al.
Publicado: (2026)
por: Xu, Zhufeng, et al.
Publicado: (2026)
Kling-MotionControl Technical Report
por: Kling Team, et al.
Publicado: (2026)
por: Kling Team, et al.
Publicado: (2026)
CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation
por: Wang, Qinghe, et al.
Publicado: (2025)
por: Wang, Qinghe, et al.
Publicado: (2025)
From Inpainting to Editing: Unlocking Robust Mask-Free Visual Dubbing via Generative Bootstrapping
por: He, Xu, et al.
Publicado: (2025)
por: He, Xu, et al.
Publicado: (2025)
MotionCharacter: Fine-Grained Motion Controllable Human Video Generation
por: Fang, Haopeng, et al.
Publicado: (2024)
por: Fang, Haopeng, et al.
Publicado: (2024)
MoViD: View-Invariant 3D Human Pose Estimation via Motion-View Disentanglement
por: Liu, Yejia, et al.
Publicado: (2026)
por: Liu, Yejia, et al.
Publicado: (2026)
OmniSync: Towards Universal Lip Synchronization via Diffusion Transformers
por: Peng, Ziqiao, et al.
Publicado: (2025)
por: Peng, Ziqiao, et al.
Publicado: (2025)
Cafe-Talk: Generating 3D Talking Face Animation with Multimodal Coarse- and Fine-grained Control
por: Chen, Hejia, et al.
Publicado: (2025)
por: Chen, Hejia, et al.
Publicado: (2025)
MIDAS: Multimodal Interactive Digital-humAn Synthesis via Real-time Autoregressive Video Generation
por: Chen, Ming, et al.
Publicado: (2025)
por: Chen, Ming, et al.
Publicado: (2025)
VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization
por: Cheng, Junhao, et al.
Publicado: (2026)
por: Cheng, Junhao, et al.
Publicado: (2026)
Scaling Image and Video Generation via Test-Time Evolutionary Search
por: He, Haoran, et al.
Publicado: (2025)
por: He, Haoran, et al.
Publicado: (2025)
MultiShotMaster: A Controllable Multi-Shot Video Generation Framework
por: Wang, Qinghe, et al.
Publicado: (2025)
por: Wang, Qinghe, et al.
Publicado: (2025)
UniVideo: Unified Understanding, Generation, and Editing for Videos
por: Wei, Cong, et al.
Publicado: (2025)
por: Wei, Cong, et al.
Publicado: (2025)
CamCloneMaster: Enabling Reference-based Camera Control for Video Generation
por: Luo, Yawen, et al.
Publicado: (2025)
por: Luo, Yawen, et al.
Publicado: (2025)
A Reason-then-Describe Instruction Interpreter for Controllable Video Generation
por: Wu, Shengqiong, et al.
Publicado: (2025)
por: Wu, Shengqiong, et al.
Publicado: (2025)
MotionFlow:Learning Implicit Motion Flow for Complex Camera Trajectory Control in Video Generation
por: Lei, Guojun, et al.
Publicado: (2025)
por: Lei, Guojun, et al.
Publicado: (2025)
Geometry-Aware Implicit Memory for Video World Models
por: Wei, Zhengxuan, et al.
Publicado: (2026)
por: Wei, Zhengxuan, et al.
Publicado: (2026)
VINO: A Unified Visual Generator with Interleaved OmniModal Context
por: Chen, Junyi, et al.
Publicado: (2026)
por: Chen, Junyi, et al.
Publicado: (2026)
Implicit Priors Editing in Stable Diffusion via Targeted Token Adjustment
por: He, Feng, et al.
Publicado: (2024)
por: He, Feng, et al.
Publicado: (2024)
Rethinking Generative Human Video Coding with Implicit Motion Transformation
por: Chen, Bolin, et al.
Publicado: (2025)
por: Chen, Bolin, et al.
Publicado: (2025)
CineScene: Implicit 3D as Effective Scene Representation for Cinematic Video Generation
por: Huang, Kaiyi, et al.
Publicado: (2026)
por: Huang, Kaiyi, et al.
Publicado: (2026)
FullDiT: Multi-Task Video Generative Foundation Model with Full Attention
por: Ju, Xuan, et al.
Publicado: (2025)
por: Ju, Xuan, et al.
Publicado: (2025)
Motion-Aware Video Frame Interpolation
por: Han, Pengfei, et al.
Publicado: (2024)
por: Han, Pengfei, et al.
Publicado: (2024)
AdaptiveFusion: Adaptive Multi-Modal Multi-View Fusion for 3D Human Body Reconstruction
por: Chen, Anjun, et al.
Publicado: (2024)
por: Chen, Anjun, et al.
Publicado: (2024)
3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation
por: Fu, Xiao, et al.
Publicado: (2024)
por: Fu, Xiao, et al.
Publicado: (2024)
A Survey of Interactive Generative Video
por: Yu, Jiwen, et al.
Publicado: (2025)
por: Yu, Jiwen, et al.
Publicado: (2025)
Generating Attribute-Aware Human Motions from Textual Prompt
por: Wang, Xinghan, et al.
Publicado: (2025)
por: Wang, Xinghan, et al.
Publicado: (2025)
Human Video Generation from a Single Image with 3D Pose and View Control
por: Wang, Tiantian, et al.
Publicado: (2026)
por: Wang, Tiantian, et al.
Publicado: (2026)
Any2Caption:Interpreting Any Condition to Caption for Controllable Video Generation
por: Wu, Shengqiong, et al.
Publicado: (2025)
por: Wu, Shengqiong, et al.
Publicado: (2025)
HumanScore: Benchmarking Human Motions in Generated Videos
por: Fang, Yusu, et al.
Publicado: (2026)
por: Fang, Yusu, et al.
Publicado: (2026)
ViewFormer: Exploring Spatiotemporal Modeling for Multi-View 3D Occupancy Perception via View-Guided Transformers
por: Li, Jinke, et al.
Publicado: (2024)
por: Li, Jinke, et al.
Publicado: (2024)
SceneExpander: Expanding 3D Scenes with Free-Form Inserted Views
por: He, Zijian, et al.
Publicado: (2026)
por: He, Zijian, et al.
Publicado: (2026)
HumanDreamer: Generating Controllable Human-Motion Videos via Decoupled Generation
por: Wang, Boyuan, et al.
Publicado: (2025)
por: Wang, Boyuan, et al.
Publicado: (2025)
Uni3C: Unifying Precisely 3D-Enhanced Camera and Human Motion Controls for Video Generation
por: Cao, Chenjie, et al.
Publicado: (2025)
por: Cao, Chenjie, et al.
Publicado: (2025)
KlingAvatar 2.0 Technical Report
por: Kling Team, et al.
Publicado: (2025)
por: Kling Team, et al.
Publicado: (2025)
PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion
por: Gao, Heyuan, et al.
Publicado: (2026)
por: Gao, Heyuan, et al.
Publicado: (2026)
Improving Video Generation with Human Feedback
por: Liu, Jie, et al.
Publicado: (2025)
por: Liu, Jie, et al.
Publicado: (2025)
DreamForge: Motion-Aware Autoregressive Video Generation for Multi-View Driving Scenes
por: Mei, Jianbiao, et al.
Publicado: (2024)
por: Mei, Jianbiao, et al.
Publicado: (2024)
UNIC: Unified In-Context Video Editing
por: Ye, Zixuan, et al.
Publicado: (2025)
por: Ye, Zixuan, et al.
Publicado: (2025)
Ejemplares similares
-
Semantic-Aware Prefix Learning for Token-Efficient Image Generation
por: Li, Qingfeng, et al.
Publicado: (2026) -
IM-Animation: An Implicit Motion Representation for Identity-decoupled Character Animation
por: Xu, Zhufeng, et al.
Publicado: (2026) -
Kling-MotionControl Technical Report
por: Kling Team, et al.
Publicado: (2026) -
CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation
por: Wang, Qinghe, et al.
Publicado: (2025) -
From Inpainting to Editing: Unlocking Robust Mask-Free Visual Dubbing via Generative Bootstrapping
por: He, Xu, et al.
Publicado: (2025)