Towards 3D-Aware Video Diffusion Models: Render-Free Human Motion Control with Mesh Tokenization
Fuente:
arXiv
Saved in:
| Main Authors: | Liang, Jingyun, Wei, Min, Li, Shikai, Han, Yizeng, Yuan, Hangjie, Sun, Lei, Chen, Weihua, Wang, Fan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
by: Liang, Jingyun, et al.
Published: (2025)
by: Liang, Jingyun, et al.
Published: (2025)
MoVideo: Motion-Aware Video Generation with Diffusion Models
by: Liang, Jingyun, et al.
Published: (2023)
by: Liang, Jingyun, et al.
Published: (2023)
Subjective and Objective Quality Assessment of Rendered Human Avatar Videos in Virtual Reality
by: Chen, Yu-Chih, et al.
Published: (2024)
by: Chen, Yu-Chih, et al.
Published: (2024)
Unsupervised Cardiac Video Translation Via Motion Feature Guided Diffusion Model
by: Deb, Swakshar, et al.
Published: (2025)
by: Deb, Swakshar, et al.
Published: (2025)
RealisVSR: Detail-enhanced Diffusion for Real-World 4K Video Super-Resolution
by: Zhao, Weisong, et al.
Published: (2025)
by: Zhao, Weisong, et al.
Published: (2025)
LRC-DHVC: Towards Local Rate Control in Neural Video Compression
by: Windsheimer, Marc, et al.
Published: (2026)
by: Windsheimer, Marc, et al.
Published: (2026)
LanPaint: Training-Free Diffusion Inpainting with Asymptotically Exact and Fast Conditional Sampling
by: Zheng, Candi, et al.
Published: (2025)
by: Zheng, Candi, et al.
Published: (2025)
TVC: Tokenized Video Compression with Ultra-Low Bit Rate
by: Zhou, Lebin, et al.
Published: (2025)
by: Zhou, Lebin, et al.
Published: (2025)
D4PM: A Dual-branch Driven Denoising Diffusion Probabilistic Model with Joint Posterior Diffusion Sampling for EEG Artifacts Removal
by: Shao, Feixue, et al.
Published: (2025)
by: Shao, Feixue, et al.
Published: (2025)
Improved Motion Plane Adaptive 360-Degree Video Compression Using Affine Motion Models
by: Ritthaler, Marina, et al.
Published: (2025)
by: Ritthaler, Marina, et al.
Published: (2025)
Encoder-Quantization-Motion-based Video Quality Metrics
by: Chen, Yixu, et al.
Published: (2024)
by: Chen, Yixu, et al.
Published: (2024)
VideoSPatS: Video SPatiotemporal Splines for Disentangled Occlusion, Appearance and Motion Modeling and Editing
by: Bello, Juan Luis Gonzalez, et al.
Published: (2025)
by: Bello, Juan Luis Gonzalez, et al.
Published: (2025)
Reinforced Rate Control for Neural Video Compression via Inter-Frame Rate-Distortion Awareness
by: Cong, Wuyang, et al.
Published: (2026)
by: Cong, Wuyang, et al.
Published: (2026)
An Energy-Efficient Edge Coprocessor for Neural Rendering with Explicit Data Reuse Strategies
by: Yuan, Binzhe, et al.
Published: (2025)
by: Yuan, Binzhe, et al.
Published: (2025)
QoE-oriented Communication Service Provision for Annotation Rendering in Mobile Augmented Reality
by: Sun, Lulu, et al.
Published: (2025)
by: Sun, Lulu, et al.
Published: (2025)
LUM-ViT: Learnable Under-sampling Mask Vision Transformer for Bandwidth Limited Optical Signal Acquisition
by: Liu, Lingfeng, et al.
Published: (2024)
by: Liu, Lingfeng, et al.
Published: (2024)
Motion Free B-frame Coding for Neural Video Compression
by: Nguyen, Van Thang
Published: (2024)
by: Nguyen, Van Thang
Published: (2024)
Accelerated Image-Aware Generative Diffusion Modeling
by: Asthana, Tanmay, et al.
Published: (2024)
by: Asthana, Tanmay, et al.
Published: (2024)
Cardiac Mesh Flow: One-Step Generation of 3D+t Cardiac Four-Chamber Meshes via Flow Matching
by: Ma, Qiang, et al.
Published: (2026)
by: Ma, Qiang, et al.
Published: (2026)
Energy-Aware Frame Rate Selection for Video Coding
by: Ramasubbu, Geetha, et al.
Published: (2026)
by: Ramasubbu, Geetha, et al.
Published: (2026)
Rethinking Generative Human Video Coding with Implicit Motion Transformation
by: Chen, Bolin, et al.
Published: (2025)
by: Chen, Bolin, et al.
Published: (2025)
Decentralized Gossip Mutual Learning (GML) for automatic head and neck tumor segmentation
by: Chen, Jingyun, et al.
Published: (2024)
by: Chen, Jingyun, et al.
Published: (2024)
Decentralized Gossip Mutual Learning (GML) for brain tumor segmentation on multi-parametric MRI
by: Chen, Jingyun, et al.
Published: (2024)
by: Chen, Jingyun, et al.
Published: (2024)
FreqPrior: Improving Video Diffusion Models with Frequency Filtering Gaussian Noise
by: Yuan, Yunlong, et al.
Published: (2025)
by: Yuan, Yunlong, et al.
Published: (2025)
A Video-Aware FEC-Based Unequal Loss Protection System for Video Streaming over RTP
by: Díaz, César, et al.
Published: (2024)
by: Díaz, César, et al.
Published: (2024)
V2M4: 4D Mesh Animation Reconstruction from a Single Monocular Video
by: Chen, Jianqi, et al.
Published: (2025)
by: Chen, Jianqi, et al.
Published: (2025)
TVMC: Time-Varying Mesh Compression via Multi-Stage Anchor Mesh Generation
by: Huang, He, et al.
Published: (2025)
by: Huang, He, et al.
Published: (2025)
Rethinking Video Super-Resolution: Towards Diffusion-Based Methods without Motion Alignment
by: Zhan, Zhihao, et al.
Published: (2025)
by: Zhan, Zhihao, et al.
Published: (2025)
Reconsider the Template Mesh in Deep Learning-based Mesh Reconstruction
by: Zhang, Fengting, et al.
Published: (2025)
by: Zhang, Fengting, et al.
Published: (2025)
Efficient Sub-pixel Motion Compensation in Learned Video Codecs
by: Ladune, Théo, et al.
Published: (2025)
by: Ladune, Théo, et al.
Published: (2025)
CASC: Condition-Aware Semantic Communication with Latent Diffusion Models
by: Chen, Weixuan, et al.
Published: (2024)
by: Chen, Weixuan, et al.
Published: (2024)
Structure-Aware Adaptive Kernel MPPCA Denoising for Diffusion MRI
by: Singhal, Ananya, et al.
Published: (2025)
by: Singhal, Ananya, et al.
Published: (2025)
Diffusion-aided Extreme Video Compression with Lightweight Semantics Guidance
by: Zhang, Maojun, et al.
Published: (2026)
by: Zhang, Maojun, et al.
Published: (2026)
Perception-Aware Video Semantic Communication
by: Huang, Yinhuan, et al.
Published: (2026)
by: Huang, Yinhuan, et al.
Published: (2026)
Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models
by: Wu, Yiming, et al.
Published: (2024)
by: Wu, Yiming, et al.
Published: (2024)
GeoDiff-SAR II: 3D-Driven Foundation Diffusion Models for SAR Generation via Decoupled Control
by: Wu, Xuanting, et al.
Published: (2026)
by: Wu, Xuanting, et al.
Published: (2026)
MeshBrush: Painting the Anatomical Mesh with Neural Stylization for Endoscopy
by: Han, John J., et al.
Published: (2024)
by: Han, John J., et al.
Published: (2024)
Energy- and Quality-Aware Video Request Policy for Wireless Adaptive Streaming Clients
by: Díaz, César, et al.
Published: (2024)
by: Díaz, César, et al.
Published: (2024)
Rate Splitting Multiple Access-Enabled Adaptive Panoramic Video Semantic Transmission
by: Gao, Haixiao, et al.
Published: (2024)
by: Gao, Haixiao, et al.
Published: (2024)
Characterizing Motion Encoding in Video Diffusion Timesteps
by: Baherwani, Vatsal, et al.
Published: (2025)
by: Baherwani, Vatsal, et al.
Published: (2025)
Similar Items
-
RealisMotion: Decomposed Human Motion Control and Video Generation in the World Space
by: Liang, Jingyun, et al.
Published: (2025) -
MoVideo: Motion-Aware Video Generation with Diffusion Models
by: Liang, Jingyun, et al.
Published: (2023) -
Subjective and Objective Quality Assessment of Rendered Human Avatar Videos in Virtual Reality
by: Chen, Yu-Chih, et al.
Published: (2024) -
Unsupervised Cardiac Video Translation Via Motion Feature Guided Diffusion Model
by: Deb, Swakshar, et al.
Published: (2025) -
RealisVSR: Detail-enhanced Diffusion for Real-World 4K Video Super-Resolution
by: Zhao, Weisong, et al.
Published: (2025)