Sync4D: Video Guided Controllable Dynamics for Physics-Based 4D Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Fu, Zhoujie, Wei, Jiacheng, Shen, Wenhao, Song, Chaoyue, Yang, Xiaofeng, Liu, Fayao, Yang, Xulei, Lin, Guosheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
REACTO: Reconstructing Articulated Objects from a Single Video
by: Song, Chaoyue, et al.
Published: (2024)
by: Song, Chaoyue, et al.
Published: (2024)
Sculpt3D: Multi-View Consistent Text-to-3D Generation with Sparse 3D Prior
by: Chen, Cheng, et al.
Published: (2024)
by: Chen, Cheng, et al.
Published: (2024)
MoDA: Modeling Deformable 3D Objects from Casual Videos
by: Song, Chaoyue, et al.
Published: (2023)
by: Song, Chaoyue, et al.
Published: (2023)
Text-to-Image Rectified Flow as Plug-and-Play Priors
by: Yang, Xiaofeng, et al.
Published: (2024)
by: Yang, Xiaofeng, et al.
Published: (2024)
OccLE: Label-Efficient 3D Semantic Occupancy Prediction
by: Fang, Naiyu, et al.
Published: (2025)
by: Fang, Naiyu, et al.
Published: (2025)
VLM-Guided Group Preference Alignment for Diffusion-based Human Mesh Recovery
by: Shen, Wenhao, et al.
Published: (2026)
by: Shen, Wenhao, et al.
Published: (2026)
Puppeteer: Rig and Animate Your 3D Models
by: Song, Chaoyue, et al.
Published: (2025)
by: Song, Chaoyue, et al.
Published: (2025)
Dense Supervision Propagation for Weakly Supervised Semantic Segmentation on 3D Point Clouds
by: Wei, Jiacheng, et al.
Published: (2021)
by: Wei, Jiacheng, et al.
Published: (2021)
PhyCustom: Towards Realistic Physical Customization in Text-to-Image Generation
by: Wu, Fan, et al.
Published: (2025)
by: Wu, Fan, et al.
Published: (2025)
Prim2Room: Layout-Controllable Room Mesh Generation from Primitives
by: Feng, Chengzeng, et al.
Published: (2024)
by: Feng, Chengzeng, et al.
Published: (2024)
ADHMR: Aligning Diffusion-based Human Mesh Recovery via Direct Preference Optimization
by: Shen, Wenhao, et al.
Published: (2025)
by: Shen, Wenhao, et al.
Published: (2025)
In-Context Learning with Unpaired Clips for Instruction-based Video Editing
by: Liao, Xinyao, et al.
Published: (2025)
by: Liao, Xinyao, et al.
Published: (2025)
MagicArticulate: Make Your 3D Models Articulation-Ready
by: Song, Chaoyue, et al.
Published: (2025)
by: Song, Chaoyue, et al.
Published: (2025)
CADCrafter: Generating Computer-Aided Design Models from Unconstrained Images
by: Chen, Cheng, et al.
Published: (2025)
by: Chen, Cheng, et al.
Published: (2025)
SyncTrack4D: Cross-Video Motion Alignment and Video Synchronization for Multi-Video 4D Gaussian Splatting
by: Lee, Yonghan, et al.
Published: (2025)
by: Lee, Yonghan, et al.
Published: (2025)
MVAnimate: Enhancing Character Animation with Multi-View Optimization
by: Sun, Tianyu, et al.
Published: (2026)
by: Sun, Tianyu, et al.
Published: (2026)
DeepVerse: 4D Autoregressive Video Generation as a World Model
by: Chen, Junyi, et al.
Published: (2025)
by: Chen, Junyi, et al.
Published: (2025)
3D-VirtFusion: Synthetic 3D Data Augmentation through Generative Diffusion Models and Controllable Editing
by: Dong, Shichao, et al.
Published: (2024)
by: Dong, Shichao, et al.
Published: (2024)
Exploring Active Learning for Label-Efficient Training of Semantic Neural Radiance Field
by: Zhu, Yuzhe, et al.
Published: (2025)
by: Zhu, Yuzhe, et al.
Published: (2025)
3D Cartoon Face Generation with Controllable Expressions from a Single GAN Image
by: Wang, Hao, et al.
Published: (2022)
by: Wang, Hao, et al.
Published: (2022)
iMontage: Unified, Versatile, Highly Dynamic Many-to-many Image Generation
by: Fu, Zhoujie, et al.
Published: (2025)
by: Fu, Zhoujie, et al.
Published: (2025)
LMM-Track4D: Eliciting 4D Dynamic Reasoning in LMMs via Trajectory-Grounded Dialogue
by: Li, Chaoyue, et al.
Published: (2026)
by: Li, Chaoyue, et al.
Published: (2026)
Integrating SAM Supervision for 3D Weakly Supervised Point Cloud Segmentation
by: You, Lechun, et al.
Published: (2025)
by: You, Lechun, et al.
Published: (2025)
HOI4D: A 4D Egocentric Dataset for Category-Level Human-Object Interaction
by: Liu, Yunze, et al.
Published: (2022)
by: Liu, Yunze, et al.
Published: (2022)
Towards Lifelong Few-Shot Customization of Text-to-Image Diffusion
by: Song, Nan, et al.
Published: (2024)
by: Song, Nan, et al.
Published: (2024)
Gaussian Mixture based Evidential Learning for Stereo Matching
by: Liu, Weide, et al.
Published: (2024)
by: Liu, Weide, et al.
Published: (2024)
IPVTON: Image-based 3D Virtual Try-on with Image Prompt Adapter
by: Zhong, Xiaojing, et al.
Published: (2025)
by: Zhong, Xiaojing, et al.
Published: (2025)
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
by: Song, Chenxi, et al.
Published: (2025)
by: Song, Chenxi, et al.
Published: (2025)
An Efficient 3D Convolutional Neural Network with Channel-wise, Spatial-grouped, and Temporal Convolutions
by: Wang, Zhe, et al.
Published: (2025)
by: Wang, Zhe, et al.
Published: (2025)
Interspatial Attention for Efficient 4D Human Video Generation
by: Shao, Ruizhi, et al.
Published: (2025)
by: Shao, Ruizhi, et al.
Published: (2025)
ViSA: 3D-Aware Video Shading for Real-Time Upper-Body Avatar Creation
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
DeCo: Decoupled Human-Centered Diffusion Video Editing with Motion Consistency
by: Zhong, Xiaojing, et al.
Published: (2024)
by: Zhong, Xiaojing, et al.
Published: (2024)
Pixel-to-4D: Camera-Controlled Image-to-Video Generation with Dynamic 3D Gaussians
by: de Almeida, Melonie, et al.
Published: (2026)
by: de Almeida, Melonie, et al.
Published: (2026)
Few-Shot Segmentation with Global and Local Contrastive Learning
by: Liu, Weide, et al.
Published: (2021)
by: Liu, Weide, et al.
Published: (2021)
GenieDrive: Towards Physics-Aware Driving World Model with 4D Occupancy Guided Video Generation
by: Yang, Zhenya, et al.
Published: (2025)
by: Yang, Zhenya, et al.
Published: (2025)
Genie 4D: Semantic-Prior-Guided 4D Dynamic Scene Reconstruction
by: Yang, Yiru, et al.
Published: (2026)
by: Yang, Yiru, et al.
Published: (2026)
SC4D: Sparse-Controlled Video-to-4D Generation and Motion Transfer
by: Wu, Zijie, et al.
Published: (2024)
by: Wu, Zijie, et al.
Published: (2024)
Efficient4D: Fast Dynamic 3D Object Generation from a Single-view Video
by: Pan, Zijie, et al.
Published: (2024)
by: Pan, Zijie, et al.
Published: (2024)
VerseCrafter: Dynamic Realistic Video World Model with 4D Geometric Control
by: Zheng, Sixiao, et al.
Published: (2026)
by: Zheng, Sixiao, et al.
Published: (2026)
Audio-Sync Video Generation with Multi-Stream Temporal Control
by: Weng, Shuchen, et al.
Published: (2025)
by: Weng, Shuchen, et al.
Published: (2025)
Similar Items
-
REACTO: Reconstructing Articulated Objects from a Single Video
by: Song, Chaoyue, et al.
Published: (2024) -
Sculpt3D: Multi-View Consistent Text-to-3D Generation with Sparse 3D Prior
by: Chen, Cheng, et al.
Published: (2024) -
MoDA: Modeling Deformable 3D Objects from Casual Videos
by: Song, Chaoyue, et al.
Published: (2023) -
Text-to-Image Rectified Flow as Plug-and-Play Priors
by: Yang, Xiaofeng, et al.
Published: (2024) -
OccLE: Label-Efficient 3D Semantic Occupancy Prediction
by: Fang, Naiyu, et al.
Published: (2025)