InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Yifan, Ren, Xuanchi, Yang, Jiawei, Shen, Tianchang, Wu, Zhangjie, Gao, Jun, Wang, Yue, Chen, Siheng, Chen, Mike, Fidler, Sanja, Huang, Jiahui |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SCube: Instant Large-Scale Scene Reconstruction using VoxSplats
by: Ren, Xuanchi, et al.
Published: (2024)
by: Ren, Xuanchi, et al.
Published: (2024)
GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control
by: Ren, Xuanchi, et al.
Published: (2025)
by: Ren, Xuanchi, et al.
Published: (2025)
Lyra: Generative 3D Scene Reconstruction via Video Diffusion Model Self-Distillation
by: Bahmani, Sherwin, et al.
Published: (2025)
by: Bahmani, Sherwin, et al.
Published: (2025)
MoRight: Motion Control Done Right
by: Liu, Shaowei, et al.
Published: (2026)
by: Liu, Shaowei, et al.
Published: (2026)
XCube: Large-Scale 3D Generative Modeling using Sparse Voxel Hierarchies
by: Ren, Xuanchi, et al.
Published: (2023)
by: Ren, Xuanchi, et al.
Published: (2023)
ViPE: Video Pose Engine for 3D Geometric Perception
by: Huang, Jiahui, et al.
Published: (2025)
by: Huang, Jiahui, et al.
Published: (2025)
Feed-Forward Bullet-Time Reconstruction of Dynamic Scenes from Monocular Videos
by: Liang, Hanxue, et al.
Published: (2024)
by: Liang, Hanxue, et al.
Published: (2024)
SpaceMesh: A Continuous Representation for Learning Manifold Surface Meshes
by: Shen, Tianchang, et al.
Published: (2024)
by: Shen, Tianchang, et al.
Published: (2024)
VideoPanda: Video Panoramic Diffusion with Multi-view Attention
by: Xie, Kevin, et al.
Published: (2025)
by: Xie, Kevin, et al.
Published: (2025)
ArtiFixer: Enhancing and Extending 3D Reconstruction with Auto-Regressive Diffusion Models
by: de Lutio, Riccardo, et al.
Published: (2026)
by: de Lutio, Riccardo, et al.
Published: (2026)
fVDB: A Deep-Learning Framework for Sparse, Large-Scale, and High-Performance Spatial Intelligence
by: Williams, Francis, et al.
Published: (2024)
by: Williams, Francis, et al.
Published: (2024)
VideoMat: Extracting PBR Materials from Video Diffusion Models
by: Munkberg, Jacob, et al.
Published: (2025)
by: Munkberg, Jacob, et al.
Published: (2025)
Asset Harvester: Extracting 3D Assets from Autonomous Driving Logs for Simulation
by: Cao, Tianshi, et al.
Published: (2026)
by: Cao, Tianshi, et al.
Published: (2026)
Controllable Weather Synthesis and Removal with Video Diffusion Models
by: Lin, Chih-Hao, et al.
Published: (2025)
by: Lin, Chih-Hao, et al.
Published: (2025)
Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models
by: Ren, Xuanchi, et al.
Published: (2025)
by: Ren, Xuanchi, et al.
Published: (2025)
3D Gaussian Ray Tracing: Fast Tracing of Particle Scenes
by: Moenne-Loccoz, Nicolas, et al.
Published: (2024)
by: Moenne-Loccoz, Nicolas, et al.
Published: (2024)
LuxDiT: Lighting Estimation with Video Diffusion Transformer
by: Liang, Ruofan, et al.
Published: (2025)
by: Liang, Ruofan, et al.
Published: (2025)
Photorealistic Object Insertion with Diffusion-Guided Inverse Rendering
by: Liang, Ruofan, et al.
Published: (2024)
by: Liang, Ruofan, et al.
Published: (2024)
SuperPADL: Scaling Language-Directed Physics-Based Control with Progressive Supervised Distillation
by: Juravsky, Jordan, et al.
Published: (2024)
by: Juravsky, Jordan, et al.
Published: (2024)
ICP-3DGS: SfM-free 3D Gaussian Splatting for Large-scale Unbounded Scenes
by: Zhang, Chenhao, et al.
Published: (2025)
by: Zhang, Chenhao, et al.
Published: (2025)
NeRF-XL: Scaling NeRFs with Multiple GPUs
by: Li, Ruilong, et al.
Published: (2024)
by: Li, Ruilong, et al.
Published: (2024)
DiffusionRenderer: Neural Inverse and Forward Rendering with Video Diffusion Models
by: Liang, Ruofan, et al.
Published: (2025)
by: Liang, Ruofan, et al.
Published: (2025)
PaMoSplat: Part-Aware Motion-Guided Gaussian Splatting for Dynamic Scene Reconstruction
by: Deng, Yinan, et al.
Published: (2026)
by: Deng, Yinan, et al.
Published: (2026)
Glad: A Streaming Scene Generator for Autonomous Driving
by: Xie, Bin, et al.
Published: (2025)
by: Xie, Bin, et al.
Published: (2025)
Neurally Integrated Finite Elements for Differentiable Elasticity on Evolving Domains
by: Daviet, Gilles, et al.
Published: (2024)
by: Daviet, Gilles, et al.
Published: (2024)
Static Scene Reconstruction from Dynamic Egocentric Videos
by: Cui, Qifei, et al.
Published: (2026)
by: Cui, Qifei, et al.
Published: (2026)
VideoFrom3D: 3D Scene Video Generation via Complementary Image and Video Diffusion Models
by: Kim, Geonung, et al.
Published: (2025)
by: Kim, Geonung, et al.
Published: (2025)
HOLODECK 2.0: Vision-Language-Guided 3D World Generation with Editing
by: Bian, Zixuan, et al.
Published: (2025)
by: Bian, Zixuan, et al.
Published: (2025)
Text‐Guided Interactive Scene Synthesis with Scene Prior Guidance
by: Shaoheng Fang, et al.
Published: (2025)
by: Shaoheng Fang, et al.
Published: (2025)
CLEAR-NeRF: Collinearity and Local-region Enhanced Accurate 3D Reconstruction in Unbounded Scenes
by: Polianskii, Vladislav, et al.
Published: (2026)
by: Polianskii, Vladislav, et al.
Published: (2026)
Online Photon Guiding with 3D Gaussians for Caustics Rendering
by: Huang, Jiawei, et al.
Published: (2024)
by: Huang, Jiawei, et al.
Published: (2024)
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players
by: Liu, Fangfu, et al.
Published: (2026)
by: Liu, Fangfu, et al.
Published: (2026)
DualPhys-GS: Dual Physically-Guided 3D Gaussian Splatting for Underwater Scene Reconstruction
by: Li, Jiachen, et al.
Published: (2025)
by: Li, Jiachen, et al.
Published: (2025)
WorldFlow3D: Flowing Through 3D Distributions for Unbounded World Generation
by: Joshi, Amogh, et al.
Published: (2026)
by: Joshi, Amogh, et al.
Published: (2026)
Causal Reasoning Elicits Controllable 3D Scene Generation
by: Chen, Shen, et al.
Published: (2025)
by: Chen, Shen, et al.
Published: (2025)
Beyond Static Scenes: Camera-controllable Background Generation for Human Motion
by: Yao, Mingshuai, et al.
Published: (2025)
by: Yao, Mingshuai, et al.
Published: (2025)
DreamDrive: Generative 4D Scene Modeling from Street View Images
by: Mao, Jiageng, et al.
Published: (2024)
by: Mao, Jiageng, et al.
Published: (2024)
PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion
by: Lu, Yifan, et al.
Published: (2026)
by: Lu, Yifan, et al.
Published: (2026)
Sketch-Guided Scene Image Generation
by: Zhang, Tianyu, et al.
Published: (2024)
by: Zhang, Tianyu, et al.
Published: (2024)
Landmark Guided 4D Facial Expression Generation
by: Lu, Xin, et al.
Published: (2026)
by: Lu, Xin, et al.
Published: (2026)
Similar Items
-
SCube: Instant Large-Scale Scene Reconstruction using VoxSplats
by: Ren, Xuanchi, et al.
Published: (2024) -
GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control
by: Ren, Xuanchi, et al.
Published: (2025) -
Lyra: Generative 3D Scene Reconstruction via Video Diffusion Model Self-Distillation
by: Bahmani, Sherwin, et al.
Published: (2025) -
MoRight: Motion Control Done Right
by: Liu, Shaowei, et al.
Published: (2026) -
XCube: Large-Scale 3D Generative Modeling using Sparse Voxel Hierarchies
by: Ren, Xuanchi, et al.
Published: (2023)