Saved in:
| Main Authors: | Wang, Pan, Liu, Yang, Wu, Guile, Corral-Soto, Eduardo R., Huang, Chengjie, Xu, Binbin, Bai, Dongfeng, Yan, Xu, Ren, Yuan, Chen, Xingxin, Wu, Yizhe, Huang, Tao, Wan, Wenjun, Wu, Xin, Zhou, Pei, Dai, Xuyang, Lv, Kangbo, Zhang, Hongbo, Fried, Yosef, Ye, Aixue, Feng, Bailan, Chen, Zhenyu, Li, Zhen, Chen, Yingcong, Liao, Yiyi, Liu, Bingbing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.00092 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Nighttime Autonomous Driving Scene Reconstruction with Physically-Based Gaussian Splatting
by: Kim, Tae-Kyeong, et al.
Published: (2026)
by: Kim, Tae-Kyeong, et al.
Published: (2026)
TurboVGGT: Fast Visual Geometry Reconstruction with Adaptive Alternating Attention
by: Huang, David, et al.
Published: (2026)
by: Huang, David, et al.
Published: (2026)
ArmGS: Composite Gaussian Appearance Refinement for Modeling Dynamic Urban Environments
by: Wu, Guile, et al.
Published: (2025)
by: Wu, Guile, et al.
Published: (2025)
MoVieDrive: Urban Scene Synthesis with Multi-Modal Multi-View Video Diffusion Transformer
by: Wu, Guile, et al.
Published: (2025)
by: Wu, Guile, et al.
Published: (2025)
Language and Geometry Grounded Sparse Voxel Representations for Holistic Scene Understanding
by: Wu, Guile, et al.
Published: (2026)
by: Wu, Guile, et al.
Published: (2026)
UniGaussian: Driving Scene Reconstruction from Multiple Camera Models via Unified Gaussian Representations
by: Ren, Yuan, et al.
Published: (2024)
by: Ren, Yuan, et al.
Published: (2024)
EVolSplat4D: Efficient Volume-based Gaussian Splatting for 4D Urban Scene Synthesis
by: Miao, Sheng, et al.
Published: (2026)
by: Miao, Sheng, et al.
Published: (2026)
HIPPo: Harnessing Image-to-3D Priors for Model-free Zero-shot 6D Pose Estimation
by: Liu, Yibo, et al.
Published: (2025)
by: Liu, Yibo, et al.
Published: (2025)
UniScale: Unified Scale-Aware 3D Reconstruction for Multi-View Understanding via Prior Injection for Robotic Perception
by: Mahdavian, Mohammad, et al.
Published: (2026)
by: Mahdavian, Mohammad, et al.
Published: (2026)
Learning Effective NeRFs and SDFs Representations with 3D Generative Adversarial Networks for 3D Object Generation
by: Yang, Zheyuan, et al.
Published: (2023)
by: Yang, Zheyuan, et al.
Published: (2023)
Monocular Visual 8D Pose Estimation for Articulated Bicycles and Cyclists
by: Corral-Soto, Eduardo R., et al.
Published: (2025)
by: Corral-Soto, Eduardo R., et al.
Published: (2025)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
by: Xu, Tianshuo, et al.
Published: (2024)
by: Xu, Tianshuo, et al.
Published: (2024)
Spatial Lifting for Dense Prediction
by: Xu, Mingzhi, et al.
Published: (2025)
by: Xu, Mingzhi, et al.
Published: (2025)
Versatile Video Tokenization with Generative 2D Gaussian Splatting
by: Chen, Zhenghao, et al.
Published: (2025)
by: Chen, Zhenghao, et al.
Published: (2025)
HUGS: Holistic Urban 3D Scene Understanding via Gaussian Splatting
by: Zhou, Hongyu, et al.
Published: (2024)
by: Zhou, Hongyu, et al.
Published: (2024)
FreeFix: Boosting 3D Gaussian Splatting via Fine-Tuning-Free Diffusion Models
by: Zhou, Hongyu, et al.
Published: (2026)
by: Zhou, Hongyu, et al.
Published: (2026)
4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding
by: Chen, Zhangquan, et al.
Published: (2026)
by: Chen, Zhangquan, et al.
Published: (2026)
VQA-Diff: Exploiting VQA and Diffusion for Zero-Shot Image-to-3D Vehicle Asset Generation in Autonomous Driving
by: Liu, Yibo, et al.
Published: (2024)
by: Liu, Yibo, et al.
Published: (2024)
Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models
by: Xu, Tianshuo, et al.
Published: (2025)
by: Xu, Tianshuo, et al.
Published: (2025)
Efficient 3D Perception on Multi-Sweep Point Cloud with Gumbel Spatial Pruning
by: Sun, Tianyu, et al.
Published: (2024)
by: Sun, Tianyu, et al.
Published: (2024)
Sonic4D: Spatial Audio Generation for Immersive 4D Scene Exploration
by: Xie, Siyi, et al.
Published: (2025)
by: Xie, Siyi, et al.
Published: (2025)
SA-LUT: Spatial Adaptive 4D Look-Up Table for Photorealistic Style Transfer
by: Gong, Zerui, et al.
Published: (2025)
by: Gong, Zerui, et al.
Published: (2025)
Efficient Depth-Guided Urban View Synthesis
by: Miao, Sheng, et al.
Published: (2024)
by: Miao, Sheng, et al.
Published: (2024)
Advancing high-fidelity 3D and Texture Generation with 2.5D latents
by: Yang, Xin, et al.
Published: (2025)
by: Yang, Xin, et al.
Published: (2025)
Free4D: Tuning-free 4D Scene Generation with Spatial-Temporal Consistency
by: Liu, Tianqi, et al.
Published: (2025)
by: Liu, Tianqi, et al.
Published: (2025)
Reconstructing 4D Spatial Intelligence: A Survey
by: Cao, Yukang, et al.
Published: (2025)
by: Cao, Yukang, et al.
Published: (2025)
EndoWave: Rational-Wavelet 4D Gaussian Splatting for Endoscopic Reconstruction
by: Wu, Taoyu, et al.
Published: (2025)
by: Wu, Taoyu, et al.
Published: (2025)
SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence
by: Wu, Haoning, et al.
Published: (2025)
by: Wu, Haoning, et al.
Published: (2025)
EVolSplat: Efficient Volume-based Gaussian Splatting for Urban View Synthesis
by: Miao, Sheng, et al.
Published: (2025)
by: Miao, Sheng, et al.
Published: (2025)
Learning to Reason in 4D: Dynamic Spatial Understanding for Vision Language Models
by: Zhou, Shengchao, et al.
Published: (2025)
by: Zhou, Shengchao, et al.
Published: (2025)
TIBR4D: Tracing-Guided Iterative Boundary Refinement for Efficient 4D Gaussian Segmentation
by: Wu, He, et al.
Published: (2026)
by: Wu, He, et al.
Published: (2026)
CF-Nil systems and convergence of two-dimensional ergodic averages
by: Ouyang, Kangbo, et al.
Published: (2025)
by: Ouyang, Kangbo, et al.
Published: (2025)
Part-Level 3D Gaussian Vehicle Generation with Joint and Hinge Axis Estimation
by: Qian, Shiyao, et al.
Published: (2026)
by: Qian, Shiyao, et al.
Published: (2026)
GenFusion: Closing the Loop between Reconstruction and Generation via Videos
by: Wu, Sibo, et al.
Published: (2025)
by: Wu, Sibo, et al.
Published: (2025)
Orthogonal Spatial-temporal Distributional Transfer for 4D Generation
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
Vivid4D: Improving 4D Reconstruction from Monocular Video by Video Inpainting
by: Huang, Jiaxin, et al.
Published: (2025)
by: Huang, Jiaxin, et al.
Published: (2025)
Spatially-Weighted CLIP for Street-View Geo-localization
by: Han, Ting, et al.
Published: (2026)
by: Han, Ting, et al.
Published: (2026)
D$^2$GSLAM: 4D Dynamic Gaussian Splatting SLAM
by: Zhu, Siting, et al.
Published: (2025)
by: Zhu, Siting, et al.
Published: (2025)
DrivingRecon: Large 4D Gaussian Reconstruction Model For Autonomous Driving
by: Lu, Hao, et al.
Published: (2024)
by: Lu, Hao, et al.
Published: (2024)
4DGen: Grounded 4D Content Generation with Spatial-temporal Consistency
by: Yin, Yuyang, et al.
Published: (2023)
by: Yin, Yuyang, et al.
Published: (2023)
Similar Items
-
Nighttime Autonomous Driving Scene Reconstruction with Physically-Based Gaussian Splatting
by: Kim, Tae-Kyeong, et al.
Published: (2026) -
TurboVGGT: Fast Visual Geometry Reconstruction with Adaptive Alternating Attention
by: Huang, David, et al.
Published: (2026) -
ArmGS: Composite Gaussian Appearance Refinement for Modeling Dynamic Urban Environments
by: Wu, Guile, et al.
Published: (2025) -
MoVieDrive: Urban Scene Synthesis with Multi-Modal Multi-View Video Diffusion Transformer
by: Wu, Guile, et al.
Published: (2025) -
Language and Geometry Grounded Sparse Voxel Representations for Holistic Scene Understanding
by: Wu, Guile, et al.
Published: (2026)