Spatial4D-Bench: A Versatile 4D Spatial Intelligence Benchmark
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Pan, Liu, Yang, Wu, Guile, Corral-Soto, Eduardo R., Huang, Chengjie, Xu, Binbin, Bai, Dongfeng, Yan, Xu, Ren, Yuan, Chen, Xingxin, Wu, Yizhe, Huang, Tao, Wan, Wenjun, Wu, Xin, Zhou, Pei, Dai, Xuyang, Lv, Kangbo, Zhang, Hongbo, Fried, Yosef, Ye, Aixue, Feng, Bailan, Chen, Zhenyu, Li, Zhen, Chen, Yingcong, Liao, Yiyi, Liu, Bingbing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Nighttime Autonomous Driving Scene Reconstruction with Physically-Based Gaussian Splatting
por: Kim, Tae-Kyeong, et al.
Publicado: (2026)
por: Kim, Tae-Kyeong, et al.
Publicado: (2026)
TurboVGGT: Fast Visual Geometry Reconstruction with Adaptive Alternating Attention
por: Huang, David, et al.
Publicado: (2026)
por: Huang, David, et al.
Publicado: (2026)
ArmGS: Composite Gaussian Appearance Refinement for Modeling Dynamic Urban Environments
por: Wu, Guile, et al.
Publicado: (2025)
por: Wu, Guile, et al.
Publicado: (2025)
MoVieDrive: Urban Scene Synthesis with Multi-Modal Multi-View Video Diffusion Transformer
por: Wu, Guile, et al.
Publicado: (2025)
por: Wu, Guile, et al.
Publicado: (2025)
Language and Geometry Grounded Sparse Voxel Representations for Holistic Scene Understanding
por: Wu, Guile, et al.
Publicado: (2026)
por: Wu, Guile, et al.
Publicado: (2026)
UniGaussian: Driving Scene Reconstruction from Multiple Camera Models via Unified Gaussian Representations
por: Ren, Yuan, et al.
Publicado: (2024)
por: Ren, Yuan, et al.
Publicado: (2024)
EVolSplat4D: Efficient Volume-based Gaussian Splatting for 4D Urban Scene Synthesis
por: Miao, Sheng, et al.
Publicado: (2026)
por: Miao, Sheng, et al.
Publicado: (2026)
HIPPo: Harnessing Image-to-3D Priors for Model-free Zero-shot 6D Pose Estimation
por: Liu, Yibo, et al.
Publicado: (2025)
por: Liu, Yibo, et al.
Publicado: (2025)
UniScale: Unified Scale-Aware 3D Reconstruction for Multi-View Understanding via Prior Injection for Robotic Perception
por: Mahdavian, Mohammad, et al.
Publicado: (2026)
por: Mahdavian, Mohammad, et al.
Publicado: (2026)
Learning Effective NeRFs and SDFs Representations with 3D Generative Adversarial Networks for 3D Object Generation
por: Yang, Zheyuan, et al.
Publicado: (2023)
por: Yang, Zheyuan, et al.
Publicado: (2023)
Versatile Video Tokenization with Generative 2D Gaussian Splatting
por: Chen, Zhenghao, et al.
Publicado: (2025)
por: Chen, Zhenghao, et al.
Publicado: (2025)
Spatial Lifting for Dense Prediction
por: Xu, Mingzhi, et al.
Publicado: (2025)
por: Xu, Mingzhi, et al.
Publicado: (2025)
4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding
por: Chen, Zhangquan, et al.
Publicado: (2026)
por: Chen, Zhangquan, et al.
Publicado: (2026)
Sonic4D: Spatial Audio Generation for Immersive 4D Scene Exploration
por: Xie, Siyi, et al.
Publicado: (2025)
por: Xie, Siyi, et al.
Publicado: (2025)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
por: Xu, Tianshuo, et al.
Publicado: (2024)
por: Xu, Tianshuo, et al.
Publicado: (2024)
Monocular Visual 8D Pose Estimation for Articulated Bicycles and Cyclists
por: Corral-Soto, Eduardo R., et al.
Publicado: (2025)
por: Corral-Soto, Eduardo R., et al.
Publicado: (2025)
SA-LUT: Spatial Adaptive 4D Look-Up Table for Photorealistic Style Transfer
por: Gong, Zerui, et al.
Publicado: (2025)
por: Gong, Zerui, et al.
Publicado: (2025)
Free4D: Tuning-free 4D Scene Generation with Spatial-Temporal Consistency
por: Liu, Tianqi, et al.
Publicado: (2025)
por: Liu, Tianqi, et al.
Publicado: (2025)
Reconstructing 4D Spatial Intelligence: A Survey
por: Cao, Yukang, et al.
Publicado: (2025)
por: Cao, Yukang, et al.
Publicado: (2025)
Efficient 3D Perception on Multi-Sweep Point Cloud with Gumbel Spatial Pruning
por: Sun, Tianyu, et al.
Publicado: (2024)
por: Sun, Tianyu, et al.
Publicado: (2024)
HUGS: Holistic Urban 3D Scene Understanding via Gaussian Splatting
por: Zhou, Hongyu, et al.
Publicado: (2024)
por: Zhou, Hongyu, et al.
Publicado: (2024)
Learning to Reason in 4D: Dynamic Spatial Understanding for Vision Language Models
por: Zhou, Shengchao, et al.
Publicado: (2025)
por: Zhou, Shengchao, et al.
Publicado: (2025)
SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence
por: Wu, Haoning, et al.
Publicado: (2025)
por: Wu, Haoning, et al.
Publicado: (2025)
FreeFix: Boosting 3D Gaussian Splatting via Fine-Tuning-Free Diffusion Models
por: Zhou, Hongyu, et al.
Publicado: (2026)
por: Zhou, Hongyu, et al.
Publicado: (2026)
TIBR4D: Tracing-Guided Iterative Boundary Refinement for Efficient 4D Gaussian Segmentation
por: Wu, He, et al.
Publicado: (2026)
por: Wu, He, et al.
Publicado: (2026)
EndoWave: Rational-Wavelet 4D Gaussian Splatting for Endoscopic Reconstruction
por: Wu, Taoyu, et al.
Publicado: (2025)
por: Wu, Taoyu, et al.
Publicado: (2025)
Advancing high-fidelity 3D and Texture Generation with 2.5D latents
por: Yang, Xin, et al.
Publicado: (2025)
por: Yang, Xin, et al.
Publicado: (2025)
VQA-Diff: Exploiting VQA and Diffusion for Zero-Shot Image-to-3D Vehicle Asset Generation in Autonomous Driving
por: Liu, Yibo, et al.
Publicado: (2024)
por: Liu, Yibo, et al.
Publicado: (2024)
Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models
por: Xu, Tianshuo, et al.
Publicado: (2025)
por: Xu, Tianshuo, et al.
Publicado: (2025)
Orthogonal Spatial-temporal Distributional Transfer for 4D Generation
por: Liu, Wei, et al.
Publicado: (2026)
por: Liu, Wei, et al.
Publicado: (2026)
D$^2$GSLAM: 4D Dynamic Gaussian Splatting SLAM
por: Zhu, Siting, et al.
Publicado: (2025)
por: Zhu, Siting, et al.
Publicado: (2025)
Efficient Depth-Guided Urban View Synthesis
por: Miao, Sheng, et al.
Publicado: (2024)
por: Miao, Sheng, et al.
Publicado: (2024)
4DGen: Grounded 4D Content Generation with Spatial-temporal Consistency
por: Yin, Yuyang, et al.
Publicado: (2023)
por: Yin, Yuyang, et al.
Publicado: (2023)
Vivid4D: Improving 4D Reconstruction from Monocular Video by Video Inpainting
por: Huang, Jiaxin, et al.
Publicado: (2025)
por: Huang, Jiaxin, et al.
Publicado: (2025)
Spatially-Weighted CLIP for Street-View Geo-localization
por: Han, Ting, et al.
Publicado: (2026)
por: Han, Ting, et al.
Publicado: (2026)
CF-Nil systems and convergence of two-dimensional ergodic averages
por: Ouyang, Kangbo, et al.
Publicado: (2025)
por: Ouyang, Kangbo, et al.
Publicado: (2025)
GenFusion: Closing the Loop between Reconstruction and Generation via Videos
por: Wu, Sibo, et al.
Publicado: (2025)
por: Wu, Sibo, et al.
Publicado: (2025)
Motion 3-to-4: 3D Motion Reconstruction for 4D Synthesis
por: Chen, Hongyuan, et al.
Publicado: (2026)
por: Chen, Hongyuan, et al.
Publicado: (2026)
Part-Level 3D Gaussian Vehicle Generation with Joint and Hinge Axis Estimation
por: Qian, Shiyao, et al.
Publicado: (2026)
por: Qian, Shiyao, et al.
Publicado: (2026)
Light4D: Training-Free Extreme Viewpoint 4D Video Relighting
por: Wu, Zhenghuang, et al.
Publicado: (2026)
por: Wu, Zhenghuang, et al.
Publicado: (2026)
Ejemplares similares
-
Nighttime Autonomous Driving Scene Reconstruction with Physically-Based Gaussian Splatting
por: Kim, Tae-Kyeong, et al.
Publicado: (2026) -
TurboVGGT: Fast Visual Geometry Reconstruction with Adaptive Alternating Attention
por: Huang, David, et al.
Publicado: (2026) -
ArmGS: Composite Gaussian Appearance Refinement for Modeling Dynamic Urban Environments
por: Wu, Guile, et al.
Publicado: (2025) -
MoVieDrive: Urban Scene Synthesis with Multi-Modal Multi-View Video Diffusion Transformer
por: Wu, Guile, et al.
Publicado: (2025) -
Language and Geometry Grounded Sparse Voxel Representations for Holistic Scene Understanding
por: Wu, Guile, et al.
Publicado: (2026)