Vision-based 3D occupancy prediction in autonomous driving: a review and outlook
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yanan, Zhang, Jinqing, Wang, Zengran, Xu, Junhao, Huang, Di |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Breaking the SSL-AL Barrier: A Synergistic Semi-Supervised Active Learning Framework for 3D Object Detection
by: Wang, Zengran, et al.
Published: (2025)
by: Wang, Zengran, et al.
Published: (2025)
Lightweight Spatial Embedding for Vision-based 3D Occupancy Prediction
by: Zhang, Jinqing, et al.
Published: (2024)
by: Zhang, Jinqing, et al.
Published: (2024)
An interactive enhanced driving dataset for autonomous driving
by: Feng, Haojie, et al.
Published: (2026)
by: Feng, Haojie, et al.
Published: (2026)
GeoBEV: Learning Geometric BEV Representation for Multi-view 3D Object Detection
by: Zhang, Jinqing, et al.
Published: (2024)
by: Zhang, Jinqing, et al.
Published: (2024)
CoSDH: Communication-Efficient Collaborative Perception via Supply-Demand Awareness and Intermediate-Late Hybridization
by: Xu, Junhao, et al.
Published: (2025)
by: Xu, Junhao, et al.
Published: (2025)
FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection
by: Jiang, Zheng, et al.
Published: (2024)
by: Jiang, Zheng, et al.
Published: (2024)
OpenTrack3D: Towards Accurate and Generalizable Open-Vocabulary 3D Instance Segmentation
by: Zhou, Zhishan, et al.
Published: (2025)
by: Zhou, Zhishan, et al.
Published: (2025)
STAL3D: Unsupervised Domain Adaptation for 3D Object Detection via Collaborating Self-Training and Adversarial Learning
by: Zhang, Yanan, et al.
Published: (2024)
by: Zhang, Yanan, et al.
Published: (2024)
OccTransformer: Improving BEVFormer for 3D camera-only occupancy prediction
by: Liu, Jian, et al.
Published: (2024)
by: Liu, Jian, et al.
Published: (2024)
On depth prediction for autonomous driving using self-supervised learning
by: Boulahbal, Houssem
Published: (2024)
by: Boulahbal, Houssem
Published: (2024)
Real-time 3D semantic occupancy prediction for autonomous vehicles using memory-efficient sparse convolution
by: Sze, Samuel, et al.
Published: (2024)
by: Sze, Samuel, et al.
Published: (2024)
Open-world machine learning: A review and new outlooks
by: Zhu, Fei, et al.
Published: (2024)
by: Zhu, Fei, et al.
Published: (2024)
Towards Unbiased Source-Free Object Detection via Vision Foundation Models
by: Cai, Zhi, et al.
Published: (2026)
by: Cai, Zhi, et al.
Published: (2026)
Implicit 3D scene reconstruction using deep learning towards efficient collision understanding in autonomous driving
by: Ramanayake, Akarshani, et al.
Published: (2025)
by: Ramanayake, Akarshani, et al.
Published: (2025)
PS-TTL: Prototype-based Soft-labels and Test-Time Learning for Few-shot Object Detection
by: Gao, Yingjie, et al.
Published: (2024)
by: Gao, Yingjie, et al.
Published: (2024)
Polar Parametrization for Vision-based Surround-View 3D Detection
by: Chen, Shaoyu, et al.
Published: (2022)
by: Chen, Shaoyu, et al.
Published: (2022)
TGP: Two-modal occupancy prediction with 3D Gaussian and sparse points for 3D Environment Awareness
by: Chen, Mu, et al.
Published: (2025)
by: Chen, Mu, et al.
Published: (2025)
Learning autonomous driving from aerial imagery
by: Murali, Varun, et al.
Published: (2024)
by: Murali, Varun, et al.
Published: (2024)
Timealign: A multi-modal object detection method for time misalignment fusing in autonomous driving
by: Song, Zhihang, et al.
Published: (2024)
by: Song, Zhihang, et al.
Published: (2024)
A re-calibration method for object detection with multi-modal alignment bias in autonomous driving
by: Song, Zhihang, et al.
Published: (2024)
by: Song, Zhihang, et al.
Published: (2024)
Towards learning-based planning:The nuPlan benchmark for real-world autonomous driving
by: Karnchanachari, Napat, et al.
Published: (2024)
by: Karnchanachari, Napat, et al.
Published: (2024)
SGV3D:Towards Scenario Generalization for Vision-based Roadside 3D Object Detection
by: Yang, Lei, et al.
Published: (2024)
by: Yang, Lei, et al.
Published: (2024)
DEGS: Deformable Event-based 3D Gaussian Splatting from RGB and Event Stream
by: He, Junhao, et al.
Published: (2025)
by: He, Junhao, et al.
Published: (2025)
ResWorld: Temporal Residual World Model for End-to-End Autonomous Driving
by: Zhang, Jinqing, et al.
Published: (2026)
by: Zhang, Jinqing, et al.
Published: (2026)
Test-Time Adaptive Object Detection with Foundation Model
by: Gao, Yingjie, et al.
Published: (2025)
by: Gao, Yingjie, et al.
Published: (2025)
Centerness-based Instance-aware Knowledge Distillation with Task-wise Mutual Lifting for Object Detection on Drone Imagery
by: Du, Bowei, et al.
Published: (2024)
by: Du, Bowei, et al.
Published: (2024)
KAN-RCBEVDepth: A multi-modal fusion algorithm in object detection for autonomous driving
by: Lai, Zhihao, et al.
Published: (2024)
by: Lai, Zhihao, et al.
Published: (2024)
Replacement Learning: Training Vision Tasks with Fewer Learnable Parameters
by: Zhang, Yuming, et al.
Published: (2024)
by: Zhang, Yuming, et al.
Published: (2024)
Intelligent driving vehicle front multi-target tracking and detection based on YOLOv5 and point cloud 3D projection
by: Liu, Dayong, et al.
Published: (2025)
by: Liu, Dayong, et al.
Published: (2025)
Vision Mamba-based autonomous crack segmentation on concrete, asphalt, and masonry surfaces
by: Chen, Zhaohui, et al.
Published: (2024)
by: Chen, Zhaohui, et al.
Published: (2024)
COLA: COarse-LAbel multi-source LiDAR semantic segmentation for autonomous driving
by: Sanchez, Jules, et al.
Published: (2023)
by: Sanchez, Jules, et al.
Published: (2023)
N3D-VLM: Native 3D Grounding Enables Accurate Spatial Reasoning in Vision-Language Models
by: Wang, Yuxin, et al.
Published: (2025)
by: Wang, Yuxin, et al.
Published: (2025)
Towards Training-free Anomaly Detection with Vision and Language Foundation Models
by: Zhang, Jinjin, et al.
Published: (2025)
by: Zhang, Jinjin, et al.
Published: (2025)
From Flatland to Space: Teaching Vision-Language Models to Perceive and Reason in 3D
by: Zhang, Jiahui, et al.
Published: (2025)
by: Zhang, Jiahui, et al.
Published: (2025)
RenderOcc: Vision-Centric 3D Occupancy Prediction with 2D Rendering Supervision
by: Pan, Mingjie, et al.
Published: (2023)
by: Pan, Mingjie, et al.
Published: (2023)
3D Weakly Supervised Semantic Segmentation with 2D Vision-Language Guidance
by: Xu, Xiaoxu, et al.
Published: (2024)
by: Xu, Xiaoxu, et al.
Published: (2024)
HCC-3D: Hierarchical Compensatory Compression for 98% 3D Token Reduction in Vision-Language Models
by: Zhang, Liheng, et al.
Published: (2025)
by: Zhang, Liheng, et al.
Published: (2025)
TALO: Pushing 3D Vision Foundation Models Towards Globally Consistent Online Reconstruction
by: Zhang, Fengyi, et al.
Published: (2025)
by: Zhang, Fengyi, et al.
Published: (2025)
Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation
by: Zhuge, Yunzhi, et al.
Published: (2025)
by: Zhuge, Yunzhi, et al.
Published: (2025)
Are NeRFs ready for autonomous driving? Towards closing the real-to-simulation gap
by: Lindström, Carl, et al.
Published: (2024)
by: Lindström, Carl, et al.
Published: (2024)
Similar Items
-
Breaking the SSL-AL Barrier: A Synergistic Semi-Supervised Active Learning Framework for 3D Object Detection
by: Wang, Zengran, et al.
Published: (2025) -
Lightweight Spatial Embedding for Vision-based 3D Occupancy Prediction
by: Zhang, Jinqing, et al.
Published: (2024) -
An interactive enhanced driving dataset for autonomous driving
by: Feng, Haojie, et al.
Published: (2026) -
GeoBEV: Learning Geometric BEV Representation for Multi-view 3D Object Detection
by: Zhang, Jinqing, et al.
Published: (2024) -
CoSDH: Communication-Efficient Collaborative Perception via Supply-Demand Awareness and Intermediate-Late Hybridization
by: Xu, Junhao, et al.
Published: (2025)