OmniScene: Attention-Augmented Multimodal 4D Scene Understanding for Autonomous Driving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Pei, Lu, Hongliang, Liu, Haichao, Liu, Haipeng, Liu, Xin, Yao, Ruoyu, Li, Shengbo Eben, Ma, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VLM-E2E: Enhancing End-to-End Autonomous Driving with Multimodal Driver Attention Fusion
von: Liu, Pei, et al.
Veröffentlicht: (2025)
von: Liu, Pei, et al.
Veröffentlicht: (2025)
OmniHD-Scenes: A Next-Generation Multimodal Dataset for Autonomous Driving
von: Zheng, Lianqing, et al.
Veröffentlicht: (2024)
von: Zheng, Lianqing, et al.
Veröffentlicht: (2024)
FSF-Net: Enhance 4D Occupancy Forecasting with Coarse BEV Scene Flow for Autonomous Driving
von: Guo, Erxin, et al.
Veröffentlicht: (2024)
von: Guo, Erxin, et al.
Veröffentlicht: (2024)
Omni-Scene: Omni-Gaussian Representation for Ego-Centric Sparse-View Scene Reconstruction
von: Wei, Dongxu, et al.
Veröffentlicht: (2024)
von: Wei, Dongxu, et al.
Veröffentlicht: (2024)
LMMCoDrive: Cooperative Driving with Large Multimodal Model
von: Liu, Haichao, et al.
Veröffentlicht: (2024)
von: Liu, Haichao, et al.
Veröffentlicht: (2024)
CALMM-Drive: Confidence-Aware Autonomous Driving with Large Multimodal Model
von: Yao, Ruoyu, et al.
Veröffentlicht: (2024)
von: Yao, Ruoyu, et al.
Veröffentlicht: (2024)
Multi-Modal Data-Efficient 3D Scene Understanding for Autonomous Driving
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving
von: Shi, Chen, et al.
Veröffentlicht: (2025)
von: Shi, Chen, et al.
Veröffentlicht: (2025)
SparseWorld: Enhancing End-to-End Autonomous Driving via World Models with Sparse Scene Representation
von: Wang, Ruoyu, et al.
Veröffentlicht: (2026)
von: Wang, Ruoyu, et al.
Veröffentlicht: (2026)
CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving
von: Liu, Pei, et al.
Veröffentlicht: (2025)
von: Liu, Pei, et al.
Veröffentlicht: (2025)
Scene-Aware Explainable Multimodal Trajectory Prediction
von: Liu, Pei, et al.
Veröffentlicht: (2024)
von: Liu, Pei, et al.
Veröffentlicht: (2024)
VEOcc: Voxel-Centric Online Semantic Occupancy Prediction For Embodied Scene Understanding
von: Wang, Ruoyu, et al.
Veröffentlicht: (2026)
von: Wang, Ruoyu, et al.
Veröffentlicht: (2026)
Information Coordination as a Bridge: A Neuro-Symbolic Architecture for Reliable Autonomous Driving Scene Understanding
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
DeFlow: Decoder of Scene Flow Network in Autonomous Driving
von: Zhang, Qingwen, et al.
Veröffentlicht: (2024)
von: Zhang, Qingwen, et al.
Veröffentlicht: (2024)
T2SG: Traffic Topology Scene Graph for Topology Reasoning in Autonomous Driving
von: Lv, Changsheng, et al.
Veröffentlicht: (2024)
von: Lv, Changsheng, et al.
Veröffentlicht: (2024)
InstDrive: Instance-Aware 3D Gaussian Splatting for Driving Scenes
von: Liu, Hongyuan, et al.
Veröffentlicht: (2025)
von: Liu, Hongyuan, et al.
Veröffentlicht: (2025)
DriveWorld: 4D Pre-trained Scene Understanding via World Models for Autonomous Driving
von: Min, Chen, et al.
Veröffentlicht: (2024)
von: Min, Chen, et al.
Veröffentlicht: (2024)
GarchingSim: An Autonomous Driving Simulator with Photorealistic Scenes and Minimalist Workflow
von: Zhou, Liguo, et al.
Veröffentlicht: (2024)
von: Zhou, Liguo, et al.
Veröffentlicht: (2024)
DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
MGNet: Monocular Geometric Scene Understanding for Autonomous Driving
von: Schön, Markus, et al.
Veröffentlicht: (2022)
von: Schön, Markus, et al.
Veröffentlicht: (2022)
PreGSU-A Generalized Traffic Scene Understanding Model for Autonomous Driving based on Pre-trained Graph Attention Network
von: Wang, Yuning, et al.
Veröffentlicht: (2024)
von: Wang, Yuning, et al.
Veröffentlicht: (2024)
Text-Scene: A Scene-to-Language Parsing Framework for 3D Scene Understanding
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
Editable Scene Simulation for Autonomous Driving via Collaborative LLM-Agents
von: Wei, Yuxi, et al.
Veröffentlicht: (2024)
von: Wei, Yuxi, et al.
Veröffentlicht: (2024)
Semantic Gaussians: Open-Vocabulary Scene Understanding with 3D Gaussian Splatting
von: Guo, Jun, et al.
Veröffentlicht: (2024)
von: Guo, Jun, et al.
Veröffentlicht: (2024)
CoDriveVLM: VLM-Enhanced Urban Cooperative Dispatching and Motion Planning for Future Autonomous Mobility on Demand Systems
von: Liu, Haichao, et al.
Veröffentlicht: (2025)
von: Liu, Haichao, et al.
Veröffentlicht: (2025)
PAT3D: Physics-Augmented Text-to-3D Scene Generation
von: Lin, Guying, et al.
Veröffentlicht: (2025)
von: Lin, Guying, et al.
Veröffentlicht: (2025)
4D Driving Scene Generation With Stereo Forcing
von: Lu, Hao, et al.
Veröffentlicht: (2025)
von: Lu, Hao, et al.
Veröffentlicht: (2025)
RE-VLM: Event-Augmented Vision-Language Model for Scene Understanding
von: Liu, Hanqing, et al.
Veröffentlicht: (2026)
von: Liu, Hanqing, et al.
Veröffentlicht: (2026)
Adaptive Visual Scene Understanding: Incremental Scene Graph Generation
von: Khandelwal, Naitik, et al.
Veröffentlicht: (2023)
von: Khandelwal, Naitik, et al.
Veröffentlicht: (2023)
Nighttime Autonomous Driving Scene Reconstruction with Physically-Based Gaussian Splatting
von: Kim, Tae-Kyeong, et al.
Veröffentlicht: (2026)
von: Kim, Tae-Kyeong, et al.
Veröffentlicht: (2026)
X-Scene: Large-Scale Driving Scene Generation with High Fidelity and Flexible Controllability
von: Yang, Yu, et al.
Veröffentlicht: (2025)
von: Yang, Yu, et al.
Veröffentlicht: (2025)
MSSF: A 4D Radar and Camera Fusion Framework With Multi-Stage Sampling for 3D Object Detection in Autonomous Driving
von: Liu, Hongsi, et al.
Veröffentlicht: (2024)
von: Liu, Hongsi, et al.
Veröffentlicht: (2024)
SemGS: Feed-Forward Semantic 3D Gaussian Splatting from Sparse Views for Generalizable Scene Understanding
von: Ye, Sheng, et al.
Veröffentlicht: (2026)
von: Ye, Sheng, et al.
Veröffentlicht: (2026)
WaterScenes: A Multi-Task 4D Radar-Camera Fusion Dataset and Benchmarks for Autonomous Driving on Water Surfaces
von: Yao, Shanliang, et al.
Veröffentlicht: (2023)
von: Yao, Shanliang, et al.
Veröffentlicht: (2023)
Enhance Planning with Physics-informed Safety Controller for End-to-end Autonomous Driving
von: Zhou, Hang, et al.
Veröffentlicht: (2024)
von: Zhou, Hang, et al.
Veröffentlicht: (2024)
4D Panoptic Scene Graph Generation
von: Yang, Jingkang, et al.
Veröffentlicht: (2024)
von: Yang, Jingkang, et al.
Veröffentlicht: (2024)
Generating Multimodal Driving Scenes via Next-Scene Prediction
von: Wu, Yanhao, et al.
Veröffentlicht: (2025)
von: Wu, Yanhao, et al.
Veröffentlicht: (2025)
OmniRe: Omni Urban Scene Reconstruction
von: Chen, Ziyu, et al.
Veröffentlicht: (2024)
von: Chen, Ziyu, et al.
Veröffentlicht: (2024)
AutoSplat: Constrained Gaussian Splatting for Autonomous Driving Scene Reconstruction
von: Khan, Mustafa, et al.
Veröffentlicht: (2024)
von: Khan, Mustafa, et al.
Veröffentlicht: (2024)
Curvature-Aware Captioning:Leveraging Geodesic Attention for 3D Scene Understanding
von: He, Ziyao, et al.
Veröffentlicht: (2026)
von: He, Ziyao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
VLM-E2E: Enhancing End-to-End Autonomous Driving with Multimodal Driver Attention Fusion
von: Liu, Pei, et al.
Veröffentlicht: (2025) -
OmniHD-Scenes: A Next-Generation Multimodal Dataset for Autonomous Driving
von: Zheng, Lianqing, et al.
Veröffentlicht: (2024) -
FSF-Net: Enhance 4D Occupancy Forecasting with Coarse BEV Scene Flow for Autonomous Driving
von: Guo, Erxin, et al.
Veröffentlicht: (2024) -
Omni-Scene: Omni-Gaussian Representation for Ego-Centric Sparse-View Scene Reconstruction
von: Wei, Dongxu, et al.
Veröffentlicht: (2024) -
LMMCoDrive: Cooperative Driving with Large Multimodal Model
von: Liu, Haichao, et al.
Veröffentlicht: (2024)