BEV-LLM: Leveraging Multimodal BEV Maps for Scene Captioning in Autonomous Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Brandstaetter, Felix, Schuetz, Erik, Winter, Katharina, Flohr, Fabian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BEVDriver: Leveraging BEV Maps in LLMs for Robust Closed-Loop Driving
by: Winter, Katharina, et al.
Published: (2025)
by: Winter, Katharina, et al.
Published: (2025)
BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving
by: Tang, Tao, et al.
Published: (2024)
by: Tang, Tao, et al.
Published: (2024)
BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents
by: Zhang, Yumeng, et al.
Published: (2024)
by: Zhang, Yumeng, et al.
Published: (2024)
Hierarchical and Decoupled BEV Perception Learning Framework for Autonomous Driving
by: Dai, Yuqi, et al.
Published: (2024)
by: Dai, Yuqi, et al.
Published: (2024)
GOLD-BEV: GrOund and aeriaL Data for Dense Semantic BEV Mapping of Dynamic Scenes
by: Niemeijer, Joshua, et al.
Published: (2026)
by: Niemeijer, Joshua, et al.
Published: (2026)
MaskBEV: Towards A Unified Framework for BEV Detection and Map Segmentation
by: Zhao, Xiao, et al.
Published: (2024)
by: Zhao, Xiao, et al.
Published: (2024)
BEVMAPMATCH: Multimodal BEV Neural Map Matching for Robust Re-Localization of Autonomous Vehicles
by: Sural, Shounak, et al.
Published: (2026)
by: Sural, Shounak, et al.
Published: (2026)
MTA: Multimodal Task Alignment for BEV Perception and Captioning
by: Ma, Yunsheng, et al.
Published: (2024)
by: Ma, Yunsheng, et al.
Published: (2024)
FSF-Net: Enhance 4D Occupancy Forecasting with Coarse BEV Scene Flow for Autonomous Driving
by: Guo, Erxin, et al.
Published: (2024)
by: Guo, Erxin, et al.
Published: (2024)
ChatBEV: A Visual Language Model that Understands BEV Maps
by: Xu, Qingyao, et al.
Published: (2025)
by: Xu, Qingyao, et al.
Published: (2025)
Context-based Motion Retrieval using Open Vocabulary Methods for Autonomous Driving
by: Englmeier, Stefan, et al.
Published: (2025)
by: Englmeier, Stefan, et al.
Published: (2025)
Zero-BEV: Zero-shot Projection of Any First-Person Modality to BEV Maps
by: Monaci, Gianluca, et al.
Published: (2024)
by: Monaci, Gianluca, et al.
Published: (2024)
BEV-VAE: Multi-view Image Generation with Spatial Consistency for Autonomous Driving
by: Chen, Zeming, et al.
Published: (2025)
by: Chen, Zeming, et al.
Published: (2025)
BEVPredFormer: Spatio-temporal Attention for BEV Instance Prediction in Autonomous Driving
by: Antunes-García, Miguel, et al.
Published: (2026)
by: Antunes-García, Miguel, et al.
Published: (2026)
Leveraging BEV Paradigm for Ground-to-Aerial Image Synthesis
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
DenseBEV: Transforming BEV Grid Cells into 3D Objects
by: Dähling, Marius, et al.
Published: (2025)
by: Dähling, Marius, et al.
Published: (2025)
DiffSemanticFusion: Semantic Raster BEV Fusion for Autonomous Driving via Online HD Map Diffusion
by: Sun, Zhigang, et al.
Published: (2025)
by: Sun, Zhigang, et al.
Published: (2025)
InstanceBEV: Unifying Instance and BEV Representation for 3D Panoptic Segmentation
by: Li, Feng, et al.
Published: (2025)
by: Li, Feng, et al.
Published: (2025)
SG-BEV: Satellite-Guided BEV Fusion for Cross-View Semantic Segmentation
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
BEV$^2$PR: BEV-Enhanced Visual Place Recognition with Structural Cues
by: Ge, Fudong, et al.
Published: (2024)
by: Ge, Fudong, et al.
Published: (2024)
WorldVLM: Combining World Model Forecasting and Vision-Language Reasoning
by: Englmeier, Stefan, et al.
Published: (2026)
by: Englmeier, Stefan, et al.
Published: (2026)
BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving
by: Diaz-Zapata, Manuel Alejandro, et al.
Published: (2024)
by: Diaz-Zapata, Manuel Alejandro, et al.
Published: (2024)
BEVMOSNet: Multimodal Fusion for BEV Moving Object Segmentation
by: Cong, Hiep Truong, et al.
Published: (2025)
by: Cong, Hiep Truong, et al.
Published: (2025)
GMF-Drive: Gated Mamba Fusion with Spatial-Aware BEV Representation for End-to-End Autonomous Driving
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
Privacy-Concealing Cooperative Perception for BEV Scene Segmentation
by: Wang, Song, et al.
Published: (2026)
by: Wang, Song, et al.
Published: (2026)
RESAR-BEV: An Explainable Progressive Residual Autoregressive Approach for Camera-Radar Fusion in BEV Segmentation
by: Zeng, Zhiwen, et al.
Published: (2025)
by: Zeng, Zhiwen, et al.
Published: (2025)
PC-BEV: An Efficient Polar-Cartesian BEV Fusion Framework for LiDAR Semantic Segmentation
by: Qiu, Shoumeng, et al.
Published: (2024)
by: Qiu, Shoumeng, et al.
Published: (2024)
GeoBEV: Learning Geometric BEV Representation for Multi-view 3D Object Detection
by: Zhang, Jinqing, et al.
Published: (2024)
by: Zhang, Jinqing, et al.
Published: (2024)
Mapping like a Skeptic: Probabilistic BEV Projection for Online HD Mapping
by: Erdoğan, Fatih, et al.
Published: (2025)
by: Erdoğan, Fatih, et al.
Published: (2025)
BEV-MAE: Bird's Eye View Masked Autoencoders for Point Cloud Pre-training in Autonomous Driving Scenarios
by: Lin, Zhiwei, et al.
Published: (2022)
by: Lin, Zhiwei, et al.
Published: (2022)
ROA-BEV: 2D Region-Oriented Attention for BEV-based 3D Object Detection
by: Chen, Jiwei, et al.
Published: (2024)
by: Chen, Jiwei, et al.
Published: (2024)
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
BEVCar: Camera-Radar Fusion for BEV Map and Object Segmentation
by: Schramm, Jonas, et al.
Published: (2024)
by: Schramm, Jonas, et al.
Published: (2024)
TempBEV: Improving Learned BEV Encoders with Combined Image and BEV Space Temporal Aggregation
by: Monninger, Thomas, et al.
Published: (2024)
by: Monninger, Thomas, et al.
Published: (2024)
End-to-End Driving with Online Trajectory Evaluation via BEV World Model
by: Li, Yingyan, et al.
Published: (2025)
by: Li, Yingyan, et al.
Published: (2025)
Dur360BEV: A Real-world 360-degree Single Camera Dataset and Benchmark for Bird-Eye View Mapping in Autonomous Driving
by: E, Wenke, et al.
Published: (2025)
by: E, Wenke, et al.
Published: (2025)
LetsMap: Unsupervised Representation Learning for Semantic BEV Mapping
by: Gosala, Nikhil, et al.
Published: (2024)
by: Gosala, Nikhil, et al.
Published: (2024)
Bridging Stereo Geometry and BEV Representation with Reliable Mutual Interaction for Semantic Scene Completion
by: Li, Bohan, et al.
Published: (2023)
by: Li, Bohan, et al.
Published: (2023)
BEV-Patch-PF: Particle Filtering with BEV-Aerial Feature Matching for Off-Road Geo-Localization
by: Lee, Dongmyeong, et al.
Published: (2025)
by: Lee, Dongmyeong, et al.
Published: (2025)
DeepUrban: Interaction-Aware Trajectory Prediction and Planning for Automated Driving by Aerial Imagery
by: Selzer, Constantin, et al.
Published: (2026)
by: Selzer, Constantin, et al.
Published: (2026)
Similar Items
-
BEVDriver: Leveraging BEV Maps in LLMs for Robust Closed-Loop Driving
by: Winter, Katharina, et al.
Published: (2025) -
BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving
by: Tang, Tao, et al.
Published: (2024) -
BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents
by: Zhang, Yumeng, et al.
Published: (2024) -
Hierarchical and Decoupled BEV Perception Learning Framework for Autonomous Driving
by: Dai, Yuqi, et al.
Published: (2024) -
GOLD-BEV: GrOund and aeriaL Data for Dense Semantic BEV Mapping of Dynamic Scenes
by: Niemeijer, Joshua, et al.
Published: (2026)