BEV-TSR: Text-Scene Retrieval in BEV Space for Autonomous Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Tao, Wei, Dafeng, Jia, Zhengyu, Gao, Tian, Cai, Changwei, Hou, Chengkai, Jia, Peng, Zhan, Kun, Sun, Haiyang, Fan, Jingchen, Zhao, Yixing, Liu, Fu, Liang, Xiaodan, Lang, Xianpeng, Wang, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BEV-LLM: Leveraging Multimodal BEV Maps for Scene Captioning in Autonomous Driving
by: Brandstaetter, Felix, et al.
Published: (2025)
by: Brandstaetter, Felix, et al.
Published: (2025)
TempBEV: Improving Learned BEV Encoders with Combined Image and BEV Space Temporal Aggregation
by: Monninger, Thomas, et al.
Published: (2024)
by: Monninger, Thomas, et al.
Published: (2024)
BEV-VLM: Trajectory Planning via Unified BEV Abstraction
by: Chen, Guancheng, et al.
Published: (2025)
by: Chen, Guancheng, et al.
Published: (2025)
Other Vehicle Trajectories Are Also Needed: A Driving World Model Unifies Ego-Other Vehicle Trajectories in Video Latent Space
by: Zhu, Jian, et al.
Published: (2025)
by: Zhu, Jian, et al.
Published: (2025)
ME$^3$-BEV: Mamba-Enhanced Deep Reinforcement Learning for End-to-End Autonomous Driving with BEV-Perception
by: Lu, Siyi, et al.
Published: (2025)
by: Lu, Siyi, et al.
Published: (2025)
DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving
by: Cai, Kaiwen, et al.
Published: (2025)
by: Cai, Kaiwen, et al.
Published: (2025)
BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents
by: Zhang, Yumeng, et al.
Published: (2024)
by: Zhang, Yumeng, et al.
Published: (2024)
Hierarchical and Decoupled BEV Perception Learning Framework for Autonomous Driving
by: Dai, Yuqi, et al.
Published: (2024)
by: Dai, Yuqi, et al.
Published: (2024)
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
by: Tian, Xiaoyu, et al.
Published: (2024)
by: Tian, Xiaoyu, et al.
Published: (2024)
FSF-Net: Enhance 4D Occupancy Forecasting with Coarse BEV Scene Flow for Autonomous Driving
by: Guo, Erxin, et al.
Published: (2024)
by: Guo, Erxin, et al.
Published: (2024)
InstanceBEV: Unifying Instance and BEV Representation for 3D Panoptic Segmentation
by: Li, Feng, et al.
Published: (2025)
by: Li, Feng, et al.
Published: (2025)
Unleashing Generalization of End-to-End Autonomous Driving with Controllable Long Video Generation
by: Ma, Enhui, et al.
Published: (2024)
by: Ma, Enhui, et al.
Published: (2024)
GOLD-BEV: GrOund and aeriaL Data for Dense Semantic BEV Mapping of Dynamic Scenes
by: Niemeijer, Joshua, et al.
Published: (2026)
by: Niemeijer, Joshua, et al.
Published: (2026)
Generalizing Motion Planners with Mixture of Experts for Autonomous Driving
by: Sun, Qiao, et al.
Published: (2024)
by: Sun, Qiao, et al.
Published: (2024)
BEV-VAE: Multi-view Image Generation with Spatial Consistency for Autonomous Driving
by: Chen, Zeming, et al.
Published: (2025)
by: Chen, Zeming, et al.
Published: (2025)
BEVPredFormer: Spatio-temporal Attention for BEV Instance Prediction in Autonomous Driving
by: Antunes-García, Miguel, et al.
Published: (2026)
by: Antunes-García, Miguel, et al.
Published: (2026)
BEV-LIO(LC): BEV Image Assisted LiDAR-Inertial Odometry with Loop Closure
by: Cai, Haoxin, et al.
Published: (2025)
by: Cai, Haoxin, et al.
Published: (2025)
Privacy-Concealing Cooperative Perception for BEV Scene Segmentation
by: Wang, Song, et al.
Published: (2026)
by: Wang, Song, et al.
Published: (2026)
GeoBEV: Learning Geometric BEV Representation for Multi-view 3D Object Detection
by: Zhang, Jinqing, et al.
Published: (2024)
by: Zhang, Jinqing, et al.
Published: (2024)
BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving
by: Diaz-Zapata, Manuel Alejandro, et al.
Published: (2024)
by: Diaz-Zapata, Manuel Alejandro, et al.
Published: (2024)
Hierarchical End-to-End Autonomous Driving: Integrating BEV Perception with Deep Reinforcement Learning
by: Lu, Siyi, et al.
Published: (2024)
by: Lu, Siyi, et al.
Published: (2024)
GMF-Drive: Gated Mamba Fusion with Spatial-Aware BEV Representation for End-to-End Autonomous Driving
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
DenseBEV: Transforming BEV Grid Cells into 3D Objects
by: Dähling, Marius, et al.
Published: (2025)
by: Dähling, Marius, et al.
Published: (2025)
ChatBEV: A Visual Language Model that Understands BEV Maps
by: Xu, Qingyao, et al.
Published: (2025)
by: Xu, Qingyao, et al.
Published: (2025)
DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking
by: Zheng, Weicheng, et al.
Published: (2025)
by: Zheng, Weicheng, et al.
Published: (2025)
BEV-ODOM: Reducing Scale Drift in Monocular Visual Odometry with BEV Representation
by: Wei, Yufei, et al.
Published: (2024)
by: Wei, Yufei, et al.
Published: (2024)
SG-BEV: Satellite-Guided BEV Fusion for Cross-View Semantic Segmentation
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
MaskBEV: Towards A Unified Framework for BEV Detection and Map Segmentation
by: Zhao, Xiao, et al.
Published: (2024)
by: Zhao, Xiao, et al.
Published: (2024)
BEV$^2$PR: BEV-Enhanced Visual Place Recognition with Structural Cues
by: Ge, Fudong, et al.
Published: (2024)
by: Ge, Fudong, et al.
Published: (2024)
BEV-ODOM2: Enhanced BEV-based Monocular Visual Odometry with PV-BEV Fusion and Dense Flow Supervision for Ground Robots
by: Wei, Yufei, et al.
Published: (2025)
by: Wei, Yufei, et al.
Published: (2025)
HiNeuS: High-fidelity Neural Surface Mitigating Low-texture and Reflective Ambiguity
by: Wang, Yida, et al.
Published: (2025)
by: Wang, Yida, et al.
Published: (2025)
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
by: Li, Pengxiang, et al.
Published: (2025)
by: Li, Pengxiang, et al.
Published: (2025)
Zero-BEV: Zero-shot Projection of Any First-Person Modality to BEV Maps
by: Monaci, Gianluca, et al.
Published: (2024)
by: Monaci, Gianluca, et al.
Published: (2024)
UniPLV: Towards Label-Efficient Open-World 3D Scene Understanding by Regional Visual Language Supervision
by: Wang, Yuru, et al.
Published: (2024)
by: Wang, Yuru, et al.
Published: (2024)
TransDiffuser: Diverse Trajectory Generation with Decorrelated Multi-modal Representation for End-to-end Autonomous Driving
by: Jiang, Xuefeng, et al.
Published: (2025)
by: Jiang, Xuefeng, et al.
Published: (2025)
PC-BEV: An Efficient Polar-Cartesian BEV Fusion Framework for LiDAR Semantic Segmentation
by: Qiu, Shoumeng, et al.
Published: (2024)
by: Qiu, Shoumeng, et al.
Published: (2024)
RESAR-BEV: An Explainable Progressive Residual Autoregressive Approach for Camera-Radar Fusion in BEV Segmentation
by: Zeng, Zhiwen, et al.
Published: (2025)
by: Zeng, Zhiwen, et al.
Published: (2025)
RopeBEV: A Multi-Camera Roadside Perception Network in Bird's-Eye-View
by: Jia, Jinrang, et al.
Published: (2024)
by: Jia, Jinrang, et al.
Published: (2024)
BEVDriver: Leveraging BEV Maps in LLMs for Robust Closed-Loop Driving
by: Winter, Katharina, et al.
Published: (2025)
by: Winter, Katharina, et al.
Published: (2025)
Similar Items
-
BEV-LLM: Leveraging Multimodal BEV Maps for Scene Captioning in Autonomous Driving
by: Brandstaetter, Felix, et al.
Published: (2025) -
TempBEV: Improving Learned BEV Encoders with Combined Image and BEV Space Temporal Aggregation
by: Monninger, Thomas, et al.
Published: (2024) -
BEV-VLM: Trajectory Planning via Unified BEV Abstraction
by: Chen, Guancheng, et al.
Published: (2025) -
Other Vehicle Trajectories Are Also Needed: A Driving World Model Unifies Ego-Other Vehicle Trajectories in Video Latent Space
by: Zhu, Jian, et al.
Published: (2025) -
ME$^3$-BEV: Mamba-Enhanced Deep Reinforcement Learning for End-to-End Autonomous Driving with BEV-Perception
by: Lu, Siyi, et al.
Published: (2025)