GraphPilot: Grounded Scene Graph Conditioning for Language-Based Autonomous Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Schmidt, Fabian, Enzweiler, Markus, Valada, Abhinav |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LAD-Drive: Bridging Language and Trajectory with Action-Aware Diffusion Transformers
by: Schmidt, Fabian, et al.
Published: (2026)
by: Schmidt, Fabian, et al.
Published: (2026)
Enhancing LLM-based Autonomous Driving with Modular Traffic Light and Sign Recognition
by: Schmidt, Fabian, et al.
Published: (2025)
by: Schmidt, Fabian, et al.
Published: (2025)
NeRF and Gaussian Splatting SLAM in the Wild
by: Schmidt, Fabian, et al.
Published: (2024)
by: Schmidt, Fabian, et al.
Published: (2024)
Visual-Inertial SLAM for Unstructured Outdoor Environments: Benchmarking the Benefits and Computational Costs of Loop Closing
by: Schmidt, Fabian, et al.
Published: (2024)
by: Schmidt, Fabian, et al.
Published: (2024)
ROVER: A Multi-Season Dataset for Visual SLAM
by: Schmidt, Fabian, et al.
Published: (2024)
by: Schmidt, Fabian, et al.
Published: (2024)
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
by: Werby, Abdelrhman, et al.
Published: (2024)
by: Werby, Abdelrhman, et al.
Published: (2024)
AnchorD: Metric Grounding of Monocular Depth Using Factor Graphs
by: Dorer, Simon, et al.
Published: (2026)
by: Dorer, Simon, et al.
Published: (2026)
Effort-Based Criticality Metrics for Evaluating 3D Perception Errors in Autonomous Driving
by: Kaul, Sharang, et al.
Published: (2026)
by: Kaul, Sharang, et al.
Published: (2026)
AmodalSynthDrive: A Synthetic Amodal Perception Dataset for Autonomous Driving
by: Sekkat, Ahmed Rida, et al.
Published: (2023)
by: Sekkat, Ahmed Rida, et al.
Published: (2023)
Learning Lane Graphs from Aerial Imagery Using Transformers
by: Büchner, Martin, et al.
Published: (2024)
by: Büchner, Martin, et al.
Published: (2024)
Multi-Modal Sensor Fusion using Hybrid Attention for Autonomous Driving
by: Mayank, Mayank, et al.
Published: (2026)
by: Mayank, Mayank, et al.
Published: (2026)
GraphAD: Interaction Scene Graph for End-to-end Autonomous Driving
by: Zhang, Yunpeng, et al.
Published: (2024)
by: Zhang, Yunpeng, et al.
Published: (2024)
The Bare Necessities: Designing Simple, Effective Open-Vocabulary Scene Graphs
by: Kassab, Christina, et al.
Published: (2024)
by: Kassab, Christina, et al.
Published: (2024)
Vision-Based Autonomous UAV Navigation and Landing for Urban Search and Rescue
by: Mittal, Mayank, et al.
Published: (2019)
by: Mittal, Mayank, et al.
Published: (2019)
SearchAD: Large-Scale Rare Image Retrieval Dataset for Autonomous Driving
by: Embacher, Felix, et al.
Published: (2026)
by: Embacher, Felix, et al.
Published: (2026)
Leveraging Previous-Traversal Point Cloud Map Priors for Camera-Based 3D Object Detection and Tracking
by: Käppeler, Markus, et al.
Published: (2026)
by: Käppeler, Markus, et al.
Published: (2026)
T2SG: Traffic Topology Scene Graph for Topology Reasoning in Autonomous Driving
by: Lv, Changsheng, et al.
Published: (2024)
by: Lv, Changsheng, et al.
Published: (2024)
MGNet: Monocular Geometric Scene Understanding for Autonomous Driving
by: Schön, Markus, et al.
Published: (2022)
by: Schön, Markus, et al.
Published: (2022)
Articulated 3D Scene Graphs for Open-World Mobile Manipulation
by: Büchner, Martin, et al.
Published: (2026)
by: Büchner, Martin, et al.
Published: (2026)
CMRNext: Camera to LiDAR Matching in the Wild for Localization and Extrinsic Calibration
by: Cattaneo, Daniele, et al.
Published: (2024)
by: Cattaneo, Daniele, et al.
Published: (2024)
Syn-Mediverse: A Multimodal Synthetic Dataset for Intelligent Scene Understanding of Healthcare Facilities
by: Mohan, Rohit, et al.
Published: (2023)
by: Mohan, Rohit, et al.
Published: (2023)
StixelNExT++: Lightweight Monocular Scene Segmentation and Representation for Collective Perception
by: Vosshans, Marcel, et al.
Published: (2025)
by: Vosshans, Marcel, et al.
Published: (2025)
A Point-Based Approach to Efficient LiDAR Multi-Task Perception
by: Lang, Christopher, et al.
Published: (2024)
by: Lang, Christopher, et al.
Published: (2024)
View-on-Graph: Zero-shot 3D Visual Grounding via Vision-Language Reasoning on Scene Graphs
by: Liu, Yuanyuan, et al.
Published: (2025)
by: Liu, Yuanyuan, et al.
Published: (2025)
ScenePilot-4K: A Large-Scale First-Person Dataset and Benchmark for Vision-Language Models in Autonomous Driving
by: Wang, Yujin, et al.
Published: (2026)
by: Wang, Yujin, et al.
Published: (2026)
FIHA: Autonomous Hallucination Evaluation in Vision-Language Models with Davidson Scene Graphs
by: Yan, Bowen, et al.
Published: (2024)
by: Yan, Bowen, et al.
Published: (2024)
BEV-LLM: Leveraging Multimodal BEV Maps for Scene Captioning in Autonomous Driving
by: Brandstaetter, Felix, et al.
Published: (2025)
by: Brandstaetter, Felix, et al.
Published: (2025)
Joint Target-Less Intrinsic and Extrinsic Camera-LiDAR Calibration using Deep Point Correspondences
by: Bultmann, Simon, et al.
Published: (2026)
by: Bultmann, Simon, et al.
Published: (2026)
Visual Loop Closure Detection Through Deep Graph Consensus
by: Büchner, Martin, et al.
Published: (2025)
by: Büchner, Martin, et al.
Published: (2025)
Graph-Based Multi-Modal Sensor Fusion for Autonomous Driving
by: Sani, Depanshu, et al.
Published: (2024)
by: Sani, Depanshu, et al.
Published: (2024)
Image Synthesis with Graph Conditioning: CLIP-Guided Diffusion Models for Scene Graphs
by: Mishra, Rameshwar, et al.
Published: (2024)
by: Mishra, Rameshwar, et al.
Published: (2024)
BCTR: Bidirectional Conditioning Transformer for Scene Graph Generation
by: Hao, Peng, et al.
Published: (2024)
by: Hao, Peng, et al.
Published: (2024)
Scene Graph Generation via Conditional Random Fields
by: Cong, Weilin, et al.
Published: (2018)
by: Cong, Weilin, et al.
Published: (2018)
SceneProp: Combining Neural Network and Markov Random Field for Scene-Graph Grounding
by: Otani, Keita, et al.
Published: (2025)
by: Otani, Keita, et al.
Published: (2025)
AdaDrive: Self-Adaptive Slow-Fast System for Language-Grounded Autonomous Driving
by: Zhang, Ruifei, et al.
Published: (2025)
by: Zhang, Ruifei, et al.
Published: (2025)
RaLF: Flow-based Global and Metric Radar Localization in LiDAR Maps
by: Nayak, Abhijeet, et al.
Published: (2023)
by: Nayak, Abhijeet, et al.
Published: (2023)
A Good Foundation is Worth Many Labels: Label-Efficient Panoptic Segmentation
by: Vödisch, Niclas, et al.
Published: (2024)
by: Vödisch, Niclas, et al.
Published: (2024)
Few-Shot Panoptic Segmentation With Foundation Models
by: Käppeler, Markus, et al.
Published: (2023)
by: Käppeler, Markus, et al.
Published: (2023)
SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching
by: Sun, Xuefei, et al.
Published: (2026)
by: Sun, Xuefei, et al.
Published: (2026)
SceneGraphVLM: Dynamic Scene Graph Generation from Video with Vision-Language Models
by: Makarov, Vladislav, et al.
Published: (2026)
by: Makarov, Vladislav, et al.
Published: (2026)
Similar Items
-
LAD-Drive: Bridging Language and Trajectory with Action-Aware Diffusion Transformers
by: Schmidt, Fabian, et al.
Published: (2026) -
Enhancing LLM-based Autonomous Driving with Modular Traffic Light and Sign Recognition
by: Schmidt, Fabian, et al.
Published: (2025) -
NeRF and Gaussian Splatting SLAM in the Wild
by: Schmidt, Fabian, et al.
Published: (2024) -
Visual-Inertial SLAM for Unstructured Outdoor Environments: Benchmarking the Benefits and Computational Costs of Loop Closing
by: Schmidt, Fabian, et al.
Published: (2024) -
ROVER: A Multi-Season Dataset for Visual SLAM
by: Schmidt, Fabian, et al.
Published: (2024)