LingoQA: Visual Question Answering for Autonomous Driving
Fuente:
arXiv
Saved in:
| Main Authors: | Marcu, Ana-Maria, Chen, Long, Hünermann, Jan, Karnsund, Alice, Hanotte, Benoit, Chidananda, Prajwal, Nair, Saurabh, Badrinarayanan, Vijay, Kendall, Alex, Shotton, Jamie, Arani, Elahe, Sinavski, Oleg |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CarLLaVA: Vision language models for camera-only closed-loop driving
by: Renz, Katrin, et al.
Published: (2024)
by: Renz, Katrin, et al.
Published: (2024)
SimLingo: Vision-Only Closed-Loop Autonomous Driving with Language-Action Alignment
by: Renz, Katrin, et al.
Published: (2025)
by: Renz, Katrin, et al.
Published: (2025)
PixTrack: Precise 6DoF Object Pose Tracking using NeRF Templates and Feature-metric Alignment
by: Chidananda, Prajwal, et al.
Published: (2022)
by: Chidananda, Prajwal, et al.
Published: (2022)
GAIA-2: A Controllable Multi-View Generative World Model for Autonomous Driving
by: Russell, Lloyd, et al.
Published: (2025)
by: Russell, Lloyd, et al.
Published: (2025)
Rig3R: Rig-Aware Conditioning for Learned 3D Reconstruction
by: Li, Samuel, et al.
Published: (2025)
by: Li, Samuel, et al.
Published: (2025)
LA-Pose: Latent Action Pretraining Meets Pose Estimation
by: Wang, Zhengqing, et al.
Published: (2026)
by: Wang, Zhengqing, et al.
Published: (2026)
PhysVid: Physics Aware Local Conditioning for Generative Video Models
by: Pathak, Saurabh, et al.
Published: (2026)
by: Pathak, Saurabh, et al.
Published: (2026)
El Vaticano II, ¿texto constitucional de la fe? Una carta de Peter Hünermann
by: Peter Hünermann
Published: (2016)
by: Peter Hünermann
Published: (2016)
Robotic Learning in your Backyard: A Neural Simulator from Open Source Components
by: Zhou, Liyou, et al.
Published: (2024)
by: Zhou, Liyou, et al.
Published: (2024)
STRIDE-QA: Visual Question Answering Dataset for Spatiotemporal Reasoning in Urban Driving Scenes
by: Ishihara, Keishi, et al.
Published: (2025)
by: Ishihara, Keishi, et al.
Published: (2025)
BEnQA: A Question Answering and Reasoning Benchmark for Bengali and English
by: Shafayat, Sheikh, et al.
Published: (2024)
by: Shafayat, Sheikh, et al.
Published: (2024)
PolQA: Polish Question Answering Dataset
by: Rybak, Piotr, et al.
Published: (2022)
by: Rybak, Piotr, et al.
Published: (2022)
VoQA: Visual-only Question Answering
by: An, Jianing, et al.
Published: (2025)
by: An, Jianing, et al.
Published: (2025)
Conserve-Update-Revise to Cure Generalization and Robustness Trade-off in Adversarial Training
by: Gowda, Shruthi, et al.
Published: (2024)
by: Gowda, Shruthi, et al.
Published: (2024)
Gradual Divergence for Seamless Adaptation: A Novel Domain Incremental Learning Method
by: Jeeveswaran, Kishaan, et al.
Published: (2024)
by: Jeeveswaran, Kishaan, et al.
Published: (2024)
Beyond Unimodal Learning: The Importance of Integrating Multiple Modalities for Lifelong Learning
by: Sarfraz, Fahad, et al.
Published: (2024)
by: Sarfraz, Fahad, et al.
Published: (2024)
Can We Break Free from Strong Data Augmentations in Self-Supervised Learning?
by: Gowda, Shruthi, et al.
Published: (2024)
by: Gowda, Shruthi, et al.
Published: (2024)
Answering Questions in Stages: Prompt Chaining for Contract QA
by: Roegiest, Adam, et al.
Published: (2024)
by: Roegiest, Adam, et al.
Published: (2024)
DebateQA: Evaluating Question Answering on Debatable Knowledge
by: Xu, Rongwu, et al.
Published: (2024)
by: Xu, Rongwu, et al.
Published: (2024)
FoQA: A Faroese Question-Answering Dataset
by: Simonsen, Annika, et al.
Published: (2025)
by: Simonsen, Annika, et al.
Published: (2025)
NuScenes-QA: A Multi-modal Visual Question Answering Benchmark for Autonomous Driving Scenario
by: Qian, Tianwen, et al.
Published: (2023)
by: Qian, Tianwen, et al.
Published: (2023)
Untangling complex ethical issues involving wildlife
by: Justine Shotton
Published: (2024)
by: Justine Shotton
Published: (2024)
Welfare and wildlife rehabilitation
by: Justine Shotton
Published: (2026)
by: Justine Shotton
Published: (2026)
MedCoT-RAG: Causal Chain-of-Thought RAG for Medical Question Answering
by: Wang, Ziyu, et al.
Published: (2025)
by: Wang, Ziyu, et al.
Published: (2025)
WaymoQA: A Multi-View Visual Question Answering Dataset for Safety-Critical Reasoning in Autonomous Driving
by: Yu, Seungjun, et al.
Published: (2025)
by: Yu, Seungjun, et al.
Published: (2025)
ClimaQA_SLO - Slovenian Climate Question-Answering Benchmark
by: Ferk Ovčjak, Monika, et al.
Published: (2025)
by: Ferk Ovčjak, Monika, et al.
Published: (2025)
MMToM-QA: Multimodal Theory of Mind Question Answering
by: Jin, Chuanyang, et al.
Published: (2024)
by: Jin, Chuanyang, et al.
Published: (2024)
SyllabusQA: A Course Logistics Question Answering Dataset
by: Fernandez, Nigel, et al.
Published: (2024)
by: Fernandez, Nigel, et al.
Published: (2024)
M2QA: Multi-domain Multilingual Question Answering
by: Engländer, Leon, et al.
Published: (2024)
by: Engländer, Leon, et al.
Published: (2024)
GRS-QA -- Graph Reasoning-Structured Question Answering Dataset
by: Pahilajani, Anish, et al.
Published: (2024)
by: Pahilajani, Anish, et al.
Published: (2024)
Memory-QA: Answering Recall Questions Based on Multimodal Memories
by: Jiang, Hongda, et al.
Published: (2025)
by: Jiang, Hongda, et al.
Published: (2025)
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
by: Abdallah, Abdelrahman, et al.
Published: (2024)
by: Abdallah, Abdelrahman, et al.
Published: (2024)
Citation Analysis of Legal Instruments in Research
by: Chidananda, M, et al.
Published: (2025)
by: Chidananda, M, et al.
Published: (2025)
The Effectiveness of Random Forgetting for Robust Generalization
by: Ramkumar, Vijaya Raghavan T, et al.
Published: (2024)
by: Ramkumar, Vijaya Raghavan T, et al.
Published: (2024)
Learn the Lingo of Finance
by: Turner, Anne M.
Published: (2004)
by: Turner, Anne M.
Published: (2004)
A comprehensive review of sensors of radiation‐induced damage, radiation‐induced proximal events, and cell death
by: Saurabh Saini, et al.
Published: (2024)
by: Saurabh Saini, et al.
Published: (2024)
DriveLM: Driving with Graph Visual Question Answering
by: Sima, Chonghao, et al.
Published: (2023)
by: Sima, Chonghao, et al.
Published: (2023)
pdfQA: Diverse, Challenging, and Realistic Question Answering over PDFs
by: Schimanski, Tobias, et al.
Published: (2026)
by: Schimanski, Tobias, et al.
Published: (2026)
RespondeoQA: a Benchmark for Bilingual Latin-English Question Answering
by: Hudspeth, Marisa, et al.
Published: (2026)
by: Hudspeth, Marisa, et al.
Published: (2026)
DashboardQA: Benchmarking Multimodal Agents for Question Answering on Interactive Dashboards
by: Kartha, Aaryaman, et al.
Published: (2025)
by: Kartha, Aaryaman, et al.
Published: (2025)
Similar Items
-
CarLLaVA: Vision language models for camera-only closed-loop driving
by: Renz, Katrin, et al.
Published: (2024) -
SimLingo: Vision-Only Closed-Loop Autonomous Driving with Language-Action Alignment
by: Renz, Katrin, et al.
Published: (2025) -
PixTrack: Precise 6DoF Object Pose Tracking using NeRF Templates and Feature-metric Alignment
by: Chidananda, Prajwal, et al.
Published: (2022) -
GAIA-2: A Controllable Multi-View Generative World Model for Autonomous Driving
by: Russell, Lloyd, et al.
Published: (2025) -
Rig3R: Rig-Aware Conditioning for Learned 3D Reconstruction
by: Li, Samuel, et al.
Published: (2025)