LAD-Drive: Bridging Language and Trajectory with Action-Aware Diffusion Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Schmidt, Fabian, Fedurko, Karol, Enzweiler, Markus, Valada, Abhinav |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
NeRF and Gaussian Splatting SLAM in the Wild
von: Schmidt, Fabian, et al.
Veröffentlicht: (2024)
von: Schmidt, Fabian, et al.
Veröffentlicht: (2024)
Visual-Inertial SLAM for Unstructured Outdoor Environments: Benchmarking the Benefits and Computational Costs of Loop Closing
von: Schmidt, Fabian, et al.
Veröffentlicht: (2024)
von: Schmidt, Fabian, et al.
Veröffentlicht: (2024)
GraphPilot: Grounded Scene Graph Conditioning for Language-Based Autonomous Driving
von: Schmidt, Fabian, et al.
Veröffentlicht: (2025)
von: Schmidt, Fabian, et al.
Veröffentlicht: (2025)
ROVER: A Multi-Season Dataset for Visual SLAM
von: Schmidt, Fabian, et al.
Veröffentlicht: (2024)
von: Schmidt, Fabian, et al.
Veröffentlicht: (2024)
Enhancing LLM-based Autonomous Driving with Modular Traffic Light and Sign Recognition
von: Schmidt, Fabian, et al.
Veröffentlicht: (2025)
von: Schmidt, Fabian, et al.
Veröffentlicht: (2025)
DITTO: Demonstration Imitation by Trajectory Transformation
von: Heppert, Nick, et al.
Veröffentlicht: (2024)
von: Heppert, Nick, et al.
Veröffentlicht: (2024)
Learning Lane Graphs from Aerial Imagery Using Transformers
von: Büchner, Martin, et al.
Veröffentlicht: (2024)
von: Büchner, Martin, et al.
Veröffentlicht: (2024)
Bridging Perspectives: Foundation Model Guided BEV Maps for 3D Object Detection and Tracking
von: Käppeler, Markus, et al.
Veröffentlicht: (2025)
von: Käppeler, Markus, et al.
Veröffentlicht: (2025)
Effort-Based Criticality Metrics for Evaluating 3D Perception Errors in Autonomous Driving
von: Kaul, Sharang, et al.
Veröffentlicht: (2026)
von: Kaul, Sharang, et al.
Veröffentlicht: (2026)
CMRNext: Camera to LiDAR Matching in the Wild for Localization and Extrinsic Calibration
von: Cattaneo, Daniele, et al.
Veröffentlicht: (2024)
von: Cattaneo, Daniele, et al.
Veröffentlicht: (2024)
Leveraging Previous-Traversal Point Cloud Map Priors for Camera-Based 3D Object Detection and Tracking
von: Käppeler, Markus, et al.
Veröffentlicht: (2026)
von: Käppeler, Markus, et al.
Veröffentlicht: (2026)
Taxonomy-Aware Continual Semantic Segmentation in Hyperbolic Spaces for Open-World Perception
von: Hindel, Julia, et al.
Veröffentlicht: (2024)
von: Hindel, Julia, et al.
Veröffentlicht: (2024)
RaLF: Flow-based Global and Metric Radar Localization in LiDAR Maps
von: Nayak, Abhijeet, et al.
Veröffentlicht: (2023)
von: Nayak, Abhijeet, et al.
Veröffentlicht: (2023)
A Good Foundation is Worth Many Labels: Label-Efficient Panoptic Segmentation
von: Vödisch, Niclas, et al.
Veröffentlicht: (2024)
von: Vödisch, Niclas, et al.
Veröffentlicht: (2024)
Few-Shot Panoptic Segmentation With Foundation Models
von: Käppeler, Markus, et al.
Veröffentlicht: (2023)
von: Käppeler, Markus, et al.
Veröffentlicht: (2023)
CenterGrasp: Object-Aware Implicit Representation Learning for Simultaneous Shape Reconstruction and 6-DoF Grasp Estimation
von: Chisari, Eugenio, et al.
Veröffentlicht: (2023)
von: Chisari, Eugenio, et al.
Veröffentlicht: (2023)
Panoptic-Depth Forecasting
von: Hurtado, Juana Valeria, et al.
Veröffentlicht: (2024)
von: Hurtado, Juana Valeria, et al.
Veröffentlicht: (2024)
Perception Matters: Enhancing Embodied AI with Uncertainty-Aware Semantic Segmentation
von: Prasanna, Sai, et al.
Veröffentlicht: (2024)
von: Prasanna, Sai, et al.
Veröffentlicht: (2024)
AnchorD: Metric Grounding of Monocular Depth Using Factor Graphs
von: Dorer, Simon, et al.
Veröffentlicht: (2026)
von: Dorer, Simon, et al.
Veröffentlicht: (2026)
Vision-Based Autonomous UAV Navigation and Landing for Urban Search and Rescue
von: Mittal, Mayank, et al.
Veröffentlicht: (2019)
von: Mittal, Mayank, et al.
Veröffentlicht: (2019)
CoDEPS: Online Continual Learning for Depth Estimation and Panoptic Segmentation
von: Vödisch, Niclas, et al.
Veröffentlicht: (2023)
von: Vödisch, Niclas, et al.
Veröffentlicht: (2023)
CenterArt: Joint Shape Reconstruction and 6-DoF Grasp Estimation of Articulated Objects
von: Mokhtar, Sassan, et al.
Veröffentlicht: (2024)
von: Mokhtar, Sassan, et al.
Veröffentlicht: (2024)
Imagine2touch: Predictive Tactile Sensing for Robotic Manipulation using Efficient Low-Dimensional Signals
von: Ayad, Abdallah, et al.
Veröffentlicht: (2024)
von: Ayad, Abdallah, et al.
Veröffentlicht: (2024)
RAD-LAD: Rule and Language Grounded Autonomous Driving in Real-Time
von: Ghosh, Anurag, et al.
Veröffentlicht: (2026)
von: Ghosh, Anurag, et al.
Veröffentlicht: (2026)
Learning Appearance and Motion Cues for Panoptic Tracking
von: Hurtado, Juana Valeria, et al.
Veröffentlicht: (2025)
von: Hurtado, Juana Valeria, et al.
Veröffentlicht: (2025)
StixelNExT++: Lightweight Monocular Scene Segmentation and Representation for Collective Perception
von: Vosshans, Marcel, et al.
Veröffentlicht: (2025)
von: Vosshans, Marcel, et al.
Veröffentlicht: (2025)
Hyp2Former: Hierarchy-Aware Hyperbolic Embeddings for Open-Set Panoptic Segmentation
von: Lu, Yao, et al.
Veröffentlicht: (2026)
von: Lu, Yao, et al.
Veröffentlicht: (2026)
Self-Supervised Online Robot-Agnostic Traversability Estimation for Open-World Environments
von: Hindel, Julia, et al.
Veröffentlicht: (2026)
von: Hindel, Julia, et al.
Veröffentlicht: (2026)
PseudoTouch: Efficiently Imaging the Surface Feel of Objects for Robotic Manipulation
von: Röfer, Adrian, et al.
Veröffentlicht: (2024)
von: Röfer, Adrian, et al.
Veröffentlicht: (2024)
Label-Efficient LiDAR Panoptic Segmentation
von: Çanakçı, Ahmet Selim, et al.
Veröffentlicht: (2025)
von: Çanakçı, Ahmet Selim, et al.
Veröffentlicht: (2025)
Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy
von: Hou, Zhi, et al.
Veröffentlicht: (2025)
von: Hou, Zhi, et al.
Veröffentlicht: (2025)
DiWA: Diffusion Policy Adaptation with World Models
von: Chandra, Akshay L, et al.
Veröffentlicht: (2025)
von: Chandra, Akshay L, et al.
Veröffentlicht: (2025)
INoD: Injected Noise Discriminator for Self-Supervised Representation Learning in Agricultural Fields
von: Hindel, Julia, et al.
Veröffentlicht: (2023)
von: Hindel, Julia, et al.
Veröffentlicht: (2023)
Multi-Scale Neighborhood Occupancy Masked Autoencoder for Self-Supervised Learning in LiDAR Point Clouds
von: Abdelsamad, Mohamed, et al.
Veröffentlicht: (2025)
von: Abdelsamad, Mohamed, et al.
Veröffentlicht: (2025)
Articulated Object Estimation in the Wild
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2025)
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2025)
Unifying Language-Action Understanding and Generation for Autonomous Driving
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
von: Wang, Xinyang, et al.
Veröffentlicht: (2026)
F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions
von: Lv, Qi, et al.
Veröffentlicht: (2025)
von: Lv, Qi, et al.
Veröffentlicht: (2025)
Sparse3DTrack: Monocular 3D Object Tracking Using Sparse Supervision
von: Gosala, Nikhil, et al.
Veröffentlicht: (2026)
von: Gosala, Nikhil, et al.
Veröffentlicht: (2026)
AutoMoT: A Unified Vision-Language-Action Model with Asynchronous Mixture-of-Transformers for End-to-End Autonomous Driving
von: Huang, Wenhui, et al.
Veröffentlicht: (2026)
von: Huang, Wenhui, et al.
Veröffentlicht: (2026)
Visual Loop Closure Detection Through Deep Graph Consensus
von: Büchner, Martin, et al.
Veröffentlicht: (2025)
von: Büchner, Martin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
NeRF and Gaussian Splatting SLAM in the Wild
von: Schmidt, Fabian, et al.
Veröffentlicht: (2024) -
Visual-Inertial SLAM for Unstructured Outdoor Environments: Benchmarking the Benefits and Computational Costs of Loop Closing
von: Schmidt, Fabian, et al.
Veröffentlicht: (2024) -
GraphPilot: Grounded Scene Graph Conditioning for Language-Based Autonomous Driving
von: Schmidt, Fabian, et al.
Veröffentlicht: (2025) -
ROVER: A Multi-Season Dataset for Visual SLAM
von: Schmidt, Fabian, et al.
Veröffentlicht: (2024) -
Enhancing LLM-based Autonomous Driving with Modular Traffic Light and Sign Recognition
von: Schmidt, Fabian, et al.
Veröffentlicht: (2025)