Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail
Fuente:
arXiv
Saved in:
| Main Authors: | NVIDIA, :, Wang, Yan, Luo, Wenjie, Bai, Junjie, Cao, Yulong, Che, Tong, Chen, Ke, Chen, Yuxiao, Diamond, Jenna, Ding, Yifan, Ding, Wenhao, Feng, Liang, Heinrich, Greg, Huang, Jack, Karkus, Peter, Li, Boyi, Li, Pinyi, Lin, Tsung-Yi, Liu, Dongran, Liu, Ming-Yu, Liu, Langechuan, Liu, Zhijian, Lu, Jason, Mao, Yunxiang, Molchanov, Pavlo, Pavao, Lindsey, Peng, Zhenghao, Ranzinger, Mike, Schmerling, Ed, Shen, Shida, Shi, Yunfei, Tariq, Sarah, Tian, Ran, Wekel, Tilman, Weng, Xinshuo, Xiao, Tianjun, Yang, Eric, Yang, Xiaodong, You, Yurong, Zeng, Xiaohui, Zhang, Wenyuan, Ivanovic, Boris, Pavone, Marco |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tokenize the World into Object-level Knowledge to Address Long-tail Events in Autonomous Driving
by: Tian, Ran, et al.
Published: (2024)
by: Tian, Ran, et al.
Published: (2024)
Language-Image Models with 3D Understanding
by: Cho, Jang Hyun, et al.
Published: (2024)
by: Cho, Jang Hyun, et al.
Published: (2024)
AM-RADIO: Agglomerative Vision Foundation Model -- Reduce All Domains Into One
by: Ranzinger, Mike, et al.
Published: (2023)
by: Ranzinger, Mike, et al.
Published: (2023)
Closed-Loop Supervised Fine-Tuning of Tokenized Traffic Models
by: Zhang, Zhejun, et al.
Published: (2024)
by: Zhang, Zhejun, et al.
Published: (2024)
Promptable Closed-loop Traffic Simulation
by: Tan, Shuhan, et al.
Published: (2024)
by: Tan, Shuhan, et al.
Published: (2024)
DTPP: Differentiable Joint Conditional Prediction and Cost Evaluation for Tree Policy Planning in Autonomous Driving
by: Huang, Zhiyu, et al.
Published: (2023)
by: Huang, Zhiyu, et al.
Published: (2023)
PHI-S: Distribution Balancing for Label-Free Multi-Teacher Distillation
by: Ranzinger, Mike, et al.
Published: (2024)
by: Ranzinger, Mike, et al.
Published: (2024)
FeatSharp: Your Vision Model Features, Sharper
by: Ranzinger, Mike, et al.
Published: (2025)
by: Ranzinger, Mike, et al.
Published: (2025)
STORM: Spatio-Temporal Reconstruction Model for Large-Scale Outdoor Scenes
by: Yang, Jiawei, et al.
Published: (2024)
by: Yang, Jiawei, et al.
Published: (2024)
Sample-Efficient Safety Assurances using Conformal Prediction
by: Luo, Rachel, et al.
Published: (2021)
by: Luo, Rachel, et al.
Published: (2021)
Accelerating Structured Chain-of-Thought in Autonomous Vehicles
by: Gu, Yi, et al.
Published: (2026)
by: Gu, Yi, et al.
Published: (2026)
C-RADIOv4 (Tech Report)
by: Ranzinger, Mike, et al.
Published: (2026)
by: Ranzinger, Mike, et al.
Published: (2026)
Extrapolated Urban View Synthesis Benchmark
by: Han, Xiangyu, et al.
Published: (2024)
by: Han, Xiangyu, et al.
Published: (2024)
Scaling Vision Pre-Training to 4K Resolution
by: Shi, Baifeng, et al.
Published: (2025)
by: Shi, Baifeng, et al.
Published: (2025)
Counterfactual VLA: Self-Reflective Vision-Language-Action Model with Adaptive Reasoning
by: Peng, Zhenghao "Mark", et al.
Published: (2025)
by: Peng, Zhenghao "Mark", et al.
Published: (2025)
Universal Deep Research: Bring Your Own Model and Strategy
by: Belcak, Peter, et al.
Published: (2025)
by: Belcak, Peter, et al.
Published: (2025)
RADIOv2.5: Improved Baselines for Agglomerative Vision Foundation Models
by: Heinrich, Greg, et al.
Published: (2024)
by: Heinrich, Greg, et al.
Published: (2024)
Efficient Multi-Camera Tokenization with Triplanes for End-to-End Driving
by: Ivanovic, Boris, et al.
Published: (2025)
by: Ivanovic, Boris, et al.
Published: (2025)
System-Level Analysis of Module Uncertainty Quantification in the Autonomy Pipeline
by: Deglurkar, Sampada, et al.
Published: (2024)
by: Deglurkar, Sampada, et al.
Published: (2024)
RoaD: Rollouts as Demonstrations for Closed-Loop Supervised Fine-Tuning of Autonomous Driving Policies
by: Garcia-Cobo, Guillermo, et al.
Published: (2025)
by: Garcia-Cobo, Guillermo, et al.
Published: (2025)
Towards Efficient and Effective Multi-Camera Encoding for End-to-End Driving
by: Yang, Jiawei, et al.
Published: (2025)
by: Yang, Jiawei, et al.
Published: (2025)
DreamDrive: Generative 4D Scene Modeling from Street View Images
by: Mao, Jiageng, et al.
Published: (2024)
by: Mao, Jiageng, et al.
Published: (2024)
Gen-Drive: Enhancing Diffusion Generative Driving Policies with Reward Modeling and Reinforcement Learning Fine-tuning
by: Huang, Zhiyu, et al.
Published: (2024)
by: Huang, Zhiyu, et al.
Published: (2024)
Can Users Specify Driving Speed? Bench2Drive-Speed: Benchmark and Baselines for Desired-Speed Conditioned Autonomous Driving
by: Shao, Yuqian, et al.
Published: (2026)
by: Shao, Yuqian, et al.
Published: (2026)
LoRD: Adapting Differentiable Driving Policies to Distribution Shifts
by: Diehl, Christopher, et al.
Published: (2024)
by: Diehl, Christopher, et al.
Published: (2024)
LoRA3D: Low-Rank Self-Calibration of 3D Geometric Foundation Models
by: Lu, Ziqi, et al.
Published: (2024)
by: Lu, Ziqi, et al.
Published: (2024)
Driving Everywhere with Large Language Model Policy Adaptation
by: Li, Boyi, et al.
Published: (2024)
by: Li, Boyi, et al.
Published: (2024)
TwinTURBO: Semi-Supervised Fine-Tuning of Foundation Models via Mutual Information Decompositions for Downstream Task and Latent Spaces
by: Quétant, Guillaume, et al.
Published: (2025)
by: Quétant, Guillaume, et al.
Published: (2025)
ZAPP! Zonotope Agreement of Prediction and Planning for Continuous-Time Collision Avoidance with Discrete-Time Dynamics
by: Paparusso, Luca, et al.
Published: (2024)
by: Paparusso, Luca, et al.
Published: (2024)
Latency Analysis and Optimization of Alpamayo 1 via Efficient Trajectory Generation
by: Jeon, Yunseong, et al.
Published: (2026)
by: Jeon, Yunseong, et al.
Published: (2026)
Trends in Motion Prediction Toward Deployable and Generalizable Autonomy: A Revisit and Perspectives
by: Wang, Letian, et al.
Published: (2025)
by: Wang, Letian, et al.
Published: (2025)
Learning Multiple Initial Solutions to Optimization Problems
by: Sharony, Elad, et al.
Published: (2024)
by: Sharony, Elad, et al.
Published: (2024)
Trifluoromethanesulfonic Acid‐Promoted Esterification of Unactivated Tertiary Amides with Tetrahydrofurans and Potassium Halides to Access Haloalkyl Esters
by: Yueyue Fan, et al.
Published: (2025)
by: Yueyue Fan, et al.
Published: (2025)
DistillNeRF: Perceiving 3D Scenes from Single-Glance Images by Distilling Neural Fields and Foundation Model Features
by: Wang, Letian, et al.
Published: (2024)
by: Wang, Letian, et al.
Published: (2024)
StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement
by: Seo, Junwon, et al.
Published: (2026)
by: Seo, Junwon, et al.
Published: (2026)
Real-Time Anomaly Detection and Reactive Planning with Large Language Models
by: Sinha, Rohan, et al.
Published: (2024)
by: Sinha, Rohan, et al.
Published: (2024)
Diagnostic Runtime Monitoring with Martingales
by: Hindy, Ali, et al.
Published: (2024)
by: Hindy, Ali, et al.
Published: (2024)
Multi-View Fusion Neural Network for Traffic Demand Prediction
by: Zhang, Dongran, et al.
Published: (2024)
by: Zhang, Dongran, et al.
Published: (2024)
Minifinetuning: Low-Data Generation Domain Adaptation through Corrective Self-Distillation
by: Belcak, Peter, et al.
Published: (2025)
by: Belcak, Peter, et al.
Published: (2025)
A Novel Neural-symbolic System under Statistical Relational Learning
by: Yu, Dongran, et al.
Published: (2023)
by: Yu, Dongran, et al.
Published: (2023)
Similar Items
-
Tokenize the World into Object-level Knowledge to Address Long-tail Events in Autonomous Driving
by: Tian, Ran, et al.
Published: (2024) -
Language-Image Models with 3D Understanding
by: Cho, Jang Hyun, et al.
Published: (2024) -
AM-RADIO: Agglomerative Vision Foundation Model -- Reduce All Domains Into One
by: Ranzinger, Mike, et al.
Published: (2023) -
Closed-Loop Supervised Fine-Tuning of Tokenized Traffic Models
by: Zhang, Zhejun, et al.
Published: (2024) -
Promptable Closed-loop Traffic Simulation
by: Tan, Shuhan, et al.
Published: (2024)