Latent Chain-of-Thought World Modeling for End-to-End Driving
Fuente:
arXiv
Guardado en:
| Autores principales: | Tan, Shuhan, Chitta, Kashyap, Chen, Yuxiao, Tian, Ran, You, Yurong, Wang, Yan, Luo, Wenjie, Cao, Yulong, Krahenbuhl, Philipp, Pavone, Marco, Ivanovic, Boris |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Efficient Multi-Camera Tokenization with Triplanes for End-to-End Driving
por: Ivanovic, Boris, et al.
Publicado: (2025)
por: Ivanovic, Boris, et al.
Publicado: (2025)
Promptable Closed-loop Traffic Simulation
por: Tan, Shuhan, et al.
Publicado: (2024)
por: Tan, Shuhan, et al.
Publicado: (2024)
Towards Efficient and Effective Multi-Camera Encoding for End-to-End Driving
por: Yang, Jiawei, et al.
Publicado: (2025)
por: Yang, Jiawei, et al.
Publicado: (2025)
Accelerating Structured Chain-of-Thought in Autonomous Vehicles
por: Gu, Yi, et al.
Publicado: (2026)
por: Gu, Yi, et al.
Publicado: (2026)
Hidden Biases of End-to-End Driving Datasets
por: Zimmerlin, Julian, et al.
Publicado: (2024)
por: Zimmerlin, Julian, et al.
Publicado: (2024)
Tokenize the World into Object-level Knowledge to Address Long-tail Events in Autonomous Driving
por: Tian, Ran, et al.
Publicado: (2024)
por: Tian, Ran, et al.
Publicado: (2024)
Data Scaling Laws for End-to-End Autonomous Driving
por: Naumann, Alexander, et al.
Publicado: (2025)
por: Naumann, Alexander, et al.
Publicado: (2025)
Gen-Drive: Enhancing Diffusion Generative Driving Policies with Reward Modeling and Reinforcement Learning Fine-tuning
por: Huang, Zhiyu, et al.
Publicado: (2024)
por: Huang, Zhiyu, et al.
Publicado: (2024)
Language-Image Models with 3D Understanding
por: Cho, Jang Hyun, et al.
Publicado: (2024)
por: Cho, Jang Hyun, et al.
Publicado: (2024)
FutureX: Enhance End-to-End Autonomous Driving via Latent Chain-of-Thought World Model
por: Lin, Hongbin, et al.
Publicado: (2025)
por: Lin, Hongbin, et al.
Publicado: (2025)
DriveCoT: Integrating Chain-of-Thought Reasoning with End-to-End Driving
por: Wang, Tianqi, et al.
Publicado: (2024)
por: Wang, Tianqi, et al.
Publicado: (2024)
LEAD: Minimizing Learner-Expert Asymmetry in End-to-End Driving
por: Nguyen, Long, et al.
Publicado: (2025)
por: Nguyen, Long, et al.
Publicado: (2025)
Centaur: Robust End-to-End Autonomous Driving with Test-Time Training
por: Sima, Chonghao, et al.
Publicado: (2025)
por: Sima, Chonghao, et al.
Publicado: (2025)
DreamDrive: Generative 4D Scene Modeling from Street View Images
por: Mao, Jiageng, et al.
Publicado: (2024)
por: Mao, Jiageng, et al.
Publicado: (2024)
End-to-end Autonomous Driving: Challenges and Frontiers
por: Chen, Li, et al.
Publicado: (2023)
por: Chen, Li, et al.
Publicado: (2023)
Enhancing End-to-End Autonomous Driving with Latent World Model
por: Li, Yingyan, et al.
Publicado: (2024)
por: Li, Yingyan, et al.
Publicado: (2024)
DTPP: Differentiable Joint Conditional Prediction and Cost Evaluation for Tree Policy Planning in Autonomous Driving
por: Huang, Zhiyu, et al.
Publicado: (2023)
por: Huang, Zhiyu, et al.
Publicado: (2023)
Counterfactual VLA: Self-Reflective Vision-Language-Action Model with Adaptive Reasoning
por: Peng, Zhenghao "Mark", et al.
Publicado: (2025)
por: Peng, Zhenghao "Mark", et al.
Publicado: (2025)
Long-term Traffic Simulation with Interleaved Autoregressive Motion and Scenario Generation
por: Yang, Xiuyu, et al.
Publicado: (2025)
por: Yang, Xiuyu, et al.
Publicado: (2025)
RealDrive: Retrieval-Augmented Driving with Diffusion Models
por: Ding, Wenhao, et al.
Publicado: (2025)
por: Ding, Wenhao, et al.
Publicado: (2025)
Latent-WAM: Latent World Action Modeling for End-to-End Autonomous Driving
por: Wang, Linbo, et al.
Publicado: (2026)
por: Wang, Linbo, et al.
Publicado: (2026)
Pseudo-Simulation for Autonomous Driving
por: Cao, Wei, et al.
Publicado: (2025)
por: Cao, Wei, et al.
Publicado: (2025)
E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models
por: Cong, Wenyan, et al.
Publicado: (2025)
por: Cong, Wenyan, et al.
Publicado: (2025)
dVLM-AD: Enhance Diffusion Vision-Language-Model for Driving via Controllable Reasoning
por: Ma, Yingzi, et al.
Publicado: (2025)
por: Ma, Yingzi, et al.
Publicado: (2025)
World4Drive: End-to-End Autonomous Driving via Intention-aware Physical Latent World Model
por: Zheng, Yupeng, et al.
Publicado: (2025)
por: Zheng, Yupeng, et al.
Publicado: (2025)
Interactive Post-Training for Vision-Language-Action Models
por: Tan, Shuhan, et al.
Publicado: (2025)
por: Tan, Shuhan, et al.
Publicado: (2025)
RoaD: Rollouts as Demonstrations for Closed-Loop Supervised Fine-Tuning of Autonomous Driving Policies
por: Garcia-Cobo, Guillermo, et al.
Publicado: (2025)
por: Garcia-Cobo, Guillermo, et al.
Publicado: (2025)
Chain-of-Thought Reasoning in Streaming Full-Duplex End-to-End Spoken Dialogue Systems
por: Arora, Siddhant, et al.
Publicado: (2025)
por: Arora, Siddhant, et al.
Publicado: (2025)
Sample Complexity of Autoregressive Reasoning: Chain-of-Thought vs. End-to-End
por: Hanneke, Steve, et al.
Publicado: (2026)
por: Hanneke, Steve, et al.
Publicado: (2026)
Closed-Loop Supervised Fine-Tuning of Tokenized Traffic Models
por: Zhang, Zhejun, et al.
Publicado: (2024)
por: Zhang, Zhejun, et al.
Publicado: (2024)
DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving
por: Jia, Xiaosong, et al.
Publicado: (2025)
por: Jia, Xiaosong, et al.
Publicado: (2025)
ExploreVLA: Dense World Modeling and Exploration for End-to-End Autonomous Driving
por: Sheng, Zihao, et al.
Publicado: (2026)
por: Sheng, Zihao, et al.
Publicado: (2026)
Driving Everywhere with Large Language Model Policy Adaptation
por: Li, Boyi, et al.
Publicado: (2024)
por: Li, Boyi, et al.
Publicado: (2024)
MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving
por: Yasarla, Rajeev, et al.
Publicado: (2026)
por: Yasarla, Rajeev, et al.
Publicado: (2026)
Surprise Potential as a Measure of Interactivity in Driving Scenarios
por: Ding, Wenhao, et al.
Publicado: (2025)
por: Ding, Wenhao, et al.
Publicado: (2025)
DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving
por: Zhang, Lingjun, et al.
Publicado: (2026)
por: Zhang, Lingjun, et al.
Publicado: (2026)
SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic
por: Chitta, Kashyap, et al.
Publicado: (2024)
por: Chitta, Kashyap, et al.
Publicado: (2024)
End-to-End Chart Summarization via Visual Chain-of-Thought in Vision-Language Models
por: Choi, Raymond, et al.
Publicado: (2025)
por: Choi, Raymond, et al.
Publicado: (2025)
ResWorld: Temporal Residual World Model for End-to-End Autonomous Driving
por: Zhang, Jinqing, et al.
Publicado: (2026)
por: Zhang, Jinqing, et al.
Publicado: (2026)
Large Spatial Model: End-to-end Unposed Images to Semantic 3D
por: Fan, Zhiwen, et al.
Publicado: (2024)
por: Fan, Zhiwen, et al.
Publicado: (2024)
Ejemplares similares
-
Efficient Multi-Camera Tokenization with Triplanes for End-to-End Driving
por: Ivanovic, Boris, et al.
Publicado: (2025) -
Promptable Closed-loop Traffic Simulation
por: Tan, Shuhan, et al.
Publicado: (2024) -
Towards Efficient and Effective Multi-Camera Encoding for End-to-End Driving
por: Yang, Jiawei, et al.
Publicado: (2025) -
Accelerating Structured Chain-of-Thought in Autonomous Vehicles
por: Gu, Yi, et al.
Publicado: (2026) -
Hidden Biases of End-to-End Driving Datasets
por: Zimmerlin, Julian, et al.
Publicado: (2024)