Robust Autonomy Emerges from Self-Play
Fuente:
arXiv
Saved in:
| Main Authors: | Cusumano-Towner, Marco, Hafner, David, Hertzberg, Alex, Huval, Brody, Petrenko, Aleksei, Vinitsky, Eugene, Wijmans, Erik, Killian, Taylor, Bowers, Stuart, Sener, Ozan, Krähenbühl, Philipp, Koltun, Vladlen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cut Your Losses in Large-Vocabulary Language Models
by: Wijmans, Erik, et al.
Published: (2024)
by: Wijmans, Erik, et al.
Published: (2024)
Reinforcement Learning for Long-Horizon Interactive LLM Agents
by: Chen, Kevin, et al.
Published: (2025)
by: Chen, Kevin, et al.
Published: (2025)
Entropy-Preserving Reinforcement Learning
by: Petrenko, Aleksei, et al.
Published: (2026)
by: Petrenko, Aleksei, et al.
Published: (2026)
Does Spatial Cognition Emerge in Frontier Models?
by: Ramakrishnan, Santhosh Kumar, et al.
Published: (2024)
by: Ramakrishnan, Santhosh Kumar, et al.
Published: (2024)
Do multimodal models imagine electric sheep?
by: Ramakrishnan, Santhosh Kumar, et al.
Published: (2026)
by: Ramakrishnan, Santhosh Kumar, et al.
Published: (2026)
Conformation Generation using Transformer Flows
by: Shah, Sohil Atul, et al.
Published: (2024)
by: Shah, Sohil Atul, et al.
Published: (2024)
OpenBot-Fleet: A System for Collective Learning with Real Robots
by: Müller, Matthias, et al.
Published: (2024)
by: Müller, Matthias, et al.
Published: (2024)
The Right to Regulate in International Economic Law
by: Petrenko, Aleksei
Published: (2025)
by: Petrenko, Aleksei
Published: (2025)
Human-compatible driving partners through data-regularized self-play reinforcement learning
by: Cornelisse, Daphne, et al.
Published: (2024)
by: Cornelisse, Daphne, et al.
Published: (2024)
Depth Pro: Sharp Monocular Metric Depth in Less Than a Second
by: Bochkovskii, Aleksei, et al.
Published: (2024)
by: Bochkovskii, Aleksei, et al.
Published: (2024)
Superhuman AI for Stratego Using Self-Play Reinforcement Learning and Test-Time Search
by: Sokota, Samuel, et al.
Published: (2025)
by: Sokota, Samuel, et al.
Published: (2025)
CoMotion: Concurrent Multi-person 3D Motion
by: Newell, Alejandro, et al.
Published: (2025)
by: Newell, Alejandro, et al.
Published: (2025)
Learning to Evict from Key-Value Cache
by: Moschella, Luca, et al.
Published: (2026)
by: Moschella, Luca, et al.
Published: (2026)
Compressed Map Priors for 3D Perception
by: Zhou, Brady, et al.
Published: (2025)
by: Zhou, Brady, et al.
Published: (2025)
Матеріали для лекційних занять з курсу "АНТРОПОГЕННЕ ПЕРЕТВОРЕННЯ РЕЛЬЄФУ І ГЕОЗАГРОЗИ"
by: Koltun, Oksana
Published: (2023)
by: Koltun, Oksana
Published: (2023)
Video Game Level Design as a Multi-Agent Reinforcement Learning Problem
by: Earle, Sam, et al.
Published: (2025)
by: Earle, Sam, et al.
Published: (2025)
GEOMETRIA ÓSSEA E ATIVIDADE FÍSICA EM CRIANÇAS E ADOLESCENTES: REVISÃO SISTEMÁTICA
by: Tathyane Krahenbühl
Published: (2018)
by: Tathyane Krahenbühl
Published: (2018)
The use of the additional field player in handball: analysis of the Rio 2016 Olympic Games
by: Tathyane Krahenbühl
Published: (2019)
by: Tathyane Krahenbühl
Published: (2019)
Fatores que influenciam a massa óssea de crianças e adolescentes saudáveis mensurada pelo ultrassom quantitativo de falanges: revisão sistemática
by: Tathyane Krahenbühl
Published: (2014)
by: Tathyane Krahenbühl
Published: (2014)
Domain Adaptation Through Task Distillation
by: Zhou, Brady, et al.
Published: (2020)
by: Zhou, Brady, et al.
Published: (2020)
Image and Video Tokenization with Binary Spherical Quantization
by: Zhao, Yue, et al.
Published: (2024)
by: Zhao, Yue, et al.
Published: (2024)
Long-term Traffic Simulation with Interleaved Autoregressive Motion and Scenario Generation
by: Yang, Xiuyu, et al.
Published: (2025)
by: Yang, Xiuyu, et al.
Published: (2025)
Beyond Self-Play and Scale: A Behavior Benchmark for Generalization in Autonomous Driving
by: Distelzweig, Aron, et al.
Published: (2026)
by: Distelzweig, Aron, et al.
Published: (2026)
An End to Innocence.
by: Towner, Lawrence W.
Published: (1988)
by: Towner, Lawrence W.
Published: (1988)
Predictors of Mid-Term AVNeo Insufficiency
by: Vladlen Bazylev
Published: (2023)
by: Vladlen Bazylev
Published: (2023)
Logos and Life: Essays on Mind, Action, Language and Ethics By RogerTeichmann, Anthem Press. 2025. pp. 234
by: Lars Hertzberg
Published: (2025)
by: Lars Hertzberg
Published: (2025)
GPUDrive: Data-driven, multi-agent driving simulation at 1 million FPS
by: Kazemkhani, Saman, et al.
Published: (2024)
by: Kazemkhani, Saman, et al.
Published: (2024)
Building reliable sim driving agents by scaling self-play
by: Cornelisse, Daphne, et al.
Published: (2025)
by: Cornelisse, Daphne, et al.
Published: (2025)
Dry ice sublimation: A computational study with experimental validation for the effects of geometry
by: Ferruh Erdogdu, et al.
Published: (2025)
by: Ferruh Erdogdu, et al.
Published: (2025)
Systemic lupus erythematosus and atherosclerosis progression risk: comment on the article by Papazoglou et al.
by: Yusuf Ziya Sener, et al.
Published: (2025)
by: Yusuf Ziya Sener, et al.
Published: (2025)
Key Points in the Association of Rheumatoid Arthritis With Major Adverse Cardiovascular Events and Malignancies
by: Seher Sener, et al.
Published: (2025)
by: Seher Sener, et al.
Published: (2025)
Rural Health Services Funding: A Resource Guide. Revised Edition. Rural Information Center Publication Series, No. 41.
by: Towner, Sarah R., Comp.
Published: (1995)
by: Towner, Sarah R., Comp.
Published: (1995)
Interactive Post-Training for Vision-Language-Action Models
by: Tan, Shuhan, et al.
Published: (2025)
by: Tan, Shuhan, et al.
Published: (2025)
A New Society Emerges in Anatolia: Bioarcheological Perspectives on the Late Neolithic and Early Chalcolithic Population of the Gökhöyük (GH)
by: Elçin Şener, et al.
Published: (2026)
by: Elçin Şener, et al.
Published: (2026)
Decentralized Aerial Manipulation of a Cable-Suspended Load using Multi-Agent Reinforcement Learning
by: Zeng, Jack, et al.
Published: (2025)
by: Zeng, Jack, et al.
Published: (2025)
The Attribution of the Children’s Book Publishing Statistics in Ukraine in the Second Half of the 19th and Beginning of the 20th Centuries
by: Oksana Petrenko
Published: (2020)
by: Oksana Petrenko
Published: (2020)
Attention to Order: Transformers Discover Phase Transitions via Learnability
by: Özönder, Şener
Published: (2025)
by: Özönder, Şener
Published: (2025)
Can Access to Health Services and Universal Health Coverage Improve the Efficiency of Health Systems in Sub‐Saharan African Countries? A Study Based on a Two‐Stage Dynamic Data Envelopment Analysis (DEA) Model
by: Mehmet Şener
Published: (2025)
by: Mehmet Şener
Published: (2025)
Sarcopenia en pacientes con y sin insuficiencia renal crónica: diagnóstico, evaluación y tratamiento
by: Ana María Cusumano
Published: (2015)
by: Ana María Cusumano
Published: (2015)
Congreso Mundial de Nefrología 2009
by: Ana María Cusumano
Published: (2009)
by: Ana María Cusumano
Published: (2009)
Similar Items
-
Cut Your Losses in Large-Vocabulary Language Models
by: Wijmans, Erik, et al.
Published: (2024) -
Reinforcement Learning for Long-Horizon Interactive LLM Agents
by: Chen, Kevin, et al.
Published: (2025) -
Entropy-Preserving Reinforcement Learning
by: Petrenko, Aleksei, et al.
Published: (2026) -
Does Spatial Cognition Emerge in Frontier Models?
by: Ramakrishnan, Santhosh Kumar, et al.
Published: (2024) -
Do multimodal models imagine electric sheep?
by: Ramakrishnan, Santhosh Kumar, et al.
Published: (2026)