Robust Autonomy Emerges from Self-Play
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Cusumano-Towner, Marco, Hafner, David, Hertzberg, Alex, Huval, Brody, Petrenko, Aleksei, Vinitsky, Eugene, Wijmans, Erik, Killian, Taylor, Bowers, Stuart, Sener, Ozan, Krähenbühl, Philipp, Koltun, Vladlen |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Cut Your Losses in Large-Vocabulary Language Models
par: Wijmans, Erik, et autres
Publié: (2024)
par: Wijmans, Erik, et autres
Publié: (2024)
Reinforcement Learning for Long-Horizon Interactive LLM Agents
par: Chen, Kevin, et autres
Publié: (2025)
par: Chen, Kevin, et autres
Publié: (2025)
Entropy-Preserving Reinforcement Learning
par: Petrenko, Aleksei, et autres
Publié: (2026)
par: Petrenko, Aleksei, et autres
Publié: (2026)
Does Spatial Cognition Emerge in Frontier Models?
par: Ramakrishnan, Santhosh Kumar, et autres
Publié: (2024)
par: Ramakrishnan, Santhosh Kumar, et autres
Publié: (2024)
Do multimodal models imagine electric sheep?
par: Ramakrishnan, Santhosh Kumar, et autres
Publié: (2026)
par: Ramakrishnan, Santhosh Kumar, et autres
Publié: (2026)
Conformation Generation using Transformer Flows
par: Shah, Sohil Atul, et autres
Publié: (2024)
par: Shah, Sohil Atul, et autres
Publié: (2024)
OpenBot-Fleet: A System for Collective Learning with Real Robots
par: Müller, Matthias, et autres
Publié: (2024)
par: Müller, Matthias, et autres
Publié: (2024)
The Right to Regulate in International Economic Law
par: Petrenko, Aleksei
Publié: (2025)
par: Petrenko, Aleksei
Publié: (2025)
Human-compatible driving partners through data-regularized self-play reinforcement learning
par: Cornelisse, Daphne, et autres
Publié: (2024)
par: Cornelisse, Daphne, et autres
Publié: (2024)
Depth Pro: Sharp Monocular Metric Depth in Less Than a Second
par: Bochkovskii, Aleksei, et autres
Publié: (2024)
par: Bochkovskii, Aleksei, et autres
Publié: (2024)
Superhuman AI for Stratego Using Self-Play Reinforcement Learning and Test-Time Search
par: Sokota, Samuel, et autres
Publié: (2025)
par: Sokota, Samuel, et autres
Publié: (2025)
CoMotion: Concurrent Multi-person 3D Motion
par: Newell, Alejandro, et autres
Publié: (2025)
par: Newell, Alejandro, et autres
Publié: (2025)
Learning to Evict from Key-Value Cache
par: Moschella, Luca, et autres
Publié: (2026)
par: Moschella, Luca, et autres
Publié: (2026)
Compressed Map Priors for 3D Perception
par: Zhou, Brady, et autres
Publié: (2025)
par: Zhou, Brady, et autres
Publié: (2025)
Матеріали для лекційних занять з курсу "АНТРОПОГЕННЕ ПЕРЕТВОРЕННЯ РЕЛЬЄФУ І ГЕОЗАГРОЗИ"
par: Koltun, Oksana
Publié: (2023)
par: Koltun, Oksana
Publié: (2023)
Video Game Level Design as a Multi-Agent Reinforcement Learning Problem
par: Earle, Sam, et autres
Publié: (2025)
par: Earle, Sam, et autres
Publié: (2025)
GEOMETRIA ÓSSEA E ATIVIDADE FÍSICA EM CRIANÇAS E ADOLESCENTES: REVISÃO SISTEMÁTICA
par: Tathyane Krahenbühl
Publié: (2018)
par: Tathyane Krahenbühl
Publié: (2018)
The use of the additional field player in handball: analysis of the Rio 2016 Olympic Games
par: Tathyane Krahenbühl
Publié: (2019)
par: Tathyane Krahenbühl
Publié: (2019)
Fatores que influenciam a massa óssea de crianças e adolescentes saudáveis mensurada pelo ultrassom quantitativo de falanges: revisão sistemática
par: Tathyane Krahenbühl
Publié: (2014)
par: Tathyane Krahenbühl
Publié: (2014)
Domain Adaptation Through Task Distillation
par: Zhou, Brady, et autres
Publié: (2020)
par: Zhou, Brady, et autres
Publié: (2020)
Image and Video Tokenization with Binary Spherical Quantization
par: Zhao, Yue, et autres
Publié: (2024)
par: Zhao, Yue, et autres
Publié: (2024)
Long-term Traffic Simulation with Interleaved Autoregressive Motion and Scenario Generation
par: Yang, Xiuyu, et autres
Publié: (2025)
par: Yang, Xiuyu, et autres
Publié: (2025)
Beyond Self-Play and Scale: A Behavior Benchmark for Generalization in Autonomous Driving
par: Distelzweig, Aron, et autres
Publié: (2026)
par: Distelzweig, Aron, et autres
Publié: (2026)
An End to Innocence.
par: Towner, Lawrence W.
Publié: (1988)
par: Towner, Lawrence W.
Publié: (1988)
Predictors of Mid-Term AVNeo Insufficiency
par: Vladlen Bazylev
Publié: (2023)
par: Vladlen Bazylev
Publié: (2023)
Logos and Life: Essays on Mind, Action, Language and Ethics By RogerTeichmann, Anthem Press. 2025. pp. 234
par: Lars Hertzberg
Publié: (2025)
par: Lars Hertzberg
Publié: (2025)
GPUDrive: Data-driven, multi-agent driving simulation at 1 million FPS
par: Kazemkhani, Saman, et autres
Publié: (2024)
par: Kazemkhani, Saman, et autres
Publié: (2024)
Building reliable sim driving agents by scaling self-play
par: Cornelisse, Daphne, et autres
Publié: (2025)
par: Cornelisse, Daphne, et autres
Publié: (2025)
Dry ice sublimation: A computational study with experimental validation for the effects of geometry
par: Ferruh Erdogdu, et autres
Publié: (2025)
par: Ferruh Erdogdu, et autres
Publié: (2025)
Systemic lupus erythematosus and atherosclerosis progression risk: comment on the article by Papazoglou et al.
par: Yusuf Ziya Sener, et autres
Publié: (2025)
par: Yusuf Ziya Sener, et autres
Publié: (2025)
Key Points in the Association of Rheumatoid Arthritis With Major Adverse Cardiovascular Events and Malignancies
par: Seher Sener, et autres
Publié: (2025)
par: Seher Sener, et autres
Publié: (2025)
Rural Health Services Funding: A Resource Guide. Revised Edition. Rural Information Center Publication Series, No. 41.
par: Towner, Sarah R., Comp.
Publié: (1995)
par: Towner, Sarah R., Comp.
Publié: (1995)
Interactive Post-Training for Vision-Language-Action Models
par: Tan, Shuhan, et autres
Publié: (2025)
par: Tan, Shuhan, et autres
Publié: (2025)
A New Society Emerges in Anatolia: Bioarcheological Perspectives on the Late Neolithic and Early Chalcolithic Population of the Gökhöyük (GH)
par: Elçin Şener, et autres
Publié: (2026)
par: Elçin Şener, et autres
Publié: (2026)
Decentralized Aerial Manipulation of a Cable-Suspended Load using Multi-Agent Reinforcement Learning
par: Zeng, Jack, et autres
Publié: (2025)
par: Zeng, Jack, et autres
Publié: (2025)
The Attribution of the Children’s Book Publishing Statistics in Ukraine in the Second Half of the 19th and Beginning of the 20th Centuries
par: Oksana Petrenko
Publié: (2020)
par: Oksana Petrenko
Publié: (2020)
Attention to Order: Transformers Discover Phase Transitions via Learnability
par: Özönder, Şener
Publié: (2025)
par: Özönder, Şener
Publié: (2025)
Can Access to Health Services and Universal Health Coverage Improve the Efficiency of Health Systems in Sub‐Saharan African Countries? A Study Based on a Two‐Stage Dynamic Data Envelopment Analysis (DEA) Model
par: Mehmet Şener
Publié: (2025)
par: Mehmet Şener
Publié: (2025)
Sarcopenia en pacientes con y sin insuficiencia renal crónica: diagnóstico, evaluación y tratamiento
par: Ana María Cusumano
Publié: (2015)
par: Ana María Cusumano
Publié: (2015)
Congreso Mundial de Nefrología 2009
par: Ana María Cusumano
Publié: (2009)
par: Ana María Cusumano
Publié: (2009)
Documents similaires
-
Cut Your Losses in Large-Vocabulary Language Models
par: Wijmans, Erik, et autres
Publié: (2024) -
Reinforcement Learning for Long-Horizon Interactive LLM Agents
par: Chen, Kevin, et autres
Publié: (2025) -
Entropy-Preserving Reinforcement Learning
par: Petrenko, Aleksei, et autres
Publié: (2026) -
Does Spatial Cognition Emerge in Frontier Models?
par: Ramakrishnan, Santhosh Kumar, et autres
Publié: (2024) -
Do multimodal models imagine electric sheep?
par: Ramakrishnan, Santhosh Kumar, et autres
Publié: (2026)