Craftax: A Lightning-Fast Benchmark for Open-Ended Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Matthews, Michael, Beukman, Michael, Ellis, Benjamin, Samvelyan, Mikayel, Jackson, Matthew, Coward, Samuel, Foerster, Jakob |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-Agent Craftax: Benchmarking Open-Ended Multi-Agent Reinforcement Learning at the Hyperscale
von: Omari, Bassel Al, et al.
Veröffentlicht: (2025)
von: Omari, Bassel Al, et al.
Veröffentlicht: (2025)
Robust Agents in Open-Ended Worlds
von: Samvelyan, Mikayel
Veröffentlicht: (2025)
von: Samvelyan, Mikayel
Veröffentlicht: (2025)
JaxUED: A simple and useable UED library in Jax
von: Coward, Samuel, et al.
Veröffentlicht: (2024)
von: Coward, Samuel, et al.
Veröffentlicht: (2024)
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
von: Matthews, Michael, et al.
Veröffentlicht: (2024)
von: Matthews, Michael, et al.
Veröffentlicht: (2024)
Refining Minimax Regret for Unsupervised Environment Design
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
JaxLife: An Open-Ended Agentic Simulator
von: Lu, Chris, et al.
Veröffentlicht: (2024)
von: Lu, Chris, et al.
Veröffentlicht: (2024)
DéjàQ: Open-Ended Evolution of Diverse, Learnable and Verifiable Problems
von: Röpke, Willem, et al.
Veröffentlicht: (2026)
von: Röpke, Willem, et al.
Veröffentlicht: (2026)
Goal-Conditioned Agents that Learn Everything All at Once
von: Matthews, Michael, et al.
Veröffentlicht: (2026)
von: Matthews, Michael, et al.
Veröffentlicht: (2026)
Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts
von: Samvelyan, Mikayel, et al.
Veröffentlicht: (2024)
von: Samvelyan, Mikayel, et al.
Veröffentlicht: (2024)
Policy-Guided Diffusion
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2024)
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2024)
SPARQ: Synthetic Problem Generation for Reasoning via Quality-Diversity Algorithms
von: Havrilla, Alex, et al.
Veröffentlicht: (2025)
von: Havrilla, Alex, et al.
Veröffentlicht: (2025)
Preventing Learning Stagnation in PPO by Scaling to 1 Million Parallel Environments
von: Beukman, Michael, et al.
Veröffentlicht: (2026)
von: Beukman, Michael, et al.
Veröffentlicht: (2026)
An Optimisation Framework for Unsupervised Environment Design
von: Monette, Nathan, et al.
Veröffentlicht: (2025)
von: Monette, Nathan, et al.
Veröffentlicht: (2025)
High entropy leads to symmetry equivariant policies in Dec-POMDPs
von: Forkel, Johannes, et al.
Veröffentlicht: (2025)
von: Forkel, Johannes, et al.
Veröffentlicht: (2025)
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
von: Rutherford, Alexander, et al.
Veröffentlicht: (2024)
von: Rutherford, Alexander, et al.
Veröffentlicht: (2024)
A Clean Slate for Offline Reinforcement Learning
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2025)
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2025)
Multi-Agent Diagnostics for Robustness via Illuminated Diversity
von: Samvelyan, Mikayel, et al.
Veröffentlicht: (2024)
von: Samvelyan, Mikayel, et al.
Veröffentlicht: (2024)
Can Learned Optimization Make Reinforcement Learning Less Difficult?
von: Goldie, Alexander David, et al.
Veröffentlicht: (2024)
von: Goldie, Alexander David, et al.
Veröffentlicht: (2024)
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
von: Lu, Chris, et al.
Veröffentlicht: (2024)
von: Lu, Chris, et al.
Veröffentlicht: (2024)
Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps
von: Ellis, Benjamin, et al.
Veröffentlicht: (2024)
von: Ellis, Benjamin, et al.
Veröffentlicht: (2024)
Discovering Temporally-Aware Reinforcement Learning Algorithms
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2024)
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2024)
Helix: Evolutionary Reinforcement Learning for Open-Ended Scientific Problem Solving
von: Su, Chang, et al.
Veröffentlicht: (2026)
von: Su, Chang, et al.
Veröffentlicht: (2026)
Simplifying Deep Temporal Difference Learning
von: Gallici, Matteo, et al.
Veröffentlicht: (2024)
von: Gallici, Matteo, et al.
Veröffentlicht: (2024)
Focusing Robot Open-Ended Reinforcement Learning Through Users' Purposes
von: Cartoni, Emilio, et al.
Veröffentlicht: (2025)
von: Cartoni, Emilio, et al.
Veröffentlicht: (2025)
Learning Multi-Agent Communication with Contrastive Learning
von: Lo, Yat Long, et al.
Veröffentlicht: (2023)
von: Lo, Yat Long, et al.
Veröffentlicht: (2023)
RobocupGym: A challenging continuous control benchmark in Robocup
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
Beyond the Boundaries of Proximal Policy Optimization
von: Tan, Charlie B., et al.
Veröffentlicht: (2024)
von: Tan, Charlie B., et al.
Veröffentlicht: (2024)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
von: Sims, Anya, et al.
Veröffentlicht: (2024)
von: Sims, Anya, et al.
Veröffentlicht: (2024)
Abstraction for Offline Goal-Conditioned Reinforcement Learning
von: Wibault, Clarisse, et al.
Veröffentlicht: (2026)
von: Wibault, Clarisse, et al.
Veröffentlicht: (2026)
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
von: Barde, Paul, et al.
Veröffentlicht: (2023)
von: Barde, Paul, et al.
Veröffentlicht: (2023)
Discovering Minimal Reinforcement Learning Environments
von: Liesen, Jarek, et al.
Veröffentlicht: (2024)
von: Liesen, Jarek, et al.
Veröffentlicht: (2024)
SOReL and TOReL: Two Methods for Fully Offline Reinforcement Learning
von: Fellows, Mattie, et al.
Veröffentlicht: (2025)
von: Fellows, Mattie, et al.
Veröffentlicht: (2025)
Recurrent Reinforcement Learning with Memoroids
von: Morad, Steven, et al.
Veröffentlicht: (2024)
von: Morad, Steven, et al.
Veröffentlicht: (2024)
Self-Rewarding Rubric-Based Reinforcement Learning for Open-Ended Reasoning
von: Ye, Zhiling, et al.
Veröffentlicht: (2025)
von: Ye, Zhiling, et al.
Veröffentlicht: (2025)
Hierarchical Behaviour Spaces
von: Matthews, Michael Tryfan, et al.
Veröffentlicht: (2026)
von: Matthews, Michael Tryfan, et al.
Veröffentlicht: (2026)
How Should We Meta-Learn Reinforcement Learning Algorithms?
von: Goldie, Alexander David, et al.
Veröffentlicht: (2025)
von: Goldie, Alexander David, et al.
Veröffentlicht: (2025)
Autotelic Reinforcement Learning: Exploring Intrinsic Motivations for Skill Acquisition in Open-Ended Environments
von: Srivastava, Prakhar, et al.
Veröffentlicht: (2025)
von: Srivastava, Prakhar, et al.
Veröffentlicht: (2025)
Open-Endedness is Essential for Artificial Superhuman Intelligence
von: Hughes, Edward, et al.
Veröffentlicht: (2024)
von: Hughes, Edward, et al.
Veröffentlicht: (2024)
GRLO: Towards Generalizable Reinforcement Learning in Open-Ended Environments from Zero
von: Yin, Shangjian, et al.
Veröffentlicht: (2026)
von: Yin, Shangjian, et al.
Veröffentlicht: (2026)
A Motivational Architecture for Open-Ended Learning Challenges in Robots
von: Romero, Alejandro, et al.
Veröffentlicht: (2025)
von: Romero, Alejandro, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multi-Agent Craftax: Benchmarking Open-Ended Multi-Agent Reinforcement Learning at the Hyperscale
von: Omari, Bassel Al, et al.
Veröffentlicht: (2025) -
Robust Agents in Open-Ended Worlds
von: Samvelyan, Mikayel
Veröffentlicht: (2025) -
JaxUED: A simple and useable UED library in Jax
von: Coward, Samuel, et al.
Veröffentlicht: (2024) -
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
von: Matthews, Michael, et al.
Veröffentlicht: (2024) -
Refining Minimax Regret for Unsupervised Environment Design
von: Beukman, Michael, et al.
Veröffentlicht: (2024)