An Optimisation Framework for Unsupervised Environment Design
Fuente:
arXiv
Saved in:
| Main Authors: | Monette, Nathan, Letcher, Alistair, Beukman, Michael, Jackson, Matthew T., Rutherford, Alexander, Goldie, Alexander D., Foerster, Jakob N. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Goal-Conditioned Agents that Learn Everything All at Once
by: Matthews, Michael, et al.
Published: (2026)
by: Matthews, Michael, et al.
Published: (2026)
Refining Minimax Regret for Unsupervised Environment Design
by: Beukman, Michael, et al.
Published: (2024)
by: Beukman, Michael, et al.
Published: (2024)
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
by: Rutherford, Alexander, et al.
Published: (2024)
by: Rutherford, Alexander, et al.
Published: (2024)
JaxUED: A simple and useable UED library in Jax
by: Coward, Samuel, et al.
Published: (2024)
by: Coward, Samuel, et al.
Published: (2024)
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
by: Matthews, Michael, et al.
Published: (2024)
by: Matthews, Michael, et al.
Published: (2024)
Craftax: A Lightning-Fast Benchmark for Open-Ended Reinforcement Learning
by: Matthews, Michael, et al.
Published: (2024)
by: Matthews, Michael, et al.
Published: (2024)
Preventing Learning Stagnation in PPO by Scaling to 1 Million Parallel Environments
by: Beukman, Michael, et al.
Published: (2026)
by: Beukman, Michael, et al.
Published: (2026)
Can Learned Optimization Make Reinforcement Learning Less Difficult?
by: Goldie, Alexander David, et al.
Published: (2024)
by: Goldie, Alexander David, et al.
Published: (2024)
Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps
by: Ellis, Benjamin, et al.
Published: (2024)
by: Ellis, Benjamin, et al.
Published: (2024)
Multi-Agent Craftax: Benchmarking Open-Ended Multi-Agent Reinforcement Learning at the Hyperscale
by: Omari, Bassel Al, et al.
Published: (2025)
by: Omari, Bassel Al, et al.
Published: (2025)
High entropy leads to symmetry equivariant policies in Dec-POMDPs
by: Forkel, Johannes, et al.
Published: (2025)
by: Forkel, Johannes, et al.
Published: (2025)
Abstraction for Offline Goal-Conditioned Reinforcement Learning
by: Wibault, Clarisse, et al.
Published: (2026)
by: Wibault, Clarisse, et al.
Published: (2026)
Improving Regret Approximation for Unsupervised Dynamic Environment Generation
by: Mead, Harry, et al.
Published: (2026)
by: Mead, Harry, et al.
Published: (2026)
Mirror Learning: A Unifying Framework of Policy Optimisation
by: Kuba, Jakub Grudzien, et al.
Published: (2022)
by: Kuba, Jakub Grudzien, et al.
Published: (2022)
How Should We Meta-Learn Reinforcement Learning Algorithms?
by: Goldie, Alexander David, et al.
Published: (2025)
by: Goldie, Alexander David, et al.
Published: (2025)
Procedural Generation of Algorithm Discovery Tasks in Machine Learning
by: Goldie, Alexander D., et al.
Published: (2026)
by: Goldie, Alexander D., et al.
Published: (2026)
Learning to Drive in New Cities Without Human Demonstrations
by: Wang, Zilin, et al.
Published: (2026)
by: Wang, Zilin, et al.
Published: (2026)
Discovering Minimal Reinforcement Learning Environments
by: Liesen, Jarek, et al.
Published: (2024)
by: Liesen, Jarek, et al.
Published: (2024)
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
by: Xiong, Zheng, et al.
Published: (2025)
by: Xiong, Zheng, et al.
Published: (2025)
A Clean Slate for Offline Reinforcement Learning
by: Jackson, Matthew Thomas, et al.
Published: (2025)
by: Jackson, Matthew Thomas, et al.
Published: (2025)
Policy-Guided Diffusion
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games
by: Lanier, JB, et al.
Published: (2026)
by: Lanier, JB, et al.
Published: (2026)
Improving Environment Novelty Quantification for Effective Unsupervised Environment Design
by: Teoh, Jayden, et al.
Published: (2025)
by: Teoh, Jayden, et al.
Published: (2025)
Gradient-based Optimisation of Modulation Effects
by: Carson, Alistair, et al.
Published: (2026)
by: Carson, Alistair, et al.
Published: (2026)
Evolution Strategies at the Hyperscale
by: Sarkar, Bidipta, et al.
Published: (2025)
by: Sarkar, Bidipta, et al.
Published: (2025)
PACE: Parameter Change for Unsupervised Environment Design
by: Yuan, Fang, et al.
Published: (2026)
by: Yuan, Fang, et al.
Published: (2026)
JaxLife: An Open-Ended Agentic Simulator
by: Lu, Chris, et al.
Published: (2024)
by: Lu, Chris, et al.
Published: (2024)
The Yokai Learning Environment: Tracking Beliefs Over Space and Time
by: Ruhdorfer, Constantin, et al.
Published: (2025)
by: Ruhdorfer, Constantin, et al.
Published: (2025)
Discovering Temporally-Aware Reinforcement Learning Algorithms
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
Unsupervised Optimisation of GNNs for Node Clustering
by: Leeney, William, et al.
Published: (2024)
by: Leeney, William, et al.
Published: (2024)
Learning Multi-Agent Communication with Contrastive Learning
by: Lo, Yat Long, et al.
Published: (2023)
by: Lo, Yat Long, et al.
Published: (2023)
Beyond Fixed Tasks: Unsupervised Environment Design for Task-Level Pairs
by: Furelos-Blanco, Daniel, et al.
Published: (2025)
by: Furelos-Blanco, Daniel, et al.
Published: (2025)
That Chip Has Sailed: A Critique of Unfounded Skepticism Around AI for Chip Design
by: Goldie, Anna, et al.
Published: (2024)
by: Goldie, Anna, et al.
Published: (2024)
RobocupGym: A challenging continuous control benchmark in Robocup
by: Beukman, Michael, et al.
Published: (2024)
by: Beukman, Michael, et al.
Published: (2024)
Towards Agent-based Test Support Systems: An Unsupervised Environment Design Approach
by: Ogbodo, Collins O., et al.
Published: (2025)
by: Ogbodo, Collins O., et al.
Published: (2025)
Imagined Autocurricula
by: Güzel, Ahmet H., et al.
Published: (2025)
by: Güzel, Ahmet H., et al.
Published: (2025)
Noisy Zero-Shot Coordination: Breaking The Common Knowledge Assumption In Zero-Shot Coordination Games
by: Anwar, Usman, et al.
Published: (2024)
by: Anwar, Usman, et al.
Published: (2024)
SOReL and TOReL: Two Methods for Fully Offline Reinforcement Learning
by: Fellows, Mattie, et al.
Published: (2025)
by: Fellows, Mattie, et al.
Published: (2025)
Comparison of Optimised Geometric Deep Learning Architectures, over Varying Toxicological Assay Data Environments
by: Kalian, Alexander D., et al.
Published: (2025)
by: Kalian, Alexander D., et al.
Published: (2025)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
by: Sims, Anya, et al.
Published: (2024)
by: Sims, Anya, et al.
Published: (2024)
Similar Items
-
Goal-Conditioned Agents that Learn Everything All at Once
by: Matthews, Michael, et al.
Published: (2026) -
Refining Minimax Regret for Unsupervised Environment Design
by: Beukman, Michael, et al.
Published: (2024) -
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
by: Rutherford, Alexander, et al.
Published: (2024) -
JaxUED: A simple and useable UED library in Jax
by: Coward, Samuel, et al.
Published: (2024) -
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
by: Matthews, Michael, et al.
Published: (2024)