Procedural Generation of Algorithm Discovery Tasks in Machine Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Goldie, Alexander D., Wang, Zilin, Hayler, Adrian, Nathani, Deepak, Toledo, Edan, Thampiratwong, Ken, Kalisz, Aleksandra, Beukman, Michael, Letcher, Alistair, Reddy, Shashank, Wibault, Clarisse, Wolf, Theo, O'Neill, Charles, Berdica, Uljad, Roberts, Nicholas, Rahmani, Saeed, Erlebach, Hannah, Raileanu, Roberta, Whiteson, Shimon, Foerster, Jakob N. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Clean Slate for Offline Reinforcement Learning
by: Jackson, Matthew Thomas, et al.
Published: (2025)
by: Jackson, Matthew Thomas, et al.
Published: (2025)
SOReL and TOReL: Two Methods for Fully Offline Reinforcement Learning
by: Fellows, Mattie, et al.
Published: (2025)
by: Fellows, Mattie, et al.
Published: (2025)
Evolution Strategies at the Hyperscale
by: Sarkar, Bidipta, et al.
Published: (2025)
by: Sarkar, Bidipta, et al.
Published: (2025)
How Should We Meta-Learn Reinforcement Learning Algorithms?
by: Goldie, Alexander David, et al.
Published: (2025)
by: Goldie, Alexander David, et al.
Published: (2025)
Abstraction for Offline Goal-Conditioned Reinforcement Learning
by: Wibault, Clarisse, et al.
Published: (2026)
by: Wibault, Clarisse, et al.
Published: (2026)
Learning to Drive in New Cities Without Human Demonstrations
by: Wang, Zilin, et al.
Published: (2026)
by: Wang, Zilin, et al.
Published: (2026)
Evolving Many Worlds: Towards Open-Ended Discovery in Petri Dish NCA via Population-Based Training
by: Berdica, Uljad, et al.
Published: (2026)
by: Berdica, Uljad, et al.
Published: (2026)
An Optimisation Framework for Unsupervised Environment Design
by: Monette, Nathan, et al.
Published: (2025)
by: Monette, Nathan, et al.
Published: (2025)
Intent Factored Generation: Unleashing the Diversity in Your Language Model
by: Ahmed, Eltayeb, et al.
Published: (2025)
by: Ahmed, Eltayeb, et al.
Published: (2025)
Reinforcement Learning Controllers for Soft Robots using Learned Environments
by: Berdica, Uljad, et al.
Published: (2024)
by: Berdica, Uljad, et al.
Published: (2024)
Can Learned Optimization Make Reinforcement Learning Less Difficult?
by: Goldie, Alexander David, et al.
Published: (2024)
by: Goldie, Alexander David, et al.
Published: (2024)
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
by: Xiong, Zheng, et al.
Published: (2025)
by: Xiong, Zheng, et al.
Published: (2025)
Goal-Conditioned Agents that Learn Everything All at Once
by: Matthews, Michael, et al.
Published: (2026)
by: Matthews, Michael, et al.
Published: (2026)
Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps
by: Ellis, Benjamin, et al.
Published: (2024)
by: Ellis, Benjamin, et al.
Published: (2024)
Counterfactual Multi-Agent Policy Gradients
by: Foerster, Jakob, et al.
Published: (2017)
by: Foerster, Jakob, et al.
Published: (2017)
GoalLadder: Incremental Goal Discovery with Vision-Language Models
by: Zakharov, Alexey, et al.
Published: (2025)
by: Zakharov, Alexey, et al.
Published: (2025)
JaxUED: A simple and useable UED library in Jax
by: Coward, Samuel, et al.
Published: (2024)
by: Coward, Samuel, et al.
Published: (2024)
When Do We Need LLMs? A Diagnostic for Language-Driven Bandits
by: Berdica, Uljad, et al.
Published: (2026)
by: Berdica, Uljad, et al.
Published: (2026)
Recurrent Structural Policy Gradient for Partially Observable Mean Field Games
by: Wibault, Clarisse, et al.
Published: (2026)
by: Wibault, Clarisse, et al.
Published: (2026)
Inflation Forecasting Post‐COVID‐19: Evidence From Germany
by: Tiphaine Wibault
Published: (2026)
by: Tiphaine Wibault
Published: (2026)
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
by: Matthews, Michael, et al.
Published: (2024)
by: Matthews, Michael, et al.
Published: (2024)
JaxLife: An Open-Ended Agentic Simulator
by: Lu, Chris, et al.
Published: (2024)
by: Lu, Chris, et al.
Published: (2024)
Equivariant Networks for Zero-Shot Coordination
by: Muglich, Darius, et al.
Published: (2022)
by: Muglich, Darius, et al.
Published: (2022)
Policy-Guided Diffusion
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
Discovering Temporally-Aware Reinforcement Learning Algorithms
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
Tight and Efficient Gradient Bounds for Parameterized Quantum Circuits
by: Letcher, Alistair, et al.
Published: (2023)
by: Letcher, Alistair, et al.
Published: (2023)
IGDrivSim: A Benchmark for the Imitation Gap in Autonomous Driving
by: Grislain, Clémence, et al.
Published: (2024)
by: Grislain, Clémence, et al.
Published: (2024)
Rate-Informed Discovery via Bayesian Adaptive Multifidelity Sampling
by: Sinha, Aman, et al.
Published: (2024)
by: Sinha, Aman, et al.
Published: (2024)
High entropy leads to symmetry equivariant policies in Dec-POMDPs
by: Forkel, Johannes, et al.
Published: (2025)
by: Forkel, Johannes, et al.
Published: (2025)
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
by: Castellini, Jacopo, et al.
Published: (2019)
by: Castellini, Jacopo, et al.
Published: (2019)
Sparks of Science: Hypothesis Generation Using Structured Paper Data
by: O'Neill, Charles, et al.
Published: (2025)
by: O'Neill, Charles, et al.
Published: (2025)
SplAgger: Split Aggregation for Meta-Reinforcement Learning
by: Beck, Jacob, et al.
Published: (2024)
by: Beck, Jacob, et al.
Published: (2024)
Bayesian Exploration Networks
by: Fellows, Mattie, et al.
Published: (2023)
by: Fellows, Mattie, et al.
Published: (2023)
A Bayesian Solution To The Imitation Gap
by: Vuorio, Risto, et al.
Published: (2024)
by: Vuorio, Risto, et al.
Published: (2024)
Epistemic Dissonance and Modal Boundaries
by: Raileanu, Dragos
Published: (2025)
by: Raileanu, Dragos
Published: (2025)
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
by: Rutherford, Alexander, et al.
Published: (2024)
by: Rutherford, Alexander, et al.
Published: (2024)
A Comparative Study of Transfer Learning for Emotion Recognition using CNN and Modified VGG16 Models
by: Nathani, Samay
Published: (2024)
by: Nathani, Samay
Published: (2024)
Giant Trees of Western America and the World, by Al Carder [Review]
by: O'Neill, Jim
Published: (2006)
by: O'Neill, Jim
Published: (2006)
Vietnam Task
by: O'Neill, Robert
Published: (2023)
by: O'Neill, Robert
Published: (2023)
OPUS Stakeholders Consultation on Research Assessment Framework
by: O'Neill, Gareth
Published: (2025)
by: O'Neill, Gareth
Published: (2025)
Similar Items
-
A Clean Slate for Offline Reinforcement Learning
by: Jackson, Matthew Thomas, et al.
Published: (2025) -
SOReL and TOReL: Two Methods for Fully Offline Reinforcement Learning
by: Fellows, Mattie, et al.
Published: (2025) -
Evolution Strategies at the Hyperscale
by: Sarkar, Bidipta, et al.
Published: (2025) -
How Should We Meta-Learn Reinforcement Learning Algorithms?
by: Goldie, Alexander David, et al.
Published: (2025) -
Abstraction for Offline Goal-Conditioned Reinforcement Learning
by: Wibault, Clarisse, et al.
Published: (2026)