minimax: Efficient Baselines for Autocurricula in JAX
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Minqi, Dennis, Michael, Grefenstette, Edward, Rocktäschel, Tim |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Imagined Autocurricula
by: Güzel, Ahmet H., et al.
Published: (2025)
by: Güzel, Ahmet H., et al.
Published: (2025)
Multi-Agent Diagnostics for Robustness via Illuminated Diversity
by: Samvelyan, Mikayel, et al.
Published: (2024)
by: Samvelyan, Mikayel, et al.
Published: (2024)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
by: Rutherford, Alexander, et al.
Published: (2023)
by: Rutherford, Alexander, et al.
Published: (2023)
Open-Endedness is Essential for Artificial Superhuman Intelligence
by: Hughes, Edward, et al.
Published: (2024)
by: Hughes, Edward, et al.
Published: (2024)
Preference-Based Alignment of Discrete Diffusion Models
by: Borso, Umberto, et al.
Published: (2025)
by: Borso, Umberto, et al.
Published: (2025)
Learning to Act without Actions
by: Schmidt, Dominik, et al.
Published: (2023)
by: Schmidt, Dominik, et al.
Published: (2023)
Interaction Dynamics as a Reward Signal for LLMs
by: Gooding, Sian, et al.
Published: (2025)
by: Gooding, Sian, et al.
Published: (2025)
Investigating Non-Transitivity in LLM-as-a-Judge
by: Xu, Yi, et al.
Published: (2025)
by: Xu, Yi, et al.
Published: (2025)
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
by: Pignatelli, Eduardo, et al.
Published: (2024)
by: Pignatelli, Eduardo, et al.
Published: (2024)
Refining Minimax Regret for Unsupervised Environment Design
by: Beukman, Michael, et al.
Published: (2024)
by: Beukman, Michael, et al.
Published: (2024)
Linear bandits with polylogarithmic minimax regret
by: Lumbreras, Josep, et al.
Published: (2024)
by: Lumbreras, Josep, et al.
Published: (2024)
PuzzleJAX: A Benchmark for Reasoning and Learning
by: Earle, Sam, et al.
Published: (2025)
by: Earle, Sam, et al.
Published: (2025)
Outliers and Calibration Sets have Diminishing Effect on Quantization of Modern LLMs
by: Paglieri, Davide, et al.
Published: (2024)
by: Paglieri, Davide, et al.
Published: (2024)
DéjàQ: Open-Ended Evolution of Diverse, Learnable and Verifiable Problems
by: Röpke, Willem, et al.
Published: (2026)
by: Röpke, Willem, et al.
Published: (2026)
laplax -- Laplace Approximations with JAX
by: Weber, Tobias, et al.
Published: (2025)
by: Weber, Tobias, et al.
Published: (2025)
TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation
by: Cook, Jonathan, et al.
Published: (2024)
by: Cook, Jonathan, et al.
Published: (2024)
CAX: Cellular Automata Accelerated in JAX
by: Faldor, Maxence, et al.
Published: (2024)
by: Faldor, Maxence, et al.
Published: (2024)
The Generalization Gap in Offline Reinforcement Learning
by: Mediratta, Ishita, et al.
Published: (2023)
by: Mediratta, Ishita, et al.
Published: (2023)
Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts
by: Samvelyan, Mikayel, et al.
Published: (2024)
by: Samvelyan, Mikayel, et al.
Published: (2024)
NAVIX: Scaling MiniGrid Environments with JAX
by: Pignatelli, Eduardo, et al.
Published: (2024)
by: Pignatelli, Eduardo, et al.
Published: (2024)
A Subgoal-driven Framework for Improving Long-Horizon LLM Agents
by: Wang, Taiyi, et al.
Published: (2026)
by: Wang, Taiyi, et al.
Published: (2026)
Learning Bug Context for PyTorch-to-JAX Translation with LLMs
by: Phan, Hung, et al.
Published: (2025)
by: Phan, Hung, et al.
Published: (2025)
Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX
by: Bonnet, Clément, et al.
Published: (2023)
by: Bonnet, Clément, et al.
Published: (2023)
Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX
by: Nishimori, Soichiro, et al.
Published: (2026)
by: Nishimori, Soichiro, et al.
Published: (2026)
Programming by Backprop: An Instruction is Worth 100 Examples When Finetuning LLMs
by: Cook, Jonathan, et al.
Published: (2025)
by: Cook, Jonathan, et al.
Published: (2025)
Jasmine: A Simple, Performant and Scalable JAX-based World Modeling Codebase
by: Mahajan, Mihir, et al.
Published: (2025)
by: Mahajan, Mihir, et al.
Published: (2025)
JaxARC: A High-Performance JAX-based Environment for Abstraction and Reasoning Research
by: Aadam, et al.
Published: (2026)
by: Aadam, et al.
Published: (2026)
Chargax: A JAX Accelerated EV Charging Simulator
by: Ponse, Koen, et al.
Published: (2025)
by: Ponse, Koen, et al.
Published: (2025)
Understanding the Effects of RLHF on LLM Generalisation and Diversity
by: Kirk, Robert, et al.
Published: (2023)
by: Kirk, Robert, et al.
Published: (2023)
Scaling Opponent Shaping to High Dimensional Games
by: Khan, Akbir, et al.
Published: (2023)
by: Khan, Akbir, et al.
Published: (2023)
Scaling Is All You Need: Autonomous Driving with JAX-Accelerated Reinforcement Learning
by: Harmel, Moritz, et al.
Published: (2023)
by: Harmel, Moritz, et al.
Published: (2023)
Learning When to Plan: Efficiently Allocating Test-Time Compute for LLM Agents
by: Paglieri, Davide, et al.
Published: (2025)
by: Paglieri, Davide, et al.
Published: (2025)
Generative Data Refinement: Just Ask for Better Data
by: Jiang, Minqi, et al.
Published: (2025)
by: Jiang, Minqi, et al.
Published: (2025)
JPC: Flexible Inference for Predictive Coding Networks in JAX
by: Innocenti, Francesco, et al.
Published: (2024)
by: Innocenti, Francesco, et al.
Published: (2024)
Worse than Random: The Importance of a Baseline for Unsupervised Feature Selection
by: Rajabinasab, Muhammad, et al.
Published: (2026)
by: Rajabinasab, Muhammad, et al.
Published: (2026)
BLADE: Block-Sparse Attention Meets Step Distillation for Efficient Video Generation
by: Gu, Youping, et al.
Published: (2025)
by: Gu, Youping, et al.
Published: (2025)
Integrated Influence: Data Attribution with Baseline
by: Yang, Linxiao, et al.
Published: (2025)
by: Yang, Linxiao, et al.
Published: (2025)
Simple Baselines are Competitive with Code Evolution
by: Gideoni, Yonatan, et al.
Published: (2026)
by: Gideoni, Yonatan, et al.
Published: (2026)
Rethinking Weak Supervision in Anomaly Detection: A Comprehensive Benchmark
by: Yao, Xu, et al.
Published: (2026)
by: Yao, Xu, et al.
Published: (2026)
Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions
by: Rosser, J, et al.
Published: (2026)
by: Rosser, J, et al.
Published: (2026)
Similar Items
-
Imagined Autocurricula
by: Güzel, Ahmet H., et al.
Published: (2025) -
Multi-Agent Diagnostics for Robustness via Illuminated Diversity
by: Samvelyan, Mikayel, et al.
Published: (2024) -
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
by: Rutherford, Alexander, et al.
Published: (2023) -
Open-Endedness is Essential for Artificial Superhuman Intelligence
by: Hughes, Edward, et al.
Published: (2024) -
Preference-Based Alignment of Discrete Diffusion Models
by: Borso, Umberto, et al.
Published: (2025)