Scaling Opponent Shaping to High Dimensional Games
Fuente:
arXiv
Saved in:
| Main Authors: | Khan, Akbir, Willi, Timon, Kwan, Newton, Tacchetti, Andrea, Lu, Chris, Grefenstette, Edward, Rocktäschel, Tim, Foerster, Jakob |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Analysing the Sample Complexity of Opponent Shaping
by: Fung, Kitty, et al.
Published: (2024)
by: Fung, Kitty, et al.
Published: (2024)
Programming by Backprop: An Instruction is Worth 100 Examples When Finetuning LLMs
by: Cook, Jonathan, et al.
Published: (2025)
by: Cook, Jonathan, et al.
Published: (2025)
The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind
by: Lupu, Andrei, et al.
Published: (2025)
by: Lupu, Andrei, et al.
Published: (2025)
minimax: Efficient Baselines for Autocurricula in JAX
by: Jiang, Minqi, et al.
Published: (2023)
by: Jiang, Minqi, et al.
Published: (2023)
Debating with More Persuasive LLMs Leads to More Truthful Answers
by: Khan, Akbir, et al.
Published: (2024)
by: Khan, Akbir, et al.
Published: (2024)
BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games
by: Paglieri, Davide, et al.
Published: (2024)
by: Paglieri, Davide, et al.
Published: (2024)
Learning When to Plan: Efficiently Allocating Test-Time Compute for LLM Agents
by: Paglieri, Davide, et al.
Published: (2025)
by: Paglieri, Davide, et al.
Published: (2025)
Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions
by: Rosser, J, et al.
Published: (2026)
by: Rosser, J, et al.
Published: (2026)
Mixture of Experts in a Mixture of RL settings
by: Willi, Timon, et al.
Published: (2024)
by: Willi, Timon, et al.
Published: (2024)
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
by: Rutherford, Alexander, et al.
Published: (2024)
by: Rutherford, Alexander, et al.
Published: (2024)
TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation
by: Cook, Jonathan, et al.
Published: (2024)
by: Cook, Jonathan, et al.
Published: (2024)
DéjàQ: Open-Ended Evolution of Diverse, Learnable and Verifiable Problems
by: Röpke, Willem, et al.
Published: (2026)
by: Röpke, Willem, et al.
Published: (2026)
ADIOS: Antibody Development via Opponent Shaping
by: Towers, Sebastian, et al.
Published: (2024)
by: Towers, Sebastian, et al.
Published: (2024)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
by: Rutherford, Alexander, et al.
Published: (2023)
by: Rutherford, Alexander, et al.
Published: (2023)
Mixtures of Experts Unlock Parameter Scaling for Deep RL
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Artificial Generational Intelligence: Cultural Accumulation in Reinforcement Learning
by: Cook, Jonathan, et al.
Published: (2024)
by: Cook, Jonathan, et al.
Published: (2024)
Differentiable Belief-based Opponent Shaping
by: Sane, Aarav G, et al.
Published: (2026)
by: Sane, Aarav G, et al.
Published: (2026)
Minibal: Balanced Game-Playing Without Opponent Modeling
by: Cohen-Solal, Quentin, et al.
Published: (2026)
by: Cohen-Solal, Quentin, et al.
Published: (2026)
Opponent Shaping in LLM Agents
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
The Danger Of Arrogance: Welfare Equilibra As A Solution To Stackelberg Self-Play In Non-Coincidental Games
by: Levi, Jake, et al.
Published: (2024)
by: Levi, Jake, et al.
Published: (2024)
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
by: Matthews, Michael, et al.
Published: (2024)
by: Matthews, Michael, et al.
Published: (2024)
JaxLife: An Open-Ended Agentic Simulator
by: Lu, Chris, et al.
Published: (2024)
by: Lu, Chris, et al.
Published: (2024)
Imagined Autocurricula
by: Güzel, Ahmet H., et al.
Published: (2025)
by: Güzel, Ahmet H., et al.
Published: (2025)
StratFormer: Adaptive Opponent Modeling and Exploitation in Imperfect-Information Games
by: Caen, Andy, et al.
Published: (2026)
by: Caen, Andy, et al.
Published: (2026)
LLM-First Search: Self-Guided Exploration of the Solution Space
by: Herr, Nathan, et al.
Published: (2025)
by: Herr, Nathan, et al.
Published: (2025)
AI & Human Co-Improvement for Safer Co-Superintelligence
by: Weston, Jason, et al.
Published: (2025)
by: Weston, Jason, et al.
Published: (2025)
Anticipating Oblivious Opponents in Stochastic Games
by: Kalat, Shadi Tasdighi, et al.
Published: (2024)
by: Kalat, Shadi Tasdighi, et al.
Published: (2024)
DecisionHoldem: Safe Depth-Limited Solving With Diverse Opponents for Imperfect-Information Games
by: Zhou, Qibin, et al.
Published: (2022)
by: Zhou, Qibin, et al.
Published: (2022)
Interaction Dynamics as a Reward Signal for LLMs
by: Gooding, Sian, et al.
Published: (2025)
by: Gooding, Sian, et al.
Published: (2025)
Opponent Modeling in Multiplayer Imperfect-Information Games
by: Ganzfried, Sam, et al.
Published: (2022)
by: Ganzfried, Sam, et al.
Published: (2022)
Consistent Opponent Modeling in Imperfect-Information Games
by: Ganzfried, Sam
Published: (2025)
by: Ganzfried, Sam
Published: (2025)
Rethinking Teacher-Student Curriculum Learning through the Cooperative Mechanics of Experience
by: Diaz, Manfred, et al.
Published: (2024)
by: Diaz, Manfred, et al.
Published: (2024)
Behaviour Distillation
by: Lupu, Andrei, et al.
Published: (2024)
by: Lupu, Andrei, et al.
Published: (2024)
Preference-Based Alignment of Discrete Diffusion Models
by: Borso, Umberto, et al.
Published: (2025)
by: Borso, Umberto, et al.
Published: (2025)
LOQA: Learning with Opponent Q-Learning Awareness
by: Aghajohari, Milad, et al.
Published: (2024)
by: Aghajohari, Milad, et al.
Published: (2024)
Recurrent Reinforcement Learning with Memoroids
by: Morad, Steven, et al.
Published: (2024)
by: Morad, Steven, et al.
Published: (2024)
Investigating Non-Transitivity in LLM-as-a-Judge
by: Xu, Yi, et al.
Published: (2025)
by: Xu, Yi, et al.
Published: (2025)
Can Learned Optimization Make Reinforcement Learning Less Difficult?
by: Goldie, Alexander David, et al.
Published: (2024)
by: Goldie, Alexander David, et al.
Published: (2024)
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
by: Lu, Chris, et al.
Published: (2024)
by: Lu, Chris, et al.
Published: (2024)
JaxUED: A simple and useable UED library in Jax
by: Coward, Samuel, et al.
Published: (2024)
by: Coward, Samuel, et al.
Published: (2024)
Similar Items
-
Analysing the Sample Complexity of Opponent Shaping
by: Fung, Kitty, et al.
Published: (2024) -
Programming by Backprop: An Instruction is Worth 100 Examples When Finetuning LLMs
by: Cook, Jonathan, et al.
Published: (2025) -
The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind
by: Lupu, Andrei, et al.
Published: (2025) -
minimax: Efficient Baselines for Autocurricula in JAX
by: Jiang, Minqi, et al.
Published: (2023) -
Debating with More Persuasive LLMs Leads to More Truthful Answers
by: Khan, Akbir, et al.
Published: (2024)