GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Fan, Zhiyuan, Farina, Gabriele |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Optimality of Dilated Entropy and Lower Bounds for Online Learning in Extensive-Form Games
by: Fan, Zhiyuan, et al.
Published: (2024)
by: Fan, Zhiyuan, et al.
Published: (2024)
A Policy-Gradient Approach to Solving Imperfect-Information Games with Best-Iterate Convergence
by: Liu, Mingyang, et al.
Published: (2024)
by: Liu, Mingyang, et al.
Published: (2024)
Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries
by: Maiti, Arnab, et al.
Published: (2025)
by: Maiti, Arnab, et al.
Published: (2025)
On the Universal Near Optimality of Hedge in Combinatorial Settings
by: Fan, Zhiyuan, et al.
Published: (2025)
by: Fan, Zhiyuan, et al.
Published: (2025)
Online Learning and Equilibrium Computation with Ranking Feedback
by: Liu, Mingyang, et al.
Published: (2026)
by: Liu, Mingyang, et al.
Published: (2026)
An Efficient Black-Box Reduction from Online Learning to Multicalibration, and a New Route to $Φ$-Regret Minimization
by: Farina, Gabriele, et al.
Published: (2026)
by: Farina, Gabriele, et al.
Published: (2026)
RL-CFR: Improving Action Abstraction for Imperfect Information Extensive-Form Games with Reinforcement Learning
by: Li, Boning, et al.
Published: (2024)
by: Li, Boning, et al.
Published: (2024)
Meta-Learning in Self-Play Regret Minimization
by: Sychrovský, David, et al.
Published: (2025)
by: Sychrovský, David, et al.
Published: (2025)
Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games
by: Lanier, JB, et al.
Published: (2026)
by: Lanier, JB, et al.
Published: (2026)
The Stability of Online Algorithms in Performative Prediction
by: Farina, Gabriele, et al.
Published: (2026)
by: Farina, Gabriele, et al.
Published: (2026)
Faster Rates for No-Regret Learning in General Games via Cautious Optimism
by: Soleymani, Ashkan, et al.
Published: (2025)
by: Soleymani, Ashkan, et al.
Published: (2025)
Learning and Computation of $Φ$-Equilibria at the Frontier of Tractability
by: Zhang, Brian Hu, et al.
Published: (2025)
by: Zhang, Brian Hu, et al.
Published: (2025)
Learning to Play Against Unknown Opponents
by: Arunachaleswaran, Eshwar Ram, et al.
Published: (2024)
by: Arunachaleswaran, Eshwar Ram, et al.
Published: (2024)
The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play
by: La Malfa, Gabriele, et al.
Published: (2026)
by: La Malfa, Gabriele, et al.
Published: (2026)
LiteEFG: An Efficient Python Library for Solving Extensive-form Games
by: Liu, Mingyang, et al.
Published: (2024)
by: Liu, Mingyang, et al.
Published: (2024)
Cautious Optimism: A Meta-Algorithm for Near-Constant Regret in General Games
by: Soleymani, Ashkan, et al.
Published: (2025)
by: Soleymani, Ashkan, et al.
Published: (2025)
Pure Exploration via Frank-Wolfe Self-Play
by: Liu, Xinyu, et al.
Published: (2025)
by: Liu, Xinyu, et al.
Published: (2025)
Efficiently Training Neural Networks for Imperfect Information Games by Sampling Information Sets
by: Bertram, Timo, et al.
Published: (2024)
by: Bertram, Timo, et al.
Published: (2024)
Last-Iterate Convergence Properties of Regret-Matching Algorithms in Games
by: Cai, Yang, et al.
Published: (2023)
by: Cai, Yang, et al.
Published: (2023)
Optimal Correlated Equilibria in General-Sum Extensive-Form Games: Fixed-Parameter Algorithms, Hardness, and Two-Sided Column-Generation
by: Zhang, Brian, et al.
Published: (2022)
by: Zhang, Brian, et al.
Published: (2022)
Fast and Furious Symmetric Learning in Zero-Sum Games: Gradient Descent as Fictitious Play
by: Lazarsfeld, John, et al.
Published: (2025)
by: Lazarsfeld, John, et al.
Published: (2025)
Safe Exploitative Play with Untrusted Type Beliefs
by: Li, Tongxin, et al.
Published: (2024)
by: Li, Tongxin, et al.
Published: (2024)
Playing Markov Games Without Observing Payoffs
by: Ablin, Daniel, et al.
Published: (2025)
by: Ablin, Daniel, et al.
Published: (2025)
A Polynomial-Time Algorithm for Variational Inequalities under the Minty Condition
by: Anagnostides, Ioannis, et al.
Published: (2025)
by: Anagnostides, Ioannis, et al.
Published: (2025)
Fast Last-Iterate Convergence of Learning in Games Requires Forgetful Algorithms
by: Cai, Yang, et al.
Published: (2024)
by: Cai, Yang, et al.
Published: (2024)
On Separation Between Best-Iterate, Random-Iterate, and Last-Iterate Convergence of Learning in Games
by: Cai, Yang, et al.
Published: (2025)
by: Cai, Yang, et al.
Published: (2025)
Performative Reinforcement Learning with Linear Markov Decision Process
by: Mandal, Debmalya, et al.
Published: (2024)
by: Mandal, Debmalya, et al.
Published: (2024)
Reinforcement Learning for Game-Theoretic Resource Allocation on Graphs
by: An, Zijian, et al.
Published: (2025)
by: An, Zijian, et al.
Published: (2025)
Differentially Private Equilibrium Finding in Polymatrix Games
by: Liu, Mingyang, et al.
Published: (2025)
by: Liu, Mingyang, et al.
Published: (2025)
Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models
by: Fang, Zeyu, et al.
Published: (2024)
by: Fang, Zeyu, et al.
Published: (2024)
Intersectional Fairness in Reinforcement Learning with Large State and Constraint Spaces
by: Eaton, Eric, et al.
Published: (2025)
by: Eaton, Eric, et al.
Published: (2025)
ReLExS: Reinforcement Learning Explanations for Stackelberg No-Regret Learners
by: Huang, Xiangge, et al.
Published: (2024)
by: Huang, Xiangge, et al.
Published: (2024)
Balancing of competitive two-player Game Levels with Reinforcement Learning
by: Rupp, Florian, et al.
Published: (2023)
by: Rupp, Florian, et al.
Published: (2023)
Robust Adversarial Reinforcement Learning in Stochastic Games via Sequence Modeling
by: Tang, Xiaohang, et al.
Published: (2025)
by: Tang, Xiaohang, et al.
Published: (2025)
Decentralized Multi-Agent Reinforcement Learning for Continuous-Space Stochastic Games
by: Altabaa, Awni, et al.
Published: (2023)
by: Altabaa, Awni, et al.
Published: (2023)
Impact of Price Inflation on Algorithmic Collusion Through Reinforcement Learning Agents
by: Tinoco, Sebastián, et al.
Published: (2025)
by: Tinoco, Sebastián, et al.
Published: (2025)
Taming Equilibrium Bias in Risk-Sensitive Multi-Agent Reinforcement Learning
by: Fei, Yingjie, et al.
Published: (2024)
by: Fei, Yingjie, et al.
Published: (2024)
Fictitious Play in Extensive-Form Games of Imperfect Information
by: Castiglione, Jason, et al.
Published: (2025)
by: Castiglione, Jason, et al.
Published: (2025)
A Reinforcement Learning Approach in Multi-Phase Second-Price Auction Design
by: Ai, Rui, et al.
Published: (2022)
by: Ai, Rui, et al.
Published: (2022)
Learning to Play Multi-Follower Bayesian Stackelberg Games
by: Personnat, Gerson, et al.
Published: (2025)
by: Personnat, Gerson, et al.
Published: (2025)
Similar Items
-
On the Optimality of Dilated Entropy and Lower Bounds for Online Learning in Extensive-Form Games
by: Fan, Zhiyuan, et al.
Published: (2024) -
A Policy-Gradient Approach to Solving Imperfect-Information Games with Best-Iterate Convergence
by: Liu, Mingyang, et al.
Published: (2024) -
Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries
by: Maiti, Arnab, et al.
Published: (2025) -
On the Universal Near Optimality of Hedge in Combinatorial Settings
by: Fan, Zhiyuan, et al.
Published: (2025) -
Online Learning and Equilibrium Computation with Ranking Feedback
by: Liu, Mingyang, et al.
Published: (2026)