Last-Iterate Convergence of Payoff-Based Independent Learning in Zero-Sum Stochastic Games
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Zaiwei, Zhang, Kaiqing, Mazumdar, Eric, Ozdaglar, Asuman, Wierman, Adam |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-Player Zero-Sum Markov Games with Networked Separable Interactions
di: Park, Chanwoo, et al.
Pubblicazione: (2023)
di: Park, Chanwoo, et al.
Pubblicazione: (2023)
The Power of Regularization in Solving Extensive-Form Games
di: Liu, Mingyang, et al.
Pubblicazione: (2022)
di: Liu, Mingyang, et al.
Pubblicazione: (2022)
Do LLM Agents Have Regret? A Case Study in Online Learning and Games
di: Park, Chanwoo, et al.
Pubblicazione: (2024)
di: Park, Chanwoo, et al.
Pubblicazione: (2024)
Finite-Sample Guarantees for Learning Dynamics in Zero-Sum Polymatrix Games
di: Faizal, Fathima Zarin, et al.
Pubblicazione: (2024)
di: Faizal, Fathima Zarin, et al.
Pubblicazione: (2024)
A Policy-Gradient Approach to Solving Imperfect-Information Games with Best-Iterate Convergence
di: Liu, Mingyang, et al.
Pubblicazione: (2024)
di: Liu, Mingyang, et al.
Pubblicazione: (2024)
Last-Iterate Convergence: Zero-Sum Games and Constrained Min-Max Optimization
di: Daskalakis, Constantinos, et al.
Pubblicazione: (2018)
di: Daskalakis, Constantinos, et al.
Pubblicazione: (2018)
Learning Zero-Sum Linear Quadratic Games with Improved Sample Complexity and Last-Iterate Convergence
di: Wu, Jiduan, et al.
Pubblicazione: (2023)
di: Wu, Jiduan, et al.
Pubblicazione: (2023)
Online Learning and Equilibrium Computation with Ranking Feedback
di: Liu, Mingyang, et al.
Pubblicazione: (2026)
di: Liu, Mingyang, et al.
Pubblicazione: (2026)
Convergent Q-Learning for Infinite-Horizon General-Sum Markov Games through Behavioral Economics
di: Zhang, Yizhou, et al.
Pubblicazione: (2025)
di: Zhang, Yizhou, et al.
Pubblicazione: (2025)
LiteEFG: An Efficient Python Library for Solving Extensive-form Games
di: Liu, Mingyang, et al.
Pubblicazione: (2024)
di: Liu, Mingyang, et al.
Pubblicazione: (2024)
Last-Iterate Convergence of No-Regret Learning for Equilibria in Bargaining Games
di: Kamp, Serafina, et al.
Pubblicazione: (2025)
di: Kamp, Serafina, et al.
Pubblicazione: (2025)
Differentially Private Equilibrium Finding in Polymatrix Games
di: Liu, Mingyang, et al.
Pubblicazione: (2025)
di: Liu, Mingyang, et al.
Pubblicazione: (2025)
Last-Iterate Convergence Properties of Regret-Matching Algorithms in Games
di: Cai, Yang, et al.
Pubblicazione: (2023)
di: Cai, Yang, et al.
Pubblicazione: (2023)
Universal Complexity Bounds Based on Value Iteration for Stochastic Mean Payoff Games and Entropy Games
di: Allamigeon, Xavier, et al.
Pubblicazione: (2022)
di: Allamigeon, Xavier, et al.
Pubblicazione: (2022)
On Separation Between Best-Iterate, Random-Iterate, and Last-Iterate Convergence of Learning in Games
di: Cai, Yang, et al.
Pubblicazione: (2025)
di: Cai, Yang, et al.
Pubblicazione: (2025)
A Payoff-Based Policy Gradient Method in Stochastic Games with Long-Run Average Payoffs
di: Zhang, Junyue, et al.
Pubblicazione: (2024)
di: Zhang, Junyue, et al.
Pubblicazione: (2024)
Learning to Steer Learners in Games
di: Zhang, Yizhou, et al.
Pubblicazione: (2025)
di: Zhang, Yizhou, et al.
Pubblicazione: (2025)
Breaking the Curse of Multiagency in Robust Multi-Agent Reinforcement Learning
di: Shi, Laixi, et al.
Pubblicazione: (2024)
di: Shi, Laixi, et al.
Pubblicazione: (2024)
Fast Last-Iterate Convergence of Learning in Games Requires Forgetful Algorithms
di: Cai, Yang, et al.
Pubblicazione: (2024)
di: Cai, Yang, et al.
Pubblicazione: (2024)
From Average-Iterate to Last-Iterate Convergence in Games: A Reduction and Its Applications
di: Cai, Yang, et al.
Pubblicazione: (2025)
di: Cai, Yang, et al.
Pubblicazione: (2025)
Stochastic Window Mean-Payoff Games
di: Doyen, Laurent, et al.
Pubblicazione: (2023)
di: Doyen, Laurent, et al.
Pubblicazione: (2023)
Scale-Invariant Regret Matching and Online Learning with Optimal Convergence: Bridging Theory and Practice in Zero-Sum Games
di: Zhang, Brian Hu, et al.
Pubblicazione: (2025)
di: Zhang, Brian Hu, et al.
Pubblicazione: (2025)
Bayesian Learning in Episodic Zero-Sum Games
di: Yueh, Chang-Wei, et al.
Pubblicazione: (2026)
di: Yueh, Chang-Wei, et al.
Pubblicazione: (2026)
Boosting Perturbed Gradient Ascent for Last-Iterate Convergence in Games
di: Abe, Kenshi, et al.
Pubblicazione: (2024)
di: Abe, Kenshi, et al.
Pubblicazione: (2024)
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
di: Yan, Yuling, et al.
Pubblicazione: (2022)
di: Yan, Yuling, et al.
Pubblicazione: (2022)
Playing Markov Games Without Observing Payoffs
di: Ablin, Daniel, et al.
Pubblicazione: (2025)
di: Ablin, Daniel, et al.
Pubblicazione: (2025)
Provably Convergent Actor-Critic for MARL through Risk-aversion
di: Zhang, Yizhou, et al.
Pubblicazione: (2026)
di: Zhang, Yizhou, et al.
Pubblicazione: (2026)
Efficient Last-Iterate Convergence in Regret Minimization via Adaptive Reward Transformation
di: Ren, Hang, et al.
Pubblicazione: (2025)
di: Ren, Hang, et al.
Pubblicazione: (2025)
Optimal Rates for Feasible Payoff Set Estimation in Games
di: Barbara, Annalisa, et al.
Pubblicazione: (2026)
di: Barbara, Annalisa, et al.
Pubblicazione: (2026)
Computing Equilibrium beyond Unilateral Deviation
di: Liu, Mingyang, et al.
Pubblicazione: (2026)
di: Liu, Mingyang, et al.
Pubblicazione: (2026)
Offline Two-Player Zero-Sum Markov Games with KL Regularization
di: Chen, Claire, et al.
Pubblicazione: (2026)
di: Chen, Claire, et al.
Pubblicazione: (2026)
Leveraging Noisy Observations in Zero-Sum Games
di: Athanasakos, Emmanouil M, et al.
Pubblicazione: (2024)
di: Athanasakos, Emmanouil M, et al.
Pubblicazione: (2024)
Two-Player Zero-Sum Games with Bandit Feedback
di: Yılmaz, Elif, et al.
Pubblicazione: (2025)
di: Yılmaz, Elif, et al.
Pubblicazione: (2025)
State-Constrained Zero-Sum Differential Games with One-Sided Information
di: Ghimire, Mukesh, et al.
Pubblicazione: (2024)
di: Ghimire, Mukesh, et al.
Pubblicazione: (2024)
Equilibrium Selection for Multi-agent Reinforcement Learning: A Unified Framework
di: Zhang, Runyu, et al.
Pubblicazione: (2024)
di: Zhang, Runyu, et al.
Pubblicazione: (2024)
Understanding Model Selection For Learning In Strategic Environments
di: Handina, Tinashe, et al.
Pubblicazione: (2024)
di: Handina, Tinashe, et al.
Pubblicazione: (2024)
Efficient Last-iterate Convergence Algorithms in Solving Games
di: Meng, Linjian, et al.
Pubblicazione: (2023)
di: Meng, Linjian, et al.
Pubblicazione: (2023)
Optimism Without Regularization: Constant Regret in Zero-Sum Games
di: Lazarsfeld, John, et al.
Pubblicazione: (2025)
di: Lazarsfeld, John, et al.
Pubblicazione: (2025)
Fast and Furious Symmetric Learning in Zero-Sum Games: Gradient Descent as Fictitious Play
di: Lazarsfeld, John, et al.
Pubblicazione: (2025)
di: Lazarsfeld, John, et al.
Pubblicazione: (2025)
Learning in Zero-Sum Markov Games: Relaxing Strong Reachability and Mixing Time Assumptions
di: Ouhamma, Reda, et al.
Pubblicazione: (2023)
di: Ouhamma, Reda, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Multi-Player Zero-Sum Markov Games with Networked Separable Interactions
di: Park, Chanwoo, et al.
Pubblicazione: (2023) -
The Power of Regularization in Solving Extensive-Form Games
di: Liu, Mingyang, et al.
Pubblicazione: (2022) -
Do LLM Agents Have Regret? A Case Study in Online Learning and Games
di: Park, Chanwoo, et al.
Pubblicazione: (2024) -
Finite-Sample Guarantees for Learning Dynamics in Zero-Sum Polymatrix Games
di: Faizal, Fathima Zarin, et al.
Pubblicazione: (2024) -
A Policy-Gradient Approach to Solving Imperfect-Information Games with Best-Iterate Convergence
di: Liu, Mingyang, et al.
Pubblicazione: (2024)