Solving Zero-Sum Convex Markov Games
Fuente:
arXiv
Saved in:
| Main Authors: | Kalogiannis, Fivos, Vlatakis-Gkaragkounis, Emmanouil-Vasileios, Gemp, Ian, Piliouras, Georgios |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Aspiration-based Perturbed Learning Automata in Games with Noisy Utility Measurements. Part A: Stochastic Stability in Non-zero-Sum Games
by: Chasparis, Georgios C.
Published: (2025)
by: Chasparis, Georgios C.
Published: (2025)
Approximating Nash Equilibria in Normal-Form Games via Stochastic Optimization
by: Gemp, Ian, et al.
Published: (2023)
by: Gemp, Ian, et al.
Published: (2023)
Convex Markov Games: A New Frontier for Multi-Agent Reinforcement Learning
by: Gemp, Ian, et al.
Published: (2024)
by: Gemp, Ian, et al.
Published: (2024)
Solving Neural Min-Max Games: The Role of Architecture, Initialization & Dynamics
by: Patel, Deep, et al.
Published: (2025)
by: Patel, Deep, et al.
Published: (2025)
A Quadratic Speedup in Finding Nash Equilibria of Quantum Zero-Sum Games
by: Vasconcelos, Francisca, et al.
Published: (2023)
by: Vasconcelos, Francisca, et al.
Published: (2023)
Policy Iteration for Two-Player General-Sum Stochastic Stackelberg Games
by: Kudo, Mikoto, et al.
Published: (2024)
by: Kudo, Mikoto, et al.
Published: (2024)
Last-Iterate Convergence of Adaptive Riemannian Gradient Descent for Equilibrium Computation
by: Cai, Yang, et al.
Published: (2023)
by: Cai, Yang, et al.
Published: (2023)
Data-Scarce Identification of Game Dynamics via Sum-of-Squares Optimization
by: Sakos, Iosif, et al.
Published: (2023)
by: Sakos, Iosif, et al.
Published: (2023)
Certifying Concavity and Monotonicity in Games via Sum-of-Squares Hierarchies
by: Leon, Vincent, et al.
Published: (2025)
by: Leon, Vincent, et al.
Published: (2025)
Learning from Delayed Feedback in Games via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2025)
by: Fujimoto, Yuma, et al.
Published: (2025)
Linear Convergence in Games with Delayed Feedback via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2026)
by: Fujimoto, Yuma, et al.
Published: (2026)
Continuous-time Discounted Mirror-Descent Dynamics in Monotone Concave Games
by: Gao, Bolin, et al.
Published: (2019)
by: Gao, Bolin, et al.
Published: (2019)
Global Behavior of Learning Dynamics in Zero-Sum Games with Memory Asymmetry
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Policy Optimization and Multi-agent Reinforcement Learning for Mean-variance Team Stochastic Games
by: Hu, Junkai, et al.
Published: (2025)
by: Hu, Junkai, et al.
Published: (2025)
Stochastic Semi-Gradient Descent for Learning Mean Field Games with Population-Aware Function Approximation
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Distributed Equilibrium-Seeking in Target Coverage Games via Self-Configurable Networks under Limited Communication
by: Bhargav, Jayanth, et al.
Published: (2026)
by: Bhargav, Jayanth, et al.
Published: (2026)
Memory Asymmetry Creates Heteroclinic Orbits to Nash Equilibrium in Learning in Zero-Sum Games
by: Fujimoto, Yuma, et al.
Published: (2023)
by: Fujimoto, Yuma, et al.
Published: (2023)
Synchronization in Learning in Periodic Zero-Sum Games Triggers Divergence from Nash Equilibrium
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
No Coin Left Behind: Maximizing Strategic Surplus Against No-Regret Dynamics
by: Su, Yiheng, et al.
Published: (2026)
by: Su, Yiheng, et al.
Published: (2026)
Faster Rates for No-Regret Learning in General Games via Cautious Optimism
by: Soleymani, Ashkan, et al.
Published: (2025)
by: Soleymani, Ashkan, et al.
Published: (2025)
Online Learning for Dynamic Vickrey-Clarke-Groves Mechanism in Unknown Environments
by: Leon, Vincent, et al.
Published: (2025)
by: Leon, Vincent, et al.
Published: (2025)
Cautious Optimism: A Meta-Algorithm for Near-Constant Regret in General Games
by: Soleymani, Ashkan, et al.
Published: (2025)
by: Soleymani, Ashkan, et al.
Published: (2025)
Algorithms and Complexity for Computing Nash Equilibria in Adversarial Team Games
by: Anagnostides, Ioannis, et al.
Published: (2023)
by: Anagnostides, Ioannis, et al.
Published: (2023)
Adaptive Incentive Design with Regret Minimization
by: Vasileiou, Georgios, et al.
Published: (2026)
by: Vasileiou, Georgios, et al.
Published: (2026)
A No-Regret Framework for Adaptive Incentive Design
by: Vasileiou, Georgios, et al.
Published: (2026)
by: Vasileiou, Georgios, et al.
Published: (2026)
Incentive Design without Hypergradients: A Social-Gradient Method
by: Vasileiou, Georgios, et al.
Published: (2026)
by: Vasileiou, Georgios, et al.
Published: (2026)
Data-Driven Behaviour Estimation in Parametric Games
by: Maddux, Anna M., et al.
Published: (2022)
by: Maddux, Anna M., et al.
Published: (2022)
Information Compression in Dynamic Information Disclosure Games
by: Tang, Dengwang, et al.
Published: (2024)
by: Tang, Dengwang, et al.
Published: (2024)
MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games
by: Hu, Anran, et al.
Published: (2024)
by: Hu, Anran, et al.
Published: (2024)
Achieving Logarithmic Regret in KL-Regularized Zero-Sum Markov Games
by: Nayak, Anupam, et al.
Published: (2025)
by: Nayak, Anupam, et al.
Published: (2025)
Game-theoretic Decentralized Coordination for Airspace Sector Overload Mitigation
by: Im, Jaehan, et al.
Published: (2025)
by: Im, Jaehan, et al.
Published: (2025)
$α$-Potential Games for Decentralized Control of Connected and Automated Vehicles
by: Di, Xuan, et al.
Published: (2025)
by: Di, Xuan, et al.
Published: (2025)
Dynamic Games among Teams with Delayed Intra-Team Information Sharing
by: Tang, Dengwang, et al.
Published: (2021)
by: Tang, Dengwang, et al.
Published: (2021)
Identity Concealment Games: How I Learned to Stop Revealing and Love the Coincidences
by: Karabag, Mustafa O., et al.
Published: (2021)
by: Karabag, Mustafa O., et al.
Published: (2021)
Prudent-Banker: No Extra Fees for Baseline Safety in Adversarial Bandits With and Without Delays
by: Hu, Ting, et al.
Published: (2026)
by: Hu, Ting, et al.
Published: (2026)
Learning Safely Without Knowing the World:COMPASS-Hedge
by: Hu, Ting, et al.
Published: (2026)
by: Hu, Ting, et al.
Published: (2026)
A Multi-Player Potential Game Approach for Sensor Network Localization with Noisy Measurements
by: Xu, Gehui, et al.
Published: (2024)
by: Xu, Gehui, et al.
Published: (2024)
Second-Order Algorithms for Finding Local Nash Equilibria in Zero-Sum Games
by: Gupta, Kushagra, et al.
Published: (2024)
by: Gupta, Kushagra, et al.
Published: (2024)
Fair Incentives for Repeated Engagement
by: Freund, Daniel, et al.
Published: (2021)
by: Freund, Daniel, et al.
Published: (2021)
Convex Markov Games and Beyond: New Proof of Existence, Characterization and Learning Algorithms for Nash Equilibria
by: Barakat, Anas, et al.
Published: (2026)
by: Barakat, Anas, et al.
Published: (2026)
Similar Items
-
Aspiration-based Perturbed Learning Automata in Games with Noisy Utility Measurements. Part A: Stochastic Stability in Non-zero-Sum Games
by: Chasparis, Georgios C.
Published: (2025) -
Approximating Nash Equilibria in Normal-Form Games via Stochastic Optimization
by: Gemp, Ian, et al.
Published: (2023) -
Convex Markov Games: A New Frontier for Multi-Agent Reinforcement Learning
by: Gemp, Ian, et al.
Published: (2024) -
Solving Neural Min-Max Games: The Role of Architecture, Initialization & Dynamics
by: Patel, Deep, et al.
Published: (2025) -
A Quadratic Speedup in Finding Nash Equilibria of Quantum Zero-Sum Games
by: Vasconcelos, Francisca, et al.
Published: (2023)