Policy Iteration for Two-Player General-Sum Stochastic Stackelberg Games
Fuente:
arXiv
Saved in:
| Main Authors: | Kudo, Mikoto, Akimoto, Youhei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sample-Efficient Hypergradient Estimation for Decentralized Bi-Level Reinforcement Learning
by: Kudo, Mikoto, et al.
Published: (2026)
by: Kudo, Mikoto, et al.
Published: (2026)
Aspiration-based Perturbed Learning Automata in Games with Noisy Utility Measurements. Part A: Stochastic Stability in Non-zero-Sum Games
by: Chasparis, Georgios C.
Published: (2025)
by: Chasparis, Georgios C.
Published: (2025)
Solving Zero-Sum Convex Markov Games
by: Kalogiannis, Fivos, et al.
Published: (2025)
by: Kalogiannis, Fivos, et al.
Published: (2025)
Policy Optimization and Multi-agent Reinforcement Learning for Mean-variance Team Stochastic Games
by: Hu, Junkai, et al.
Published: (2025)
by: Hu, Junkai, et al.
Published: (2025)
Stochastic Semi-Gradient Descent for Learning Mean Field Games with Population-Aware Function Approximation
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
A Multi-Player Potential Game Approach for Sensor Network Localization with Noisy Measurements
by: Xu, Gehui, et al.
Published: (2024)
by: Xu, Gehui, et al.
Published: (2024)
Certifying Concavity and Monotonicity in Games via Sum-of-Squares Hierarchies
by: Leon, Vincent, et al.
Published: (2025)
by: Leon, Vincent, et al.
Published: (2025)
Learning from Delayed Feedback in Games via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2025)
by: Fujimoto, Yuma, et al.
Published: (2025)
Linear Convergence in Games with Delayed Feedback via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2026)
by: Fujimoto, Yuma, et al.
Published: (2026)
Continuous-time Discounted Mirror-Descent Dynamics in Monotone Concave Games
by: Gao, Bolin, et al.
Published: (2019)
by: Gao, Bolin, et al.
Published: (2019)
Nash Equilibrium and Learning Dynamics in Three-Player Matching $m$-Action Games
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Smooth Information Gathering in Two-Player Noncooperative Games
by: Palafox, Fernando, et al.
Published: (2024)
by: Palafox, Fernando, et al.
Published: (2024)
Global Behavior of Learning Dynamics in Zero-Sum Games with Memory Asymmetry
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Online Learning for Dynamic Vickrey-Clarke-Groves Mechanism in Unknown Environments
by: Leon, Vincent, et al.
Published: (2025)
by: Leon, Vincent, et al.
Published: (2025)
Information Compression in Dynamic Information Disclosure Games
by: Tang, Dengwang, et al.
Published: (2024)
by: Tang, Dengwang, et al.
Published: (2024)
Data-Driven Behaviour Estimation in Parametric Games
by: Maddux, Anna M., et al.
Published: (2022)
by: Maddux, Anna M., et al.
Published: (2022)
Synchronization in Learning in Periodic Zero-Sum Games Triggers Divergence from Nash Equilibrium
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Memory Asymmetry Creates Heteroclinic Orbits to Nash Equilibrium in Learning in Zero-Sum Games
by: Fujimoto, Yuma, et al.
Published: (2023)
by: Fujimoto, Yuma, et al.
Published: (2023)
$α$-Potential Games for Decentralized Control of Connected and Automated Vehicles
by: Di, Xuan, et al.
Published: (2025)
by: Di, Xuan, et al.
Published: (2025)
MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games
by: Hu, Anran, et al.
Published: (2024)
by: Hu, Anran, et al.
Published: (2024)
Dynamic Games among Teams with Delayed Intra-Team Information Sharing
by: Tang, Dengwang, et al.
Published: (2021)
by: Tang, Dengwang, et al.
Published: (2021)
Game-theoretic Decentralized Coordination for Airspace Sector Overload Mitigation
by: Im, Jaehan, et al.
Published: (2025)
by: Im, Jaehan, et al.
Published: (2025)
Identity Concealment Games: How I Learned to Stop Revealing and Love the Coincidences
by: Karabag, Mustafa O., et al.
Published: (2021)
by: Karabag, Mustafa O., et al.
Published: (2021)
Independent Policy Mirror Descent for Markov Potential Games: Scaling to Large Number of Players
by: Alatur, Pragnya, et al.
Published: (2024)
by: Alatur, Pragnya, et al.
Published: (2024)
Distributed Equilibrium-Seeking in Target Coverage Games via Self-Configurable Networks under Limited Communication
by: Bhargav, Jayanth, et al.
Published: (2026)
by: Bhargav, Jayanth, et al.
Published: (2026)
General Performance Evaluation for Competitive Resource Allocation Games via Unseen Payoff Estimation
by: Diamond, N'yoma, et al.
Published: (2024)
by: Diamond, N'yoma, et al.
Published: (2024)
Fair Incentives for Repeated Engagement
by: Freund, Daniel, et al.
Published: (2021)
by: Freund, Daniel, et al.
Published: (2021)
Optimal Modified Feedback Strategies in LQ Games under Control Imperfections
by: Rabbani, Mahdis, et al.
Published: (2025)
by: Rabbani, Mahdis, et al.
Published: (2025)
A Data Driven Structural Decomposition of Dynamic Games via Best Response Maps
by: Rabbani, Mahdis, et al.
Published: (2026)
by: Rabbani, Mahdis, et al.
Published: (2026)
Last-Iterate Guarantees for Learning in Co-coercive Games
by: Chandak, Siddharth, et al.
Published: (2026)
by: Chandak, Siddharth, et al.
Published: (2026)
Policy Optimization finds Nash Equilibrium in Regularized General-Sum LQ Games
by: Zaman, Muhammad Aneeq uz, et al.
Published: (2024)
by: Zaman, Muhammad Aneeq uz, et al.
Published: (2024)
Differentiable Bilevel Programming for Stackelberg Congestion Games
by: Li, Jiayang, et al.
Published: (2022)
by: Li, Jiayang, et al.
Published: (2022)
Aggregate Fictitious Play for Learning in Anonymous Polymatrix Games (Extended Version)
by: Kara, Semih, et al.
Published: (2025)
by: Kara, Semih, et al.
Published: (2025)
Near-Optimal Policy Optimization for Correlated Equilibrium in General-Sum Markov Games
by: Cai, Yang, et al.
Published: (2024)
by: Cai, Yang, et al.
Published: (2024)
Choose Your Battles: Distributed Learning Over Multiple Tug of War Games
by: Chandak, Siddharth, et al.
Published: (2025)
by: Chandak, Siddharth, et al.
Published: (2025)
On the Convergence of Tâtonnement for Linear Fisher Markets
by: Nan, Tianlong, et al.
Published: (2024)
by: Nan, Tianlong, et al.
Published: (2024)
When is Mean-Field Reinforcement Learning Tractable and Relevant?
by: Yardim, Batuhan, et al.
Published: (2024)
by: Yardim, Batuhan, et al.
Published: (2024)
Strategic Negotiations in Endogenous Network Formation
by: Jalan, Akhil, et al.
Published: (2024)
by: Jalan, Akhil, et al.
Published: (2024)
Global solution to sensor network localization: A non-convex potential game approach and its distributed implementation
by: Xu, Gehui, et al.
Published: (2024)
by: Xu, Gehui, et al.
Published: (2024)
A Stochastic Surveillance Stackelberg Game: Co-Optimizing Defense Placement and Patrol Strategy
by: John, Yohan, et al.
Published: (2023)
by: John, Yohan, et al.
Published: (2023)
Similar Items
-
Sample-Efficient Hypergradient Estimation for Decentralized Bi-Level Reinforcement Learning
by: Kudo, Mikoto, et al.
Published: (2026) -
Aspiration-based Perturbed Learning Automata in Games with Noisy Utility Measurements. Part A: Stochastic Stability in Non-zero-Sum Games
by: Chasparis, Georgios C.
Published: (2025) -
Solving Zero-Sum Convex Markov Games
by: Kalogiannis, Fivos, et al.
Published: (2025) -
Policy Optimization and Multi-agent Reinforcement Learning for Mean-variance Team Stochastic Games
by: Hu, Junkai, et al.
Published: (2025) -
Stochastic Semi-Gradient Descent for Learning Mean Field Games with Population-Aware Function Approximation
by: Zhang, Chenyu, et al.
Published: (2024)