Policy Optimization and Multi-agent Reinforcement Learning for Mean-variance Team Stochastic Games
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Junkai, Xia, Li |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stochastic Semi-Gradient Descent for Learning Mean Field Games with Population-Aware Function Approximation
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Policy Iteration for Two-Player General-Sum Stochastic Stackelberg Games
by: Kudo, Mikoto, et al.
Published: (2024)
by: Kudo, Mikoto, et al.
Published: (2024)
MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games
by: Hu, Anran, et al.
Published: (2024)
by: Hu, Anran, et al.
Published: (2024)
Aspiration-based Perturbed Learning Automata in Games with Noisy Utility Measurements. Part A: Stochastic Stability in Non-zero-Sum Games
by: Chasparis, Georgios C.
Published: (2025)
by: Chasparis, Georgios C.
Published: (2025)
Dynamic Games among Teams with Delayed Intra-Team Information Sharing
by: Tang, Dengwang, et al.
Published: (2021)
by: Tang, Dengwang, et al.
Published: (2021)
Learning from Delayed Feedback in Games via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2025)
by: Fujimoto, Yuma, et al.
Published: (2025)
When is Mean-Field Reinforcement Learning Tractable and Relevant?
by: Yardim, Batuhan, et al.
Published: (2024)
by: Yardim, Batuhan, et al.
Published: (2024)
Solving Zero-Sum Convex Markov Games
by: Kalogiannis, Fivos, et al.
Published: (2025)
by: Kalogiannis, Fivos, et al.
Published: (2025)
Linear Convergence in Games with Delayed Feedback via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2026)
by: Fujimoto, Yuma, et al.
Published: (2026)
Continuous-time Discounted Mirror-Descent Dynamics in Monotone Concave Games
by: Gao, Bolin, et al.
Published: (2019)
by: Gao, Bolin, et al.
Published: (2019)
Population-aware Online Mirror Descent for Mean-Field Games by Deep Reinforcement Learning
by: Wu, Zida, et al.
Published: (2024)
by: Wu, Zida, et al.
Published: (2024)
Online Learning for Dynamic Vickrey-Clarke-Groves Mechanism in Unknown Environments
by: Leon, Vincent, et al.
Published: (2025)
by: Leon, Vincent, et al.
Published: (2025)
DelAC: A Multi-agent Reinforcement Learning of Team-Symmetric Stochastic Games
by: Lee, Duan-Shin, et al.
Published: (2026)
by: Lee, Duan-Shin, et al.
Published: (2026)
$α$-Potential Games for Decentralized Control of Connected and Automated Vehicles
by: Di, Xuan, et al.
Published: (2025)
by: Di, Xuan, et al.
Published: (2025)
A Multi-Player Potential Game Approach for Sensor Network Localization with Noisy Measurements
by: Xu, Gehui, et al.
Published: (2024)
by: Xu, Gehui, et al.
Published: (2024)
Identity Concealment Games: How I Learned to Stop Revealing and Love the Coincidences
by: Karabag, Mustafa O., et al.
Published: (2021)
by: Karabag, Mustafa O., et al.
Published: (2021)
Data-Driven Behaviour Estimation in Parametric Games
by: Maddux, Anna M., et al.
Published: (2022)
by: Maddux, Anna M., et al.
Published: (2022)
Information Compression in Dynamic Information Disclosure Games
by: Tang, Dengwang, et al.
Published: (2024)
by: Tang, Dengwang, et al.
Published: (2024)
Certifying Concavity and Monotonicity in Games via Sum-of-Squares Hierarchies
by: Leon, Vincent, et al.
Published: (2025)
by: Leon, Vincent, et al.
Published: (2025)
Game-theoretic Decentralized Coordination for Airspace Sector Overload Mitigation
by: Im, Jaehan, et al.
Published: (2025)
by: Im, Jaehan, et al.
Published: (2025)
Efficient and Scalable Deep Reinforcement Learning for Mean Field Control Games
by: Peng, Nianli, et al.
Published: (2024)
by: Peng, Nianli, et al.
Published: (2024)
Distributed Equilibrium-Seeking in Target Coverage Games via Self-Configurable Networks under Limited Communication
by: Bhargav, Jayanth, et al.
Published: (2026)
by: Bhargav, Jayanth, et al.
Published: (2026)
Fair Incentives for Repeated Engagement
by: Freund, Daniel, et al.
Published: (2021)
by: Freund, Daniel, et al.
Published: (2021)
Aggregate Fictitious Play for Learning in Anonymous Polymatrix Games (Extended Version)
by: Kara, Semih, et al.
Published: (2025)
by: Kara, Semih, et al.
Published: (2025)
Optimal Modified Feedback Strategies in LQ Games under Control Imperfections
by: Rabbani, Mahdis, et al.
Published: (2025)
by: Rabbani, Mahdis, et al.
Published: (2025)
Choose Your Battles: Distributed Learning Over Multiple Tug of War Games
by: Chandak, Siddharth, et al.
Published: (2025)
by: Chandak, Siddharth, et al.
Published: (2025)
Networked Communication for Mean-Field Games with Function Approximation and Empirical Mean-Field Estimation
by: Benjamin, Patrick, et al.
Published: (2024)
by: Benjamin, Patrick, et al.
Published: (2024)
Multi-agent Reach-avoid MDP via Potential Games and Low-rank Policy Structure
by: Casselman, Adam, et al.
Published: (2024)
by: Casselman, Adam, et al.
Published: (2024)
A Data Driven Structural Decomposition of Dynamic Games via Best Response Maps
by: Rabbani, Mahdis, et al.
Published: (2026)
by: Rabbani, Mahdis, et al.
Published: (2026)
General Performance Evaluation for Competitive Resource Allocation Games via Unseen Payoff Estimation
by: Diamond, N'yoma, et al.
Published: (2024)
by: Diamond, N'yoma, et al.
Published: (2024)
Global Behavior of Learning Dynamics in Zero-Sum Games with Memory Asymmetry
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Nash Equilibrium and Learning Dynamics in Three-Player Matching $m$-Action Games
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Satisficing Paths and Independent Multi-Agent Reinforcement Learning in Stochastic Games
by: Yongacoglu, Bora, et al.
Published: (2021)
by: Yongacoglu, Bora, et al.
Published: (2021)
Memory Asymmetry Creates Heteroclinic Orbits to Nash Equilibrium in Learning in Zero-Sum Games
by: Fujimoto, Yuma, et al.
Published: (2023)
by: Fujimoto, Yuma, et al.
Published: (2023)
Synchronization in Learning in Periodic Zero-Sum Games Triggers Divergence from Nash Equilibrium
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Risk Sensitivity in Markov Games and Multi-Agent Reinforcement Learning: A Systematic Review
by: Ghaemi, Hafez, et al.
Published: (2024)
by: Ghaemi, Hafez, et al.
Published: (2024)
Non-convex potential games for finding global solutions to sensor network localization
by: Xu, Gehui, et al.
Published: (2023)
by: Xu, Gehui, et al.
Published: (2023)
On the Convergence of Tâtonnement for Linear Fisher Markets
by: Nan, Tianlong, et al.
Published: (2024)
by: Nan, Tianlong, et al.
Published: (2024)
Strategic Negotiations in Endogenous Network Formation
by: Jalan, Akhil, et al.
Published: (2024)
by: Jalan, Akhil, et al.
Published: (2024)
Global solution to sensor network localization: A non-convex potential game approach and its distributed implementation
by: Xu, Gehui, et al.
Published: (2024)
by: Xu, Gehui, et al.
Published: (2024)
Similar Items
-
Stochastic Semi-Gradient Descent for Learning Mean Field Games with Population-Aware Function Approximation
by: Zhang, Chenyu, et al.
Published: (2024) -
Policy Iteration for Two-Player General-Sum Stochastic Stackelberg Games
by: Kudo, Mikoto, et al.
Published: (2024) -
MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games
by: Hu, Anran, et al.
Published: (2024) -
Aspiration-based Perturbed Learning Automata in Games with Noisy Utility Measurements. Part A: Stochastic Stability in Non-zero-Sum Games
by: Chasparis, Georgios C.
Published: (2025) -
Dynamic Games among Teams with Delayed Intra-Team Information Sharing
by: Tang, Dengwang, et al.
Published: (2021)