Stochastic Semi-Gradient Descent for Learning Mean Field Games with Population-Aware Function Approximation
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Chenyu, Chen, Xu, Di, Xuan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Policy Optimization and Multi-agent Reinforcement Learning for Mean-variance Team Stochastic Games
por: Hu, Junkai, et al.
Publicado: (2025)
por: Hu, Junkai, et al.
Publicado: (2025)
Population-aware Online Mirror Descent for Mean-Field Games by Deep Reinforcement Learning
por: Wu, Zida, et al.
Publicado: (2024)
por: Wu, Zida, et al.
Publicado: (2024)
MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games
por: Hu, Anran, et al.
Publicado: (2024)
por: Hu, Anran, et al.
Publicado: (2024)
Continuous-time Discounted Mirror-Descent Dynamics in Monotone Concave Games
por: Gao, Bolin, et al.
Publicado: (2019)
por: Gao, Bolin, et al.
Publicado: (2019)
$α$-Potential Games for Decentralized Control of Connected and Automated Vehicles
por: Di, Xuan, et al.
Publicado: (2025)
por: Di, Xuan, et al.
Publicado: (2025)
When is Mean-Field Reinforcement Learning Tractable and Relevant?
por: Yardim, Batuhan, et al.
Publicado: (2024)
por: Yardim, Batuhan, et al.
Publicado: (2024)
A Single Online Agent Can Efficiently Learn Mean Field Games
por: Zhang, Chenyu, et al.
Publicado: (2024)
por: Zhang, Chenyu, et al.
Publicado: (2024)
Aspiration-based Perturbed Learning Automata in Games with Noisy Utility Measurements. Part A: Stochastic Stability in Non-zero-Sum Games
por: Chasparis, Georgios C.
Publicado: (2025)
por: Chasparis, Georgios C.
Publicado: (2025)
Policy Iteration for Two-Player General-Sum Stochastic Stackelberg Games
por: Kudo, Mikoto, et al.
Publicado: (2024)
por: Kudo, Mikoto, et al.
Publicado: (2024)
Networked Communication for Mean-Field Games with Function Approximation and Empirical Mean-Field Estimation
por: Benjamin, Patrick, et al.
Publicado: (2024)
por: Benjamin, Patrick, et al.
Publicado: (2024)
Learning from Delayed Feedback in Games via Extra Prediction
por: Fujimoto, Yuma, et al.
Publicado: (2025)
por: Fujimoto, Yuma, et al.
Publicado: (2025)
A Multi-Player Potential Game Approach for Sensor Network Localization with Noisy Measurements
por: Xu, Gehui, et al.
Publicado: (2024)
por: Xu, Gehui, et al.
Publicado: (2024)
Identity Concealment Games: How I Learned to Stop Revealing and Love the Coincidences
por: Karabag, Mustafa O., et al.
Publicado: (2021)
por: Karabag, Mustafa O., et al.
Publicado: (2021)
Solving Zero-Sum Convex Markov Games
por: Kalogiannis, Fivos, et al.
Publicado: (2025)
por: Kalogiannis, Fivos, et al.
Publicado: (2025)
Distributed Equilibrium-Seeking in Target Coverage Games via Self-Configurable Networks under Limited Communication
por: Bhargav, Jayanth, et al.
Publicado: (2026)
por: Bhargav, Jayanth, et al.
Publicado: (2026)
Linear Convergence in Games with Delayed Feedback via Extra Prediction
por: Fujimoto, Yuma, et al.
Publicado: (2026)
por: Fujimoto, Yuma, et al.
Publicado: (2026)
Data-Driven Behaviour Estimation in Parametric Games
por: Maddux, Anna M., et al.
Publicado: (2022)
por: Maddux, Anna M., et al.
Publicado: (2022)
Information Compression in Dynamic Information Disclosure Games
por: Tang, Dengwang, et al.
Publicado: (2024)
por: Tang, Dengwang, et al.
Publicado: (2024)
Game-theoretic Decentralized Coordination for Airspace Sector Overload Mitigation
por: Im, Jaehan, et al.
Publicado: (2025)
por: Im, Jaehan, et al.
Publicado: (2025)
Certifying Concavity and Monotonicity in Games via Sum-of-Squares Hierarchies
por: Leon, Vincent, et al.
Publicado: (2025)
por: Leon, Vincent, et al.
Publicado: (2025)
Incentive Design without Hypergradients: A Social-Gradient Method
por: Vasileiou, Georgios, et al.
Publicado: (2026)
por: Vasileiou, Georgios, et al.
Publicado: (2026)
Dynamic Games among Teams with Delayed Intra-Team Information Sharing
por: Tang, Dengwang, et al.
Publicado: (2021)
por: Tang, Dengwang, et al.
Publicado: (2021)
Graphon Mean Field Games with a Representative Player: Analysis and Learning Algorithm
por: Zhou, Fuzhong, et al.
Publicado: (2024)
por: Zhou, Fuzhong, et al.
Publicado: (2024)
Online Learning for Dynamic Vickrey-Clarke-Groves Mechanism in Unknown Environments
por: Leon, Vincent, et al.
Publicado: (2025)
por: Leon, Vincent, et al.
Publicado: (2025)
Optimal Modified Feedback Strategies in LQ Games under Control Imperfections
por: Rabbani, Mahdis, et al.
Publicado: (2025)
por: Rabbani, Mahdis, et al.
Publicado: (2025)
Non-convex potential games for finding global solutions to sensor network localization
por: Xu, Gehui, et al.
Publicado: (2023)
por: Xu, Gehui, et al.
Publicado: (2023)
Global solution to sensor network localization: A non-convex potential game approach and its distributed implementation
por: Xu, Gehui, et al.
Publicado: (2024)
por: Xu, Gehui, et al.
Publicado: (2024)
A Data Driven Structural Decomposition of Dynamic Games via Best Response Maps
por: Rabbani, Mahdis, et al.
Publicado: (2026)
por: Rabbani, Mahdis, et al.
Publicado: (2026)
General Performance Evaluation for Competitive Resource Allocation Games via Unseen Payoff Estimation
por: Diamond, N'yoma, et al.
Publicado: (2024)
por: Diamond, N'yoma, et al.
Publicado: (2024)
Global Behavior of Learning Dynamics in Zero-Sum Games with Memory Asymmetry
por: Fujimoto, Yuma, et al.
Publicado: (2024)
por: Fujimoto, Yuma, et al.
Publicado: (2024)
Nash Equilibrium and Learning Dynamics in Three-Player Matching $m$-Action Games
por: Fujimoto, Yuma, et al.
Publicado: (2024)
por: Fujimoto, Yuma, et al.
Publicado: (2024)
Memory Asymmetry Creates Heteroclinic Orbits to Nash Equilibrium in Learning in Zero-Sum Games
por: Fujimoto, Yuma, et al.
Publicado: (2023)
por: Fujimoto, Yuma, et al.
Publicado: (2023)
Synchronization in Learning in Periodic Zero-Sum Games Triggers Divergence from Nash Equilibrium
por: Fujimoto, Yuma, et al.
Publicado: (2024)
por: Fujimoto, Yuma, et al.
Publicado: (2024)
A Fixed Point Framework for the Existence of EFX Allocations
por: Etesami, S. Rasoul
Publicado: (2025)
por: Etesami, S. Rasoul
Publicado: (2025)
Harnessing Information in Incentive Design
por: Velicheti, Raj Kiriti, et al.
Publicado: (2025)
por: Velicheti, Raj Kiriti, et al.
Publicado: (2025)
Adaptive Incentive Design with Regret Minimization
por: Vasileiou, Georgios, et al.
Publicado: (2026)
por: Vasileiou, Georgios, et al.
Publicado: (2026)
How competitive are pay-as-bid auction games?
por: Vanelli, Martina, et al.
Publicado: (2025)
por: Vanelli, Martina, et al.
Publicado: (2025)
A No-Regret Framework for Adaptive Incentive Design
por: Vasileiou, Georgios, et al.
Publicado: (2026)
por: Vasileiou, Georgios, et al.
Publicado: (2026)
On the Convergence of Tâtonnement for Linear Fisher Markets
por: Nan, Tianlong, et al.
Publicado: (2024)
por: Nan, Tianlong, et al.
Publicado: (2024)
Strategic Negotiations in Endogenous Network Formation
por: Jalan, Akhil, et al.
Publicado: (2024)
por: Jalan, Akhil, et al.
Publicado: (2024)
Ejemplares similares
-
Policy Optimization and Multi-agent Reinforcement Learning for Mean-variance Team Stochastic Games
por: Hu, Junkai, et al.
Publicado: (2025) -
Population-aware Online Mirror Descent for Mean-Field Games by Deep Reinforcement Learning
por: Wu, Zida, et al.
Publicado: (2024) -
MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games
por: Hu, Anran, et al.
Publicado: (2024) -
Continuous-time Discounted Mirror-Descent Dynamics in Monotone Concave Games
por: Gao, Bolin, et al.
Publicado: (2019) -
$α$-Potential Games for Decentralized Control of Connected and Automated Vehicles
por: Di, Xuan, et al.
Publicado: (2025)