Internal State-Based Policy Gradient Methods for Partially Observable Markov Potential Games
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Wonseok, Doan, Thinh T. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Independent Learning of Nash Equilibria in Partially Observable Markov Potential Games with Decoupled Dynamics
by: Jordan, Philip, et al.
Published: (2026)
by: Jordan, Philip, et al.
Published: (2026)
Multi-agent Off-policy Actor-Critic Reinforcement Learning for Partially Observable Environments
by: Zhaikhan, Ainur, et al.
Published: (2024)
by: Zhaikhan, Ainur, et al.
Published: (2024)
Observation Interference in Partially Observable Assistance Games
by: Emmons, Scott, et al.
Published: (2024)
by: Emmons, Scott, et al.
Published: (2024)
Independent Policy Mirror Descent for Markov Potential Games: Scaling to Large Number of Players
by: Alatur, Pragnya, et al.
Published: (2024)
by: Alatur, Pragnya, et al.
Published: (2024)
Independent Learning in Constrained Markov Potential Games
by: Jordan, Philip, et al.
Published: (2024)
by: Jordan, Philip, et al.
Published: (2024)
Nash Approximation Gap in Truncated Infinite-horizon Partially Observable Markov Games
by: Sang, Lan, et al.
Published: (2026)
by: Sang, Lan, et al.
Published: (2026)
Optimistic Multi-Agent Policy Gradient
by: Zhao, Wenshuai, et al.
Published: (2023)
by: Zhao, Wenshuai, et al.
Published: (2023)
Linear Convergence of Independent Natural Policy Gradient in Games with Entropy Regularization
by: Sun, Youbang, et al.
Published: (2024)
by: Sun, Youbang, et al.
Published: (2024)
Learning Decentralized Partially Observable Mean Field Control for Artificial Collective Behavior
by: Cui, Kai, et al.
Published: (2023)
by: Cui, Kai, et al.
Published: (2023)
Complex Instruction Following with Diverse Style Policies in Football Games
by: Sun, Chenglu, et al.
Published: (2025)
by: Sun, Chenglu, et al.
Published: (2025)
Generating Local Shields for Decentralised Partially Observable Markov Decision Processes
by: Yang, Haoran, et al.
Published: (2026)
by: Yang, Haoran, et al.
Published: (2026)
Collaborative State Fusion in Partially Known Multi-agent Environments
by: Zhou, Tianlong, et al.
Published: (2024)
by: Zhou, Tianlong, et al.
Published: (2024)
RecBayes: Recurrent Bayesian Ad Hoc Teamwork in Large Partially Observable Domains
by: Ribeiro, João G., et al.
Published: (2025)
by: Ribeiro, João G., et al.
Published: (2025)
PIANIST: Learning Partially Observable World Models with LLMs for Multi-Agent Decision Making
by: Light, Jonathan, et al.
Published: (2024)
by: Light, Jonathan, et al.
Published: (2024)
Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning
by: Yang, Shan, et al.
Published: (2026)
by: Yang, Shan, et al.
Published: (2026)
Partially Observable Multi-Agent Reinforcement Learning with Information Sharing
by: Liu, Xiangyu, et al.
Published: (2023)
by: Liu, Xiangyu, et al.
Published: (2023)
Remembering the Markov Property in Cooperative MARL
by: Tessera, Kale-ab Abebe, et al.
Published: (2025)
by: Tessera, Kale-ab Abebe, et al.
Published: (2025)
Open Ad Hoc Teamwork with Cooperative Game Theory
by: Wang, Jianhong, et al.
Published: (2024)
by: Wang, Jianhong, et al.
Published: (2024)
Distributed Policy Gradient for Linear Quadratic Networked Control with Limited Communication Range
by: Yan, Yuzi, et al.
Published: (2024)
by: Yan, Yuzi, et al.
Published: (2024)
Ranking Joint Policies in Dynamic Games using Evolutionary Dynamics
by: Koliou, Natalia, et al.
Published: (2025)
by: Koliou, Natalia, et al.
Published: (2025)
Conformal Off-Policy Prediction for Multi-Agent Systems
by: Kuipers, Tom, et al.
Published: (2024)
by: Kuipers, Tom, et al.
Published: (2024)
Approximate Linear Programming for Decentralized Policy Iteration in Cooperative Multi-agent Markov Decision Processes
by: Mandal, Lakshmi, et al.
Published: (2023)
by: Mandal, Lakshmi, et al.
Published: (2023)
Policy Optimization in Multi-Agent Settings under Partially Observable Environments
by: Zhaikhan, Ainur, et al.
Published: (2025)
by: Zhaikhan, Ainur, et al.
Published: (2025)
Scalable Offline Reinforcement Learning for Mean Field Games
by: Brunnbauer, Axel, et al.
Published: (2024)
by: Brunnbauer, Axel, et al.
Published: (2024)
Light Aircraft Game : Basic Implementation and training results analysis
by: Cao, Hanzhong
Published: (2025)
by: Cao, Hanzhong
Published: (2025)
Tractable Equilibrium Computation in Markov Games through Risk Aversion
by: Mazumdar, Eric, et al.
Published: (2024)
by: Mazumdar, Eric, et al.
Published: (2024)
Multi-agent Cooperative Games Using Belief Map Assisted Training
by: Huang, Qinwei, et al.
Published: (2024)
by: Huang, Qinwei, et al.
Published: (2024)
Multi-Agent Model-Based Reinforcement Learning with Joint State-Action Learned Embeddings
by: Wang, Zhizun, et al.
Published: (2026)
by: Wang, Zhizun, et al.
Published: (2026)
Automatic Gradient Estimation for Calibrating Crowd Models with Discrete Decision Making
by: Andelfinger, Philipp, et al.
Published: (2024)
by: Andelfinger, Philipp, et al.
Published: (2024)
Shapley Value Based Multi-Agent Reinforcement Learning: Theory, Method and Its Application to Energy Network
by: Wang, Jianhong
Published: (2024)
by: Wang, Jianhong
Published: (2024)
Classification with a Network of Partially Informative Agents: Enabling Wise Crowds from Individually Myopic Classifiers
by: Yao, Tong, et al.
Published: (2024)
by: Yao, Tong, et al.
Published: (2024)
Learning to Control Unknown Strongly Monotone Games
by: Chandak, Siddharth, et al.
Published: (2024)
by: Chandak, Siddharth, et al.
Published: (2024)
Assigning Credit with Partial Reward Decoupling in Multi-Agent Proximal Policy Optimization
by: Kapoor, Aditya, et al.
Published: (2024)
by: Kapoor, Aditya, et al.
Published: (2024)
Fairness Aware Reinforcement Learning via Proximal Policy Optimization
by: La Malfa, Gabriele, et al.
Published: (2025)
by: La Malfa, Gabriele, et al.
Published: (2025)
Hierarchical Policy-Gradient Reinforcement Learning for Multi-Agent Shepherding Control of Non-Cohesive Targets
by: Covone, Stefano, et al.
Published: (2025)
by: Covone, Stefano, et al.
Published: (2025)
Co-Optimizing Reconfigurable Environments and Policies for Decentralized Multi-Agent Navigation
by: Gao, Zhan, et al.
Published: (2024)
by: Gao, Zhan, et al.
Published: (2024)
Federated Offline Reinforcement Learning: Collaborative Single-Policy Coverage Suffices
by: Woo, Jiin, et al.
Published: (2024)
by: Woo, Jiin, et al.
Published: (2024)
Cooperative Game-Theoretic Credit Assignment for Multi-Agent Policy Gradients via the Core
by: Ji, Mengda, et al.
Published: (2025)
by: Ji, Mengda, et al.
Published: (2025)
GNN-Empowered Effective Partial Observation MARL Method for AoI Management in Multi-UAV Network
by: Pan, Yuhao, et al.
Published: (2024)
by: Pan, Yuhao, et al.
Published: (2024)
Convergence of Byzantine-Resilient Gradient Tracking via Probabilistic Edge Dropout
by: Dezhboro, Amirhossein, et al.
Published: (2026)
by: Dezhboro, Amirhossein, et al.
Published: (2026)
Similar Items
-
Independent Learning of Nash Equilibria in Partially Observable Markov Potential Games with Decoupled Dynamics
by: Jordan, Philip, et al.
Published: (2026) -
Multi-agent Off-policy Actor-Critic Reinforcement Learning for Partially Observable Environments
by: Zhaikhan, Ainur, et al.
Published: (2024) -
Observation Interference in Partially Observable Assistance Games
by: Emmons, Scott, et al.
Published: (2024) -
Independent Policy Mirror Descent for Markov Potential Games: Scaling to Large Number of Players
by: Alatur, Pragnya, et al.
Published: (2024) -
Independent Learning in Constrained Markov Potential Games
by: Jordan, Philip, et al.
Published: (2024)