Approximate Linear Programming for Decentralized Policy Iteration in Cooperative Multi-agent Markov Decision Processes
Fuente:
arXiv
Saved in:
| Main Authors: | Mandal, Lakshmi, Lakshminarayanan, Chandrashekar, Bhatnagar, Shalabh |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Discrete GCBF Proximal Policy Optimization for Multi-agent Safe Optimal Control
by: Zhang, Songyuan, et al.
Published: (2025)
by: Zhang, Songyuan, et al.
Published: (2025)
Linear Convergence of Independent Natural Policy Gradient in Games with Entropy Regularization
by: Sun, Youbang, et al.
Published: (2024)
by: Sun, Youbang, et al.
Published: (2024)
Structured Cooperative Multi-Agent Reinforcement Learning: a Bayesian Network Perspective
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
Hierarchical Decentralized Stochastic Control for Cyber-Physical Systems
by: Kaza, Kesav, et al.
Published: (2025)
by: Kaza, Kesav, et al.
Published: (2025)
High-Probability Convergence Guarantees of Decentralized SGD
by: Armacki, Aleksandar, et al.
Published: (2025)
by: Armacki, Aleksandar, et al.
Published: (2025)
Decentralized Online Riemannian Optimization Beyond Hadamard Manifolds
by: Sahinoglu, Emre, et al.
Published: (2025)
by: Sahinoglu, Emre, et al.
Published: (2025)
Near-Optimal Online Learning for Multi-Agent Submodular Coordination: Tight Approximation and Communication Efficiency
by: Zhang, Qixin, et al.
Published: (2025)
by: Zhang, Qixin, et al.
Published: (2025)
High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking
by: Armacki, Aleksandar, et al.
Published: (2026)
by: Armacki, Aleksandar, et al.
Published: (2026)
Adaptive Decentralized Composite Optimization via Three-Operator Splitting
by: Chen, Xiaokai, et al.
Published: (2026)
by: Chen, Xiaokai, et al.
Published: (2026)
Effective Policy Learning for Multi-Agent Online Coordination Beyond Submodular Objectives
by: Zhang, Qixin, et al.
Published: (2025)
by: Zhang, Qixin, et al.
Published: (2025)
Learning Decentralized Partially Observable Mean Field Control for Artificial Collective Behavior
by: Cui, Kai, et al.
Published: (2023)
by: Cui, Kai, et al.
Published: (2023)
Scalable spectral representations for multi-agent reinforcement learning in network MDPs
by: Ren, Zhaolin, et al.
Published: (2024)
by: Ren, Zhaolin, et al.
Published: (2024)
Exploiting inter-agent coupling information for efficient reinforcement learning of cooperative LQR
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
An active learning method for solving competitive multi-agent decision-making and control problems
by: Fabiani, Filippo, et al.
Published: (2022)
by: Fabiani, Filippo, et al.
Published: (2022)
Distributed Online Submodular Maximization under Communication Delays: A Simultaneous Decision-Making Approach
by: Xu, Zirui, et al.
Published: (2026)
by: Xu, Zirui, et al.
Published: (2026)
Deep Distributed Optimization for Large-Scale Quadratic Programming
by: Saravanos, Augustinos D., et al.
Published: (2024)
by: Saravanos, Augustinos D., et al.
Published: (2024)
Locally Interdependent Multi-Agent MDP: Theoretical Framework for Decentralized Agents with Dynamic Dependencies
by: DeWeese, Alex, et al.
Published: (2024)
by: DeWeese, Alex, et al.
Published: (2024)
Stochastic Mirror Descent under Iterate-Dependent Markov Noise: Analysis in the Asymptotic and Finite Time Regimes
by: Paul, Anik Kumar, et al.
Published: (2026)
by: Paul, Anik Kumar, et al.
Published: (2026)
Optimization and Learning in Open Multi-Agent Systems
by: Deplano, Diego, et al.
Published: (2025)
by: Deplano, Diego, et al.
Published: (2025)
Logarithmically Quantized Distributed Optimization over Dynamic Multi-Agent Networks
by: Doostmohammadian, Mohammadreza, et al.
Published: (2024)
by: Doostmohammadian, Mohammadreza, et al.
Published: (2024)
DeMuon: A Decentralized Muon for Matrix Optimization over Graphs
by: He, Chuan, et al.
Published: (2025)
by: He, Chuan, et al.
Published: (2025)
Distributed Markov Chain Monte Carlo Sampling based on the Alternating Direction Method of Multipliers
by: Tzikas, Alexandros E., et al.
Published: (2024)
by: Tzikas, Alexandros E., et al.
Published: (2024)
On the Reliability Limits of LLM-Based Multi-Agent Planning
by: Ao, Ruicheng, et al.
Published: (2026)
by: Ao, Ruicheng, et al.
Published: (2026)
Decentralized Contingency MPC based on Safe Sets for Nonlinear Multi-agent Collision Avoidance
by: Studt, Max, et al.
Published: (2026)
by: Studt, Max, et al.
Published: (2026)
Online Reinforcement Learning in Markov Decision Process Using Linear Programming
by: Leon, Vincent, et al.
Published: (2023)
by: Leon, Vincent, et al.
Published: (2023)
Major-Minor Mean Field Multi-Agent Reinforcement Learning
by: Cui, Kai, et al.
Published: (2023)
by: Cui, Kai, et al.
Published: (2023)
Community-based Multi-Agent Reinforcement Learning with Transfer and Active Exploration
by: Shi, Zhaoyang
Published: (2025)
by: Shi, Zhaoyang
Published: (2025)
Policy Optimization and Multi-agent Reinforcement Learning for Mean-variance Team Stochastic Games
by: Hu, Junkai, et al.
Published: (2025)
by: Hu, Junkai, et al.
Published: (2025)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
by: Adibi, Arman, et al.
Published: (2024)
by: Adibi, Arman, et al.
Published: (2024)
Policy Iteration for Two-Player General-Sum Stochastic Stackelberg Games
by: Kudo, Mikoto, et al.
Published: (2024)
by: Kudo, Mikoto, et al.
Published: (2024)
Thinking Beyond Visibility: A Near-Optimal Policy Framework for Locally Interdependent Multi-Agent MDPs
by: DeWeese, Alex, et al.
Published: (2025)
by: DeWeese, Alex, et al.
Published: (2025)
SAFE--MA--RRT: Multi-Agent Motion Planning with Data-Driven Safety Certificates
by: Esmaeili, Babak, et al.
Published: (2025)
by: Esmaeili, Babak, et al.
Published: (2025)
Robust Online Learning over Networks
by: Bastianello, Nicola, et al.
Published: (2023)
by: Bastianello, Nicola, et al.
Published: (2023)
A Generalized Sinkhorn Algorithm for Mean-Field Schrödinger Bridge
by: Eldesoukey, Asmaa, et al.
Published: (2026)
by: Eldesoukey, Asmaa, et al.
Published: (2026)
Formation Shape Control using the Gromov-Wasserstein Metric
by: Nakashima, Haruto, et al.
Published: (2025)
by: Nakashima, Haruto, et al.
Published: (2025)
Principled Learning-to-Communicate with Quasi-Classical Information Structures
by: Liu, Xiangyu, et al.
Published: (2026)
by: Liu, Xiangyu, et al.
Published: (2026)
Distributed Random Reshuffling Methods with Improved Convergence
by: Huang, Kun, et al.
Published: (2023)
by: Huang, Kun, et al.
Published: (2023)
Analysis of Multiscale Reinforcement Q-Learning Algorithms for Mean Field Control Games
by: Angiuli, Andrea, et al.
Published: (2024)
by: Angiuli, Andrea, et al.
Published: (2024)
Continuous-time q-learning for mean-field control with common noise, part-I: Theoretical foundations
by: Ren, Zhenjie, et al.
Published: (2026)
by: Ren, Zhenjie, et al.
Published: (2026)
Data/moment-driven approaches for fast predictive control of collective dynamics
by: Albi, Giacomo, et al.
Published: (2024)
by: Albi, Giacomo, et al.
Published: (2024)
Similar Items
-
Discrete GCBF Proximal Policy Optimization for Multi-agent Safe Optimal Control
by: Zhang, Songyuan, et al.
Published: (2025) -
Linear Convergence of Independent Natural Policy Gradient in Games with Entropy Regularization
by: Sun, Youbang, et al.
Published: (2024) -
Structured Cooperative Multi-Agent Reinforcement Learning: a Bayesian Network Perspective
by: Syed, Shahbaz P Qadri, et al.
Published: (2025) -
Hierarchical Decentralized Stochastic Control for Cyber-Physical Systems
by: Kaza, Kesav, et al.
Published: (2025) -
High-Probability Convergence Guarantees of Decentralized SGD
by: Armacki, Aleksandar, et al.
Published: (2025)