Linear Convergence of Independent Natural Policy Gradient in Games with Entropy Regularization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Youbang, Liu, Tao, Kumar, P. R., Shahrampour, Shahin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Decentralized Online Riemannian Optimization Beyond Hadamard Manifolds
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025)
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025)
Local Linear Convergence of Infeasible Optimization with Orthogonal Constraints
von: Sun, Youbang, et al.
Veröffentlicht: (2024)
von: Sun, Youbang, et al.
Veröffentlicht: (2024)
Regret Analysis of Policy Optimization over Submanifolds for Linearly Constrained Online LQG
von: Chang, Ting-Jui, et al.
Veröffentlicht: (2024)
von: Chang, Ting-Jui, et al.
Veröffentlicht: (2024)
Finite-Time Analysis of Stochastic Nonconvex Nonsmooth Optimization on the Riemannian Manifolds
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025)
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025)
High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking
von: Armacki, Aleksandar, et al.
Veröffentlicht: (2026)
von: Armacki, Aleksandar, et al.
Veröffentlicht: (2026)
Retraction-Free Decentralized Non-convex Optimization with Orthogonal Constraints
von: Sun, Youbang, et al.
Veröffentlicht: (2024)
von: Sun, Youbang, et al.
Veröffentlicht: (2024)
Online Optimization Perspective on First-Order and Zero-Order Decentralized Nonsmooth Nonconvex Stochastic Optimization
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2024)
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2024)
Approximate Linear Programming for Decentralized Policy Iteration in Cooperative Multi-agent Markov Decision Processes
von: Mandal, Lakshmi, et al.
Veröffentlicht: (2023)
von: Mandal, Lakshmi, et al.
Veröffentlicht: (2023)
Effective Policy Learning for Multi-Agent Online Coordination Beyond Submodular Objectives
von: Zhang, Qixin, et al.
Veröffentlicht: (2025)
von: Zhang, Qixin, et al.
Veröffentlicht: (2025)
Linear Convergence in Games with Delayed Feedback via Extra Prediction
von: Fujimoto, Yuma, et al.
Veröffentlicht: (2026)
von: Fujimoto, Yuma, et al.
Veröffentlicht: (2026)
High-Probability Convergence Guarantees of Decentralized SGD
von: Armacki, Aleksandar, et al.
Veröffentlicht: (2025)
von: Armacki, Aleksandar, et al.
Veröffentlicht: (2025)
Distributed Random Reshuffling Methods with Improved Convergence
von: Huang, Kun, et al.
Veröffentlicht: (2023)
von: Huang, Kun, et al.
Veröffentlicht: (2023)
Linear Convergence of Entropy-Regularized Natural Policy Gradient with Linear Function Approximation
von: Cayci, Semih, et al.
Veröffentlicht: (2021)
von: Cayci, Semih, et al.
Veröffentlicht: (2021)
Analysis of Multiscale Reinforcement Q-Learning Algorithms for Mean Field Control Games
von: Angiuli, Andrea, et al.
Veröffentlicht: (2024)
von: Angiuli, Andrea, et al.
Veröffentlicht: (2024)
Securing Equal Share: A Principled Approach for Learning Multiplayer Symmetric Games
von: Ge, Jiawei, et al.
Veröffentlicht: (2024)
von: Ge, Jiawei, et al.
Veröffentlicht: (2024)
Online Optimization on Hadamard Manifolds: Curvature Independent Regret Bounds on Horospherically Convex Objectives
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025)
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025)
Near-Optimal Online Learning for Multi-Agent Submodular Coordination: Tight Approximation and Communication Efficiency
von: Zhang, Qixin, et al.
Veröffentlicht: (2025)
von: Zhang, Qixin, et al.
Veröffentlicht: (2025)
Exploiting inter-agent coupling information for efficient reinforcement learning of cooperative LQR
von: Syed, Shahbaz P Qadri, et al.
Veröffentlicht: (2025)
von: Syed, Shahbaz P Qadri, et al.
Veröffentlicht: (2025)
Structured Cooperative Multi-Agent Reinforcement Learning: a Bayesian Network Perspective
von: Syed, Shahbaz P Qadri, et al.
Veröffentlicht: (2025)
von: Syed, Shahbaz P Qadri, et al.
Veröffentlicht: (2025)
Principled Learning-to-Communicate with Quasi-Classical Information Structures
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
CEDAS: A Compressed Decentralized Stochastic Gradient Method with Improved Convergence
von: Huang, Kun, et al.
Veröffentlicht: (2023)
von: Huang, Kun, et al.
Veröffentlicht: (2023)
Discrete GCBF Proximal Policy Optimization for Multi-agent Safe Optimal Control
von: Zhang, Songyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Songyuan, et al.
Veröffentlicht: (2025)
Scalable spectral representations for multi-agent reinforcement learning in network MDPs
von: Ren, Zhaolin, et al.
Veröffentlicht: (2024)
von: Ren, Zhaolin, et al.
Veröffentlicht: (2024)
Logarithmically Quantized Distributed Optimization over Dynamic Multi-Agent Networks
von: Doostmohammadian, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Doostmohammadian, Mohammadreza, et al.
Veröffentlicht: (2024)
A Generalized Sinkhorn Algorithm for Mean-Field Schrödinger Bridge
von: Eldesoukey, Asmaa, et al.
Veröffentlicht: (2026)
von: Eldesoukey, Asmaa, et al.
Veröffentlicht: (2026)
Distributed Online Submodular Maximization under Communication Delays: A Simultaneous Decision-Making Approach
von: Xu, Zirui, et al.
Veröffentlicht: (2026)
von: Xu, Zirui, et al.
Veröffentlicht: (2026)
Hierarchical Decentralized Stochastic Control for Cyber-Physical Systems
von: Kaza, Kesav, et al.
Veröffentlicht: (2025)
von: Kaza, Kesav, et al.
Veröffentlicht: (2025)
Robust Online Learning over Networks
von: Bastianello, Nicola, et al.
Veröffentlicht: (2023)
von: Bastianello, Nicola, et al.
Veröffentlicht: (2023)
An active learning method for solving competitive multi-agent decision-making and control problems
von: Fabiani, Filippo, et al.
Veröffentlicht: (2022)
von: Fabiani, Filippo, et al.
Veröffentlicht: (2022)
Formation Shape Control using the Gromov-Wasserstein Metric
von: Nakashima, Haruto, et al.
Veröffentlicht: (2025)
von: Nakashima, Haruto, et al.
Veröffentlicht: (2025)
Optimization and Learning in Open Multi-Agent Systems
von: Deplano, Diego, et al.
Veröffentlicht: (2025)
von: Deplano, Diego, et al.
Veröffentlicht: (2025)
Policy Iteration for Two-Player General-Sum Stochastic Stackelberg Games
von: Kudo, Mikoto, et al.
Veröffentlicht: (2024)
von: Kudo, Mikoto, et al.
Veröffentlicht: (2024)
Bench-MFG: A Benchmark Suite for Learning in Stationary Mean Field Games
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2026)
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2026)
Stochastic Semi-Gradient Descent for Learning Mean Field Games with Population-Aware Function Approximation
von: Zhang, Chenyu, et al.
Veröffentlicht: (2024)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2024)
Data/moment-driven approaches for fast predictive control of collective dynamics
von: Albi, Giacomo, et al.
Veröffentlicht: (2024)
von: Albi, Giacomo, et al.
Veröffentlicht: (2024)
Deep Distributed Optimization for Large-Scale Quadratic Programming
von: Saravanos, Augustinos D., et al.
Veröffentlicht: (2024)
von: Saravanos, Augustinos D., et al.
Veröffentlicht: (2024)
Momentum for the Win: Collaborative Federated Reinforcement Learning across Heterogeneous Environments
von: Wang, Han, et al.
Veröffentlicht: (2024)
von: Wang, Han, et al.
Veröffentlicht: (2024)
Solving Continuous Mean Field Games: Deep Reinforcement Learning for Non-Stationary Dynamics
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2025)
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2025)
Community-based Multi-Agent Reinforcement Learning with Transfer and Active Exploration
von: Shi, Zhaoyang
Veröffentlicht: (2025)
von: Shi, Zhaoyang
Veröffentlicht: (2025)
On the Reliability Limits of LLM-Based Multi-Agent Planning
von: Ao, Ruicheng, et al.
Veröffentlicht: (2026)
von: Ao, Ruicheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Decentralized Online Riemannian Optimization Beyond Hadamard Manifolds
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025) -
Local Linear Convergence of Infeasible Optimization with Orthogonal Constraints
von: Sun, Youbang, et al.
Veröffentlicht: (2024) -
Regret Analysis of Policy Optimization over Submanifolds for Linearly Constrained Online LQG
von: Chang, Ting-Jui, et al.
Veröffentlicht: (2024) -
Finite-Time Analysis of Stochastic Nonconvex Nonsmooth Optimization on the Riemannian Manifolds
von: Sahinoglu, Emre, et al.
Veröffentlicht: (2025) -
High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking
von: Armacki, Aleksandar, et al.
Veröffentlicht: (2026)