Convergence of Natural Policy Gradient for a Family of Infinite-State Queueing MDPs
Fuente:
arXiv
Saved in:
| Main Authors: | Grosof, Isaac, Maguluri, Siva Theja, Srikant, R. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Performance of NPG in Countable State-Space Average-Cost RL
by: Murthy, Yashaswini, et al.
Published: (2024)
by: Murthy, Yashaswini, et al.
Published: (2024)
Non-Asymptotic Convergence of Stochastic Iterative Algorithms: A Lyapunov Framework
by: Chen, Zaiwei, et al.
Published: (2026)
by: Chen, Zaiwei, et al.
Published: (2026)
Convergence Rate of the Join-the-Shortest-Queue System
by: Ma, Yuanzhe, et al.
Published: (2025)
by: Ma, Yuanzhe, et al.
Published: (2025)
Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem
by: Haque, Shaan Ul, et al.
Published: (2024)
by: Haque, Shaan Ul, et al.
Published: (2024)
Concentration of Contractive Stochastic Approximation: Additive and Multiplicative Noise
by: Chen, Zaiwei, et al.
Published: (2023)
by: Chen, Zaiwei, et al.
Published: (2023)
Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise
by: Agrawal, Shubhada, et al.
Published: (2026)
by: Agrawal, Shubhada, et al.
Published: (2026)
Tail Bounds for Queues with Abandonment: Constant, Moderate, Large Deviations, and Efficient Concentration
by: Wang, Zedong, et al.
Published: (2026)
by: Wang, Zedong, et al.
Published: (2026)
A Heavy Traffic Theory of Matching Queues
by: Varma, Sushil Mahavir, et al.
Published: (2021)
by: Varma, Sushil Mahavir, et al.
Published: (2021)
Federated Stochastic Approximation under Markov Noise and Heterogeneity: Applications in Reinforcement Learning
by: Khodadadian, Sajad, et al.
Published: (2022)
by: Khodadadian, Sajad, et al.
Published: (2022)
Markov Chain Variance Estimation: A Stochastic Approximation Approach
by: Agrawal, Shubhada, et al.
Published: (2024)
by: Agrawal, Shubhada, et al.
Published: (2024)
Linear Convergence of Entropy-Regularized Natural Policy Gradient with Linear Function Approximation
by: Cayci, Semih, et al.
Published: (2021)
by: Cayci, Semih, et al.
Published: (2021)
Tight Finite Time Bounds of Two-Time-Scale Linear Stochastic Approximation with Markovian Noise
by: Haque, Shaan Ul, et al.
Published: (2023)
by: Haque, Shaan Ul, et al.
Published: (2023)
How Accurately Can a Gaussian Approximate Stochastic Approximation Iterates?
by: Haque, Shaan Ul, et al.
Published: (2026)
by: Haque, Shaan Ul, et al.
Published: (2026)
Steady-State Behavior of Constant-Stepsize Stochastic Approximation: Gaussian Approximation and Tail Bounds
by: Wang, Zedong, et al.
Published: (2026)
by: Wang, Zedong, et al.
Published: (2026)
Higher-Order Approximations of Sojourn Times in M/G/1 Queues via Stein's Method
by: Chatterjee, Bihan, et al.
Published: (2026)
by: Chatterjee, Bihan, et al.
Published: (2026)
Policy Gradient in Robust MDPs with Global Convergence Guarantee
by: Wang, Qiuhao, et al.
Published: (2022)
by: Wang, Qiuhao, et al.
Published: (2022)
A Non-Asymptotic Theory of Seminorm Lyapunov Stability: From Deterministic to Stochastic Iterative Algorithms
by: Chen, Zaiwei, et al.
Published: (2025)
by: Chen, Zaiwei, et al.
Published: (2025)
Dynamic Pricing and Matching for Two-Sided Queues
by: Varma, Sushil Mahavir, et al.
Published: (2019)
by: Varma, Sushil Mahavir, et al.
Published: (2019)
Order-Optimal Regret with Novel Policy Gradient Approaches in Infinite-Horizon Average Reward MDPs
by: Ganesh, Swetha, et al.
Published: (2024)
by: Ganesh, Swetha, et al.
Published: (2024)
Finite-Time Behavior of Erlang-C Model: Mixing Time, Mean Queue Length and Tail Bounds
by: Nguyen, Hoang Huy, et al.
Published: (2025)
by: Nguyen, Hoang Huy, et al.
Published: (2025)
On the Global Convergence of Policy Gradient in Average Reward Markov Decision Processes
by: Kumar, Navdeep, et al.
Published: (2024)
by: Kumar, Navdeep, et al.
Published: (2024)
Confident Natural Policy Gradient for Local Planning in $q_π$-realizable Constrained MDPs
by: Tian, Tian, et al.
Published: (2024)
by: Tian, Tian, et al.
Published: (2024)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Learning General Parameterized Policies for Infinite Horizon Average Reward Constrained MDPs via Primal-Dual Policy Gradient Algorithm
by: Bai, Qinbo, et al.
Published: (2024)
by: Bai, Qinbo, et al.
Published: (2024)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
by: Srikant, R.
Published: (2024)
by: Srikant, R.
Published: (2024)
Stochastic Approximation for Nonlinear Discrete Stochastic Control: Finite-Sample Bounds
by: Nguyen, Hoang Huy, et al.
Published: (2023)
by: Nguyen, Hoang Huy, et al.
Published: (2023)
On the Convergence of Modified Policy Iteration in Risk Sensitive Exponential Cost Markov Decision Processes
by: Murthy, Yashaswini, et al.
Published: (2023)
by: Murthy, Yashaswini, et al.
Published: (2023)
Soft Robust MDPs and Risk-Sensitive MDPs: Equivalence, Policy Gradient, and Sample Complexity
by: Zhang, Runyu, et al.
Published: (2023)
by: Zhang, Runyu, et al.
Published: (2023)
Towards Principled, Practical Policy Gradient for Bandits and Tabular MDPs
by: Lu, Michael, et al.
Published: (2024)
by: Lu, Michael, et al.
Published: (2024)
Last-Iterate Convergence of General Parameterized Policies in Constrained MDPs
by: Mondal, Washim Uddin, et al.
Published: (2024)
by: Mondal, Washim Uddin, et al.
Published: (2024)
Why Policy Gradient Algorithms Work for Undiscounted Total-Reward MDPs
by: Lee, Jongmin, et al.
Published: (2025)
by: Lee, Jongmin, et al.
Published: (2025)
Provably Convergent Primal-Dual DPO for Constrained LLM Alignment
by: Du, Yihan, et al.
Published: (2025)
by: Du, Yihan, et al.
Published: (2025)
On Value Iteration Convergence in Connected MDPs
by: Mustafin, Arsenii, et al.
Published: (2024)
by: Mustafin, Arsenii, et al.
Published: (2024)
Reinforcement Learning for Infinite-Horizon Average-Reward Linear MDPs via Approximation by Discounted-Reward MDPs
by: Hong, Kihyuk, et al.
Published: (2024)
by: Hong, Kihyuk, et al.
Published: (2024)
Global Convergence of Average Reward Constrained MDPs with Neural Critic and General Policy Parameterization
by: Satheesh, Anirudh, et al.
Published: (2026)
by: Satheesh, Anirudh, et al.
Published: (2026)
Linear Convergence of Independent Natural Policy Gradient in Games with Entropy Regularization
by: Sun, Youbang, et al.
Published: (2024)
by: Sun, Youbang, et al.
Published: (2024)
Convergence of Distributionally Robust Q-Learning with Linear Function Approximation
by: Mandal, Saptarshi, et al.
Published: (2025)
by: Mandal, Saptarshi, et al.
Published: (2025)
A Policy Gradient Primal-Dual Algorithm for Constrained MDPs with Uniform PAC Guarantees
by: Kitamura, Toshinori, et al.
Published: (2024)
by: Kitamura, Toshinori, et al.
Published: (2024)
Reusing Historical Trajectories in Natural Policy Gradient via Importance Sampling: Convergence and Convergence Rate
by: Lin, Yifan, et al.
Published: (2024)
by: Lin, Yifan, et al.
Published: (2024)
Scalable Policy-Based RL Algorithms for POMDPs
by: Anjarlekar, Ameya, et al.
Published: (2025)
by: Anjarlekar, Ameya, et al.
Published: (2025)
Similar Items
-
Performance of NPG in Countable State-Space Average-Cost RL
by: Murthy, Yashaswini, et al.
Published: (2024) -
Non-Asymptotic Convergence of Stochastic Iterative Algorithms: A Lyapunov Framework
by: Chen, Zaiwei, et al.
Published: (2026) -
Convergence Rate of the Join-the-Shortest-Queue System
by: Ma, Yuanzhe, et al.
Published: (2025) -
Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem
by: Haque, Shaan Ul, et al.
Published: (2024) -
Concentration of Contractive Stochastic Approximation: Additive and Multiplicative Noise
by: Chen, Zaiwei, et al.
Published: (2023)