Performance of NPG in Countable State-Space Average-Cost RL
Fuente:
arXiv
Saved in:
| Main Authors: | Murthy, Yashaswini, Grosof, Isaac, Maguluri, Siva Theja, Srikant, R. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Convergence of Natural Policy Gradient for a Family of Infinite-State Queueing MDPs
by: Grosof, Isaac, et al.
Published: (2024)
by: Grosof, Isaac, et al.
Published: (2024)
Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem
by: Haque, Shaan Ul, et al.
Published: (2024)
by: Haque, Shaan Ul, et al.
Published: (2024)
Concentration of Contractive Stochastic Approximation: Additive and Multiplicative Noise
by: Chen, Zaiwei, et al.
Published: (2023)
by: Chen, Zaiwei, et al.
Published: (2023)
Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise
by: Agrawal, Shubhada, et al.
Published: (2026)
by: Agrawal, Shubhada, et al.
Published: (2026)
Tight Finite Time Bounds of Two-Time-Scale Linear Stochastic Approximation with Markovian Noise
by: Haque, Shaan Ul, et al.
Published: (2023)
by: Haque, Shaan Ul, et al.
Published: (2023)
A Non-Asymptotic Theory of Seminorm Lyapunov Stability: From Deterministic to Stochastic Iterative Algorithms
by: Chen, Zaiwei, et al.
Published: (2025)
by: Chen, Zaiwei, et al.
Published: (2025)
On the Convergence of Modified Policy Iteration in Risk Sensitive Exponential Cost Markov Decision Processes
by: Murthy, Yashaswini, et al.
Published: (2023)
by: Murthy, Yashaswini, et al.
Published: (2023)
Electric Vehicle Fleet and Charging Infrastructure Planning
by: Varma, Sushil Mahavir, et al.
Published: (2023)
by: Varma, Sushil Mahavir, et al.
Published: (2023)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
by: Srikant, R.
Published: (2024)
by: Srikant, R.
Published: (2024)
On the Global Convergence of Policy Gradient in Average Reward Markov Decision Processes
by: Kumar, Navdeep, et al.
Published: (2024)
by: Kumar, Navdeep, et al.
Published: (2024)
Dynamic Pricing and Matching for Two-Sided Queues
by: Varma, Sushil Mahavir, et al.
Published: (2019)
by: Varma, Sushil Mahavir, et al.
Published: (2019)
Optimal Pricing in Multi Server Systems
by: S, Ashok Krishnan K., et al.
Published: (2021)
by: S, Ashok Krishnan K., et al.
Published: (2021)
Linear Convergence of Entropy-Regularized Natural Policy Gradient with Linear Function Approximation
by: Cayci, Semih, et al.
Published: (2021)
by: Cayci, Semih, et al.
Published: (2021)
Stochastic Approximation for Nonlinear Discrete Stochastic Control: Finite-Sample Bounds
by: Nguyen, Hoang Huy, et al.
Published: (2023)
by: Nguyen, Hoang Huy, et al.
Published: (2023)
Almost Sure Convergence of Stochastic Approximation: An Interplay of Noise and Step Size
by: Nguyen, Quang Dinh Thien, et al.
Published: (2026)
by: Nguyen, Quang Dinh Thien, et al.
Published: (2026)
Non-Asymptotic Convergence of Stochastic Iterative Algorithms: A Lyapunov Framework
by: Chen, Zaiwei, et al.
Published: (2026)
by: Chen, Zaiwei, et al.
Published: (2026)
Incentive-Aware Federated Averaging with Performance Guarantees under Strategic Participation
by: Maleki, Fateme, et al.
Published: (2026)
by: Maleki, Fateme, et al.
Published: (2026)
Span-Agnostic Optimal Sample Complexity and Oracle Inequalities for Average-Reward RL
by: Zurek, Matthew, et al.
Published: (2025)
by: Zurek, Matthew, et al.
Published: (2025)
Q-Measure-Learning for Continuous State RL: Efficient Implementation and Convergence
by: Wang, Shengbo
Published: (2026)
by: Wang, Shengbo
Published: (2026)
Optimal Single-Policy Sample Complexity and Transient Coverage for Average-Reward Offline RL
by: Zurek, Matthew, et al.
Published: (2025)
by: Zurek, Matthew, et al.
Published: (2025)
Going from a Representative Agent to Counterfactuals in Combinatorial Choice
by: Ruan, Yanqiu, et al.
Published: (2025)
by: Ruan, Yanqiu, et al.
Published: (2025)
Convergence of Distributionally Robust Q-Learning with Linear Function Approximation
by: Mandal, Saptarshi, et al.
Published: (2025)
by: Mandal, Saptarshi, et al.
Published: (2025)
Efficient Trajectory Inference in Wasserstein Space Using Consecutive Averaging
by: Banerjee, Amartya, et al.
Published: (2024)
by: Banerjee, Amartya, et al.
Published: (2024)
DADA: Dual Averaging with Distance Adaptation
by: Moshtaghifar, Mohammad, et al.
Published: (2025)
by: Moshtaghifar, Mohammad, et al.
Published: (2025)
Robust Regression over Averaged Uncertainty
by: Bertsimas, Dimitris, et al.
Published: (2023)
by: Bertsimas, Dimitris, et al.
Published: (2023)
Regularized Q-learning through Robust Averaging
by: Schmitt-Förster, Peter, et al.
Published: (2024)
by: Schmitt-Förster, Peter, et al.
Published: (2024)
A Unified Analysis for Finite Weight Averaging
by: Wang, Peng, et al.
Published: (2024)
by: Wang, Peng, et al.
Published: (2024)
Planning and Learning in Average Risk-aware MDPs
by: Wang, Weikai, et al.
Published: (2025)
by: Wang, Weikai, et al.
Published: (2025)
Data Deletion Can Help in Adaptive RL
by: Budhraja, Param, et al.
Published: (2026)
by: Budhraja, Param, et al.
Published: (2026)
Refined Analysis of Federated Averaging and Federated Richardson-Romberg
by: Mangold, Paul, et al.
Published: (2024)
by: Mangold, Paul, et al.
Published: (2024)
Central Limit Theorems for Asynchronous Averaged Q-Learning
by: Liu, Xingtu
Published: (2025)
by: Liu, Xingtu
Published: (2025)
Composite Optimization with Error Feedback: the Dual Averaging Approach
by: Gao, Yuan, et al.
Published: (2025)
by: Gao, Yuan, et al.
Published: (2025)
Layer-wise Quantization for Quantized Optimistic Dual Averaging
by: Nguyen, Anh Duc, et al.
Published: (2025)
by: Nguyen, Anh Duc, et al.
Published: (2025)
A Simplified Analysis of SGD for Linear Regression with Weight Averaging
by: Meterez, Alexandru, et al.
Published: (2025)
by: Meterez, Alexandru, et al.
Published: (2025)
Unified Convergence Analysis for Adaptive Optimization with Moving Average Estimator
by: Guo, Zhishuai, et al.
Published: (2021)
by: Guo, Zhishuai, et al.
Published: (2021)
Optimal Sample Complexity for Average Reward Markov Decision Processes
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
by: Chen, Zijun, et al.
Published: (2025)
by: Chen, Zijun, et al.
Published: (2025)
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
by: Boone, Victor, et al.
Published: (2024)
by: Boone, Victor, et al.
Published: (2024)
Efficient Private SCO for Heavy-Tailed Data via Averaged Clipping
by: Jin, Chenhan, et al.
Published: (2022)
by: Jin, Chenhan, et al.
Published: (2022)
A Heuristic Approach for Performance Tuning in RL-based Quadrotor Control via Reward Design and Termination Conditions
by: Suarez, Fausto Mauricio Lagos, et al.
Published: (2026)
by: Suarez, Fausto Mauricio Lagos, et al.
Published: (2026)
Similar Items
-
Convergence of Natural Policy Gradient for a Family of Infinite-State Queueing MDPs
by: Grosof, Isaac, et al.
Published: (2024) -
Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem
by: Haque, Shaan Ul, et al.
Published: (2024) -
Concentration of Contractive Stochastic Approximation: Additive and Multiplicative Noise
by: Chen, Zaiwei, et al.
Published: (2023) -
Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise
by: Agrawal, Shubhada, et al.
Published: (2026) -
Tight Finite Time Bounds of Two-Time-Scale Linear Stochastic Approximation with Markovian Noise
by: Haque, Shaan Ul, et al.
Published: (2023)