Analysis of On-policy Policy Gradient Methods under the Distribution Mismatch
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Weizhen, He, Jianping, Duan, Xiaoming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Elementary Analysis of Policy Gradient Methods
por: Liu, Jiacai, et al.
Publicado: (2024)
por: Liu, Jiacai, et al.
Publicado: (2024)
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
por: Klein, Sara, et al.
Publicado: (2023)
por: Klein, Sara, et al.
Publicado: (2023)
A New Convergence Analysis of Plug-and-Play Proximal Gradient Descent Under Prior Mismatch
por: Xu, Guixian, et al.
Publicado: (2026)
por: Xu, Guixian, et al.
Publicado: (2026)
Model-Free Output Feedback Stabilization via Policy Gradient Methods
por: Zhang, Ankang, et al.
Publicado: (2026)
por: Zhang, Ankang, et al.
Publicado: (2026)
Delightful Distributed Policy Gradient
por: Osband, Ian
Publicado: (2026)
por: Osband, Ian
Publicado: (2026)
Accelerated Convergence of Stochastic Heavy Ball Method under Anisotropic Gradient Noise
por: Pan, Rui, et al.
Publicado: (2023)
por: Pan, Rui, et al.
Publicado: (2023)
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
por: Xiao, Minheng, et al.
Publicado: (2024)
por: Xiao, Minheng, et al.
Publicado: (2024)
Communication-Efficient Gradient Descent-Accent Methods for Distributed Variational Inequalities: Unified Analysis and Local Updates
por: Zhang, Siqi, et al.
Publicado: (2023)
por: Zhang, Siqi, et al.
Publicado: (2023)
Clipped Gradient Methods for Nonsmooth Convex Optimization under Heavy-Tailed Noise: A Refined Analysis
por: Liu, Zijian
Publicado: (2025)
por: Liu, Zijian
Publicado: (2025)
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
por: Carmona, René, et al.
Publicado: (2019)
por: Carmona, René, et al.
Publicado: (2019)
Policy Gradient Methods for Discrete Time Linear Quadratic Regulator With Random Parameters
por: Li, Deyue
Publicado: (2023)
por: Li, Deyue
Publicado: (2023)
A Weighted Gradient Tracking Privacy-Preserving Method for Distributed Optimization
por: Xie, Furan, et al.
Publicado: (2025)
por: Xie, Furan, et al.
Publicado: (2025)
In-Expectation Convergence of Stochastic Gradient Methods under Heavy-Tailed Noise
por: Liu, Zijian
Publicado: (2026)
por: Liu, Zijian
Publicado: (2026)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
por: Ding, Dongsheng, et al.
Publicado: (2023)
por: Ding, Dongsheng, et al.
Publicado: (2023)
Multi-level Monte-Carlo Gradient Methods for Stochastic Optimization with Biased Oracles
por: Hu, Yifan, et al.
Publicado: (2024)
por: Hu, Yifan, et al.
Publicado: (2024)
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2024)
por: Zeng, Sihan, et al.
Publicado: (2024)
Linear Convergence of Entropy-Regularized Natural Policy Gradient with Linear Function Approximation
por: Cayci, Semih, et al.
Publicado: (2021)
por: Cayci, Semih, et al.
Publicado: (2021)
Communication-Efficient Adaptive Batch Size Strategies for Distributed Local Gradient Methods
por: Lau, Tim Tsz-Kit, et al.
Publicado: (2024)
por: Lau, Tim Tsz-Kit, et al.
Publicado: (2024)
Fill-and-Spill: Deep Reinforcement Learning Policy Gradient Methods for Reservoir Operation Decision and Control
por: Tabas, Sadegh Sadeghi, et al.
Publicado: (2024)
por: Tabas, Sadegh Sadeghi, et al.
Publicado: (2024)
On the Global Convergence of Risk-Averse Natural Policy Gradient Methods with Expected Conditional Risk Measures
por: Yu, Xian, et al.
Publicado: (2023)
por: Yu, Xian, et al.
Publicado: (2023)
A Concise Lyapunov Analysis of Nesterov's Accelerated Gradient Method
por: Liu, Jun
Publicado: (2025)
por: Liu, Jun
Publicado: (2025)
Almost Sure Convergence Analysis of Differentially Private Stochastic Gradient Methods
por: Mukherjee, Amartya, et al.
Publicado: (2025)
por: Mukherjee, Amartya, et al.
Publicado: (2025)
Recurrent Natural Policy Gradient for POMDPs
por: Cayci, Semih, et al.
Publicado: (2024)
por: Cayci, Semih, et al.
Publicado: (2024)
Greedy Low-Rank Gradient Compression for Distributed Learning with Convergence Guarantees
por: Chen, Chuyan, et al.
Publicado: (2025)
por: Chen, Chuyan, et al.
Publicado: (2025)
Stochastic Gradients under Nuisances
por: Yu, Facheng, et al.
Publicado: (2025)
por: Yu, Facheng, et al.
Publicado: (2025)
Directional Smoothness and Gradient Methods: Convergence and Adaptivity
por: Mishkin, Aaron, et al.
Publicado: (2024)
por: Mishkin, Aaron, et al.
Publicado: (2024)
Robustness of Iteratively Pre-Conditioned Gradient-Descent Method: The Case of Distributed Linear Regression Problem
por: Chakrabarti, Kushal, et al.
Publicado: (2021)
por: Chakrabarti, Kushal, et al.
Publicado: (2021)
Iterative Pre-Conditioning for Expediting the Gradient-Descent Method: The Distributed Linear Least-Squares Problem
por: Chakrabarti, Kushal, et al.
Publicado: (2020)
por: Chakrabarti, Kushal, et al.
Publicado: (2020)
Gradient Methods with Online Scaling
por: Gao, Wenzhi, et al.
Publicado: (2024)
por: Gao, Wenzhi, et al.
Publicado: (2024)
Near-Optimal Convergence of Accelerated Gradient Methods under Generalized and $(L_0, L_1)$-Smoothness
por: Tyurin, Alexander
Publicado: (2025)
por: Tyurin, Alexander
Publicado: (2025)
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
por: Han, Yinbin, et al.
Publicado: (2023)
por: Han, Yinbin, et al.
Publicado: (2023)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
por: Ganesh, Swetha, et al.
Publicado: (2024)
por: Ganesh, Swetha, et al.
Publicado: (2024)
Stochastic Gradient Methods with Preconditioned Updates
por: Sadiev, Abdurakhmon, et al.
Publicado: (2022)
por: Sadiev, Abdurakhmon, et al.
Publicado: (2022)
Can Adaptive Gradient Methods Converge under Heavy-Tailed Noise? A Case Study of AdaGrad
por: Liu, Zijian
Publicado: (2026)
por: Liu, Zijian
Publicado: (2026)
Natural Policy Gradient as Doubly Smoothed Policy Iteration: A Bellman-Operator Framework
por: Nanda, Phalguni, et al.
Publicado: (2026)
por: Nanda, Phalguni, et al.
Publicado: (2026)
Decentralized Riemannian Conjugate Gradient Method on the Stiefel Manifold
por: Chen, Jun, et al.
Publicado: (2023)
por: Chen, Jun, et al.
Publicado: (2023)
Quadratic Gradient: A Unified Framework Bridging Gradient Descent and Newton-Type Methods by Synthesizing Hessians and Gradients
por: Chiang, John
Publicado: (2022)
por: Chiang, John
Publicado: (2022)
On the Last-Iterate Convergence of Shuffling Gradient Methods
por: Liu, Zijian, et al.
Publicado: (2024)
por: Liu, Zijian, et al.
Publicado: (2024)
On the Convergence of Adaptive Gradient Methods for Nonconvex Optimization
por: Zhou, Dongruo, et al.
Publicado: (2018)
por: Zhou, Dongruo, et al.
Publicado: (2018)
On Penalty-based Bilevel Gradient Descent Method
por: Shen, Han, et al.
Publicado: (2023)
por: Shen, Han, et al.
Publicado: (2023)
Ejemplares similares
-
Elementary Analysis of Policy Gradient Methods
por: Liu, Jiacai, et al.
Publicado: (2024) -
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
por: Klein, Sara, et al.
Publicado: (2023) -
A New Convergence Analysis of Plug-and-Play Proximal Gradient Descent Under Prior Mismatch
por: Xu, Guixian, et al.
Publicado: (2026) -
Model-Free Output Feedback Stabilization via Policy Gradient Methods
por: Zhang, Ankang, et al.
Publicado: (2026) -
Delightful Distributed Policy Gradient
por: Osband, Ian
Publicado: (2026)