On the Convergence of Projected Policy Gradient for Any Constant Step Sizes
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Jiacai, Li, Wenye, Lin, Dachao, Wei, Ke, Zhang, Zhihua |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Elementary Analysis of Policy Gradient Methods
by: Liu, Jiacai, et al.
Published: (2024)
by: Liu, Jiacai, et al.
Published: (2024)
On the Convergence of Policy Mirror Descent with Temporal Difference Evaluation
by: Liu, Jiacai, et al.
Published: (2025)
by: Liu, Jiacai, et al.
Published: (2025)
On the Convergence of Policy in Unregularized Policy Mirror Descent
by: Lin, Dachao, et al.
Published: (2022)
by: Lin, Dachao, et al.
Published: (2022)
Anderson Acceleration Without Restart: A Novel Method with $n$-Step Super Quadratic Convergence Rate
by: Ye, Haishan, et al.
Published: (2024)
by: Ye, Haishan, et al.
Published: (2024)
Policy Mirror Descent with Temporal Difference Learning: Sample Complexity under Online Markov Data
by: Li, Wenye, et al.
Published: (2025)
by: Li, Wenye, et al.
Published: (2025)
Global Convergence of Natural Policy Gradient with Hessian-aided Momentum Variance Reduction
by: Feng, Jie, et al.
Published: (2024)
by: Feng, Jie, et al.
Published: (2024)
A Theoretical and Empirical Study on the Convergence of Adam with an "Exact" Constant Step Size in Non-Convex Settings
by: Mazumder, Alokendu, et al.
Published: (2023)
by: Mazumder, Alokendu, et al.
Published: (2023)
Accelerated Objective Gap and Gradient Norm Convergence for Gradient Descent via Long Steps
by: Grimmer, Benjamin, et al.
Published: (2024)
by: Grimmer, Benjamin, et al.
Published: (2024)
Monitoring the Convergence Speed of PDHG to Find Better Primal and Dual Step Sizes
by: Fercoq, Olivier
Published: (2024)
by: Fercoq, Olivier
Published: (2024)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Analysis of Gradient Descent with Varying Step Sizes using Integral Quadratic Constraints
by: Padmanabhan, Ram, et al.
Published: (2022)
by: Padmanabhan, Ram, et al.
Published: (2022)
Convergence of Policy Gradient for Stochastic Linear-Quadratic Control Problem in Infinite Horizon
by: Zhang, Xinpei, et al.
Published: (2024)
by: Zhang, Xinpei, et al.
Published: (2024)
Almost Sure Convergence of Stochastic Approximation: An Interplay of Noise and Step Size
by: Nguyen, Quang Dinh Thien, et al.
Published: (2026)
by: Nguyen, Quang Dinh Thien, et al.
Published: (2026)
Adaptive Open-Loop Step-Sizes for Accelerated Convergence Rates of the Frank-Wolfe Algorithm
by: Wirth, Elias, et al.
Published: (2025)
by: Wirth, Elias, et al.
Published: (2025)
Reusing Historical Trajectories in Natural Policy Gradient via Importance Sampling: Convergence and Convergence Rate
by: Lin, Yifan, et al.
Published: (2024)
by: Lin, Yifan, et al.
Published: (2024)
Convergence and Inference of Stream SGD, with Applications to Queueing Systems and Inventory Control
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
Adaptive Step Sizes for Preconditioned Stochastic Gradient Descent
by: Köhne, Frederik, et al.
Published: (2023)
by: Köhne, Frederik, et al.
Published: (2023)
Convergence Rate Analysis of SOAP with Arbitrary Orthogonal Projection Matrices
by: Li, Huan, et al.
Published: (2026)
by: Li, Huan, et al.
Published: (2026)
Accelerated Affine-Invariant Convergence Rates of the Frank-Wolfe Algorithm with Open-Loop Step-Sizes
by: Wirth, Elias, et al.
Published: (2023)
by: Wirth, Elias, et al.
Published: (2023)
Optimal Convergence Rate for Mirror Descent Methods with special Time-Varying Step Sizes Rules
by: Alkousa, Mohammad, et al.
Published: (2024)
by: Alkousa, Mohammad, et al.
Published: (2024)
Learning Provably Improves the Convergence of Gradient Descent
by: Song, Qingyu, et al.
Published: (2025)
by: Song, Qingyu, et al.
Published: (2025)
Policy Gradient Methods for the Cost-Constrained LQR: Strong Duality and Global Convergence
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
Adam Converges Without Any Modification On Update Rules
by: Zhang, Yushun, et al.
Published: (2026)
by: Zhang, Yushun, et al.
Published: (2026)
On Convergence and Stability of Two Extended BB-like Step Sizes
by: Xu, Xin
Published: (2025)
by: Xu, Xin
Published: (2025)
A Globally Convergent Policy Gradient Method for Linear Quadratic Gaussian (LQG) Control
by: Sadamoto, Tomonori, et al.
Published: (2023)
by: Sadamoto, Tomonori, et al.
Published: (2023)
AGDA+: Proximal Alternating Gradient Descent Ascent Method with a Nonmonotone Adaptive Step-Size Search for Nonconvex Minimax Problems
by: Zhang, Xuan, et al.
Published: (2024)
by: Zhang, Xuan, et al.
Published: (2024)
Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation
by: Rodriguez-Gil, Jhojan A., et al.
Published: (2026)
by: Rodriguez-Gil, Jhojan A., et al.
Published: (2026)
Finite-Time Decoupled Convergence in Nonlinear Two-Time-Scale Stochastic Approximation
by: Han, Yuze, et al.
Published: (2024)
by: Han, Yuze, et al.
Published: (2024)
Faster Convergence of Riemannian Stochastic Gradient Descent with Increasing Batch Size
by: Oowada, Kanata, et al.
Published: (2025)
by: Oowada, Kanata, et al.
Published: (2025)
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
by: Han, Yinbin, et al.
Published: (2023)
by: Han, Yinbin, et al.
Published: (2023)
Gradient Descent on Logistic Regression with Non-Separable Data and Large Step Sizes
by: Meng, Si Yi, et al.
Published: (2024)
by: Meng, Si Yi, et al.
Published: (2024)
Convergence Analysis of Stochastic Accelerated Gradient Methods for Generalized Smooth Optimizations
by: Yu, Chenhao, et al.
Published: (2025)
by: Yu, Chenhao, et al.
Published: (2025)
Almost Sure Convergence of Networked Policy Gradient over Time-Varying Networks in Markov Potential Games
by: Aydin, Sarper, et al.
Published: (2024)
by: Aydin, Sarper, et al.
Published: (2024)
Gradient Descent on Logistic Regression: Do Large Step-Sizes Work with Data on the Sphere?
by: Meng, Si Yi, et al.
Published: (2025)
by: Meng, Si Yi, et al.
Published: (2025)
Global Stability and Step Size Robustness of RMSProp
by: Dimitrieski, Naum, et al.
Published: (2026)
by: Dimitrieski, Naum, et al.
Published: (2026)
Last-Iterate Convergence of Randomized Kaczmarz and SGD with Greedy Step Size
by: Dereziński, Michał, et al.
Published: (2026)
by: Dereziński, Michał, et al.
Published: (2026)
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
by: Klein, Sara, et al.
Published: (2023)
by: Klein, Sara, et al.
Published: (2023)
GANs as Gradient Flows that Converge
by: Huang, Yu-Jui, et al.
Published: (2022)
by: Huang, Yu-Jui, et al.
Published: (2022)
Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation
by: Zhang, Yixuan, et al.
Published: (2024)
by: Zhang, Yixuan, et al.
Published: (2024)
Policy Gradient Method for LQG Control via Input-Output-History Representation: Convergence to $O(ε)$-Stationary Points
by: Sadamoto, Tomonori, et al.
Published: (2025)
by: Sadamoto, Tomonori, et al.
Published: (2025)
Similar Items
-
Elementary Analysis of Policy Gradient Methods
by: Liu, Jiacai, et al.
Published: (2024) -
On the Convergence of Policy Mirror Descent with Temporal Difference Evaluation
by: Liu, Jiacai, et al.
Published: (2025) -
On the Convergence of Policy in Unregularized Policy Mirror Descent
by: Lin, Dachao, et al.
Published: (2022) -
Anderson Acceleration Without Restart: A Novel Method with $n$-Step Super Quadratic Convergence Rate
by: Ye, Haishan, et al.
Published: (2024) -
Policy Mirror Descent with Temporal Difference Learning: Sample Complexity under Online Markov Data
by: Li, Wenye, et al.
Published: (2025)