On the Convergence of Policy in Unregularized Policy Mirror Descent
Fuente:
arXiv
Guardado en:
| Autores principales: | Lin, Dachao, Zhang, Zhihua |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Convergence of Gradient Descent with Small Initialization for Unregularized Matrix Completion
por: Ma, Jianhao, et al.
Publicado: (2024)
por: Ma, Jianhao, et al.
Publicado: (2024)
On the Convergence of Policy Mirror Descent with Temporal Difference Evaluation
por: Liu, Jiacai, et al.
Publicado: (2025)
por: Liu, Jiacai, et al.
Publicado: (2025)
Convergence of Policy Mirror Descent Beyond Compatible Function Approximation
por: Sherman, Uri, et al.
Publicado: (2025)
por: Sherman, Uri, et al.
Publicado: (2025)
A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence
por: Alfano, Carlo, et al.
Publicado: (2023)
por: Alfano, Carlo, et al.
Publicado: (2023)
On the Convergence of Projected Policy Gradient for Any Constant Step Sizes
por: Liu, Jiacai, et al.
Publicado: (2023)
por: Liu, Jiacai, et al.
Publicado: (2023)
Implicit Bias and Convergence of Matrix Stochastic Mirror Descent
por: Akhtiamov, Danil, et al.
Publicado: (2026)
por: Akhtiamov, Danil, et al.
Publicado: (2026)
Policy Mirror Descent with Temporal Difference Learning: Sample Complexity under Online Markov Data
por: Li, Wenye, et al.
Publicado: (2025)
por: Li, Wenye, et al.
Publicado: (2025)
Parameter-free Mirror Descent
por: Jacobsen, Andrew, et al.
Publicado: (2022)
por: Jacobsen, Andrew, et al.
Publicado: (2022)
Unregularized limit of stochastic gradient method for Wasserstein distributionally robust optimization
por: Le, Tam
Publicado: (2025)
por: Le, Tam
Publicado: (2025)
Mirror Descent on Riemannian Manifolds
por: Jiang, Jiaxin, et al.
Publicado: (2026)
por: Jiang, Jiaxin, et al.
Publicado: (2026)
A Mirror Descent Perspective of Smoothed Sign Descent
por: Wang, Shuyang, et al.
Publicado: (2024)
por: Wang, Shuyang, et al.
Publicado: (2024)
Learning Over-Relaxation Policies for ADMM with Convergence Guarantees
por: Lin, Junan, et al.
Publicado: (2026)
por: Lin, Junan, et al.
Publicado: (2026)
Mirror Descent on Reproducing Kernel Banach Spaces
por: Kumar, Akash, et al.
Publicado: (2024)
por: Kumar, Akash, et al.
Publicado: (2024)
Mirror and Preconditioned Gradient Descent in Wasserstein Space
por: Bonet, Clément, et al.
Publicado: (2024)
por: Bonet, Clément, et al.
Publicado: (2024)
Learning Provably Improves the Convergence of Gradient Descent
por: Song, Qingyu, et al.
Publicado: (2025)
por: Song, Qingyu, et al.
Publicado: (2025)
Reusing Historical Trajectories in Natural Policy Gradient via Importance Sampling: Convergence and Convergence Rate
por: Lin, Yifan, et al.
Publicado: (2024)
por: Lin, Yifan, et al.
Publicado: (2024)
Value Mirror Descent for Reinforcement Learning
por: Jia, Zhichao, et al.
Publicado: (2026)
por: Jia, Zhichao, et al.
Publicado: (2026)
Anderson Acceleration Without Restart: A Novel Method with $n$-Step Super Quadratic Convergence Rate
por: Ye, Haishan, et al.
Publicado: (2024)
por: Ye, Haishan, et al.
Publicado: (2024)
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
por: Han, Yinbin, et al.
Publicado: (2023)
por: Han, Yinbin, et al.
Publicado: (2023)
Mirror Descent-Type Algorithms for the Variational Inequality Problem with Functional Constraints
por: Alkousa, Mohammad S., et al.
Publicado: (2026)
por: Alkousa, Mohammad S., et al.
Publicado: (2026)
On the Convergence of Stochastic Gradient Descent with Perturbed Forward-Backward Passes
por: Kong, Boao, et al.
Publicado: (2026)
por: Kong, Boao, et al.
Publicado: (2026)
Finite-Time Decoupled Convergence in Nonlinear Two-Time-Scale Stochastic Approximation
por: Han, Yuze, et al.
Publicado: (2024)
por: Han, Yuze, et al.
Publicado: (2024)
Convergence of Spectral Descent for Non-smooth Optimization
por: Yang, Yixuan, et al.
Publicado: (2026)
por: Yang, Yixuan, et al.
Publicado: (2026)
Convergence of Alternating Gradient Descent for Matrix Factorization
por: Ward, Rachel, et al.
Publicado: (2023)
por: Ward, Rachel, et al.
Publicado: (2023)
Zeroth-Order Stochastic Mirror Descent Algorithms for Minimax Excess Risk Optimization
por: Gu, Zhihao, et al.
Publicado: (2024)
por: Gu, Zhihao, et al.
Publicado: (2024)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
por: Ding, Dongsheng, et al.
Publicado: (2023)
por: Ding, Dongsheng, et al.
Publicado: (2023)
Mirror Descent-Ascent for mean-field min-max problems
por: Lascu, Razvan-Andrei, et al.
Publicado: (2024)
por: Lascu, Razvan-Andrei, et al.
Publicado: (2024)
Natural Hypergradient Descent: Algorithm Design, Convergence Analysis, and Parallel Implementation
por: Kong, Deyi, et al.
Publicado: (2026)
por: Kong, Deyi, et al.
Publicado: (2026)
Second-Order Mirror Descent: Convergence in Games Beyond Averaging and Discounting
por: Gao, Bolin, et al.
Publicado: (2021)
por: Gao, Bolin, et al.
Publicado: (2021)
Convergence Analysis of Stochastic Gradient Descent with MCMC Estimators
por: Li, Tianyou, et al.
Publicado: (2023)
por: Li, Tianyou, et al.
Publicado: (2023)
Open Problem: Anytime Convergence Rate of Gradient Descent
por: Kornowski, Guy, et al.
Publicado: (2024)
por: Kornowski, Guy, et al.
Publicado: (2024)
On the Convergence of Gradient Descent on Learning Transformers with Residual Connections
por: Qin, Zhen, et al.
Publicado: (2025)
por: Qin, Zhen, et al.
Publicado: (2025)
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
por: Klein, Sara, et al.
Publicado: (2023)
por: Klein, Sara, et al.
Publicado: (2023)
Convergence Properties of Natural Gradient Descent for Minimizing KL Divergence
por: Datar, Adwait, et al.
Publicado: (2025)
por: Datar, Adwait, et al.
Publicado: (2025)
Exponential Convergence of (Stochastic) Gradient Descent for Separable Logistic Regression
por: Kale, Sacchit, et al.
Publicado: (2026)
por: Kale, Sacchit, et al.
Publicado: (2026)
Faster Convergence of Stochastic Accelerated Gradient Descent under Interpolation
por: Mishkin, Aaron, et al.
Publicado: (2024)
por: Mishkin, Aaron, et al.
Publicado: (2024)
Convergence and Implicit Bias of Gradient Descent on Continual Linear Classification
por: Jung, Hyunji, et al.
Publicado: (2025)
por: Jung, Hyunji, et al.
Publicado: (2025)
Convergence of Steepest Descent and Adam under Non-Uniform Smoothness
por: Vaswani, Sharan, et al.
Publicado: (2026)
por: Vaswani, Sharan, et al.
Publicado: (2026)
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
por: Carmona, René, et al.
Publicado: (2019)
por: Carmona, René, et al.
Publicado: (2019)
Linear Convergence of Entropy-Regularized Natural Policy Gradient with Linear Function Approximation
por: Cayci, Semih, et al.
Publicado: (2021)
por: Cayci, Semih, et al.
Publicado: (2021)
Ejemplares similares
-
Convergence of Gradient Descent with Small Initialization for Unregularized Matrix Completion
por: Ma, Jianhao, et al.
Publicado: (2024) -
On the Convergence of Policy Mirror Descent with Temporal Difference Evaluation
por: Liu, Jiacai, et al.
Publicado: (2025) -
Convergence of Policy Mirror Descent Beyond Compatible Function Approximation
por: Sherman, Uri, et al.
Publicado: (2025) -
A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence
por: Alfano, Carlo, et al.
Publicado: (2023) -
On the Convergence of Projected Policy Gradient for Any Constant Step Sizes
por: Liu, Jiacai, et al.
Publicado: (2023)