Geometry and convergence of natural policy gradient methods
Fuente:
arXiv
Salvato in:
| Autori principali: | Müller, Johannes, Montúfar, Guido |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Fisher-Rao Gradient Flows of Linear Programs and State-Action Natural Policy Gradients
di: Müller, Johannes, et al.
Pubblicazione: (2024)
di: Müller, Johannes, et al.
Pubblicazione: (2024)
Optimal Rates of Convergence for Entropy Regularization in Discounted Markov Decision Processes
di: Müller, Johannes, et al.
Pubblicazione: (2024)
di: Müller, Johannes, et al.
Pubblicazione: (2024)
A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces
di: Kerimkulov, Bekzhan, et al.
Pubblicazione: (2023)
di: Kerimkulov, Bekzhan, et al.
Pubblicazione: (2023)
SPIRAL: A superlinearly convergent incremental proximal algorithm for nonconvex finite sum minimization
di: Behmandpoor, Pourya, et al.
Pubblicazione: (2022)
di: Behmandpoor, Pourya, et al.
Pubblicazione: (2022)
Safeguarding adaptive methods: global convergence of Barzilai-Borwein and other stepsize choices
di: Ou, Hongjia, et al.
Pubblicazione: (2024)
di: Ou, Hongjia, et al.
Pubblicazione: (2024)
Quasi-Newton methods for minimizing a quadratic function subject to uncertainty
di: Peng, Shen, et al.
Pubblicazione: (2021)
di: Peng, Shen, et al.
Pubblicazione: (2021)
Nonsmooth exact penalty methods for equality-constrained optimization: complexity and implementation
di: Diouane, Youssef, et al.
Pubblicazione: (2024)
di: Diouane, Youssef, et al.
Pubblicazione: (2024)
Soft quasi-Newton: Guaranteed positive definiteness by relaxing the secant constraint
di: Berglund, Erik, et al.
Pubblicazione: (2024)
di: Berglund, Erik, et al.
Pubblicazione: (2024)
Proximal Quasi-Newton Method for Composite Optimization over the Stiefel Manifold
di: Wang, Qinsi, et al.
Pubblicazione: (2025)
di: Wang, Qinsi, et al.
Pubblicazione: (2025)
Novel Limited Memory Quasi-Newton Methods Based On Optimal Matrix Approximation
di: Berglund, Erik, et al.
Pubblicazione: (2024)
di: Berglund, Erik, et al.
Pubblicazione: (2024)
Riemannian Interior Point Methods for Constrained Optimization on Manifolds
di: Lai, Zhijian, et al.
Pubblicazione: (2022)
di: Lai, Zhijian, et al.
Pubblicazione: (2022)
A new envelope function for nonsmooth DC optimization
di: Themelis, Andreas, et al.
Pubblicazione: (2020)
di: Themelis, Andreas, et al.
Pubblicazione: (2020)
(Adaptive) Scaled gradient methods beyond locally Holder smoothness: Lyapunov analysis, convergence rate and complexity
di: Ghaderi, Susan, et al.
Pubblicazione: (2025)
di: Ghaderi, Susan, et al.
Pubblicazione: (2025)
Alternating Iteratively Reweighted $\ell_1$ and Subspace Newton Algorithms for Nonconvex Sparse Optimization
di: Wang, Hao, et al.
Pubblicazione: (2024)
di: Wang, Hao, et al.
Pubblicazione: (2024)
Improving Quasi-Newton Methods via Image and Projection Operators
di: Ji, Zhenyuan
Pubblicazione: (2025)
di: Ji, Zhenyuan
Pubblicazione: (2025)
A Stochastic Quasi-Newton Method in the Absence of Common Random Numbers
di: Menickelly, Matt, et al.
Pubblicazione: (2023)
di: Menickelly, Matt, et al.
Pubblicazione: (2023)
Riemannian Gradient Method with Momentum
di: Leggio, Filippo, et al.
Pubblicazione: (2026)
di: Leggio, Filippo, et al.
Pubblicazione: (2026)
A non-monotone trust-region method with noisy oracles and additional sampling
di: Krejic, Natasa, et al.
Pubblicazione: (2023)
di: Krejic, Natasa, et al.
Pubblicazione: (2023)
Lipschitz upper semicontinuity of linear inequality systems under full perturbations
di: Camacho, Jesús, et al.
Pubblicazione: (2025)
di: Camacho, Jesús, et al.
Pubblicazione: (2025)
A linesearch-type normal map-based semismooth Newton method for nonsmooth nonconvex composite optimization
di: Zeng, Hanfeng, et al.
Pubblicazione: (2026)
di: Zeng, Hanfeng, et al.
Pubblicazione: (2026)
Block-coordinate and incremental aggregated proximal gradient methods for nonsmooth nonconvex problems
di: Latafat, Puya, et al.
Pubblicazione: (2019)
di: Latafat, Puya, et al.
Pubblicazione: (2019)
Nonmonotone subgradient methods based on a local descent lemma
di: Aragón-Artacho, Francisco J., et al.
Pubblicazione: (2025)
di: Aragón-Artacho, Francisco J., et al.
Pubblicazione: (2025)
Foundations of Riemannian Geometry for Riemannian Optimization: A Monograph with Detailed Derivations
di: Ghojogh, Benyamin
Pubblicazione: (2026)
di: Ghojogh, Benyamin
Pubblicazione: (2026)
Modified limited memory BFGS with displacement aggregation and its application to the largest eigenvalue problem
di: Sahu, Manish Kumar, et al.
Pubblicazione: (2023)
di: Sahu, Manish Kumar, et al.
Pubblicazione: (2023)
Revisiting implicit variables in mathematical optimization: simplified modeling and a numerical evidence
di: Mehlitz, Patrick
Pubblicazione: (2025)
di: Mehlitz, Patrick
Pubblicazione: (2025)
Non-subdifferentiability optimality and mean value theorems via new relative subdifferentials
di: Thinh, Vo Duc, et al.
Pubblicazione: (2024)
di: Thinh, Vo Duc, et al.
Pubblicazione: (2024)
Nonlinear Conjugate Gradient Methods for Optimization of Set-Valued Mappings of Finite Cardinality
di: Ghosh, Debdas, et al.
Pubblicazione: (2024)
di: Ghosh, Debdas, et al.
Pubblicazione: (2024)
Stability for Nash Equilibrium Problems
di: Diao, Ruoyu, et al.
Pubblicazione: (2024)
di: Diao, Ruoyu, et al.
Pubblicazione: (2024)
Quasi-Newton Method for Set Optimization Problems with Set-Valued Mapping Given by Finitely Many Vector-Valued Functions
di: Ghosh, Debdas, et al.
Pubblicazione: (2024)
di: Ghosh, Debdas, et al.
Pubblicazione: (2024)
Generalized gradient flows in Hadamard manifolds and convex optimization on entanglement polytopes
di: Hirai, Hiroshi
Pubblicazione: (2025)
di: Hirai, Hiroshi
Pubblicazione: (2025)
A scaling characterization of nc-rank via unbounded gradient flow
di: Hirai, Hiroshi
Pubblicazione: (2025)
di: Hirai, Hiroshi
Pubblicazione: (2025)
Robust Graph-Based Semi-Supervised Learning via $p$-Conductances
di: Robertson, Sawyer Jack, et al.
Pubblicazione: (2025)
di: Robertson, Sawyer Jack, et al.
Pubblicazione: (2025)
A Riemannian quasi-Newton algorithm for optimization with Euclidean bounds
di: Baran, Mateusz, et al.
Pubblicazione: (2026)
di: Baran, Mateusz, et al.
Pubblicazione: (2026)
Mixed Newton Method for Optimization in Complex Spaces
di: Yudin, Nikita, et al.
Pubblicazione: (2024)
di: Yudin, Nikita, et al.
Pubblicazione: (2024)
Approximate stationarity in disjunctive optimization: concepts, qualification conditions, and application to MPCCs
di: Käming, Isabella, et al.
Pubblicazione: (2025)
di: Käming, Isabella, et al.
Pubblicazione: (2025)
A stochastic use of the Kurdyka-Lojasiewicz property: Investigation of optimization algorithms behaviours in a non-convex differentiable framework
di: Fest, Jean-Baptiste, et al.
Pubblicazione: (2023)
di: Fest, Jean-Baptiste, et al.
Pubblicazione: (2023)
Proximal Limited-Memory Quasi-Newton Methods for Nonsmooth Nonconvex Optimization
di: Dahl, Simeon vom, et al.
Pubblicazione: (2026)
di: Dahl, Simeon vom, et al.
Pubblicazione: (2026)
Variational Properties of Decomposable Functions. Part I: Strict Epi-Calculus and Applications
di: Ouyang, Wenqing, et al.
Pubblicazione: (2023)
di: Ouyang, Wenqing, et al.
Pubblicazione: (2023)
Characterizations of the Aubin Property of the Solution Mapping for Nonlinear Semidefinite Programming
di: Chen, Liang, et al.
Pubblicazione: (2024)
di: Chen, Liang, et al.
Pubblicazione: (2024)
Variational Properties of Decomposable Functions. Part II: Strong Second-Order Theory
di: Ouyang, Wenqing, et al.
Pubblicazione: (2023)
di: Ouyang, Wenqing, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Fisher-Rao Gradient Flows of Linear Programs and State-Action Natural Policy Gradients
di: Müller, Johannes, et al.
Pubblicazione: (2024) -
Optimal Rates of Convergence for Entropy Regularization in Discounted Markov Decision Processes
di: Müller, Johannes, et al.
Pubblicazione: (2024) -
A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces
di: Kerimkulov, Bekzhan, et al.
Pubblicazione: (2023) -
SPIRAL: A superlinearly convergent incremental proximal algorithm for nonconvex finite sum minimization
di: Behmandpoor, Pourya, et al.
Pubblicazione: (2022) -
Safeguarding adaptive methods: global convergence of Barzilai-Borwein and other stepsize choices
di: Ou, Hongjia, et al.
Pubblicazione: (2024)