A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces
Fuente:
arXiv
Saved in:
| Main Authors: | Kerimkulov, Bekzhan, Leahy, James-Michael, Siska, David, Szpruch, Lukasz, Zhang, Yufei |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Riemannian Gradient Method with Momentum
by: Leggio, Filippo, et al.
Published: (2026)
by: Leggio, Filippo, et al.
Published: (2026)
Stochastic Gradient Descent Revisited
by: Louzi, Azar
Published: (2024)
by: Louzi, Azar
Published: (2024)
Blackwell optimality and policy stability for long-run risk sensitive stochastic control
by: Bäuerle, Nicole, et al.
Published: (2024)
by: Bäuerle, Nicole, et al.
Published: (2024)
Policy stability and ultimate stationarity in discounted risk-sensitive stochastic control
by: Bäuerle, Nicole, et al.
Published: (2026)
by: Bäuerle, Nicole, et al.
Published: (2026)
Convex Regularization and Convergence of Policy Gradient Flows under Safety Constraints
by: Malo, Pekka, et al.
Published: (2024)
by: Malo, Pekka, et al.
Published: (2024)
A Stochastic Quasi-Newton Method in the Absence of Common Random Numbers
by: Menickelly, Matt, et al.
Published: (2023)
by: Menickelly, Matt, et al.
Published: (2023)
A new envelope function for nonsmooth DC optimization
by: Themelis, Andreas, et al.
Published: (2020)
by: Themelis, Andreas, et al.
Published: (2020)
Existence of bounded solutions to multiplicative Poisson equations under mixing property
by: Pitera, Marcin, et al.
Published: (2023)
by: Pitera, Marcin, et al.
Published: (2023)
Norm-induced Cuts: Outer Approximation for Lipschitzian Constraint Functions
by: Göß, Adrian, et al.
Published: (2024)
by: Göß, Adrian, et al.
Published: (2024)
Depth-first directional search for nonconvex optimization
by: Zhang, Yuxuan, et al.
Published: (2024)
by: Zhang, Yuxuan, et al.
Published: (2024)
A Globally Convergent Gradient Method with Momentum
by: Lapucci, Matteo, et al.
Published: (2024)
by: Lapucci, Matteo, et al.
Published: (2024)
Projected Gradient Methods with Momentum
by: Lapucci, Matteo, et al.
Published: (2026)
by: Lapucci, Matteo, et al.
Published: (2026)
PANOC-lite: A simpler and more efficient algorithm for composite minimization
by: Bodard, Alexander, et al.
Published: (2026)
by: Bodard, Alexander, et al.
Published: (2026)
Global Optimization of Gaussian processes
by: Schweidtmann, Artur M., et al.
Published: (2020)
by: Schweidtmann, Artur M., et al.
Published: (2020)
Convex quadratic sets and the complexity of mixed integer convex quadratic programming
by: Del Pia, Alberto
Published: (2023)
by: Del Pia, Alberto
Published: (2023)
A stochastic use of the Kurdyka-Lojasiewicz property: Investigation of optimization algorithms behaviours in a non-convex differentiable framework
by: Fest, Jean-Baptiste, et al.
Published: (2023)
by: Fest, Jean-Baptiste, et al.
Published: (2023)
Proximal Limited-Memory Quasi-Newton Methods for Nonsmooth Nonconvex Optimization
by: Dahl, Simeon vom, et al.
Published: (2026)
by: Dahl, Simeon vom, et al.
Published: (2026)
Optimization with Trained Machine Learning Models Embedded
by: Schweidtmann, Artur M., et al.
Published: (2022)
by: Schweidtmann, Artur M., et al.
Published: (2022)
Block-coordinate and incremental aggregated proximal gradient methods for nonsmooth nonconvex problems
by: Latafat, Puya, et al.
Published: (2019)
by: Latafat, Puya, et al.
Published: (2019)
Bregman Finito/MISO for nonconvex regularized finite sum minimization without Lipschitz gradient continuity
by: Latafat, Puya, et al.
Published: (2021)
by: Latafat, Puya, et al.
Published: (2021)
(Adaptive) Scaled gradient methods beyond locally Holder smoothness: Lyapunov analysis, convergence rate and complexity
by: Ghaderi, Susan, et al.
Published: (2025)
by: Ghaderi, Susan, et al.
Published: (2025)
What is the long-run distribution of stochastic gradient descent? A large deviations analysis
by: Azizian, Waïss, et al.
Published: (2024)
by: Azizian, Waïss, et al.
Published: (2024)
SPIRAL: A superlinearly convergent incremental proximal algorithm for nonconvex finite sum minimization
by: Behmandpoor, Pourya, et al.
Published: (2022)
by: Behmandpoor, Pourya, et al.
Published: (2022)
cuHALLaR: A GPU Accelerated Low-Rank Augmented Lagrangian Method for Large-Scale Semidefinite Programming
by: Aguirre, Jacob M., et al.
Published: (2025)
by: Aguirre, Jacob M., et al.
Published: (2025)
Deep neural networks can provably solve Bellman equations for Markov decision processes without the curse of dimensionality
by: Jentzen, Arnulf, et al.
Published: (2025)
by: Jentzen, Arnulf, et al.
Published: (2025)
Fisher-Rao Gradient Flows of Linear Programs and State-Action Natural Policy Gradients
by: Müller, Johannes, et al.
Published: (2024)
by: Müller, Johannes, et al.
Published: (2024)
Structure, Analysis, and Synthesis of First-Order Algorithms
by: Miller, Jared, et al.
Published: (2026)
by: Miller, Jared, et al.
Published: (2026)
A globalization of L-BFGS and the Barzilai-Borwein method for nonconvex unconstrained optimization
by: Mannel, Florian
Published: (2024)
by: Mannel, Florian
Published: (2024)
A Generalized Version of Chung's Lemma and its Applications
by: Jiang, Li, et al.
Published: (2024)
by: Jiang, Li, et al.
Published: (2024)
A structured L-BFGS method and its application to inverse problems
by: Mannel, Florian, et al.
Published: (2023)
by: Mannel, Florian, et al.
Published: (2023)
A structured L-BFGS method with diagonal scaling and its application to image registration
by: Mannel, Florian, et al.
Published: (2024)
by: Mannel, Florian, et al.
Published: (2024)
Projected gradient descent accumulates at Bouligand stationary points
by: Olikier, Guillaume, et al.
Published: (2024)
by: Olikier, Guillaume, et al.
Published: (2024)
Variational analysis of unbounded and discontinuous generalized eigenvalue functions with application to topology optimization
by: Nishioka, Akatsuki, et al.
Published: (2024)
by: Nishioka, Akatsuki, et al.
Published: (2024)
A low-rank augmented Lagrangian method for large-scale semidefinite programming based on a hybrid convex-nonconvex approach
by: Monteiro, Renato D. C., et al.
Published: (2024)
by: Monteiro, Renato D. C., et al.
Published: (2024)
Deep Relaxation of Controlled Stochastic Gradient Descent via Singular Perturbations
by: Bardi, Martino, et al.
Published: (2022)
by: Bardi, Martino, et al.
Published: (2022)
Optimization in complex spaces with the Mixed Newton Method
by: Bakhurin, Sergey, et al.
Published: (2022)
by: Bakhurin, Sergey, et al.
Published: (2022)
A new dual spectral projected gradient method for log-determinant semidefinite programming with hidden clustering structures
by: Namchaisiri, Charles, et al.
Published: (2024)
by: Namchaisiri, Charles, et al.
Published: (2024)
Empirical risk minimization for risk-neutral composite optimal control with applications to bang-bang control
by: Milz, Johannes, et al.
Published: (2024)
by: Milz, Johannes, et al.
Published: (2024)
Solving Sparse MIQCQPs: Application to the Unit Commitment Problem with ACOPF Constraints
by: Gómez-Casares, Ignacio, et al.
Published: (2025)
by: Gómez-Casares, Ignacio, et al.
Published: (2025)
Long run control of nonhomogeneous Markov processes
by: Stettner, Łukasz
Published: (2025)
by: Stettner, Łukasz
Published: (2025)
Similar Items
-
Riemannian Gradient Method with Momentum
by: Leggio, Filippo, et al.
Published: (2026) -
Stochastic Gradient Descent Revisited
by: Louzi, Azar
Published: (2024) -
Blackwell optimality and policy stability for long-run risk sensitive stochastic control
by: Bäuerle, Nicole, et al.
Published: (2024) -
Policy stability and ultimate stationarity in discounted risk-sensitive stochastic control
by: Bäuerle, Nicole, et al.
Published: (2026) -
Convex Regularization and Convergence of Policy Gradient Flows under Safety Constraints
by: Malo, Pekka, et al.
Published: (2024)