Learning mirror maps in policy mirror descent
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Alfano, Carlo, Towers, Sebastian, Sapora, Silvia, Lu, Chris, Rebeschini, Patrick |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence
par: Alfano, Carlo, et autres
Publié: (2023)
par: Alfano, Carlo, et autres
Publié: (2023)
On exploration of an interior mirror descent flow for stochastic nonconvex constrained problem
par: Ding, Kuangyu, et autres
Publié: (2025)
par: Ding, Kuangyu, et autres
Publié: (2025)
Meta-Learning Objectives for Preference Optimization
par: Alfano, Carlo, et autres
Publié: (2024)
par: Alfano, Carlo, et autres
Publié: (2024)
Entropy annealing for policy mirror descent in continuous time and space
par: Sethi, Deven, et autres
Publié: (2024)
par: Sethi, Deven, et autres
Publié: (2024)
Global minimisation of nonconvex functions by generalising the mirror descent method
par: Millán, Reinier Díaz, et autres
Publié: (2024)
par: Millán, Reinier Díaz, et autres
Publié: (2024)
Iterative regularization in classification via hinge loss diagonal descent
par: Apidopoulos, Vassilis, et autres
Publié: (2022)
par: Apidopoulos, Vassilis, et autres
Publié: (2022)
Convergence of stochastic gradient descent under a local Lojasiewicz condition for deep neural networks
par: An, Jing, et autres
Publié: (2023)
par: An, Jing, et autres
Publié: (2023)
Manifold constrained steepest descent
par: Yang, Kaiwei, et autres
Publié: (2026)
par: Yang, Kaiwei, et autres
Publié: (2026)
Riemannian coordinate descent algorithms on matrix manifolds
par: Han, Andi, et autres
Publié: (2024)
par: Han, Andi, et autres
Publié: (2024)
New logarithmic step size for stochastic gradient descent
par: Shamaee, M. Soheil, et autres
Publié: (2024)
par: Shamaee, M. Soheil, et autres
Publié: (2024)
Gradient descent in matrix factorization: Understanding large initialization
par: Chen, Hengchao, et autres
Publié: (2023)
par: Chen, Hengchao, et autres
Publié: (2023)
On the global convergence of gradient descent for wide shallow models with bounded nonlinearities
par: Petit, Romain, et autres
Publié: (2026)
par: Petit, Romain, et autres
Publié: (2026)
The duality structure gradient descent algorithm: analysis and applications to neural networks
par: Flynn, Thomas
Publié: (2017)
par: Flynn, Thomas
Publié: (2017)
On the stability of gradient descent with second order dynamics for time-varying cost functions
par: Gibson, Travis E., et autres
Publié: (2024)
par: Gibson, Travis E., et autres
Publié: (2024)
Convergence of continuous-time stochastic gradient descent with applications to deep neural networks
par: Lugosi, Gabor, et autres
Publié: (2024)
par: Lugosi, Gabor, et autres
Publié: (2024)
The Mixing method: low-rank coordinate descent for semidefinite programming with diagonal constraints
par: Wang, Po-Wei, et autres
Publié: (2017)
par: Wang, Po-Wei, et autres
Publié: (2017)
Convergence of two-timescale gradient descent ascent dynamics: finite-dimensional and mean-field perspectives
par: An, Jing, et autres
Publié: (2025)
par: An, Jing, et autres
Publié: (2025)
Error dynamics of mini-batch gradient descent with random reshuffling for least squares regression
par: Lok, Jackie, et autres
Publié: (2024)
par: Lok, Jackie, et autres
Publié: (2024)
Linear convergence of proximal descent schemes on the Wasserstein space
par: Lascu, Razvan-Andrei, et autres
Publié: (2024)
par: Lascu, Razvan-Andrei, et autres
Publié: (2024)
A stochastic gradient descent algorithm with random search directions
par: Gbaguidi, Eméric
Publié: (2025)
par: Gbaguidi, Eméric
Publié: (2025)
Convergence of flow-based generative models via proximal gradient descent in Wasserstein space
par: Cheng, Xiuyuan, et autres
Publié: (2023)
par: Cheng, Xiuyuan, et autres
Publié: (2023)
An accelerated first-order regularized momentum descent ascent algorithm for stochastic nonconvex-concave minimax problems
par: Zhang, Huiling, et autres
Publié: (2023)
par: Zhang, Huiling, et autres
Publié: (2023)
A short proof of near-linear convergence of adaptive gradient descent under fourth-order growth and convexity
par: Davis, Damek, et autres
Publié: (2026)
par: Davis, Damek, et autres
Publié: (2026)
A block-coordinate descent framework for non-convex composite optimization. Application to sparse precision matrix estimation
par: Lauga, Guillaume
Publié: (2026)
par: Lauga, Guillaume
Publié: (2026)
Aiding Global Convergence in Federated Learning via Local Perturbation and Mutual Similarity Information
par: Buttaci, Emanuel, et autres
Publié: (2024)
par: Buttaci, Emanuel, et autres
Publié: (2024)
Non-convergence to global minimizers for Adam and stochastic gradient descent optimization and constructions of local minimizers in the training of artificial neural networks
par: Jentzen, Arnulf, et autres
Publié: (2024)
par: Jentzen, Arnulf, et autres
Publié: (2024)
Operator World Models for Reinforcement Learning
par: Novelli, Pietro, et autres
Publié: (2024)
par: Novelli, Pietro, et autres
Publié: (2024)
Monotone Optimisation with Learned Projections
par: Rashwan, Ahmed, et autres
Publié: (2026)
par: Rashwan, Ahmed, et autres
Publié: (2026)
Scalable Decentralized Learning with Teleportation
par: Takezawa, Yuki, et autres
Publié: (2025)
par: Takezawa, Yuki, et autres
Publié: (2025)
When majority rules, minority loses: bias amplification of gradient descent
par: Bachoc, François, et autres
Publié: (2025)
par: Bachoc, François, et autres
Publié: (2025)
Global convergence of gradient descent for phase retrieval
par: Fougereux, Théodore, et autres
Publié: (2024)
par: Fougereux, Théodore, et autres
Publié: (2024)
Deep Learning Model Predictive Control for Deep Brain Stimulation in Parkinson's Disease
par: Steffen, Sebastian, et autres
Publié: (2025)
par: Steffen, Sebastian, et autres
Publié: (2025)
Multi-objective Deep Learning: Taxonomy and Survey of the State of the Art
par: Peitz, Sebastian, et autres
Publié: (2024)
par: Peitz, Sebastian, et autres
Publié: (2024)
Locally Adaptive Federated Learning
par: Mukherjee, Sohom, et autres
Publié: (2023)
par: Mukherjee, Sohom, et autres
Publié: (2023)
A note on convergence of Wasserstein policy optimization
par: Šiška, David, et autres
Publié: (2026)
par: Šiška, David, et autres
Publié: (2026)
A General Continuous-Time Formulation of Stochastic ADMM and Its Variants
par: Li, Chris Junchi
Publié: (2024)
par: Li, Chris Junchi
Publié: (2024)
Accelerated Fully First-Order Methods for Bilevel and Minimax Optimization
par: Li, Chris Junchi
Publié: (2024)
par: Li, Chris Junchi
Publié: (2024)
Enhancing Stochastic Optimization for Statistical Efficiency Using ROOT-SGD with Diminishing Stepsize
par: Li, Chris Junchi
Publié: (2024)
par: Li, Chris Junchi
Publié: (2024)
Distributed Event-Based Learning via ADMM
par: Er, Guner Dilsad, et autres
Publié: (2024)
par: Er, Guner Dilsad, et autres
Publié: (2024)
CV@R penalized portfolio optimization with biased stochastic mirror descent
par: Costa, Manon, et autres
Publié: (2024)
par: Costa, Manon, et autres
Publié: (2024)
Documents similaires
-
A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence
par: Alfano, Carlo, et autres
Publié: (2023) -
On exploration of an interior mirror descent flow for stochastic nonconvex constrained problem
par: Ding, Kuangyu, et autres
Publié: (2025) -
Meta-Learning Objectives for Preference Optimization
par: Alfano, Carlo, et autres
Publié: (2024) -
Entropy annealing for policy mirror descent in continuous time and space
par: Sethi, Deven, et autres
Publié: (2024) -
Global minimisation of nonconvex functions by generalising the mirror descent method
par: Millán, Reinier Díaz, et autres
Publié: (2024)