Learning mirror maps in policy mirror descent
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Alfano, Carlo, Towers, Sebastian, Sapora, Silvia, Lu, Chris, Rebeschini, Patrick |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence
von: Alfano, Carlo, et al.
Veröffentlicht: (2023)
von: Alfano, Carlo, et al.
Veröffentlicht: (2023)
On exploration of an interior mirror descent flow for stochastic nonconvex constrained problem
von: Ding, Kuangyu, et al.
Veröffentlicht: (2025)
von: Ding, Kuangyu, et al.
Veröffentlicht: (2025)
Meta-Learning Objectives for Preference Optimization
von: Alfano, Carlo, et al.
Veröffentlicht: (2024)
von: Alfano, Carlo, et al.
Veröffentlicht: (2024)
Entropy annealing for policy mirror descent in continuous time and space
von: Sethi, Deven, et al.
Veröffentlicht: (2024)
von: Sethi, Deven, et al.
Veröffentlicht: (2024)
Global minimisation of nonconvex functions by generalising the mirror descent method
von: Millán, Reinier Díaz, et al.
Veröffentlicht: (2024)
von: Millán, Reinier Díaz, et al.
Veröffentlicht: (2024)
Iterative regularization in classification via hinge loss diagonal descent
von: Apidopoulos, Vassilis, et al.
Veröffentlicht: (2022)
von: Apidopoulos, Vassilis, et al.
Veröffentlicht: (2022)
Convergence of stochastic gradient descent under a local Lojasiewicz condition for deep neural networks
von: An, Jing, et al.
Veröffentlicht: (2023)
von: An, Jing, et al.
Veröffentlicht: (2023)
Manifold constrained steepest descent
von: Yang, Kaiwei, et al.
Veröffentlicht: (2026)
von: Yang, Kaiwei, et al.
Veröffentlicht: (2026)
Riemannian coordinate descent algorithms on matrix manifolds
von: Han, Andi, et al.
Veröffentlicht: (2024)
von: Han, Andi, et al.
Veröffentlicht: (2024)
New logarithmic step size for stochastic gradient descent
von: Shamaee, M. Soheil, et al.
Veröffentlicht: (2024)
von: Shamaee, M. Soheil, et al.
Veröffentlicht: (2024)
Gradient descent in matrix factorization: Understanding large initialization
von: Chen, Hengchao, et al.
Veröffentlicht: (2023)
von: Chen, Hengchao, et al.
Veröffentlicht: (2023)
On the global convergence of gradient descent for wide shallow models with bounded nonlinearities
von: Petit, Romain, et al.
Veröffentlicht: (2026)
von: Petit, Romain, et al.
Veröffentlicht: (2026)
The duality structure gradient descent algorithm: analysis and applications to neural networks
von: Flynn, Thomas
Veröffentlicht: (2017)
von: Flynn, Thomas
Veröffentlicht: (2017)
On the stability of gradient descent with second order dynamics for time-varying cost functions
von: Gibson, Travis E., et al.
Veröffentlicht: (2024)
von: Gibson, Travis E., et al.
Veröffentlicht: (2024)
Convergence of continuous-time stochastic gradient descent with applications to deep neural networks
von: Lugosi, Gabor, et al.
Veröffentlicht: (2024)
von: Lugosi, Gabor, et al.
Veröffentlicht: (2024)
The Mixing method: low-rank coordinate descent for semidefinite programming with diagonal constraints
von: Wang, Po-Wei, et al.
Veröffentlicht: (2017)
von: Wang, Po-Wei, et al.
Veröffentlicht: (2017)
Convergence of two-timescale gradient descent ascent dynamics: finite-dimensional and mean-field perspectives
von: An, Jing, et al.
Veröffentlicht: (2025)
von: An, Jing, et al.
Veröffentlicht: (2025)
Error dynamics of mini-batch gradient descent with random reshuffling for least squares regression
von: Lok, Jackie, et al.
Veröffentlicht: (2024)
von: Lok, Jackie, et al.
Veröffentlicht: (2024)
Linear convergence of proximal descent schemes on the Wasserstein space
von: Lascu, Razvan-Andrei, et al.
Veröffentlicht: (2024)
von: Lascu, Razvan-Andrei, et al.
Veröffentlicht: (2024)
A stochastic gradient descent algorithm with random search directions
von: Gbaguidi, Eméric
Veröffentlicht: (2025)
von: Gbaguidi, Eméric
Veröffentlicht: (2025)
Convergence of flow-based generative models via proximal gradient descent in Wasserstein space
von: Cheng, Xiuyuan, et al.
Veröffentlicht: (2023)
von: Cheng, Xiuyuan, et al.
Veröffentlicht: (2023)
An accelerated first-order regularized momentum descent ascent algorithm for stochastic nonconvex-concave minimax problems
von: Zhang, Huiling, et al.
Veröffentlicht: (2023)
von: Zhang, Huiling, et al.
Veröffentlicht: (2023)
A short proof of near-linear convergence of adaptive gradient descent under fourth-order growth and convexity
von: Davis, Damek, et al.
Veröffentlicht: (2026)
von: Davis, Damek, et al.
Veröffentlicht: (2026)
A block-coordinate descent framework for non-convex composite optimization. Application to sparse precision matrix estimation
von: Lauga, Guillaume
Veröffentlicht: (2026)
von: Lauga, Guillaume
Veröffentlicht: (2026)
Aiding Global Convergence in Federated Learning via Local Perturbation and Mutual Similarity Information
von: Buttaci, Emanuel, et al.
Veröffentlicht: (2024)
von: Buttaci, Emanuel, et al.
Veröffentlicht: (2024)
Non-convergence to global minimizers for Adam and stochastic gradient descent optimization and constructions of local minimizers in the training of artificial neural networks
von: Jentzen, Arnulf, et al.
Veröffentlicht: (2024)
von: Jentzen, Arnulf, et al.
Veröffentlicht: (2024)
Operator World Models for Reinforcement Learning
von: Novelli, Pietro, et al.
Veröffentlicht: (2024)
von: Novelli, Pietro, et al.
Veröffentlicht: (2024)
Monotone Optimisation with Learned Projections
von: Rashwan, Ahmed, et al.
Veröffentlicht: (2026)
von: Rashwan, Ahmed, et al.
Veröffentlicht: (2026)
Scalable Decentralized Learning with Teleportation
von: Takezawa, Yuki, et al.
Veröffentlicht: (2025)
von: Takezawa, Yuki, et al.
Veröffentlicht: (2025)
When majority rules, minority loses: bias amplification of gradient descent
von: Bachoc, François, et al.
Veröffentlicht: (2025)
von: Bachoc, François, et al.
Veröffentlicht: (2025)
Global convergence of gradient descent for phase retrieval
von: Fougereux, Théodore, et al.
Veröffentlicht: (2024)
von: Fougereux, Théodore, et al.
Veröffentlicht: (2024)
Deep Learning Model Predictive Control for Deep Brain Stimulation in Parkinson's Disease
von: Steffen, Sebastian, et al.
Veröffentlicht: (2025)
von: Steffen, Sebastian, et al.
Veröffentlicht: (2025)
Multi-objective Deep Learning: Taxonomy and Survey of the State of the Art
von: Peitz, Sebastian, et al.
Veröffentlicht: (2024)
von: Peitz, Sebastian, et al.
Veröffentlicht: (2024)
Locally Adaptive Federated Learning
von: Mukherjee, Sohom, et al.
Veröffentlicht: (2023)
von: Mukherjee, Sohom, et al.
Veröffentlicht: (2023)
A note on convergence of Wasserstein policy optimization
von: Šiška, David, et al.
Veröffentlicht: (2026)
von: Šiška, David, et al.
Veröffentlicht: (2026)
A General Continuous-Time Formulation of Stochastic ADMM and Its Variants
von: Li, Chris Junchi
Veröffentlicht: (2024)
von: Li, Chris Junchi
Veröffentlicht: (2024)
Accelerated Fully First-Order Methods for Bilevel and Minimax Optimization
von: Li, Chris Junchi
Veröffentlicht: (2024)
von: Li, Chris Junchi
Veröffentlicht: (2024)
Enhancing Stochastic Optimization for Statistical Efficiency Using ROOT-SGD with Diminishing Stepsize
von: Li, Chris Junchi
Veröffentlicht: (2024)
von: Li, Chris Junchi
Veröffentlicht: (2024)
Distributed Event-Based Learning via ADMM
von: Er, Guner Dilsad, et al.
Veröffentlicht: (2024)
von: Er, Guner Dilsad, et al.
Veröffentlicht: (2024)
Flattened one-bit stochastic gradient descent: compressed distributed optimization with controlled variance
von: Stollenwerk, Alexander, et al.
Veröffentlicht: (2024)
von: Stollenwerk, Alexander, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence
von: Alfano, Carlo, et al.
Veröffentlicht: (2023) -
On exploration of an interior mirror descent flow for stochastic nonconvex constrained problem
von: Ding, Kuangyu, et al.
Veröffentlicht: (2025) -
Meta-Learning Objectives for Preference Optimization
von: Alfano, Carlo, et al.
Veröffentlicht: (2024) -
Entropy annealing for policy mirror descent in continuous time and space
von: Sethi, Deven, et al.
Veröffentlicht: (2024) -
Global minimisation of nonconvex functions by generalising the mirror descent method
von: Millán, Reinier Díaz, et al.
Veröffentlicht: (2024)