A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence
Fuente:
arXiv
Saved in:
| Main Authors: | Alfano, Carlo, Yuan, Rui, Rebeschini, Patrick |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A connection between Tempering and Entropic Mirror Descent
by: Chopin, Nicolas, et al.
Published: (2023)
by: Chopin, Nicolas, et al.
Published: (2023)
High-probability Convergence Bounds for Nonlinear Stochastic Gradient Descent Under Heavy-tailed Noise
by: Armacki, Aleksandar, et al.
Published: (2023)
by: Armacki, Aleksandar, et al.
Published: (2023)
Decentralized Sparse Linear Regression via Gradient-Tracking: Linear Convergence and Statistical Guarantees
by: Maros, Marie, et al.
Published: (2022)
by: Maros, Marie, et al.
Published: (2022)
Learning mirror maps in policy mirror descent
by: Alfano, Carlo, et al.
Published: (2024)
by: Alfano, Carlo, et al.
Published: (2024)
On the Convergence of Policy in Unregularized Policy Mirror Descent
by: Lin, Dachao, et al.
Published: (2022)
by: Lin, Dachao, et al.
Published: (2022)
Trajectory-Restricted Optimization Conditions and Geometry-Aware Linear Convergence
by: Chaudhry, Faris, et al.
Published: (2026)
by: Chaudhry, Faris, et al.
Published: (2026)
On the Convergence of Policy Mirror Descent with Temporal Difference Evaluation
by: Liu, Jiacai, et al.
Published: (2025)
by: Liu, Jiacai, et al.
Published: (2025)
Implicit Regularization for Tubal Tensor Factorizations via Gradient Descent
by: Karnik, Santhosh, et al.
Published: (2024)
by: Karnik, Santhosh, et al.
Published: (2024)
Stopping Rules for Stochastic Gradient Descent via Anytime-Valid Confidence Sequences
by: Aolaritei, Liviu, et al.
Published: (2025)
by: Aolaritei, Liviu, et al.
Published: (2025)
Time-sensitive anytime-valid testing
by: Clerico, Eugenio, et al.
Published: (2026)
by: Clerico, Eugenio, et al.
Published: (2026)
Convergence of Policy Mirror Descent Beyond Compatible Function Approximation
by: Sherman, Uri, et al.
Published: (2025)
by: Sherman, Uri, et al.
Published: (2025)
Robustly Learning Monotone Generalized Linear Models via Data Augmentation
by: Zarifis, Nikos, et al.
Published: (2025)
by: Zarifis, Nikos, et al.
Published: (2025)
Variational Transport: A Convergent Particle-BasedAlgorithm for Distributional Optimization
by: Yang, Zhuoran, et al.
Published: (2020)
by: Yang, Zhuoran, et al.
Published: (2020)
On the Uniform Convergence of Subdifferentials in Stochastic Optimization and Learning
by: Ruan, Feng
Published: (2024)
by: Ruan, Feng
Published: (2024)
Blessings and Curses of Covariate Shifts: Adversarial Learning Dynamics, Directional Convergence, and Equilibria
by: Liang, Tengyuan
Published: (2022)
by: Liang, Tengyuan
Published: (2022)
Convergence of flow-based generative models via proximal gradient descent in Wasserstein space
by: Cheng, Xiuyuan, et al.
Published: (2023)
by: Cheng, Xiuyuan, et al.
Published: (2023)
High-dimensional Limit of SGD for Diagonal Linear Networks
by: Malaxechebarría, Begoña García, et al.
Published: (2026)
by: Malaxechebarría, Begoña García, et al.
Published: (2026)
Beyond Maximum Likelihood: Variational Inequality Estimation for Generalized Linear Models
by: Zhu, Linglingzhi, et al.
Published: (2025)
by: Zhu, Linglingzhi, et al.
Published: (2025)
Learning an Optimal Assortment Policy under Observational Data
by: Han, Yuxuan, et al.
Published: (2025)
by: Han, Yuxuan, et al.
Published: (2025)
On the Sample Complexity of Set Membership Estimation for Linear Systems with Disturbances Bounded by Convex Sets
by: Xu, Haonan, et al.
Published: (2024)
by: Xu, Haonan, et al.
Published: (2024)
A Spectral Framework for Closed-Form Relative Density Estimation
by: Bach, Francis
Published: (2026)
by: Bach, Francis
Published: (2026)
Statistical Inference for Linear Functionals of Online SGD in High-dimensional Linear Regression
by: Agrawalla, Bhavya, et al.
Published: (2023)
by: Agrawalla, Bhavya, et al.
Published: (2023)
Joint Learning of Linear Dynamical Systems under Smoothness Constraints
by: Tyagi, Hemant
Published: (2024)
by: Tyagi, Hemant
Published: (2024)
Implicit Bias and Convergence of Matrix Stochastic Mirror Descent
by: Akhtiamov, Danil, et al.
Published: (2026)
by: Akhtiamov, Danil, et al.
Published: (2026)
Convergence of coordinate ascent variational inference for log-concave measures via optimal transport
by: Arnese, Manuel, et al.
Published: (2024)
by: Arnese, Manuel, et al.
Published: (2024)
$K$-Nearest-Neighbor Resampling for Off-Policy Evaluation in Stochastic Control
by: Giegrich, Michael, et al.
Published: (2023)
by: Giegrich, Michael, et al.
Published: (2023)
Second-Order Mirror Descent: Convergence in Games Beyond Averaging and Discounting
by: Gao, Bolin, et al.
Published: (2021)
by: Gao, Bolin, et al.
Published: (2021)
Markov Kernels, Distances and Optimal Control: A Parable of Linear Quadratic Non-Gaussian Distribution Steering
by: Teter, Alexis M. H., et al.
Published: (2025)
by: Teter, Alexis M. H., et al.
Published: (2025)
Gradient Projection onto Historical Descent Directions for Communication-Efficient Federated Learning
by: Descours, Arnaud, et al.
Published: (2025)
by: Descours, Arnaud, et al.
Published: (2025)
Training Dynamics of Multi-Head Softmax Attention for In-Context Learning: Emergence, Convergence, and Optimality
by: Chen, Siyu, et al.
Published: (2024)
by: Chen, Siyu, et al.
Published: (2024)
Convergence rate of random scan Coordinate Ascent Variational Inference under log-concavity
by: Lavenant, Hugo, et al.
Published: (2024)
by: Lavenant, Hugo, et al.
Published: (2024)
A Theory of Feature Learning in Kernel Models
by: Chen, Yunlu, et al.
Published: (2023)
by: Chen, Yunlu, et al.
Published: (2023)
A New Perspective On Denoising Based On Optimal Transport
by: Trillos, Nicolas Garcia, et al.
Published: (2023)
by: Trillos, Nicolas Garcia, et al.
Published: (2023)
A review of NMF, PLSA, LBA, EMA, and LCA with a focus on the identifiability issue
by: Qi, Qianqian, et al.
Published: (2025)
by: Qi, Qianqian, et al.
Published: (2025)
Computation of Least Trimmed Squares: A Branch-and-Bound framework with Hyperplane Arrangement Enhancements
by: Meng, Xiang, et al.
Published: (2026)
by: Meng, Xiang, et al.
Published: (2026)
Learning the Uncertainty Sets for Control Dynamics via Set Membership: A Non-Asymptotic Analysis
by: Li, Yingying, et al.
Published: (2023)
by: Li, Yingying, et al.
Published: (2023)
A Mirror Descent Perspective of Smoothed Sign Descent
by: Wang, Shuyang, et al.
Published: (2024)
by: Wang, Shuyang, et al.
Published: (2024)
Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model
by: Li, Gen, et al.
Published: (2020)
by: Li, Gen, et al.
Published: (2020)
Algorithms for mean-field variational inference via polyhedral optimization in the Wasserstein space
by: Jiang, Yiheng, et al.
Published: (2023)
by: Jiang, Yiheng, et al.
Published: (2023)
Certified Multi-Fidelity Zeroth-Order Optimization
by: de Montbrun, Étienne, et al.
Published: (2023)
by: de Montbrun, Étienne, et al.
Published: (2023)
Similar Items
-
A connection between Tempering and Entropic Mirror Descent
by: Chopin, Nicolas, et al.
Published: (2023) -
High-probability Convergence Bounds for Nonlinear Stochastic Gradient Descent Under Heavy-tailed Noise
by: Armacki, Aleksandar, et al.
Published: (2023) -
Decentralized Sparse Linear Regression via Gradient-Tracking: Linear Convergence and Statistical Guarantees
by: Maros, Marie, et al.
Published: (2022) -
Learning mirror maps in policy mirror descent
by: Alfano, Carlo, et al.
Published: (2024) -
On the Convergence of Policy in Unregularized Policy Mirror Descent
by: Lin, Dachao, et al.
Published: (2022)