Convergence and Sample Complexity of First-Order Methods for Agnostic Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sherman, Uri, Koren, Tomer, Mansour, Yishay |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Convergence of Policy Mirror Descent Beyond Compatible Function Approximation
von: Sherman, Uri, et al.
Veröffentlicht: (2025)
von: Sherman, Uri, et al.
Veröffentlicht: (2025)
Fast Last-Iterate Convergence of SGD in the Smooth Interpolation Regime
von: Attia, Amit, et al.
Veröffentlicht: (2025)
von: Attia, Amit, et al.
Veröffentlicht: (2025)
The Hidden Cost of Approximation in Online Mirror Descent
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025)
Rate-Optimal Policy Optimization for Linear Markov Decision Processes
von: Sherman, Uri, et al.
Veröffentlicht: (2023)
von: Sherman, Uri, et al.
Veröffentlicht: (2023)
Learning Rate Annealing Improves Tuning Robustness in Stochastic Optimization
von: Attia, Amit, et al.
Veröffentlicht: (2025)
von: Attia, Amit, et al.
Veröffentlicht: (2025)
How Free is Parameter-Free Stochastic Optimization?
von: Attia, Amit, et al.
Veröffentlicht: (2024)
von: Attia, Amit, et al.
Veröffentlicht: (2024)
Regret Minimization and Convergence to Equilibria in General-sum Markov Games
von: Erez, Liad, et al.
Veröffentlicht: (2022)
von: Erez, Liad, et al.
Veröffentlicht: (2022)
Faster Stochastic Optimization with Arbitrary Delays via Asynchronous Mini-Batching
von: Attia, Amit, et al.
Veröffentlicht: (2024)
von: Attia, Amit, et al.
Veröffentlicht: (2024)
Rapid Overfitting of Multi-Pass Stochastic Gradient Descent in Stochastic Convex Optimization
von: Vansover-Hager, Shira, et al.
Veröffentlicht: (2025)
von: Vansover-Hager, Shira, et al.
Veröffentlicht: (2025)
On the Complexity of First-Order Methods in Stochastic Bilevel Optimization
von: Kwon, Jeongyeol, et al.
Veröffentlicht: (2024)
von: Kwon, Jeongyeol, et al.
Veröffentlicht: (2024)
A New First-Order Meta-Learning Algorithm with Convergence Guarantees
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2024)
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2024)
Optimal Local Convergence Rates of Stochastic First-Order Methods under Local $α$-PL
von: Masiha, Saeed, et al.
Veröffentlicht: (2024)
von: Masiha, Saeed, et al.
Veröffentlicht: (2024)
Sample Complexity of Agnostic Multiclass Classification: Natarajan Dimension Strikes Back
von: Cohen, Alon, et al.
Veröffentlicht: (2025)
von: Cohen, Alon, et al.
Veröffentlicht: (2025)
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
von: Chen, Zijun, et al.
Veröffentlicht: (2025)
von: Chen, Zijun, et al.
Veröffentlicht: (2025)
Sample-Efficient Agnostic Boosting
von: Ghai, Udaya, et al.
Veröffentlicht: (2024)
von: Ghai, Udaya, et al.
Veröffentlicht: (2024)
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
von: Carmona, René, et al.
Veröffentlicht: (2019)
von: Carmona, René, et al.
Veröffentlicht: (2019)
Convergence and Complexity Guarantee for Inexact First-order Riemannian Optimization Algorithms
von: Li, Yuchen, et al.
Veröffentlicht: (2024)
von: Li, Yuchen, et al.
Veröffentlicht: (2024)
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
von: Muehlebach, Michael, et al.
Veröffentlicht: (2025)
von: Muehlebach, Michael, et al.
Veröffentlicht: (2025)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
von: Moghaddam, Amirreza Neshaei, et al.
Veröffentlicht: (2024)
von: Moghaddam, Amirreza Neshaei, et al.
Veröffentlicht: (2024)
Towards Fully Parameter-Free Stochastic Optimization: Grid Search with Self-Bounding Analysis
von: Zhao, Yuheng, et al.
Veröffentlicht: (2026)
von: Zhao, Yuheng, et al.
Veröffentlicht: (2026)
First-Order Methods for Linearly Constrained Bilevel Optimization
von: Kornowski, Guy, et al.
Veröffentlicht: (2024)
von: Kornowski, Guy, et al.
Veröffentlicht: (2024)
Improved Convergence in Parameter-Agnostic Error Feedback through Momentum
von: Sadiev, Abdurakhmon, et al.
Veröffentlicht: (2025)
von: Sadiev, Abdurakhmon, et al.
Veröffentlicht: (2025)
Minimax Excess Risk of First-Order Methods for Statistical Learning with Data-Dependent Oracles
von: Scaman, Kevin, et al.
Veröffentlicht: (2023)
von: Scaman, Kevin, et al.
Veröffentlicht: (2023)
Biased Stochastic First-Order Methods for Conditional Stochastic Optimization and Applications in Meta Learning
von: Hu, Yifan, et al.
Veröffentlicht: (2020)
von: Hu, Yifan, et al.
Veröffentlicht: (2020)
Revisiting Subgradient Method: Complexity and Convergence Beyond Lipschitz Continuity
von: Li, Xiao, et al.
Veröffentlicht: (2023)
von: Li, Xiao, et al.
Veröffentlicht: (2023)
Accelerated Fully First-Order Methods for Bilevel and Minimax Optimization
von: Li, Chris Junchi
Veröffentlicht: (2024)
von: Li, Chris Junchi
Veröffentlicht: (2024)
Batched First-Order Methods for Parallel LP Solving in MIP
von: Blin, Nicolas, et al.
Veröffentlicht: (2026)
von: Blin, Nicolas, et al.
Veröffentlicht: (2026)
The Dimension Strikes Back with Gradients: Generalization of Gradient Methods in Stochastic Convex Optimization
von: Schliserman, Matan, et al.
Veröffentlicht: (2024)
von: Schliserman, Matan, et al.
Veröffentlicht: (2024)
Private Online Learning via Lazy Algorithms
von: Asi, Hilal, et al.
Veröffentlicht: (2024)
von: Asi, Hilal, et al.
Veröffentlicht: (2024)
First Order Methods with Markovian Noise: from Acceleration to Variational Inequalities
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2023)
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2023)
On Penalty Methods for Nonconvex Bilevel Optimization and First-Order Stochastic Approximation
von: Kwon, Jeongyeol, et al.
Veröffentlicht: (2023)
von: Kwon, Jeongyeol, et al.
Veröffentlicht: (2023)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
von: Manenti, Massimiliano, et al.
Veröffentlicht: (2025)
von: Manenti, Massimiliano, et al.
Veröffentlicht: (2025)
Lower Complexity Bounds for Nonconvex-Strongly-Convex Bilevel Optimization with First-Order Oracles
von: Ji, Kaiyi
Veröffentlicht: (2025)
von: Ji, Kaiyi
Veröffentlicht: (2025)
First and Second Order Approximations to Stochastic Gradient Descent Methods with Momentum Terms
von: Lu, Eric
Veröffentlicht: (2025)
von: Lu, Eric
Veröffentlicht: (2025)
Span-Agnostic Optimal Sample Complexity and Oracle Inequalities for Average-Reward RL
von: Zurek, Matthew, et al.
Veröffentlicht: (2025)
von: Zurek, Matthew, et al.
Veröffentlicht: (2025)
Krylov Cubic Regularized Newton: A Subspace Second-Order Method with Dimension-Free Convergence Rate
von: Jiang, Ruichen, et al.
Veröffentlicht: (2024)
von: Jiang, Ruichen, et al.
Veröffentlicht: (2024)
GPU-friendly and Linearly Convergent First-order Methods for Certifying Optimal $k$-sparse GLMs
von: Liu, Jiachang, et al.
Veröffentlicht: (2026)
von: Liu, Jiachang, et al.
Veröffentlicht: (2026)
Decoupling Learning and Decision-Making: Breaking the $\mathcal{O}(\sqrt{T})$ Barrier in Online Resource Allocation with First-Order Methods
von: Gao, Wenzhi, et al.
Veröffentlicht: (2024)
von: Gao, Wenzhi, et al.
Veröffentlicht: (2024)
Last Iterate Convergence of Incremental Methods and Applications in Continual Learning
von: Cai, Xufeng, et al.
Veröffentlicht: (2024)
von: Cai, Xufeng, et al.
Veröffentlicht: (2024)
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
von: Xiao, Minheng, et al.
Veröffentlicht: (2024)
von: Xiao, Minheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Convergence of Policy Mirror Descent Beyond Compatible Function Approximation
von: Sherman, Uri, et al.
Veröffentlicht: (2025) -
Fast Last-Iterate Convergence of SGD in the Smooth Interpolation Regime
von: Attia, Amit, et al.
Veröffentlicht: (2025) -
The Hidden Cost of Approximation in Online Mirror Descent
von: Schlisselberg, Ofir, et al.
Veröffentlicht: (2025) -
Rate-Optimal Policy Optimization for Linear Markov Decision Processes
von: Sherman, Uri, et al.
Veröffentlicht: (2023) -
Learning Rate Annealing Improves Tuning Robustness in Stochastic Optimization
von: Attia, Amit, et al.
Veröffentlicht: (2025)