Stochastic Halpern iteration in normed spaces and applications to reinforcement learning
Fuente:
arXiv
Saved in:
| Main Authors: | Bravo, Mario, Contreras, Juan Pablo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Minimax-optimal Halpern iterations for Lipschitz maps
by: Bravo, Mario, et al.
Published: (2026)
by: Bravo, Mario, et al.
Published: (2026)
Asymptotic regularity of a generalised stochastic Halpern scheme
by: Pischke, Nicholas, et al.
Published: (2024)
by: Pischke, Nicholas, et al.
Published: (2024)
Randomized algorithms and PAC bounds for inverse reinforcement learning in continuous spaces
by: Kamoutsi, Angeliki, et al.
Published: (2024)
by: Kamoutsi, Angeliki, et al.
Published: (2024)
Stochastic set-valued optimization and its application to robust learning
by: Giovannelli, Tommaso, et al.
Published: (2026)
by: Giovannelli, Tommaso, et al.
Published: (2026)
Convergence analysis of the Halpern iteration with adaptive anchoring parameters
by: He, Songnian, et al.
Published: (2025)
by: He, Songnian, et al.
Published: (2025)
What is the objective of reasoning with reinforcement learning?
by: Davis, Damek, et al.
Published: (2025)
by: Davis, Damek, et al.
Published: (2025)
Meta-reinforcement learning with minimum attention
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
Taming "data-hungry" reinforcement learning? Stability in continuous state-action spaces
by: Duan, Yaqi, et al.
Published: (2024)
by: Duan, Yaqi, et al.
Published: (2024)
Flowsheet synthesis through hierarchical reinforcement learning and graph neural networks
by: Stops, Laura, et al.
Published: (2022)
by: Stops, Laura, et al.
Published: (2022)
A successive approximation method in functional spaces for hierarchical optimal control problems and its application to learning
by: Befekadu, Getachew K.
Published: (2024)
by: Befekadu, Getachew K.
Published: (2024)
Online reinforcement learning via sparse Gaussian mixture model Q-functions
by: Vu, Minh, et al.
Published: (2025)
by: Vu, Minh, et al.
Published: (2025)
A reinforcement learning agent for maintenance of deteriorating systems with increasingly imperfect repairs
by: Marugán, Alberto Pliego, et al.
Published: (2025)
by: Marugán, Alberto Pliego, et al.
Published: (2025)
Preconditioned Halpern iteration with adaptive anchoring parameters and an acceleration to Chambolle--Pock algorithm
by: Lv, Fangbing, et al.
Published: (2025)
by: Lv, Fangbing, et al.
Published: (2025)
l1-norm regularized l1-norm best-fit lines
by: Ling, Xiao, et al.
Published: (2024)
by: Ling, Xiao, et al.
Published: (2024)
On decomposability and subdifferential of the tensor nuclear norm
by: Guan, Jiewen, et al.
Published: (2025)
by: Guan, Jiewen, et al.
Published: (2025)
Independent policy gradient-based reinforcement learning for economic and reliable energy management of multi-microgrid systems
by: Hu, Junkai, et al.
Published: (2025)
by: Hu, Junkai, et al.
Published: (2025)
Sum-of-norms regularized Nonnegative Matrix Factorization
by: Ang, Andersen, et al.
Published: (2024)
by: Ang, Andersen, et al.
Published: (2024)
On bounds for norms of reparameterized ReLU artificial neural network parameters: sums of fractional powers of the Lipschitz norm control the network parameter vector
by: Jentzen, Arnulf, et al.
Published: (2022)
by: Jentzen, Arnulf, et al.
Published: (2022)
Learning to accelerate Krasnosel'skii-Mann fixed-point iterations with guarantees
by: Martin, Andrea, et al.
Published: (2026)
by: Martin, Andrea, et al.
Published: (2026)
Scalable spectral representations for multi-agent reinforcement learning in network MDPs
by: Ren, Zhaolin, et al.
Published: (2024)
by: Ren, Zhaolin, et al.
Published: (2024)
An adaptively inexact first-order method for bilevel optimization with application to hyperparameter learning
by: Salehi, Mohammad Sadegh, et al.
Published: (2023)
by: Salehi, Mohammad Sadegh, et al.
Published: (2023)
Mini-Batch Stochastic Halpern Algorithm for Nonexpansive Fixed point Problems
by: Iiduka, Hideaki
Published: (2026)
by: Iiduka, Hideaki
Published: (2026)
Mixing Times and Privacy Analysis for the Projected Langevin Algorithm under a Modulus of Continuity
by: Bravo, Mario, et al.
Published: (2025)
by: Bravo, Mario, et al.
Published: (2025)
Adaptive multi-gradient methods for quasiconvex vector optimization and applications to multi-task learning
by: Minh, Nguyen Anh, et al.
Published: (2024)
by: Minh, Nguyen Anh, et al.
Published: (2024)
Continuous-time reinforcement learning for optimal switching over multiple regimes
by: Huang, Yijie, et al.
Published: (2025)
by: Huang, Yijie, et al.
Published: (2025)
Exploiting inter-agent coupling information for efficient reinforcement learning of cooperative LQR
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
by: Syed, Shahbaz P Qadri, et al.
Published: (2025)
Stochastic-Constrained Stochastic Optimization with Markovian Data
by: Kim, Yeongjong, et al.
Published: (2023)
by: Kim, Yeongjong, et al.
Published: (2023)
Momentum Does Not Reduce Stochastic Noise in Stochastic Gradient Descent
by: Sato, Naoki, et al.
Published: (2024)
by: Sato, Naoki, et al.
Published: (2024)
Accelerating nuclear-norm regularized low-rank matrix optimization through Burer-Monteiro decomposition
by: Lee, Ching-pei, et al.
Published: (2022)
by: Lee, Ching-pei, et al.
Published: (2022)
Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems
by: Yang, Xianjin, et al.
Published: (2025)
by: Yang, Xianjin, et al.
Published: (2025)
Bilevel reinforcement learning via the development of hyper-gradient without lower-level convexity
by: Yang, Yan, et al.
Published: (2024)
by: Yang, Yan, et al.
Published: (2024)
Continuous-time reinforcement learning: ellipticity enables model-free value function approximation
by: Mou, Wenlong
Published: (2026)
by: Mou, Wenlong
Published: (2026)
Stabilizing reinforcement learning control: A modular framework for optimizing over all stable behavior
by: Lawrence, Nathan P., et al.
Published: (2023)
by: Lawrence, Nathan P., et al.
Published: (2023)
humancompatible.train: Implementing Optimization Algorithms for Stochastically-Constrained Stochastic Optimization Problems
by: Kliachkin, Andrii, et al.
Published: (2025)
by: Kliachkin, Andrii, et al.
Published: (2025)
Rapid Overfitting of Multi-Pass Stochastic Gradient Descent in Stochastic Convex Optimization
by: Vansover-Hager, Shira, et al.
Published: (2025)
by: Vansover-Hager, Shira, et al.
Published: (2025)
Learning a local trading strategy: deep reinforcement learning for grid-scale renewable energy integration
by: Ju, Caleb, et al.
Published: (2024)
by: Ju, Caleb, et al.
Published: (2024)
Biased Stochastic First-Order Methods for Conditional Stochastic Optimization and Applications in Meta Learning
by: Hu, Yifan, et al.
Published: (2020)
by: Hu, Yifan, et al.
Published: (2020)
$\ell_1$-norm rank-one symmetric matrix factorization has no spurious second-order stationary points
by: Guan, Jiewen, et al.
Published: (2024)
by: Guan, Jiewen, et al.
Published: (2024)
Tuning-Free Stochastic Optimization
by: Khaled, Ahmed, et al.
Published: (2024)
by: Khaled, Ahmed, et al.
Published: (2024)
Stochastic Gradients under Nuisances
by: Yu, Facheng, et al.
Published: (2025)
by: Yu, Facheng, et al.
Published: (2025)
Similar Items
-
Minimax-optimal Halpern iterations for Lipschitz maps
by: Bravo, Mario, et al.
Published: (2026) -
Asymptotic regularity of a generalised stochastic Halpern scheme
by: Pischke, Nicholas, et al.
Published: (2024) -
Randomized algorithms and PAC bounds for inverse reinforcement learning in continuous spaces
by: Kamoutsi, Angeliki, et al.
Published: (2024) -
Stochastic set-valued optimization and its application to robust learning
by: Giovannelli, Tommaso, et al.
Published: (2026) -
Convergence analysis of the Halpern iteration with adaptive anchoring parameters
by: He, Songnian, et al.
Published: (2025)