Stochastic first-order methods for average-reward Markov decision processes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Tianjiao, Wu, Feiyang, Lan, Guanghui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Auto-conditioned primal-dual hybrid gradient method and alternating direction method of multipliers
von: Lan, Guanghui, et al.
Veröffentlicht: (2024)
von: Lan, Guanghui, et al.
Veröffentlicht: (2024)
Logarithmic regret bounds for continuous-time average-reward Markov decision processes
von: Gao, Xuefeng, et al.
Veröffentlicht: (2022)
von: Gao, Xuefeng, et al.
Veröffentlicht: (2022)
A simple uniformly optimal method without line search for convex optimization
von: Li, Tianjiao, et al.
Veröffentlicht: (2023)
von: Li, Tianjiao, et al.
Veröffentlicht: (2023)
Projected gradient methods for nonconvex and stochastic smooth optimization: new complexities and auto-conditioned stepsizes
von: Lan, Guanghui, et al.
Veröffentlicht: (2024)
von: Lan, Guanghui, et al.
Veröffentlicht: (2024)
Accelerated stochastic approximation with state-dependent noise
von: Ilandarideva, Sasila, et al.
Veröffentlicht: (2023)
von: Ilandarideva, Sasila, et al.
Veröffentlicht: (2023)
Stochastic Auto-conditioned Fast Gradient Methods with Optimal Rates
von: Ji, Yao, et al.
Veröffentlicht: (2026)
von: Ji, Yao, et al.
Veröffentlicht: (2026)
A safe exploration approach to constrained Markov decision processes
von: Ni, Tingting, et al.
Veröffentlicht: (2023)
von: Ni, Tingting, et al.
Veröffentlicht: (2023)
Data-driven robust Markov decision processes on Borel spaces: performance guarantees via an axiomatic approach
von: Ramani, Sivaramakrishnan
Veröffentlicht: (2026)
von: Ramani, Sivaramakrishnan
Veröffentlicht: (2026)
Value Mirror Descent for Reinforcement Learning
von: Jia, Zhichao, et al.
Veröffentlicht: (2026)
von: Jia, Zhichao, et al.
Veröffentlicht: (2026)
Average-reward reinforcement learning in semi-Markov decision processes via relative value iteration
von: Yu, Huizhen, et al.
Veröffentlicht: (2025)
von: Yu, Huizhen, et al.
Veröffentlicht: (2025)
First-order methods for Stochastic Variational Inequality problems with Function Constraints
von: Boob, Digvijay, et al.
Veröffentlicht: (2023)
von: Boob, Digvijay, et al.
Veröffentlicht: (2023)
Can SGD Handle Heavy-Tailed Noise?
von: Fatkhullin, Ilyas, et al.
Veröffentlicht: (2025)
von: Fatkhullin, Ilyas, et al.
Veröffentlicht: (2025)
An adaptively inexact first-order method for bilevel optimization with application to hyperparameter learning
von: Salehi, Mohammad Sadegh, et al.
Veröffentlicht: (2023)
von: Salehi, Mohammad Sadegh, et al.
Veröffentlicht: (2023)
Toward generalizable learning of all (linear) first-order methods via memory augmented Transformers
von: Dutta, Sanchayan, et al.
Veröffentlicht: (2024)
von: Dutta, Sanchayan, et al.
Veröffentlicht: (2024)
Multiscale replay: A robust algorithm for stochastic variational inequalities with a Markovian buffer
von: Nakul, Milind, et al.
Veröffentlicht: (2026)
von: Nakul, Milind, et al.
Veröffentlicht: (2026)
Universal Online Convex Optimization Meets Second-order Bounds
von: Zhang, Lijun, et al.
Veröffentlicht: (2021)
von: Zhang, Lijun, et al.
Veröffentlicht: (2021)
Projection-Free Functional Constrained Optimization for Risk Aversion and Sparsity Control
von: Cheng, Yi, et al.
Veröffentlicht: (2022)
von: Cheng, Yi, et al.
Veröffentlicht: (2022)
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
Stochastic Zeroth order Descent with Structured Directions
von: Rando, Marco, et al.
Veröffentlicht: (2022)
von: Rando, Marco, et al.
Veröffentlicht: (2022)
One-Sided Matrix Completion from Ultra-Sparse Samples
von: Zhang, Hongyang R., et al.
Veröffentlicht: (2026)
von: Zhang, Hongyang R., et al.
Veröffentlicht: (2026)
End-to-End Training of High-Dimensional Optimal Control with Implicit Hamiltonians via Jacobian-Free Backpropagation
von: Gelphman, Eric, et al.
Veröffentlicht: (2025)
von: Gelphman, Eric, et al.
Veröffentlicht: (2025)
State evolution beyond first-order methods I: Rigorous predictions and finite-sample guarantees
von: Celentano, Michael, et al.
Veröffentlicht: (2025)
von: Celentano, Michael, et al.
Veröffentlicht: (2025)
Adaptive Batch Size and Learning Rate Scheduler for Stochastic Gradient Descent Based on Minimization of Stochastic First-order Oracle Complexity
von: Umeda, Hikaru, et al.
Veröffentlicht: (2025)
von: Umeda, Hikaru, et al.
Veröffentlicht: (2025)
Tight analyses of first-order methods with error feedback
von: Thomsen, Daniel Berg, et al.
Veröffentlicht: (2025)
von: Thomsen, Daniel Berg, et al.
Veröffentlicht: (2025)
Bregman Linearized Augmented Lagrangian Method for Nonconvex Constrained Stochastic Zeroth-order Optimization
von: Shi, Qiankun, et al.
Veröffentlicht: (2025)
von: Shi, Qiankun, et al.
Veröffentlicht: (2025)
Parameter Symmetry and Noise Equilibrium of Stochastic Gradient Descent
von: Ziyin, Liu, et al.
Veröffentlicht: (2024)
von: Ziyin, Liu, et al.
Veröffentlicht: (2024)
Risk-sensitive Markov Decision Process and Learning under General Utility Functions
von: Wu, Zhengqi, et al.
Veröffentlicht: (2023)
von: Wu, Zhengqi, et al.
Veröffentlicht: (2023)
An accelerated first-order regularized momentum descent ascent algorithm for stochastic nonconvex-concave minimax problems
von: Zhang, Huiling, et al.
Veröffentlicht: (2023)
von: Zhang, Huiling, et al.
Veröffentlicht: (2023)
Robust stochastic first order methods in heavy-tailed noise via medoid mini-batch gradient sampling
von: Vukovic, Manojlo, et al.
Veröffentlicht: (2026)
von: Vukovic, Manojlo, et al.
Veröffentlicht: (2026)
Explicit and Non-asymptotic Query Complexities of Rank-Based Zeroth-order Algorithm on Stochastic Smooth Functions
von: Ye, Haishan
Veröffentlicht: (2025)
von: Ye, Haishan
Veröffentlicht: (2025)
Markov decision processes: on the convergence of the Monte-Carlo first visit algorithm
von: Delattre, Sylvain, et al.
Veröffentlicht: (2025)
von: Delattre, Sylvain, et al.
Veröffentlicht: (2025)
Stochastic Hessian Fittings with Lie Groups
von: Li, Xi-Lin
Veröffentlicht: (2024)
von: Li, Xi-Lin
Veröffentlicht: (2024)
Optimal Asynchronous Stochastic Nonconvex Optimization under Heavy-Tailed Noise
von: Wu, Yidong, et al.
Veröffentlicht: (2026)
von: Wu, Yidong, et al.
Veröffentlicht: (2026)
Adam with model exponential moving average is effective for nonconvex optimization
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2024)
von: Ahn, Kwangjun, et al.
Veröffentlicht: (2024)
An Energy-Based Self-Adaptive Learning Rate for Stochastic Gradient Descent: Enhancing Unconstrained Optimization with VAV method
von: Zhang, Jiahao, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2024)
Soft decision trees for survival analysis
von: Consolo, Antonio, et al.
Veröffentlicht: (2025)
von: Consolo, Antonio, et al.
Veröffentlicht: (2025)
Asynchronous and Stochastic Distributed Resource Allocation
von: Li, Qiang, et al.
Veröffentlicht: (2025)
von: Li, Qiang, et al.
Veröffentlicht: (2025)
Improved Learning Rates for Stochastic Optimization
von: Li, Shaojie, et al.
Veröffentlicht: (2021)
von: Li, Shaojie, et al.
Veröffentlicht: (2021)
Stochastic-Constrained Stochastic Optimization with Markovian Data
von: Kim, Yeongjong, et al.
Veröffentlicht: (2023)
von: Kim, Yeongjong, et al.
Veröffentlicht: (2023)
Robust Out-of-Distribution Stochastic Optimization
von: Li, Xianyu, et al.
Veröffentlicht: (2026)
von: Li, Xianyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Auto-conditioned primal-dual hybrid gradient method and alternating direction method of multipliers
von: Lan, Guanghui, et al.
Veröffentlicht: (2024) -
Logarithmic regret bounds for continuous-time average-reward Markov decision processes
von: Gao, Xuefeng, et al.
Veröffentlicht: (2022) -
A simple uniformly optimal method without line search for convex optimization
von: Li, Tianjiao, et al.
Veröffentlicht: (2023) -
Projected gradient methods for nonconvex and stochastic smooth optimization: new complexities and auto-conditioned stepsizes
von: Lan, Guanghui, et al.
Veröffentlicht: (2024) -
Accelerated stochastic approximation with state-dependent noise
von: Ilandarideva, Sasila, et al.
Veröffentlicht: (2023)