Extensions of Robbins-Siegmund Theorem with Applications in Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xinyu, Xie, Zixuan, Zhang, Shangtong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Almost Sure Convergence Rates of Stochastic Approximation and Reinforcement Learning via a Poisson-Moreau Drift
by: Liu, Xinyu, et al.
Published: (2026)
by: Liu, Xinyu, et al.
Published: (2026)
A quantitative Robbins-Siegmund theorem
by: Neri, Morenikeji, et al.
Published: (2024)
by: Neri, Morenikeji, et al.
Published: (2024)
Almost Sure Convergence Rates and Concentration of Stochastic Approximation and Reinforcement Learning with Markovian Noise
by: Qian, Xiaochi, et al.
Published: (2024)
by: Qian, Xiaochi, et al.
Published: (2024)
Asymptotic and Finite Sample Analysis of Nonexpansive Stochastic Approximations with Markovian Noise
by: Blaser, Ethan, et al.
Published: (2024)
by: Blaser, Ethan, et al.
Published: (2024)
Central Limit Theorems for Asynchronous Averaged Q-Learning
by: Liu, Xingtu
Published: (2025)
by: Liu, Xingtu
Published: (2025)
Deep Generative Demand Learning for Newsvendor and Pricing
by: Gong, Shijin, et al.
Published: (2024)
by: Gong, Shijin, et al.
Published: (2024)
Mitigating Covariate Shift in Misspecified Regression with Applications to Reinforcement Learning
by: Amortila, Philip, et al.
Published: (2024)
by: Amortila, Philip, et al.
Published: (2024)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
by: Srikant, R.
Published: (2024)
by: Srikant, R.
Published: (2024)
Risk-Sensitive Q-Learning in Continuous Time with Application to Dynamic Portfolio Selection
by: Xie, Chuhan
Published: (2025)
by: Xie, Chuhan
Published: (2025)
Multi-Agent Reinforcement Learning for Joint Police Patrol and Dispatch
by: Repasky, Matthew, et al.
Published: (2024)
by: Repasky, Matthew, et al.
Published: (2024)
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
Reinforcement Learning for a Discrete-Time Linear-Quadratic Control Problem with an Application
by: Li, Lucky
Published: (2024)
by: Li, Lucky
Published: (2024)
Reinforcement Learning for Intensity Control: An Application to Choice-Based Network Revenue Management
by: Meng, Huiling, et al.
Published: (2024)
by: Meng, Huiling, et al.
Published: (2024)
Central Limit Theorem for Two-Timescale Stochastic Approximation with Markovian Noise: Theory and Applications
by: Hu, Jie, et al.
Published: (2024)
by: Hu, Jie, et al.
Published: (2024)
Grower-in-the-Loop Interactive Reinforcement Learning for Greenhouse Climate Control
by: Xiao, Maxiu, et al.
Published: (2025)
by: Xiao, Maxiu, et al.
Published: (2025)
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2021)
by: Zeng, Sihan, et al.
Published: (2021)
Offline-Online Reinforcement Learning for Linear Mixture MDPs
by: Zhang, Zhongjun, et al.
Published: (2026)
by: Zhang, Zhongjun, et al.
Published: (2026)
A Hessian-Free Actor-Critic Algorithm for Bi-Level Reinforcement Learning with Applications to LLM Fine-Tuning
by: Zeng, Sihan, et al.
Published: (2026)
by: Zeng, Sihan, et al.
Published: (2026)
Automated Proof of Polynomial Inequalities via Reinforcement Learning
by: Liu, Banglong, et al.
Published: (2025)
by: Liu, Banglong, et al.
Published: (2025)
Reinforcement Learning for Jump-Diffusions, with Financial Applications
by: Gao, Xuefeng, et al.
Published: (2024)
by: Gao, Xuefeng, et al.
Published: (2024)
Central Limit Theorem for Two-Time-Scale Approximate Distributionally Robust RL
by: Wang, Shengbo, et al.
Published: (2026)
by: Wang, Shengbo, et al.
Published: (2026)
Offline Reinforcement Learning via Linear-Programming with Error-Bound Induced Constraints
by: Ozdaglar, Asuman, et al.
Published: (2022)
by: Ozdaglar, Asuman, et al.
Published: (2022)
Fast Policy Learning for Linear Quadratic Control with Entropy Regularization
by: Guo, Xin, et al.
Published: (2023)
by: Guo, Xin, et al.
Published: (2023)
Optimism as Risk-Seeking in Multi-Agent Reinforcement Learning
by: Zhang, Runyu, et al.
Published: (2025)
by: Zhang, Runyu, et al.
Published: (2025)
Traversing Pareto Optimal Policies: Provably Efficient Multi-Objective Reinforcement Learning
by: Qiu, Shuang, et al.
Published: (2024)
by: Qiu, Shuang, et al.
Published: (2024)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Resilient Constrained Reinforcement Learning
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Single- vs. Dual-Policy Reinforcement Learning for Dynamic Bike Rebalancing
by: Liang, Jiaqi, et al.
Published: (2024)
by: Liang, Jiaqi, et al.
Published: (2024)
Sampling-based Safe Reinforcement Learning for Nonlinear Dynamical Systems
by: Suttle, Wesley A., et al.
Published: (2024)
by: Suttle, Wesley A., et al.
Published: (2024)
Multinoulli Extension: A Lossless Continuous Relaxation for Partition-Constrained Subset Selection
by: Zhang, Qixin, et al.
Published: (2026)
by: Zhang, Qixin, et al.
Published: (2026)
Convergence and Emergence of In-Context Reinforcement Learning with Chain of Thought
by: Xie, Zixuan, et al.
Published: (2026)
by: Xie, Zixuan, et al.
Published: (2026)
Residuals-based Offline Reinforcement Learning
by: Zhu, Qing, et al.
Published: (2026)
by: Zhu, Qing, et al.
Published: (2026)
A Pontryagin Perspective on Reinforcement Learning
by: Eberhard, Onno, et al.
Published: (2024)
by: Eberhard, Onno, et al.
Published: (2024)
Foundations of Multivariate Distributional Reinforcement Learning
by: Wiltzer, Harley, et al.
Published: (2024)
by: Wiltzer, Harley, et al.
Published: (2024)
Operator World Models for Reinforcement Learning
by: Novelli, Pietro, et al.
Published: (2024)
by: Novelli, Pietro, et al.
Published: (2024)
KOALA++: Efficient Kalman-Based Optimization with Gradient-Covariance Products
by: Xia, Zixuan, et al.
Published: (2025)
by: Xia, Zixuan, et al.
Published: (2025)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
On the Foundation of Distributionally Robust Reinforcement Learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Tail Distribution of Regret in Optimistic Reinforcement Learning
by: Khodadadian, Sajad, et al.
Published: (2025)
by: Khodadadian, Sajad, et al.
Published: (2025)
Structured Reinforcement Learning for Combinatorial Decision-Making
by: Hoppe, Heiko, et al.
Published: (2025)
by: Hoppe, Heiko, et al.
Published: (2025)
Similar Items
-
Almost Sure Convergence Rates of Stochastic Approximation and Reinforcement Learning via a Poisson-Moreau Drift
by: Liu, Xinyu, et al.
Published: (2026) -
A quantitative Robbins-Siegmund theorem
by: Neri, Morenikeji, et al.
Published: (2024) -
Almost Sure Convergence Rates and Concentration of Stochastic Approximation and Reinforcement Learning with Markovian Noise
by: Qian, Xiaochi, et al.
Published: (2024) -
Asymptotic and Finite Sample Analysis of Nonexpansive Stochastic Approximations with Markovian Noise
by: Blaser, Ethan, et al.
Published: (2024) -
Central Limit Theorems for Asynchronous Averaged Q-Learning
by: Liu, Xingtu
Published: (2025)