Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
Fuente:
arXiv
Saved in:
| Main Authors: | Bai, Yitao, Doan, Thinh T., Romberg, Justin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Non-Expansive Mappings in Two-Time-Scale Stochastic Approximation: Finite-Time Analysis
by: Chandak, Siddharth
Published: (2025)
by: Chandak, Siddharth
Published: (2025)
Sample Complexity of Policy Gradient for Log-Growth Control
by: Pan, Qiuhua, et al.
Published: (2026)
by: Pan, Qiuhua, et al.
Published: (2026)
Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
by: Sakha, Masoud S., et al.
Published: (2026)
by: Sakha, Masoud S., et al.
Published: (2026)
Optimistic Training and Convergence of Q-Learning -- Extended Version
by: Mehta, Prashant, et al.
Published: (2026)
by: Mehta, Prashant, et al.
Published: (2026)
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty
by: Jia, Yanwei
Published: (2024)
by: Jia, Yanwei
Published: (2024)
On the Rate of Gaussian Approximation for Linear Regression Problems
by: Khusainov, Marat, et al.
Published: (2025)
by: Khusainov, Marat, et al.
Published: (2025)
A Note on Stability in Asynchronous Stochastic Approximation without Communication Delays
by: Yu, Huizhen, et al.
Published: (2023)
by: Yu, Huizhen, et al.
Published: (2023)
Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent
by: Sheshukova, Marina, et al.
Published: (2025)
by: Sheshukova, Marina, et al.
Published: (2025)
Asynchronous Stochastic Approximation with Applications to Average-Reward Reinforcement Learning
by: Yu, Huizhen, et al.
Published: (2024)
by: Yu, Huizhen, et al.
Published: (2024)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
by: Zheng, Yaowei, et al.
Published: (2026)
by: Zheng, Yaowei, et al.
Published: (2026)
A Continuous-Time Ensemble Kalman-Bucy Smoother for Causal Inference and Model Discovery
by: Jiang, Zhang, et al.
Published: (2026)
by: Jiang, Zhang, et al.
Published: (2026)
Stochastic Control with Signatures
by: Bank, P., et al.
Published: (2024)
by: Bank, P., et al.
Published: (2024)
Average-reward reinforcement learning in semi-Markov decision processes via relative value iteration
by: Yu, Huizhen, et al.
Published: (2025)
by: Yu, Huizhen, et al.
Published: (2025)
Policy Gradient for Continuous-Time Mean-Field Control
by: Bayraktar, Erhan, et al.
Published: (2026)
by: Bayraktar, Erhan, et al.
Published: (2026)
Revisiting Stochastic Approximation and Stochastic Gradient Descent
by: Karandikar, Rajeeva Laxman, et al.
Published: (2025)
by: Karandikar, Rajeeva Laxman, et al.
Published: (2025)
Reinforcement Learning, Optimal Control, and Bayesian Filtering in Data Assimilation
by: Hammoud, Abed
Published: (2026)
by: Hammoud, Abed
Published: (2026)
Convergence of Batch Asynchronous Stochastic Approximation With Applications to Reinforcement Learning
by: Karandikar, Rajeeva L., et al.
Published: (2021)
by: Karandikar, Rajeeva L., et al.
Published: (2021)
The Score Kalman Filter
by: Iwasaki, Kaito, et al.
Published: (2026)
by: Iwasaki, Kaito, et al.
Published: (2026)
Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence
by: Feng, Qi, et al.
Published: (2025)
by: Feng, Qi, et al.
Published: (2025)
Statistical inference for Linear Stochastic Approximation with Markovian Noise
by: Samsonov, Sergey, et al.
Published: (2025)
by: Samsonov, Sergey, et al.
Published: (2025)
Learning Optimal Filters Using Variational Inference
by: Bach, Eviatar, et al.
Published: (2024)
by: Bach, Eviatar, et al.
Published: (2024)
Mean--Variance Portfolio Selection by Continuous-Time Reinforcement Learning: Algorithms, Regret Analysis, and Empirical Study
by: Huang, Yilie, et al.
Published: (2024)
by: Huang, Yilie, et al.
Published: (2024)
Improved Central Limit Theorem and Bootstrap Approximations for Linear Stochastic Approximation
by: Butyrin, Bogdan, et al.
Published: (2025)
by: Butyrin, Bogdan, et al.
Published: (2025)
Convergence Rates for Stochastic Approximation: Biased Noise with Unbounded Variance, and Applications
by: Karandikar, Rajeeva L., et al.
Published: (2023)
by: Karandikar, Rajeeva L., et al.
Published: (2023)
Accelerating Multi-Task Temporal Difference Learning under Low-Rank Representation
by: Bai, Yitao, et al.
Published: (2025)
by: Bai, Yitao, et al.
Published: (2025)
Extended Dynamic Programming Principle and Applications to Time-Inconsistent Control
by: Xu, Yuhong, et al.
Published: (2022)
by: Xu, Yuhong, et al.
Published: (2022)
A measure-valued HJB perspective on Bayesian optimal adaptive control
by: Cox, Alexander M. G., et al.
Published: (2025)
by: Cox, Alexander M. G., et al.
Published: (2025)
Flow Matching for Efficient and Scalable Data Assimilation
by: Transue, Taos, et al.
Published: (2025)
by: Transue, Taos, et al.
Published: (2025)
Gaussian Approximation and Multiplier Bootstrap for Polyak-Ruppert Averaged Linear Stochastic Approximation with Applications to TD Learning
by: Samsonov, Sergey, et al.
Published: (2024)
by: Samsonov, Sergey, et al.
Published: (2024)
Koopman Kalman Filter (KKF): An asymptotically optimal nonlinear filtering algorithm with error bounds and its application to parameter estimation
by: Olguín, Diego, et al.
Published: (2025)
by: Olguín, Diego, et al.
Published: (2025)
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2021)
by: Zeng, Sihan, et al.
Published: (2021)
Ensemble Kalman-Bucy filtering for nonlinear model predictive control
by: Reich, Sebastian
Published: (2025)
by: Reich, Sebastian
Published: (2025)
Causal Hamilton-Jacobi-Bellman Equations for Anticipative Stochastic Optimal Control
by: Bank, Peter, et al.
Published: (2025)
by: Bank, Peter, et al.
Published: (2025)
A Probabilistic Approach to Trajectory-Based Optimal Experimental Design
by: Attia, Ahmed
Published: (2026)
by: Attia, Ahmed
Published: (2026)
Robust Probability Hypothesis Density Filtering: Theory and Algorithms
by: Lei, Ming, et al.
Published: (2025)
by: Lei, Ming, et al.
Published: (2025)
A Policy Gradient Framework for Stochastic Optimal Control Problems with Global Convergence Guarantee
by: Zhou, Mo, et al.
Published: (2023)
by: Zhou, Mo, et al.
Published: (2023)
Robust Recurrence of Discrete-Time Infinite-Horizon Stochastic Optimal Control with Discounted Cost
by: Moldenhauer, Robert H., et al.
Published: (2025)
by: Moldenhauer, Robert H., et al.
Published: (2025)
Neural Actor-Critic Methods for Hamilton-Jacobi-Bellman PDEs: Asymptotic Analysis and Numerical Studies
by: Cohen, Samuel N., et al.
Published: (2025)
by: Cohen, Samuel N., et al.
Published: (2025)
Confidence intervals for causal effects in sequential decision making
by: Vovk, Vladimir, et al.
Published: (2026)
by: Vovk, Vladimir, et al.
Published: (2026)
Markovian Foundations for Quasi-Stochastic Approximation with Applications to Extremum Seeking Control
by: Lauand, Caio Kalil, et al.
Published: (2022)
by: Lauand, Caio Kalil, et al.
Published: (2022)
Similar Items
-
Non-Expansive Mappings in Two-Time-Scale Stochastic Approximation: Finite-Time Analysis
by: Chandak, Siddharth
Published: (2025) -
Sample Complexity of Policy Gradient for Log-Growth Control
by: Pan, Qiuhua, et al.
Published: (2026) -
Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
by: Sakha, Masoud S., et al.
Published: (2026) -
Optimistic Training and Convergence of Q-Learning -- Extended Version
by: Mehta, Prashant, et al.
Published: (2026) -
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty
by: Jia, Yanwei
Published: (2024)