Nonasymptotic CLT and Error Bounds for Two-Time-Scale Stochastic Approximation
Fuente:
arXiv
Saved in:
| Main Authors: | Kong, Seo Taek, Zeng, Sihan, Doan, Thinh T., Srikant, R. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Finite-Sample Wasserstein Error Bounds and Concentration Inequalities for Nonlinear Stochastic Approximation
by: Kong, Seo Taek, et al.
Published: (2026)
by: Kong, Seo Taek, et al.
Published: (2026)
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2021)
by: Zeng, Sihan, et al.
Published: (2021)
Spectral Clustering for Crowdsourcing with Inherently Distinct Task Types
by: Mandal, Saptarshi, et al.
Published: (2023)
by: Mandal, Saptarshi, et al.
Published: (2023)
Fast Nonlinear Two-Time-Scale Stochastic Approximation: Achieving $O(1/k)$ Finite-Sample Complexity
by: Doan, Thinh T.
Published: (2024)
by: Doan, Thinh T.
Published: (2024)
Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
by: Bai, Yitao, et al.
Published: (2026)
by: Bai, Yitao, et al.
Published: (2026)
Accelerated Multi-Time-Scale Stochastic Approximation: Optimal Complexity and Applications in Reinforcement Learning and Multi-Agent Games
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
Finite-Time Complexity of Online Primal-Dual Natural Actor-Critic Algorithm for Constrained Markov Decision Processes
by: Zeng, Sihan, et al.
Published: (2021)
by: Zeng, Sihan, et al.
Published: (2021)
Provably Convergent Primal-Dual DPO for Constrained LLM Alignment
by: Du, Yihan, et al.
Published: (2025)
by: Du, Yihan, et al.
Published: (2025)
Noise Schedule Design for Diffusion Models: An Optimal Control Perspective
by: Kong, Seo Taek, et al.
Published: (2026)
by: Kong, Seo Taek, et al.
Published: (2026)
Misspecified $Q$-Learning with Sparse Linear Function Approximation: Tight Bounds on Approximation Error
by: Du, Ally Yalei, et al.
Published: (2024)
by: Du, Ally Yalei, et al.
Published: (2024)
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
Scalable Policy-Based RL Algorithms for POMDPs
by: Anjarlekar, Ameya, et al.
Published: (2025)
by: Anjarlekar, Ameya, et al.
Published: (2025)
A Nonasymptotic Theory of Gain-Dependent Error Dynamics in Behavior Cloning
by: Seo, Junghoon
Published: (2026)
by: Seo, Junghoon
Published: (2026)
Position: Don't Use the CLT in LLM Evals With Fewer Than a Few Hundred Datapoints
by: Bowyer, Sam, et al.
Published: (2025)
by: Bowyer, Sam, et al.
Published: (2025)
TurboAttention: Efficient Attention Approximation For High Throughputs LLMs
by: Kang, Hao, et al.
Published: (2024)
by: Kang, Hao, et al.
Published: (2024)
Accelerating Multi-Task Temporal Difference Learning under Low-Rank Representation
by: Bai, Yitao, et al.
Published: (2025)
by: Bai, Yitao, et al.
Published: (2025)
Beyond the Frontier: Stochastic Backtracking for Efficient Test-Time Scaling
by: Tran, Dao, et al.
Published: (2026)
by: Tran, Dao, et al.
Published: (2026)
A Theoretical Analysis of Soft-Label vs Hard-Label Training in Neural Networks
by: Mandal, Saptarshi, et al.
Published: (2024)
by: Mandal, Saptarshi, et al.
Published: (2024)
On the Convergence of Modified Policy Iteration in Risk Sensitive Exponential Cost Markov Decision Processes
by: Murthy, Yashaswini, et al.
Published: (2023)
by: Murthy, Yashaswini, et al.
Published: (2023)
On the Error-Correcting Effects of Stochasticity in Discrete Diffusion
by: Yuan, William, et al.
Published: (2026)
by: Yuan, William, et al.
Published: (2026)
Swift Hydra: Self-Reinforcing Generative Framework for Anomaly Detection with Multiple Mamba Models
by: Do, Nguyen, et al.
Published: (2025)
by: Do, Nguyen, et al.
Published: (2025)
Bounding the Worst-class Error: A Boosting Approach
by: Saito, Yuya, et al.
Published: (2023)
by: Saito, Yuya, et al.
Published: (2023)
Change of Thought: Adaptive Test-Time Computation
by: Mathur, Mrinal, et al.
Published: (2025)
by: Mathur, Mrinal, et al.
Published: (2025)
MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs
by: Nguyen, Quang H., et al.
Published: (2024)
by: Nguyen, Quang H., et al.
Published: (2024)
Second Order Bounds for Contextual Bandits with Function Approximation
by: Pacchiano, Aldo
Published: (2024)
by: Pacchiano, Aldo
Published: (2024)
Path Regularization: A Near-Complete and Optimal Nonasymptotic Generalization Theory for Multilayer Neural Networks and Double Descent Phenomenon
by: Yu, Hao
Published: (2025)
by: Yu, Hao
Published: (2025)
JustDense: Just using Dense instead of Sequence Mixer for Time Series analysis
by: Park, TaekHyun, et al.
Published: (2025)
by: Park, TaekHyun, et al.
Published: (2025)
ST-MTM: Masked Time Series Modeling with Seasonal-Trend Decomposition for Time Series Forecasting
by: Seo, Hyunwoo, et al.
Published: (2025)
by: Seo, Hyunwoo, et al.
Published: (2025)
The ODE Method for Stochastic Approximation and Reinforcement Learning with Markovian Noise
by: Liu, Shuze Daniel, et al.
Published: (2024)
by: Liu, Shuze Daniel, et al.
Published: (2024)
Bounding-Box Inference for Error-Aware Model-Based Reinforcement Learning
by: Talvitie, Erin J., et al.
Published: (2024)
by: Talvitie, Erin J., et al.
Published: (2024)
Stochastic Deep Graph Clustering for Practical Group Formation
by: Park, Junhyung, et al.
Published: (2025)
by: Park, Junhyung, et al.
Published: (2025)
A Predictive Model Based on Transformer with Statistical Feature Embedding in Manufacturing Sensor Dataset
by: Lee, Gyeong Taek, et al.
Published: (2024)
by: Lee, Gyeong Taek, et al.
Published: (2024)
The Role of Inherent Bellman Error in Offline Reinforcement Learning with Linear Function Approximation
by: Golowich, Noah, et al.
Published: (2024)
by: Golowich, Noah, et al.
Published: (2024)
Rethinking Langevin Thompson Sampling from A Stochastic Approximation Perspective
by: Wang, Weixin, et al.
Published: (2025)
by: Wang, Weixin, et al.
Published: (2025)
Joint Optimal Transport and Embedding for Network Alignment
by: Yu, Qi, et al.
Published: (2025)
by: Yu, Qi, et al.
Published: (2025)
Predictive AI Can Support Human Learning while Preserving Error Diversity
by: He, Vivianna Fang, et al.
Published: (2025)
by: He, Vivianna Fang, et al.
Published: (2025)
Deep Reinforcement Learning and The Tale of Two Temporal Difference Errors
by: Rojas, Juan Sebastian, et al.
Published: (2026)
by: Rojas, Juan Sebastian, et al.
Published: (2026)
Runtime-Certified Bounded-Error Quantized Attention
by: Calver, Dean
Published: (2026)
by: Calver, Dean
Published: (2026)
Bridging Neural ODE and ResNet: A Formal Error Bound for Safety Verification
by: Sayed, Abdelrahman Sayed, et al.
Published: (2025)
by: Sayed, Abdelrahman Sayed, et al.
Published: (2025)
Similar Items
-
Finite-Sample Wasserstein Error Bounds and Concentration Inequalities for Nonlinear Stochastic Approximation
by: Kong, Seo Taek, et al.
Published: (2026) -
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024) -
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2021) -
Spectral Clustering for Crowdsourcing with Inherently Distinct Task Types
by: Mandal, Saptarshi, et al.
Published: (2023) -
Fast Nonlinear Two-Time-Scale Stochastic Approximation: Achieving $O(1/k)$ Finite-Sample Complexity
by: Doan, Thinh T.
Published: (2024)