Finite-Time Accuracy of Temporal-Difference Learning Under Schur-Stable Recursions
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Donghwan, Kim, Do Wan |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Finite-Time Analysis of Simultaneous Double Q-learning
by: Na, Hyunjun, et al.
Published: (2024)
by: Na, Hyunjun, et al.
Published: (2024)
Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games
by: Jeong, Narim, et al.
Published: (2026)
by: Jeong, Narim, et al.
Published: (2026)
Analysis of Off-Policy Multi-Step TD-Learning with Linear Function Approximation
by: Lee, Donghwan
Published: (2024)
by: Lee, Donghwan
Published: (2024)
Lyapunov-Certified Direct Switching Theory for Q-Learning
by: Lee, Donghwan
Published: (2026)
by: Lee, Donghwan
Published: (2026)
Harnessing Membership Function Dynamics for Stability Analysis of T-S Fuzzy Systems
by: Lee, Donghwan, et al.
Published: (2024)
by: Lee, Donghwan, et al.
Published: (2024)
Deep Q-Learning with Gradient Target Tracking
by: Park, Bum Geun, et al.
Published: (2025)
by: Park, Bum Geun, et al.
Published: (2025)
Bellman Residual Minimization for Control: Geometry, Stationarity, and Convergence
by: Lee, Donghwan, et al.
Published: (2026)
by: Lee, Donghwan, et al.
Published: (2026)
Continuous-Time Distributed Dynamic Programming for Networked Multi-Agent Markov Decision Processes
by: Lee, Donghwan, et al.
Published: (2023)
by: Lee, Donghwan, et al.
Published: (2023)
Finite-Time Analysis of Temporal Difference Learning with Experience Replay
by: Lim, Han-Dong, et al.
Published: (2023)
by: Lim, Han-Dong, et al.
Published: (2023)
Benchmarking Recursive-Collapse Warning Claims Under Matched False-Positive Control
by: Mullett, David
Published: (2026)
by: Mullett, David
Published: (2026)
Scalable Learning of Intrusion Responses through Recursive Decomposition
by: Hammar, Kim, et al.
Published: (2023)
by: Hammar, Kim, et al.
Published: (2023)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Accelerating Multi-Task Temporal Difference Learning under Low-Rank Representation
by: Bai, Yitao, et al.
Published: (2025)
by: Bai, Yitao, et al.
Published: (2025)
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
by: Mitra, Aritra
Published: (2024)
by: Mitra, Aritra
Published: (2024)
Recursively Feasible Probabilistic Safe Online Learning with Control Barrier Functions
by: Castañeda, Fernando, et al.
Published: (2022)
by: Castañeda, Fernando, et al.
Published: (2022)
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
by: Maity, Sreejeet, et al.
Published: (2025)
by: Maity, Sreejeet, et al.
Published: (2025)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
by: Mitra, Aritra, et al.
Published: (2023)
by: Mitra, Aritra, et al.
Published: (2023)
PID Accelerated Temporal Difference Algorithms
by: Bedaywi, Mark, et al.
Published: (2024)
by: Bedaywi, Mark, et al.
Published: (2024)
Recursive Gaussian Process State Space Model
by: Zheng, Tengjie, et al.
Published: (2024)
by: Zheng, Tengjie, et al.
Published: (2024)
ORFit: One-Pass Learning via Bridging Orthogonal Gradient Descent and Recursive Least-Squares
by: Min, Youngjae, et al.
Published: (2022)
by: Min, Youngjae, et al.
Published: (2022)
Tight Finite Time Bounds of Two-Time-Scale Linear Stochastic Approximation with Markovian Noise
by: Haque, Shaan Ul, et al.
Published: (2023)
by: Haque, Shaan Ul, et al.
Published: (2023)
$O(1/k)$ Finite-Time Bound for Non-Linear Two-Time-Scale Stochastic Approximation
by: Chandak, Siddharth
Published: (2025)
by: Chandak, Siddharth
Published: (2025)
Finite-Time Bounds for Two-Time-Scale Stochastic Approximation with Arbitrary Norm Contractions and Markovian Noise
by: Chandak, Siddharth, et al.
Published: (2025)
by: Chandak, Siddharth, et al.
Published: (2025)
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
by: Adams, Katherine B., et al.
Published: (2025)
by: Adams, Katherine B., et al.
Published: (2025)
Heavy-Tailed and Long-Range Dependent Noise in Stochastic Approximation: A Finite-Time Analysis
by: Chandak, Siddharth, et al.
Published: (2026)
by: Chandak, Siddharth, et al.
Published: (2026)
Koopman Data-Driven Predictive Control with Robust Stability and Recursive Feasibility Guarantees
by: de Jong, Thomas, et al.
Published: (2024)
by: de Jong, Thomas, et al.
Published: (2024)
Recursively Feasible Shrinking-Horizon MPC in Dynamic Environments with Conformal Prediction Guarantees
by: Stamouli, Charis, et al.
Published: (2024)
by: Stamouli, Charis, et al.
Published: (2024)
Finite Sample Analysis of Tensor Decomposition for Learning Mixtures of Linear Systems
by: Rui, Maryann, et al.
Published: (2024)
by: Rui, Maryann, et al.
Published: (2024)
Learning to Route Electric Trucks Under Operational Uncertainty
by: Orfanoudakis, Stavros, et al.
Published: (2026)
by: Orfanoudakis, Stavros, et al.
Published: (2026)
Approximate Gradient Coding for Distributed Learning with Heterogeneous Stragglers
by: Song, Heekang, et al.
Published: (2025)
by: Song, Heekang, et al.
Published: (2025)
Learning to Boost the Performance of Stable Nonlinear Systems
by: Furieri, Luca, et al.
Published: (2024)
by: Furieri, Luca, et al.
Published: (2024)
Learning Stable and Passive Neural Differential Equations
by: Cheng, Jing, et al.
Published: (2024)
by: Cheng, Jing, et al.
Published: (2024)
Recursive Inference for Heterogeneous Multi-Output GP State-Space Models with Arbitrary Moment Matching
by: Zheng, Tengjie, et al.
Published: (2025)
by: Zheng, Tengjie, et al.
Published: (2025)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
by: Adibi, Arman, et al.
Published: (2024)
by: Adibi, Arman, et al.
Published: (2024)
Learning Linearized Models from Nonlinear Systems under Initialization Constraints with Finite Data
by: Xin, Lei, et al.
Published: (2025)
by: Xin, Lei, et al.
Published: (2025)
Route Recommendations for Traffic Management Under Learned Partial Driver Compliance
by: Bang, Heeseung, et al.
Published: (2025)
by: Bang, Heeseung, et al.
Published: (2025)
Stable Linear Subspace Identification: A Machine Learning Approach
by: Di Natale, Loris, et al.
Published: (2023)
by: Di Natale, Loris, et al.
Published: (2023)
Finite Sample Frequency Domain Identification
by: Tsiamis, Anastasios, et al.
Published: (2024)
by: Tsiamis, Anastasios, et al.
Published: (2024)
Machine Learning for Scalable and Optimal Load Shedding Under Power System Contingency
by: Zhou, Yuqi, et al.
Published: (2024)
by: Zhou, Yuqi, et al.
Published: (2024)
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
by: Lee, Donghwan
Published: (2023)
by: Lee, Donghwan
Published: (2023)
Similar Items
-
Finite-Time Analysis of Simultaneous Double Q-learning
by: Na, Hyunjun, et al.
Published: (2024) -
Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games
by: Jeong, Narim, et al.
Published: (2026) -
Analysis of Off-Policy Multi-Step TD-Learning with Linear Function Approximation
by: Lee, Donghwan
Published: (2024) -
Lyapunov-Certified Direct Switching Theory for Q-Learning
by: Lee, Donghwan
Published: (2026) -
Harnessing Membership Function Dynamics for Stability Analysis of T-S Fuzzy Systems
by: Lee, Donghwan, et al.
Published: (2024)