Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Maity, Sreejeet, Mitra, Aritra |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
von: Maity, Sreejeet, et al.
Veröffentlicht: (2025)
von: Maity, Sreejeet, et al.
Veröffentlicht: (2025)
Robust Q-Learning under Corrupted Rewards
von: Maity, Sreejeet, et al.
Veröffentlicht: (2024)
von: Maity, Sreejeet, et al.
Veröffentlicht: (2024)
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
von: Mitra, Aritra
Veröffentlicht: (2024)
von: Mitra, Aritra
Veröffentlicht: (2024)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
von: Zhu, Feng, et al.
Veröffentlicht: (2025)
von: Zhu, Feng, et al.
Veröffentlicht: (2025)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
von: Zhang, Chenyu, et al.
Veröffentlicht: (2024)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2024)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
von: Adibi, Arman, et al.
Veröffentlicht: (2024)
von: Adibi, Arman, et al.
Veröffentlicht: (2024)
Outlier-Robust Linear System Identification Under Heavy-tailed Noise
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2024)
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2024)
Towards Fast Rates for Federated and Multi-Task Reinforcement Learning
von: Zhu, Feng, et al.
Veröffentlicht: (2024)
von: Zhu, Feng, et al.
Veröffentlicht: (2024)
Boosting-Enabled Robust System Identification of Partially Observed LTI Systems Under Heavy-Tailed Noise
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2025)
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2025)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
von: Srikant, R.
Veröffentlicht: (2024)
von: Srikant, R.
Veröffentlicht: (2024)
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2025)
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2025)
Tight Finite Time Bounds of Two-Time-Scale Linear Stochastic Approximation with Markovian Noise
von: Haque, Shaan Ul, et al.
Veröffentlicht: (2023)
von: Haque, Shaan Ul, et al.
Veröffentlicht: (2023)
Finite-Time Bounds for Two-Time-Scale Stochastic Approximation with Arbitrary Norm Contractions and Markovian Noise
von: Chandak, Siddharth, et al.
Veröffentlicht: (2025)
von: Chandak, Siddharth, et al.
Veröffentlicht: (2025)
A Short and Unified Convergence Analysis of the SAG, SAGA, and IAG Algorithms
von: Zhu, Feng, et al.
Veröffentlicht: (2026)
von: Zhu, Feng, et al.
Veröffentlicht: (2026)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
von: Mitra, Aritra, et al.
Veröffentlicht: (2023)
von: Mitra, Aritra, et al.
Veröffentlicht: (2023)
Adversarially Robust Multitask Adaptive Control
von: Fallah, Kasra, et al.
Veröffentlicht: (2025)
von: Fallah, Kasra, et al.
Veröffentlicht: (2025)
End-to-End Learning Framework for Solving Non-Markovian Optimal Control
von: Zhang, Xiaole, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaole, et al.
Veröffentlicht: (2025)
Adversarially and Distributionally Robust Virtual Energy Storage Systems via the Scenario Approach
von: Pantazis, Georgios, et al.
Veröffentlicht: (2025)
von: Pantazis, Georgios, et al.
Veröffentlicht: (2025)
Data-Driven Adversarial Online Control for Unknown Linear Systems
von: Liu, Zishun, et al.
Veröffentlicht: (2023)
von: Liu, Zishun, et al.
Veröffentlicht: (2023)
$O(1/k)$ Finite-Time Bound for Non-Linear Two-Time-Scale Stochastic Approximation
von: Chandak, Siddharth
Veröffentlicht: (2025)
von: Chandak, Siddharth
Veröffentlicht: (2025)
The Silence that Speaks: Neural Estimation via Communication Gaps
von: Aggarwal, Shubham, et al.
Veröffentlicht: (2025)
von: Aggarwal, Shubham, et al.
Veröffentlicht: (2025)
InterQ: A DQN Framework for Optimal Intermittent Control
von: Aggarwal, Shubham, et al.
Veröffentlicht: (2025)
von: Aggarwal, Shubham, et al.
Veröffentlicht: (2025)
Bayesian Ambiguity Contraction-based Adaptive Robust Markov Decision Processes for Adversarial Surveillance Missions
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem
von: Haque, Shaan Ul, et al.
Veröffentlicht: (2024)
von: Haque, Shaan Ul, et al.
Veröffentlicht: (2024)
Data-driven Reachable Set Estimation with Tunable Adversarial and Wasserstein Distributional Guarantees
von: Pantazis, Georgios, et al.
Veröffentlicht: (2026)
von: Pantazis, Georgios, et al.
Veröffentlicht: (2026)
High-Probability Bounds for SGD under the Polyak-Lojasiewicz Condition with Markovian Noise
von: Kar, Avik, et al.
Veröffentlicht: (2026)
von: Kar, Avik, et al.
Veröffentlicht: (2026)
Heavy-Tailed and Long-Range Dependent Noise in Stochastic Approximation: A Finite-Time Analysis
von: Chandak, Siddharth, et al.
Veröffentlicht: (2026)
von: Chandak, Siddharth, et al.
Veröffentlicht: (2026)
A Finite-Time Analysis of TD Learning with Linear Function Approximation without Projections or Strong Convexity
von: Lee, Wei-Cheng, et al.
Veröffentlicht: (2025)
von: Lee, Wei-Cheng, et al.
Veröffentlicht: (2025)
On the Foundation of Distributionally Robust Reinforcement Learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Non-Parametric Learning of Stochastic Differential Equations with Non-asymptotic Fast Rates of Convergence
von: Bonalli, Riccardo, et al.
Veröffentlicht: (2023)
von: Bonalli, Riccardo, et al.
Veröffentlicht: (2023)
Finite Sample Frequency Domain Identification
von: Tsiamis, Anastasios, et al.
Veröffentlicht: (2024)
von: Tsiamis, Anastasios, et al.
Veröffentlicht: (2024)
Koopman Data-Driven Predictive Control with Robust Stability and Recursive Feasibility Guarantees
von: de Jong, Thomas, et al.
Veröffentlicht: (2024)
von: de Jong, Thomas, et al.
Veröffentlicht: (2024)
Wasserstein Distributionally Robust Nash Equilibrium Seeking with Heterogeneous Data: A Lagrangian Approach
von: Wang, Zifan, et al.
Veröffentlicht: (2025)
von: Wang, Zifan, et al.
Veröffentlicht: (2025)
Robust Online Learning over Networks
von: Bastianello, Nicola, et al.
Veröffentlicht: (2023)
von: Bastianello, Nicola, et al.
Veröffentlicht: (2023)
Operator Models for Continuous-Time Offline Reinforcement Learning
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2025)
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2025)
Finite Sample Identification of Partially Observed Bilinear Dynamical Systems
von: Sattar, Yahya, et al.
Veröffentlicht: (2025)
von: Sattar, Yahya, et al.
Veröffentlicht: (2025)
Finite-sample guarantees for data-driven forward-backward operator methods
von: Fabiani, Filippo, et al.
Veröffentlicht: (2025)
von: Fabiani, Filippo, et al.
Veröffentlicht: (2025)
Finite-Sample-Based Reachability for Safe Control with Gaussian Process Dynamics
von: Prajapat, Manish, et al.
Veröffentlicht: (2025)
von: Prajapat, Manish, et al.
Veröffentlicht: (2025)
DASA: Delay-Adaptive Multi-Agent Stochastic Approximation
von: Fabbro, Nicolò Dal, et al.
Veröffentlicht: (2024)
von: Fabbro, Nicolò Dal, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
von: Maity, Sreejeet, et al.
Veröffentlicht: (2025) -
Robust Q-Learning under Corrupted Rewards
von: Maity, Sreejeet, et al.
Veröffentlicht: (2024) -
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
von: Mitra, Aritra
Veröffentlicht: (2024) -
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
von: Zhu, Feng, et al.
Veröffentlicht: (2025) -
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
von: Ye, Lintao, et al.
Veröffentlicht: (2024)