Robust Q-Learning under Corrupted Rewards
Fuente:
arXiv
Saved in:
| Main Authors: | Maity, Sreejeet, Mitra, Aritra |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
by: Maity, Sreejeet, et al.
Published: (2025)
by: Maity, Sreejeet, et al.
Published: (2025)
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
by: Maity, Sreejeet, et al.
Published: (2025)
by: Maity, Sreejeet, et al.
Published: (2025)
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
by: Mitra, Aritra
Published: (2024)
by: Mitra, Aritra
Published: (2024)
Outlier-Robust Linear System Identification Under Heavy-tailed Noise
by: Kanakeri, Vinay, et al.
Published: (2024)
by: Kanakeri, Vinay, et al.
Published: (2024)
Boosting-Enabled Robust System Identification of Partially Observed LTI Systems Under Heavy-Tailed Noise
by: Kanakeri, Vinay, et al.
Published: (2025)
by: Kanakeri, Vinay, et al.
Published: (2025)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
by: Ye, Lintao, et al.
Published: (2024)
by: Ye, Lintao, et al.
Published: (2024)
Towards Fast Rates for Federated and Multi-Task Reinforcement Learning
by: Zhu, Feng, et al.
Published: (2024)
by: Zhu, Feng, et al.
Published: (2024)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
by: Zhu, Feng, et al.
Published: (2025)
by: Zhu, Feng, et al.
Published: (2025)
A Short and Unified Convergence Analysis of the SAG, SAGA, and IAG Algorithms
by: Zhu, Feng, et al.
Published: (2026)
by: Zhu, Feng, et al.
Published: (2026)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
by: Mitra, Aritra, et al.
Published: (2023)
by: Mitra, Aritra, et al.
Published: (2023)
InterQ: A DQN Framework for Optimal Intermittent Control
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
by: Kanakeri, Vinay, et al.
Published: (2025)
by: Kanakeri, Vinay, et al.
Published: (2025)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
by: Manenti, Massimiliano, et al.
Published: (2025)
by: Manenti, Massimiliano, et al.
Published: (2025)
The Silence that Speaks: Neural Estimation via Communication Gaps
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
by: Adibi, Arman, et al.
Published: (2024)
by: Adibi, Arman, et al.
Published: (2024)
Robust Neural IDA-PBC: passivity-based stabilization under approximations
by: Sanchez-Escalonilla, Santiago, et al.
Published: (2024)
by: Sanchez-Escalonilla, Santiago, et al.
Published: (2024)
Optimistic Online LQR via Intrinsic Rewards
by: Bartos, Marcell, et al.
Published: (2026)
by: Bartos, Marcell, et al.
Published: (2026)
On the Foundation of Distributionally Robust Reinforcement Learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
by: Boone, Victor, et al.
Published: (2024)
by: Boone, Victor, et al.
Published: (2024)
Federated Temporal Difference Learning with Linear Function Approximation under Environmental Heterogeneity
by: Wang, Han, et al.
Published: (2023)
by: Wang, Han, et al.
Published: (2023)
Variance-Reduced Cascade Q-learning: Algorithms and Sample Complexity
by: Boveiri, Mohammad, et al.
Published: (2024)
by: Boveiri, Mohammad, et al.
Published: (2024)
Robust Online Learning over Networks
by: Bastianello, Nicola, et al.
Published: (2023)
by: Bastianello, Nicola, et al.
Published: (2023)
Hybrid Energy-Aware Reward Shaping: A Unified Lightweight Physics-Guided Methodology for Policy Optimization
by: Liao, Qijun, et al.
Published: (2026)
by: Liao, Qijun, et al.
Published: (2026)
DASA: Delay-Adaptive Multi-Agent Stochastic Approximation
by: Fabbro, Nicolò Dal, et al.
Published: (2024)
by: Fabbro, Nicolò Dal, et al.
Published: (2024)
Convex Chance-Constrained Stochastic Control under Uncertain Specifications with Application to Learning-Based Hybrid Powertrain Control
by: Kato, Teruki, et al.
Published: (2026)
by: Kato, Teruki, et al.
Published: (2026)
Distributionally Robust Policy and Lyapunov-Certificate Learning
by: Long, Kehan, et al.
Published: (2024)
by: Long, Kehan, et al.
Published: (2024)
On Model Protection in Federated Learning against Eavesdropping Attacks
by: Maity, Dipankar, et al.
Published: (2025)
by: Maity, Dipankar, et al.
Published: (2025)
Adversarially Robust Multitask Adaptive Control
by: Fallah, Kasra, et al.
Published: (2025)
by: Fallah, Kasra, et al.
Published: (2025)
Distributed Thompson sampling under constrained communication
by: Zerefa, Saba, et al.
Published: (2024)
by: Zerefa, Saba, et al.
Published: (2024)
Probabilistic Safety Guarantee for Stochastic Control Systems Using Average Reward MDPs
by: Omidi, Saber, et al.
Published: (2025)
by: Omidi, Saber, et al.
Published: (2025)
Certified Robust Invariant Polytope Training in Neural Controlled ODEs
by: Harapanahalli, Akash, et al.
Published: (2024)
by: Harapanahalli, Akash, et al.
Published: (2024)
Online Control of Linear Systems under Unbounded Noise
by: Ito, Kaito, et al.
Published: (2024)
by: Ito, Kaito, et al.
Published: (2024)
Enhancing Convergence of Decentralized Gradient Tracking under the KL Property
by: Chen, Xiaokai, et al.
Published: (2024)
by: Chen, Xiaokai, et al.
Published: (2024)
Koopman Data-Driven Predictive Control with Robust Stability and Recursive Feasibility Guarantees
by: de Jong, Thomas, et al.
Published: (2024)
by: de Jong, Thomas, et al.
Published: (2024)
Adversarially and Distributionally Robust Virtual Energy Storage Systems via the Scenario Approach
by: Pantazis, Georgios, et al.
Published: (2025)
by: Pantazis, Georgios, et al.
Published: (2025)
Distributionally Robust Regret Optimal Control Under Moment-Based Ambiguity Sets
by: Taha, Feras Al, et al.
Published: (2025)
by: Taha, Feras Al, et al.
Published: (2025)
On Linear Convergence of PI Consensus Algorithm under the Restricted Secant Inequality
by: Chakrabarti, Kushal, et al.
Published: (2023)
by: Chakrabarti, Kushal, et al.
Published: (2023)
Robust stabilization of polytopic systems via fast and reliable neural network-based approximations
by: Fabiani, Filippo, et al.
Published: (2022)
by: Fabiani, Filippo, et al.
Published: (2022)
Wasserstein Distributionally Robust Nash Equilibrium Seeking with Heterogeneous Data: A Lagrangian Approach
by: Wang, Zifan, et al.
Published: (2025)
by: Wang, Zifan, et al.
Published: (2025)
Similar Items
-
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
by: Maity, Sreejeet, et al.
Published: (2025) -
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
by: Maity, Sreejeet, et al.
Published: (2025) -
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
by: Mitra, Aritra
Published: (2024) -
Outlier-Robust Linear System Identification Under Heavy-tailed Noise
by: Kanakeri, Vinay, et al.
Published: (2024) -
Boosting-Enabled Robust System Identification of Partially Observed LTI Systems Under Heavy-Tailed Noise
by: Kanakeri, Vinay, et al.
Published: (2025)