Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Manenti, Massimiliano, Iannelli, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Online Convex Optimization and Integral Quadratic Constraints: An automated approach to regret analysis
by: Jakob, Fabian, et al.
Published: (2025)
by: Jakob, Fabian, et al.
Published: (2025)
End-to-end guarantees for indirect data-driven control of bilinear systems with finite stochastic data
by: Chatzikiriakos, Nicolas, et al.
Published: (2024)
by: Chatzikiriakos, Nicolas, et al.
Published: (2024)
Offline Hierarchical Reinforcement Learning via Inverse Optimization
by: Schmidt, Carolin, et al.
Published: (2024)
by: Schmidt, Carolin, et al.
Published: (2024)
Hierarchical Reinforcement Learning Framework for Stochastic Spaceflight Campaign Design
by: Takubo, Yuji, et al.
Published: (2021)
by: Takubo, Yuji, et al.
Published: (2021)
Adaptive control mechanisms in gradient descent algorithms
by: Iannelli, Andrea
Published: (2025)
by: Iannelli, Andrea
Published: (2025)
Variance-Reduced Cascade Q-learning: Algorithms and Sample Complexity
by: Boveiri, Mohammad, et al.
Published: (2024)
by: Boveiri, Mohammad, et al.
Published: (2024)
Robust Q-Learning under Corrupted Rewards
by: Maity, Sreejeet, et al.
Published: (2024)
by: Maity, Sreejeet, et al.
Published: (2024)
Accelerated ADMM: Automated Parameter Tuning and Improved Linear Convergence
by: Tavakoli, Meisam, et al.
Published: (2025)
by: Tavakoli, Meisam, et al.
Published: (2025)
Analysis of Multiscale Reinforcement Q-Learning Algorithms for Mean Field Control Games
by: Angiuli, Andrea, et al.
Published: (2024)
by: Angiuli, Andrea, et al.
Published: (2024)
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
by: Maity, Sreejeet, et al.
Published: (2025)
by: Maity, Sreejeet, et al.
Published: (2025)
Resilient Constrained Reinforcement Learning
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
GreenLight-Gym: Reinforcement learning benchmark environment for control of greenhouse production systems
by: van Laatum, Bart, et al.
Published: (2024)
by: van Laatum, Bart, et al.
Published: (2024)
Non-Parametric Learning of Stochastic Differential Equations with Non-asymptotic Fast Rates of Convergence
by: Bonalli, Riccardo, et al.
Published: (2023)
by: Bonalli, Riccardo, et al.
Published: (2023)
On the Foundation of Distributionally Robust Reinforcement Learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Offline Reinforcement Learning via Inverse Optimization
by: Dimanidis, Ioannis, et al.
Published: (2025)
by: Dimanidis, Ioannis, et al.
Published: (2025)
Optimism as Risk-Seeking in Multi-Agent Reinforcement Learning
by: Zhang, Runyu, et al.
Published: (2025)
by: Zhang, Runyu, et al.
Published: (2025)
Operator Models for Continuous-Time Offline Reinforcement Learning
by: Hoischen, Nicolas, et al.
Published: (2025)
by: Hoischen, Nicolas, et al.
Published: (2025)
Towards Fast Rates for Federated and Multi-Task Reinforcement Learning
by: Zhu, Feng, et al.
Published: (2024)
by: Zhu, Feng, et al.
Published: (2024)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Hierarchical Deep Reinforcement Learning Framework for Multi-Year Asset Management Under Budget Constraints
by: Fard, Amir, et al.
Published: (2025)
by: Fard, Amir, et al.
Published: (2025)
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
by: Muehlebach, Michael, et al.
Published: (2025)
by: Muehlebach, Michael, et al.
Published: (2025)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
Online Reinforcement Learning in Markov Decision Process Using Linear Programming
by: Leon, Vincent, et al.
Published: (2023)
by: Leon, Vincent, et al.
Published: (2023)
Enhancing Convergence of Decentralized Gradient Tracking under the KL Property
by: Chen, Xiaokai, et al.
Published: (2024)
by: Chen, Xiaokai, et al.
Published: (2024)
InterQ: A DQN Framework for Optimal Intermittent Control
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
A Short and Unified Convergence Analysis of the SAG, SAGA, and IAG Algorithms
by: Zhu, Feng, et al.
Published: (2026)
by: Zhu, Feng, et al.
Published: (2026)
On Linear Convergence of PI Consensus Algorithm under the Restricted Secant Inequality
by: Chakrabarti, Kushal, et al.
Published: (2023)
by: Chakrabarti, Kushal, et al.
Published: (2023)
Hidden Convexity in Active Learning: A Convexified Online Input Design for ARX Systems
by: Chatzikiriakos, Nicolas, et al.
Published: (2025)
by: Chatzikiriakos, Nicolas, et al.
Published: (2025)
Reinforcement Learning-based Control via Y-wise Affine Neural Networks (YANNs)
by: Braniff, Austin, et al.
Published: (2025)
by: Braniff, Austin, et al.
Published: (2025)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
by: Srikant, R.
Published: (2024)
by: Srikant, R.
Published: (2024)
Reliably-stabilizing piecewise-affine neural network controllers
by: Fabiani, Filippo, et al.
Published: (2021)
by: Fabiani, Filippo, et al.
Published: (2021)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Learning to accelerate Krasnosel'skii-Mann fixed-point iterations with guarantees
by: Martin, Andrea, et al.
Published: (2026)
by: Martin, Andrea, et al.
Published: (2026)
Temporal-Aware Deep Reinforcement Learning for Energy Storage Bidding in Energy and Contingency Reserve Markets
by: Li, Jinhao, et al.
Published: (2024)
by: Li, Jinhao, et al.
Published: (2024)
Semi-on-Demand Transit Feeders with Shared Autonomous Vehicles and Reinforcement-Learning-Based Zonal Dispatching Control
by: Ng, Max T. M., et al.
Published: (2025)
by: Ng, Max T. M., et al.
Published: (2025)
Precise Insulin Delivery for Artificial Pancreas: A Reinforcement Learning Optimized Adaptive Fuzzy Control Approach
by: Mameche, Omar, et al.
Published: (2025)
by: Mameche, Omar, et al.
Published: (2025)
Attentive Convolutional Deep Reinforcement Learning for Optimizing Solar-Storage Systems in Real-Time Electricity Markets
by: Li, Jinhao, et al.
Published: (2024)
by: Li, Jinhao, et al.
Published: (2024)
Randomized Transport Plans via Hierarchical Fully Probabilistic Design
by: Y., Sarah Boufelja, et al.
Published: (2024)
by: Y., Sarah Boufelja, et al.
Published: (2024)
Robust Neural IDA-PBC: passivity-based stabilization under approximations
by: Sanchez-Escalonilla, Santiago, et al.
Published: (2024)
by: Sanchez-Escalonilla, Santiago, et al.
Published: (2024)
SINDy-RL: Interpretable and Efficient Model-Based Reinforcement Learning
by: Zolman, Nicholas, et al.
Published: (2024)
by: Zolman, Nicholas, et al.
Published: (2024)
Similar Items
-
Online Convex Optimization and Integral Quadratic Constraints: An automated approach to regret analysis
by: Jakob, Fabian, et al.
Published: (2025) -
End-to-end guarantees for indirect data-driven control of bilinear systems with finite stochastic data
by: Chatzikiriakos, Nicolas, et al.
Published: (2024) -
Offline Hierarchical Reinforcement Learning via Inverse Optimization
by: Schmidt, Carolin, et al.
Published: (2024) -
Hierarchical Reinforcement Learning Framework for Stochastic Spaceflight Campaign Design
by: Takubo, Yuji, et al.
Published: (2021) -
Adaptive control mechanisms in gradient descent algorithms
by: Iannelli, Andrea
Published: (2025)