Final Iteration Convergence Bound of Q-Learning: Switching System Approach
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Lee, Donghwna |
|---|---|
| Format: | Preprint |
| Publié: |
2022
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
par: Lee, Donghwan
Publié: (2023)
par: Lee, Donghwan
Publié: (2023)
Converse Lyapunov Results for Switched Systems with Lower and Upper Bounds on Switching Intervals
par: Della Rossa, Matteo
Publié: (2024)
par: Della Rossa, Matteo
Publié: (2024)
Lyapunov-Certified Direct Switching Theory for Q-Learning
par: Lee, Donghwan
Publié: (2026)
par: Lee, Donghwan
Publié: (2026)
Contouring Error Bounded Control for Biaxial Switched Linear Systems
par: Yuan, Meng, et autres
Publié: (2024)
par: Yuan, Meng, et autres
Publié: (2024)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
par: Manenti, Massimiliano, et autres
Publié: (2025)
par: Manenti, Massimiliano, et autres
Publié: (2025)
On Convergence of the Iteratively Preconditioned Gradient-Descent (IPG) Observer
par: Chakrabarti, Kushal, et autres
Publié: (2024)
par: Chakrabarti, Kushal, et autres
Publié: (2024)
A Path-Complete Approach for Optimal Control of Switched Systems
par: Ninite, Léa, et autres
Publié: (2026)
par: Ninite, Léa, et autres
Publié: (2026)
Mean Field Game and Control for Switching Hybrid Systems
par: C., Tejaswi K., et autres
Publié: (2024)
par: C., Tejaswi K., et autres
Publié: (2024)
A Robust Data-Driven Iterative Control Method for Linear Systems with Bounded Disturbances
par: Hu, Kaijian, et autres
Publié: (2024)
par: Hu, Kaijian, et autres
Publié: (2024)
Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games
par: Jeong, Narim, et autres
Publié: (2026)
par: Jeong, Narim, et autres
Publié: (2026)
Physics-Informed Neural Network Policy Iteration: Algorithms, Convergence, and Verification
par: Meng, Yiming, et autres
Publié: (2024)
par: Meng, Yiming, et autres
Publié: (2024)
Density-Driven Optimal Control: Convergence Guarantees for Stochastic LTI Multi-Agent Systems
par: Lee, Kooktae
Publié: (2026)
par: Lee, Kooktae
Publié: (2026)
Output-Feedback Stabilizing Policy Iteration for Convergence Assurance of Unknown Discrete-Time Systems with Unmeasurable States
par: Li, Dongdong, et autres
Publié: (2025)
par: Li, Dongdong, et autres
Publié: (2025)
Convergence and Robustness Bounds for Distributed Asynchronous Shortest-Path
par: Miller, Jared, et autres
Publié: (2025)
par: Miller, Jared, et autres
Publié: (2025)
Scalable Iterative Algorithm for Solving Optimal Transmission Switching with De-energization
par: Jeanson, Benoît, et autres
Publié: (2025)
par: Jeanson, Benoît, et autres
Publié: (2025)
A Hybrid Algorithm for Iterative Adaptation of Feedforward Controllers: an Application on Electromechanical Switches
par: Serrano-Seco, Eloy, et autres
Publié: (2024)
par: Serrano-Seco, Eloy, et autres
Publié: (2024)
Optimal Control of Switched Systems Governed by Logical Switching Dynamics
par: Zhang, Xiao, et autres
Publié: (2026)
par: Zhang, Xiao, et autres
Publié: (2026)
Adaptive Optimal Control of Linear Periodic Systems: An Off-Policy Value Iteration Approach
par: Pang, Bo, et autres
Publié: (2019)
par: Pang, Bo, et autres
Publié: (2019)
On the Convergence of an Opinion-Action Coevolution Model with Bounded Confidence
par: Song, Chen, et autres
Publié: (2026)
par: Song, Chen, et autres
Publié: (2026)
Iterative Learning Predictive Control for Constrained Uncertain Systems
par: Zuliani, Riccardo, et autres
Publié: (2025)
par: Zuliani, Riccardo, et autres
Publié: (2025)
Convergence in On-line Learning of Static and Dynamic Systems
par: Wigren, Torbjörn, et autres
Publié: (2025)
par: Wigren, Torbjörn, et autres
Publié: (2025)
Lyapunov Characterization for ISS of Impulsive Switched Systems
par: Ahmed, Saeed, et autres
Publié: (2024)
par: Ahmed, Saeed, et autres
Publié: (2024)
Fully Byzantine-Resilient Distributed Multi-Agent Q-Learning
par: Lee, Haejoon, et autres
Publié: (2026)
par: Lee, Haejoon, et autres
Publié: (2026)
Toward Value-oriented Renewable Energy Forecasting: An Iterative Learning Approach
par: Zhang, Yufan, et autres
Publié: (2023)
par: Zhang, Yufan, et autres
Publié: (2023)
Analysis of Discrete-Time Switched Linear Systems under Logic Dynamic Switchings
par: Zhang, Xiao, et autres
Publié: (2022)
par: Zhang, Xiao, et autres
Publié: (2022)
Bus Type Switching to Reduce Bound Violations in AC Power Flow
par: Van Boven, Anna, et autres
Publié: (2025)
par: Van Boven, Anna, et autres
Publié: (2025)
Robust MPC for Uncertain Linear Systems -- Combining Model Adaptation and Iterative Learning
par: Petrenz, Hannes, et autres
Publié: (2025)
par: Petrenz, Hannes, et autres
Publié: (2025)
A Switched Systems Approach to Image-Based Feature Tracking for Autonomous Satellite Inspection
par: Ogri, Tochukwu Elijah, et autres
Publié: (2025)
par: Ogri, Tochukwu Elijah, et autres
Publié: (2025)
Global Regulation of Feedforward Nonlinear Systems: A Logic-Based Switching Gain Approach
par: Fan, Debao, et autres
Publié: (2024)
par: Fan, Debao, et autres
Publié: (2024)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
par: Ding, Dongsheng, et autres
Publié: (2023)
par: Ding, Dongsheng, et autres
Publié: (2023)
Switching-Geometry Analysis of Deflated Q-Value Iteration
par: Lee, Donghwan
Publié: (2026)
par: Lee, Donghwan
Publié: (2026)
Data-Driven Unknown Input Reconstruction for MIMO Systems with Convergence Guarantees
par: Breukelman, Enno, et autres
Publié: (2026)
par: Breukelman, Enno, et autres
Publié: (2026)
Switched Linear Ensemble Systems and Structural Controllability
par: Yin, Haoyu, et autres
Publié: (2025)
par: Yin, Haoyu, et autres
Publié: (2025)
Reinforcement Learning-Based Controlled Switching Approach for Inrush Current Minimization in Power Transformers
par: Valdivielso, Jone Ugarte, et autres
Publié: (2025)
par: Valdivielso, Jone Ugarte, et autres
Publié: (2025)
Fitted Q-Iteration via Max-Plus-Linear Approximation
par: Liu, Y., et autres
Publié: (2024)
par: Liu, Y., et autres
Publié: (2024)
Deep Q-Learning with Gradient Target Tracking
par: Park, Bum Geun, et autres
Publié: (2025)
par: Park, Bum Geun, et autres
Publié: (2025)
A Fundamental Convergence Rate Bound for Gradient Based Online Optimization Algorithms with Exact Tracking
par: Wu, Alex Xinting, et autres
Publié: (2025)
par: Wu, Alex Xinting, et autres
Publié: (2025)
Deep Koopman Iterative Learning and Stability-Guaranteed Control for Unknown Nonlinear Time-Varying Systems
par: Zhang, Hengde, et autres
Publié: (2026)
par: Zhang, Hengde, et autres
Publié: (2026)
Novel Conditions for the Finite-Region Stability of 2D-Systems with Application to Iterative Learning Control
par: Liang, Chao, et autres
Publié: (2024)
par: Liang, Chao, et autres
Publié: (2024)
Disturbance Attenuation Regulator I-B: Signal Bound Convergence and Steady-State
par: Mannini, Davide, et autres
Publié: (2026)
par: Mannini, Davide, et autres
Publié: (2026)
Documents similaires
-
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
par: Lee, Donghwan
Publié: (2023) -
Converse Lyapunov Results for Switched Systems with Lower and Upper Bounds on Switching Intervals
par: Della Rossa, Matteo
Publié: (2024) -
Lyapunov-Certified Direct Switching Theory for Q-Learning
par: Lee, Donghwan
Publié: (2026) -
Contouring Error Bounded Control for Biaxial Switched Linear Systems
par: Yuan, Meng, et autres
Publié: (2024) -
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
par: Manenti, Massimiliano, et autres
Publié: (2025)