Lyapunov-Certified Direct Switching Theory for Q-Learning
Fuente:
arXiv
Saved in:
| Main Author: | Lee, Donghwan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms
by: Lee, Donghwan, et al.
Published: (2024)
by: Lee, Donghwan, et al.
Published: (2024)
Formal Synthesis of Certifiably Robust Neural Lyapunov-Barrier Certificates
by: Wang, Chengxiao, et al.
Published: (2026)
by: Wang, Chengxiao, et al.
Published: (2026)
Certified Training with Branch-and-Bound for Lyapunov-stable Neural Control
by: Shi, Zhouxing, et al.
Published: (2024)
by: Shi, Zhouxing, et al.
Published: (2024)
Formally Verifying Deep Reinforcement Learning Controllers with Lyapunov Barrier Certificates
by: Mandal, Udayan, et al.
Published: (2024)
by: Mandal, Udayan, et al.
Published: (2024)
Safe Deep Model-Based Reinforcement Learning with Lyapunov Functions
by: Zhang, Harry
Published: (2024)
by: Zhang, Harry
Published: (2024)
Deep Q-Learning with Gradient Target Tracking
by: Park, Bum Geun, et al.
Published: (2025)
by: Park, Bum Geun, et al.
Published: (2025)
Learning Geometrically-Informed Lyapunov Functions with Deep Diffeomorphic RBF Networks
by: Tesfazgi, Samuel, et al.
Published: (2025)
by: Tesfazgi, Samuel, et al.
Published: (2025)
Finite-Time Analysis of Simultaneous Double Q-learning
by: Na, Hyunjun, et al.
Published: (2024)
by: Na, Hyunjun, et al.
Published: (2024)
Switching-Geometry Analysis of Deflated Q-Value Iteration
by: Lee, Donghwan
Published: (2026)
by: Lee, Donghwan
Published: (2026)
A finite time analysis of distributed Q-learning
by: Lim, Han-Dong, et al.
Published: (2024)
by: Lim, Han-Dong, et al.
Published: (2024)
Lyapunov Function Consistent Adaptive Network Signal Control with Back Pressure and Reinforcement Learning
by: Ma, Chaolun, et al.
Published: (2022)
by: Ma, Chaolun, et al.
Published: (2022)
Runtime-Certified Bounded-Error Quantized Attention
by: Calver, Dean
Published: (2026)
by: Calver, Dean
Published: (2026)
Certifiably Robust Policies for Uncertain Parametric Environments
by: Schnitzer, Yannik, et al.
Published: (2024)
by: Schnitzer, Yannik, et al.
Published: (2024)
Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games
by: Jeong, Narim, et al.
Published: (2026)
by: Jeong, Narim, et al.
Published: (2026)
Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers
by: Shen, Keyi, et al.
Published: (2026)
by: Shen, Keyi, et al.
Published: (2026)
Analysis of Off-Policy Multi-Step TD-Learning with Linear Function Approximation
by: Lee, Donghwan
Published: (2024)
by: Lee, Donghwan
Published: (2024)
MSACL: Multi-Step Actor-Critic Learning with Lyapunov Certificates for Exponentially Stabilizing Control
by: Zhang, Yongwei, et al.
Published: (2025)
by: Zhang, Yongwei, et al.
Published: (2025)
Certifiable Safe RLHF: Fixed-Penalty Constraint Optimization for Safer Language Models
by: Pandit, Kartik, et al.
Published: (2025)
by: Pandit, Kartik, et al.
Published: (2025)
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
by: Lee, Donghwan
Published: (2026)
by: Lee, Donghwan
Published: (2026)
A Deep Q-Learning based Smart Scheduling of EVs for Demand Response in Smart Grids
by: Chifu, Viorica Rozina, et al.
Published: (2024)
by: Chifu, Viorica Rozina, et al.
Published: (2024)
Predictive Safety Shield for Dyna-Q Reinforcement Learning
by: Pin, Jin, et al.
Published: (2025)
by: Pin, Jin, et al.
Published: (2025)
Lyapunov-stable Neural Control for State and Output Feedback: A Novel Formulation
by: Yang, Lujie, et al.
Published: (2024)
by: Yang, Lujie, et al.
Published: (2024)
Multi-Year Maintenance Planning for Large-Scale Infrastructure Systems: A Novel Network Deep Q-Learning Approach
by: Fard, Amir, et al.
Published: (2025)
by: Fard, Amir, et al.
Published: (2025)
Lyapunov-Aware Quantum-Inspired Reinforcement Learning for Continuous-Time Vehicle Control: A Feasibility Study
by: Kraipatthanapong, Nutkritta, et al.
Published: (2025)
by: Kraipatthanapong, Nutkritta, et al.
Published: (2025)
Fitted Q-Iteration via Max-Plus-Linear Approximation
by: Liu, Y., et al.
Published: (2024)
by: Liu, Y., et al.
Published: (2024)
Analytical Lyapunov Function Discovery: An RL-based Generative Approach
by: Zou, Haohan, et al.
Published: (2025)
by: Zou, Haohan, et al.
Published: (2025)
Suppressing Overestimation in Q-Learning through Adversarial Behaviors
by: Lee, HyeAnn, et al.
Published: (2023)
by: Lee, HyeAnn, et al.
Published: (2023)
Certifying Stability of Reinforcement Learning Policies using Generalized Lyapunov Functions
by: Long, Kehan, et al.
Published: (2025)
by: Long, Kehan, et al.
Published: (2025)
A Discrete-Time Switching System Analysis of Q-learning
by: Lee, Donghwan, et al.
Published: (2021)
by: Lee, Donghwan, et al.
Published: (2021)
Finite-Time Accuracy of Temporal-Difference Learning Under Schur-Stable Recursions
by: Lee, Donghwan, et al.
Published: (2022)
by: Lee, Donghwan, et al.
Published: (2022)
Learning-enabled Flexible Job-shop Scheduling for Scalable Smart Manufacturing
by: Moon, Sihoon, et al.
Published: (2024)
by: Moon, Sihoon, et al.
Published: (2024)
Inertial Navigation Meets Deep Learning: A Survey of Current Trends and Future Directions
by: Cohen, Nadav, et al.
Published: (2023)
by: Cohen, Nadav, et al.
Published: (2023)
Infinite-Horizon Reach-Avoid Zero-Sum Games via Deep Reinforcement Learning
by: Li, Jingqi, et al.
Published: (2022)
by: Li, Jingqi, et al.
Published: (2022)
Dynamic Decision Making in Engineering System Design: A Deep Q-Learning Approach
by: Giahi, Ramin, et al.
Published: (2023)
by: Giahi, Ramin, et al.
Published: (2023)
Periodic Regularized Q-Learning
by: Yang, Hyukjun, et al.
Published: (2026)
by: Yang, Hyukjun, et al.
Published: (2026)
Safe-Support Q-Learning: Learning without Unsafe Exploration
by: Lim, Yeeun, et al.
Published: (2026)
by: Lim, Yeeun, et al.
Published: (2026)
Bellman Residual Minimization for Control: Geometry, Stationarity, and Convergence
by: Lee, Donghwan, et al.
Published: (2026)
by: Lee, Donghwan, et al.
Published: (2026)
Online Model-based Anomaly Detection in Multivariate Time Series: Taxonomy, Survey, Research Challenges and Future Directions
by: Correia, Lucas, et al.
Published: (2024)
by: Correia, Lucas, et al.
Published: (2024)
Robust Recovery Controller for a Quadrupedal Robot using Deep Reinforcement Learning
by: Lee, Joonho, et al.
Published: (2019)
by: Lee, Joonho, et al.
Published: (2019)
A Solvable Molecular Switch Model for Stable Temporal Information Processing
by: Nurdin, H. I., et al.
Published: (2025)
by: Nurdin, H. I., et al.
Published: (2025)
Similar Items
-
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms
by: Lee, Donghwan, et al.
Published: (2024) -
Formal Synthesis of Certifiably Robust Neural Lyapunov-Barrier Certificates
by: Wang, Chengxiao, et al.
Published: (2026) -
Certified Training with Branch-and-Bound for Lyapunov-stable Neural Control
by: Shi, Zhouxing, et al.
Published: (2024) -
Formally Verifying Deep Reinforcement Learning Controllers with Lyapunov Barrier Certificates
by: Mandal, Udayan, et al.
Published: (2024) -
Safe Deep Model-Based Reinforcement Learning with Lyapunov Functions
by: Zhang, Harry
Published: (2024)