A Discrete-Time Switching System Analysis of Q-learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Donghwan, Hu, Jianghai, He, Niao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2021
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Switching-Geometry Analysis of Deflated Q-Value Iteration
di: Lee, Donghwan
Pubblicazione: (2026)
di: Lee, Donghwan
Pubblicazione: (2026)
Stochastic Primal-Dual Q-Learning
di: Jeong, Narim, et al.
Pubblicazione: (2018)
di: Jeong, Narim, et al.
Pubblicazione: (2018)
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
di: Lee, Donghwan
Pubblicazione: (2026)
di: Lee, Donghwan
Pubblicazione: (2026)
Select-then-differentiate: Solving Bilevel Optimization with Manifold Lower-level Solution Sets
di: Masiha, Saeed, et al.
Pubblicazione: (2026)
di: Masiha, Saeed, et al.
Pubblicazione: (2026)
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
di: Lee, Donghwan
Pubblicazione: (2023)
di: Lee, Donghwan
Pubblicazione: (2023)
Lyapunov-Certified Direct Switching Theory for Q-Learning
di: Lee, Donghwan
Pubblicazione: (2026)
di: Lee, Donghwan
Pubblicazione: (2026)
Temporal Robustness in Discrete Time Linear Dynamical Systems
di: Metya, Nilava, et al.
Pubblicazione: (2025)
di: Metya, Nilava, et al.
Pubblicazione: (2025)
PoLAR: Polar-Decomposed Low-Rank Adapter Representation
di: Lion, Kai, et al.
Pubblicazione: (2025)
di: Lion, Kai, et al.
Pubblicazione: (2025)
Zeroth-Order Optimization Finds Flat Minima
di: Zhang, Liang, et al.
Pubblicazione: (2025)
di: Zhang, Liang, et al.
Pubblicazione: (2025)
Analysis of Discrete-Time Switched Linear Systems under Logic Dynamic Switchings
di: Zhang, Xiao, et al.
Pubblicazione: (2022)
di: Zhang, Xiao, et al.
Pubblicazione: (2022)
Offline Learning of Decision Functions in Multiplayer Games with Expectation Constraints
di: Huang, Yuanhanqing, et al.
Pubblicazione: (2024)
di: Huang, Yuanhanqing, et al.
Pubblicazione: (2024)
Harnessing Membership Function Dynamics for Stability Analysis of T-S Fuzzy Systems
di: Lee, Donghwan, et al.
Pubblicazione: (2024)
di: Lee, Donghwan, et al.
Pubblicazione: (2024)
Optimal Output Feedback Learning Control for Discrete-Time Linear Quadratic Regulation
di: Xie, Kedi, et al.
Pubblicazione: (2025)
di: Xie, Kedi, et al.
Pubblicazione: (2025)
Universal Approximation Theorem for Deep Q-Learning via FBSDE System
di: Qi, Qian
Pubblicazione: (2025)
di: Qi, Qian
Pubblicazione: (2025)
Lossless Convexification and Duality
di: Lee, Donghwan
Pubblicazione: (2021)
di: Lee, Donghwan
Pubblicazione: (2021)
Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments
di: Kara, Ali Devran, et al.
Pubblicazione: (2023)
di: Kara, Ali Devran, et al.
Pubblicazione: (2023)
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
di: Neufeld, Ariel, et al.
Pubblicazione: (2022)
di: Neufeld, Ariel, et al.
Pubblicazione: (2022)
Interconnection of (Q,S,R)-Dissipative Systems in Discrete Time
di: Martinelli, Andrea, et al.
Pubblicazione: (2023)
di: Martinelli, Andrea, et al.
Pubblicazione: (2023)
Multi-Year Maintenance Planning for Large-Scale Infrastructure Systems: A Novel Network Deep Q-Learning Approach
di: Fard, Amir, et al.
Pubblicazione: (2025)
di: Fard, Amir, et al.
Pubblicazione: (2025)
A primal-dual perspective for distributed TD-learning
di: Lim, Han-Dong, et al.
Pubblicazione: (2023)
di: Lim, Han-Dong, et al.
Pubblicazione: (2023)
A deep learning method for solving stochastic optimal control problems driven by fully-coupled FBSDEs
di: Ji, Shaolin, et al.
Pubblicazione: (2022)
di: Ji, Shaolin, et al.
Pubblicazione: (2022)
Using deep learning to construct stochastic local search SAT solvers with performance bounds
di: Kramer, Maximilian J., et al.
Pubblicazione: (2023)
di: Kramer, Maximilian J., et al.
Pubblicazione: (2023)
Optimal Chaining of Vehicle Plans with Time Windows
di: Fiedler, David, et al.
Pubblicazione: (2024)
di: Fiedler, David, et al.
Pubblicazione: (2024)
Efficient Algorithms for A Class of Stochastic Hidden Convex Optimization and Its Applications in Network Revenue Management
di: Chen, Xin, et al.
Pubblicazione: (2022)
di: Chen, Xin, et al.
Pubblicazione: (2022)
A Robust Algorithm for Non-IID Machine Learning Problems with Convergence Analysis
di: Xu, Qing, et al.
Pubblicazione: (2025)
di: Xu, Qing, et al.
Pubblicazione: (2025)
TiAda: A Time-scale Adaptive Algorithm for Nonconvex Minimax Optimization
di: Li, Xiang, et al.
Pubblicazione: (2022)
di: Li, Xiang, et al.
Pubblicazione: (2022)
Pickup & Delivery with Time Windows and Transfers: combining decomposition with metaheuristics
di: Avgerinos, Ioannis, et al.
Pubblicazione: (2025)
di: Avgerinos, Ioannis, et al.
Pubblicazione: (2025)
Fixed-Time and Arbitrarily Fast Exponential Stabilization of Discrete-Time Switched Linear Systems
di: Flavio, Picchiotti, et al.
Pubblicazione: (2026)
di: Flavio, Picchiotti, et al.
Pubblicazione: (2026)
Fitted Q-Iteration via Max-Plus-Linear Approximation
di: Liu, Y., et al.
Pubblicazione: (2024)
di: Liu, Y., et al.
Pubblicazione: (2024)
Resource-constrained Project Scheduling with Time-of-Use Energy Tariffs and Machine States: A Logic-based Benders Decomposition Approach
di: Juvigny, Corentin, et al.
Pubblicazione: (2026)
di: Juvigny, Corentin, et al.
Pubblicazione: (2026)
Finite-Time Analysis of Gradient Descent for Shallow Transformers
di: Arda, Enes, et al.
Pubblicazione: (2026)
di: Arda, Enes, et al.
Pubblicazione: (2026)
The Principle of Proportional Duty: A Knowledge-Duty Framework for Ethical Equilibrium in Human and Artificial Systems
di: Prescher, Timothy
Pubblicazione: (2025)
di: Prescher, Timothy
Pubblicazione: (2025)
A Context-Free Smart Grid Model Using Complex System Approach
di: Amor, Soufian Ben, et al.
Pubblicazione: (2025)
di: Amor, Soufian Ben, et al.
Pubblicazione: (2025)
Test-Time Search in Neural Graph Coarsening Procedures for the Capacitated Vehicle Routing Problem
di: Sim, Yoonju, et al.
Pubblicazione: (2025)
di: Sim, Yoonju, et al.
Pubblicazione: (2025)
BenLOC: A Benchmark for Learning to Configure MIP Optimizers
di: Li, Hongpei, et al.
Pubblicazione: (2025)
di: Li, Hongpei, et al.
Pubblicazione: (2025)
Controllable Sequences of Minimal Length for Discrete-Time Switched Linear Control Systems
di: Mason, Paolo, et al.
Pubblicazione: (2025)
di: Mason, Paolo, et al.
Pubblicazione: (2025)
A Deep Q-Network Based on Radial Basis Functions for Multi-Echelon Inventory Management
di: Cheng, Liqiang, et al.
Pubblicazione: (2024)
di: Cheng, Liqiang, et al.
Pubblicazione: (2024)
Sign-Separated Finite-Time Error Analysis of Q-Learning
di: Lee, Donghwan
Pubblicazione: (2026)
di: Lee, Donghwan
Pubblicazione: (2026)
Feasible Pairings for Decentralized Integral Controllability of Non-Square Systems
di: Tong, Yuhao, et al.
Pubblicazione: (2026)
di: Tong, Yuhao, et al.
Pubblicazione: (2026)
OR-R1: Automating Modeling and Solving of Operations Research Optimization Problem via Test-Time Reinforcement Learning
di: Ding, Zezhen, et al.
Pubblicazione: (2025)
di: Ding, Zezhen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Switching-Geometry Analysis of Deflated Q-Value Iteration
di: Lee, Donghwan
Pubblicazione: (2026) -
Stochastic Primal-Dual Q-Learning
di: Jeong, Narim, et al.
Pubblicazione: (2018) -
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
di: Lee, Donghwan
Pubblicazione: (2026) -
Select-then-differentiate: Solving Bilevel Optimization with Manifold Lower-level Solution Sets
di: Masiha, Saeed, et al.
Pubblicazione: (2026) -
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
di: Lee, Donghwan
Pubblicazione: (2023)