Bellman Residual Minimization for Control: Geometry, Stationarity, and Convergence
Fuente:
arXiv
Guardado en:
| Autores principales: | Lee, Donghwan, Yang, Hyukjun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Contraction-Aligned Analysis of Soft Bellman Residual Minimization with Weighted Lp-Norm for Markov Decision Problem
por: Yang, Hyukjun, et al.
Publicado: (2026)
por: Yang, Hyukjun, et al.
Publicado: (2026)
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
por: Lee, Donghwan
Publicado: (2026)
por: Lee, Donghwan
Publicado: (2026)
Analysis of Off-Policy Multi-Step TD-Learning with Linear Function Approximation
por: Lee, Donghwan
Publicado: (2024)
por: Lee, Donghwan
Publicado: (2024)
Lyapunov-Certified Direct Switching Theory for Q-Learning
por: Lee, Donghwan
Publicado: (2026)
por: Lee, Donghwan
Publicado: (2026)
Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games
por: Jeong, Narim, et al.
Publicado: (2026)
por: Jeong, Narim, et al.
Publicado: (2026)
Finite-Time Analysis of Simultaneous Double Q-learning
por: Na, Hyunjun, et al.
Publicado: (2024)
por: Na, Hyunjun, et al.
Publicado: (2024)
Finite-Time Accuracy of Temporal-Difference Learning Under Schur-Stable Recursions
por: Lee, Donghwan, et al.
Publicado: (2022)
por: Lee, Donghwan, et al.
Publicado: (2022)
Deep Q-Learning with Gradient Target Tracking
por: Park, Bum Geun, et al.
Publicado: (2025)
por: Park, Bum Geun, et al.
Publicado: (2025)
Verifiable Error Bounds for Physics-Informed Neural Network Solutions of Lyapunov and Hamilton-Jacobi-Bellman Equations
por: Liu, Jun
Publicado: (2026)
por: Liu, Jun
Publicado: (2026)
Signatures Meet Dynamic Programming: Generalizing Bellman Equations for Trajectory Following
por: Ohnishi, Motoya, et al.
Publicado: (2023)
por: Ohnishi, Motoya, et al.
Publicado: (2023)
Computationally Efficient RL under Linear Bellman Completeness for Deterministic Dynamics
por: Wu, Runzhe, et al.
Publicado: (2024)
por: Wu, Runzhe, et al.
Publicado: (2024)
Universal Approximation Power of Deep Residual Neural Networks via Nonlinear Control Theory
por: Tabuada, Paulo, et al.
Publicado: (2020)
por: Tabuada, Paulo, et al.
Publicado: (2020)
Periodic Regularized Q-Learning
por: Yang, Hyukjun, et al.
Publicado: (2026)
por: Yang, Hyukjun, et al.
Publicado: (2026)
Residual Power Flow for Neural Solvers
por: Stiasny, Jochen, et al.
Publicado: (2026)
por: Stiasny, Jochen, et al.
Publicado: (2026)
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
por: Klein, Sara, et al.
Publicado: (2023)
por: Klein, Sara, et al.
Publicado: (2023)
ORN-CBF: Learning Observation-conditioned Residual Neural Control Barrier Functions via Hypernetworks
por: Derajić, Bojan, et al.
Publicado: (2025)
por: Derajić, Bojan, et al.
Publicado: (2025)
Online Residual Learning from Offline Experts for Pedestrian Tracking
por: Vlachos, Anastasios, et al.
Publicado: (2024)
por: Vlachos, Anastasios, et al.
Publicado: (2024)
On the Convergence of Overlapping Schwarz Decomposition for Nonlinear Optimal Control
por: Na, Sen, et al.
Publicado: (2020)
por: Na, Sen, et al.
Publicado: (2020)
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
por: Lee, Donghwan
Publicado: (2023)
por: Lee, Donghwan
Publicado: (2023)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
por: Manenti, Massimiliano, et al.
Publicado: (2025)
por: Manenti, Massimiliano, et al.
Publicado: (2025)
Geometry-Preserving Neural Architectures on Manifolds with Boundary
por: Elamvazhuthi, Karthik, et al.
Publicado: (2026)
por: Elamvazhuthi, Karthik, et al.
Publicado: (2026)
Enhancing Convergence of Decentralized Gradient Tracking under the KL Property
por: Chen, Xiaokai, et al.
Publicado: (2024)
por: Chen, Xiaokai, et al.
Publicado: (2024)
Is Bellman Equation Enough for Learning Control?
por: You, Haoxiang, et al.
Publicado: (2025)
por: You, Haoxiang, et al.
Publicado: (2025)
ControlSynth Neural ODEs: Modeling Dynamical Systems with Guaranteed Convergence
por: Mei, Wenjie, et al.
Publicado: (2024)
por: Mei, Wenjie, et al.
Publicado: (2024)
A Short and Unified Convergence Analysis of the SAG, SAGA, and IAG Algorithms
por: Zhu, Feng, et al.
Publicado: (2026)
por: Zhu, Feng, et al.
Publicado: (2026)
On Linear Convergence of PI Consensus Algorithm under the Restricted Secant Inequality
por: Chakrabarti, Kushal, et al.
Publicado: (2023)
por: Chakrabarti, Kushal, et al.
Publicado: (2023)
Nonuniqueness and Convergence to Equivalent Solutions in Observer-based Inverse Reinforcement Learning
por: Town, Jared, et al.
Publicado: (2022)
por: Town, Jared, et al.
Publicado: (2022)
On the Global Convergence of Policy Gradient in Average Reward Markov Decision Processes
por: Kumar, Navdeep, et al.
Publicado: (2024)
por: Kumar, Navdeep, et al.
Publicado: (2024)
Safe Multi-Agent Reinforcement Learning with Convergence to Generalized Nash Equilibrium
por: Li, Zeyang, et al.
Publicado: (2024)
por: Li, Zeyang, et al.
Publicado: (2024)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
por: Ding, Dongsheng, et al.
Publicado: (2023)
por: Ding, Dongsheng, et al.
Publicado: (2023)
Control in Stochastic Environment with Delays: A Model-based Reinforcement Learning Approach
por: Yao, Zhiyuan, et al.
Publicado: (2024)
por: Yao, Zhiyuan, et al.
Publicado: (2024)
Dynamic Controlled Variables Based Dynamic Self-Optimizing Control
por: Zhou, Chenchen, et al.
Publicado: (2026)
por: Zhou, Chenchen, et al.
Publicado: (2026)
Nonasymptotic Regret Analysis of Adaptive Linear Quadratic Control with Model Misspecification
por: Lee, Bruce D., et al.
Publicado: (2023)
por: Lee, Bruce D., et al.
Publicado: (2023)
Reinforcement Learning Based Traffic Signal Design to Minimize Queue Lengths
por: Nandakumar, Anirud, et al.
Publicado: (2025)
por: Nandakumar, Anirud, et al.
Publicado: (2025)
Non-Parametric Learning of Stochastic Differential Equations with Non-asymptotic Fast Rates of Convergence
por: Bonalli, Riccardo, et al.
Publicado: (2023)
por: Bonalli, Riccardo, et al.
Publicado: (2023)
Convergence of Byzantine-Resilient Gradient Tracking via Probabilistic Edge Dropout
por: Dezhboro, Amirhossein, et al.
Publicado: (2026)
por: Dezhboro, Amirhossein, et al.
Publicado: (2026)
Active Learning for Control-Oriented Identification of Nonlinear Systems
por: Lee, Bruce D., et al.
Publicado: (2024)
por: Lee, Bruce D., et al.
Publicado: (2024)
ReACT: Reinforcement Learning for Controller Parametrization using B-Spline Geometries
por: Rudolf, Thomas, et al.
Publicado: (2024)
por: Rudolf, Thomas, et al.
Publicado: (2024)
Data-Driven Predictive Control of Nonholonomic Robots Based on a Bilinear Koopman Realization: Data Does Not Replace Geometry
por: Rosenfelder, Mario, et al.
Publicado: (2024)
por: Rosenfelder, Mario, et al.
Publicado: (2024)
Active Learning-Based Optimization of Hydroelectric Turbine Startup to Minimize Fatigue Damage
por: Mai, Vincent, et al.
Publicado: (2024)
por: Mai, Vincent, et al.
Publicado: (2024)
Ejemplares similares
-
Contraction-Aligned Analysis of Soft Bellman Residual Minimization with Weighted Lp-Norm for Markov Decision Problem
por: Yang, Hyukjun, et al.
Publicado: (2026) -
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
por: Lee, Donghwan
Publicado: (2026) -
Analysis of Off-Policy Multi-Step TD-Learning with Linear Function Approximation
por: Lee, Donghwan
Publicado: (2024) -
Lyapunov-Certified Direct Switching Theory for Q-Learning
por: Lee, Donghwan
Publicado: (2026) -
Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games
por: Jeong, Narim, et al.
Publicado: (2026)