Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
Fuente:
arXiv
Saved in:
| Main Author: | Lee, Donghwan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Switching-Geometry Analysis of Deflated Q-Value Iteration
by: Lee, Donghwan
Published: (2026)
by: Lee, Donghwan
Published: (2026)
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
by: Lee, Donghwan
Published: (2023)
by: Lee, Donghwan
Published: (2023)
Policy Optimization for PDE Control with a Warm Start
by: Zhang, Xiangyuan, et al.
Published: (2024)
by: Zhang, Xiangyuan, et al.
Published: (2024)
Harnessing Membership Function Dynamics for Stability Analysis of T-S Fuzzy Systems
by: Lee, Donghwan, et al.
Published: (2024)
by: Lee, Donghwan, et al.
Published: (2024)
A Discrete-Time Switching System Analysis of Q-learning
by: Lee, Donghwan, et al.
Published: (2021)
by: Lee, Donghwan, et al.
Published: (2021)
Fitted Q-Iteration via Max-Plus-Linear Approximation
by: Liu, Y., et al.
Published: (2024)
by: Liu, Y., et al.
Published: (2024)
PhiBE: A PDE-based Bellman Equation for Continuous Time Policy Evaluation
by: Zhu, Yuhua
Published: (2024)
by: Zhu, Yuhua
Published: (2024)
The Cesàro Value Iteration
by: Mair, Jonas, et al.
Published: (2025)
by: Mair, Jonas, et al.
Published: (2025)
Adaptive Optimal Control of Linear Periodic Systems: An Off-Policy Value Iteration Approach
by: Pang, Bo, et al.
Published: (2019)
by: Pang, Bo, et al.
Published: (2019)
Data-Enabled Policy and Value Iteration for Continuous-Time Linear Quadratic Output Feedback Control
by: Xie, Jun, et al.
Published: (2026)
by: Xie, Jun, et al.
Published: (2026)
WARP: A Benchmark for Primal-Dual Warm-Starting of Interior-Point Solvers
by: Suri, Dhruv, et al.
Published: (2026)
by: Suri, Dhruv, et al.
Published: (2026)
On-Line Policy Iteration with Trajectory-Driven Policy Generation
by: Li, Yuchao, et al.
Published: (2026)
by: Li, Yuchao, et al.
Published: (2026)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
by: Anand, Akhil S, et al.
Published: (2025)
by: Anand, Akhil S, et al.
Published: (2025)
Revisiting LQR Control from the Perspective of Receding-Horizon Policy Gradient
by: Zhang, Xiangyuan, et al.
Published: (2023)
by: Zhang, Xiangyuan, et al.
Published: (2023)
Bellman Residual Minimization for Control: Geometry, Stationarity, and Convergence
by: Lee, Donghwan, et al.
Published: (2026)
by: Lee, Donghwan, et al.
Published: (2026)
Neural Policy Composition from Free Energy Minimization
by: Rossi, Francesca, et al.
Published: (2025)
by: Rossi, Francesca, et al.
Published: (2025)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
by: Ibrahim, Sinan, et al.
Published: (2026)
by: Ibrahim, Sinan, et al.
Published: (2026)
Using Laplace Transform To Optimize the Hallucination of Generation Models
by: Kang, Cheng, et al.
Published: (2026)
by: Kang, Cheng, et al.
Published: (2026)
Safe Decentralized Operation of EV Virtual Power Plant with Limited Network Visibility via Multi-Agent Reinforcement Learning
by: Huang, Chenghao, et al.
Published: (2026)
by: Huang, Chenghao, et al.
Published: (2026)
Intent-aligned Autonomous Spacecraft Guidance via Reasoning Models
by: Takubo, Yuji, et al.
Published: (2026)
by: Takubo, Yuji, et al.
Published: (2026)
Stability-Preserving Online Adaptation of Neural Closed-loop Maps
by: Saccani, Danilo, et al.
Published: (2026)
by: Saccani, Danilo, et al.
Published: (2026)
Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning
by: Merati, Mohammad, et al.
Published: (2026)
by: Merati, Mohammad, et al.
Published: (2026)
Agent-Based Decentralized Energy Management of EV Charging Station with Solar Photovoltaics via Multi-Agent Reinforcement Learning
by: Fan, Jiarong, et al.
Published: (2025)
by: Fan, Jiarong, et al.
Published: (2025)
Distributionally Robust Control with End-to-End Statistically Guaranteed Metric Learning
by: Wu, Jingyi, et al.
Published: (2025)
by: Wu, Jingyi, et al.
Published: (2025)
Safe Control and Learning Using Generalized Action Governor
by: Fang, Peiyuan, et al.
Published: (2022)
by: Fang, Peiyuan, et al.
Published: (2022)
Approximate Model Predictive Control for Microgrid Energy Management via Imitation Learning
by: Liu, Changrui, et al.
Published: (2025)
by: Liu, Changrui, et al.
Published: (2025)
Towards a constructive framework for control theory
by: Osinenko, Pavel
Published: (2025)
by: Osinenko, Pavel
Published: (2025)
Nonlinear Control Allocation: A Learning Based Approach
by: Khan, Hafiz Zeeshan Iqbal, et al.
Published: (2022)
by: Khan, Hafiz Zeeshan Iqbal, et al.
Published: (2022)
Planning a Community Approach to Diabetes Care in Low- and Middle-Income Countries Using Optimization
by: Adams, Katherine B., et al.
Published: (2023)
by: Adams, Katherine B., et al.
Published: (2023)
Approximate Information States for Worst-Case Control and Learning in Uncertain Systems
by: Dave, Aditya, et al.
Published: (2023)
by: Dave, Aditya, et al.
Published: (2023)
Model Predictive Control and Reinforcement Learning: A Unified Framework Based on Dynamic Programming
by: Bertsekas, Dimitri P.
Published: (2024)
by: Bertsekas, Dimitri P.
Published: (2024)
Deep Reinforcement Learning for Multi-Objective Optimization: Enhancing Wind Turbine Energy Generation while Mitigating Noise Emissions
by: de Frutos, Martín, et al.
Published: (2024)
by: de Frutos, Martín, et al.
Published: (2024)
Unifying Controller Design for Stabilizing Nonlinear Systems with Norm-Bounded Control Inputs
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
Human-in-the-Loop AI for HVAC Management Enhancing Comfort and Energy Efficiency
by: Liang, Xinyu, et al.
Published: (2025)
by: Liang, Xinyu, et al.
Published: (2025)
Large Language Model-Assisted Planning of Electric Vehicle Charging Infrastructure with Real-World Case Study
by: Zheng, Xinda, et al.
Published: (2025)
by: Zheng, Xinda, et al.
Published: (2025)
Remarks on the Polyak-Lojasiewicz inequality and the convergence of gradient systems
by: de Oliveira, Arthur Castello B., et al.
Published: (2025)
by: de Oliveira, Arthur Castello B., et al.
Published: (2025)
Deep Reinforcement Learning Optimization for Uncertain Nonlinear Systems via Event-Triggered Robust Adaptive Dynamic Programming
by: Bai, Ningwei, et al.
Published: (2025)
by: Bai, Ningwei, et al.
Published: (2025)
Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments
by: Kara, Ali Devran, et al.
Published: (2023)
by: Kara, Ali Devran, et al.
Published: (2023)
A Context-Free Smart Grid Model Using Complex System Approach
by: Amor, Soufian Ben, et al.
Published: (2025)
by: Amor, Soufian Ben, et al.
Published: (2025)
Energy Management for Renewable-Colocated Artificial Intelligence Data Centers
by: Li, Siying, et al.
Published: (2025)
by: Li, Siying, et al.
Published: (2025)
Similar Items
-
Switching-Geometry Analysis of Deflated Q-Value Iteration
by: Lee, Donghwan
Published: (2026) -
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
by: Lee, Donghwan
Published: (2023) -
Policy Optimization for PDE Control with a Warm Start
by: Zhang, Xiangyuan, et al.
Published: (2024) -
Harnessing Membership Function Dynamics for Stability Analysis of T-S Fuzzy Systems
by: Lee, Donghwan, et al.
Published: (2024) -
A Discrete-Time Switching System Analysis of Q-learning
by: Lee, Donghwan, et al.
Published: (2021)