Switching-Geometry Analysis of Deflated Q-Value Iteration
Fuente:
arXiv
Saved in:
| Main Author: | Lee, Donghwan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
by: Lee, Donghwan
Published: (2026)
by: Lee, Donghwan
Published: (2026)
A Discrete-Time Switching System Analysis of Q-learning
by: Lee, Donghwan, et al.
Published: (2021)
by: Lee, Donghwan, et al.
Published: (2021)
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
by: Lee, Donghwan
Published: (2023)
by: Lee, Donghwan
Published: (2023)
Deflated Dynamics Value Iteration
by: Lee, Jongmin, et al.
Published: (2024)
by: Lee, Jongmin, et al.
Published: (2024)
Lyapunov-Certified Direct Switching Theory for Q-Learning
by: Lee, Donghwan
Published: (2026)
by: Lee, Donghwan
Published: (2026)
Fitted Q-Iteration via Max-Plus-Linear Approximation
by: Liu, Y., et al.
Published: (2024)
by: Liu, Y., et al.
Published: (2024)
Exact Dual Geometry of SOC-ICNN Value Functions
by: Liu, Kang, et al.
Published: (2026)
by: Liu, Kang, et al.
Published: (2026)
Stochastic Primal-Dual Q-Learning
by: Jeong, Narim, et al.
Published: (2018)
by: Jeong, Narim, et al.
Published: (2018)
An Improved Last-Iterate Convergence Rate for Anchored Gradient Descent Ascent
by: Surina, Anja, et al.
Published: (2026)
by: Surina, Anja, et al.
Published: (2026)
Towards Efficient Risk-Sensitive Policy Gradient: An Iteration Complexity Analysis
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
A Heuristic Algorithm Based on Beam Search and Iterated Local Search for the Maritime Inventory Routing Problem
by: Sanghikian, Nathalie, et al.
Published: (2025)
by: Sanghikian, Nathalie, et al.
Published: (2025)
First-Order Geometry, Spectral Compression, and Structural Compatibility under Bounded Computation
by: Li, Changkai
Published: (2026)
by: Li, Changkai
Published: (2026)
Lossless Convexification and Duality
by: Lee, Donghwan
Published: (2021)
by: Lee, Donghwan
Published: (2021)
Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments
by: Kara, Ali Devran, et al.
Published: (2023)
by: Kara, Ali Devran, et al.
Published: (2023)
Harnessing Membership Function Dynamics for Stability Analysis of T-S Fuzzy Systems
by: Lee, Donghwan, et al.
Published: (2024)
by: Lee, Donghwan, et al.
Published: (2024)
Universal Approximation Theorem for Deep Q-Learning via FBSDE System
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
A Robust Algorithm for Non-IID Machine Learning Problems with Convergence Analysis
by: Xu, Qing, et al.
Published: (2025)
by: Xu, Qing, et al.
Published: (2025)
Algorithmic Detection of Rank Reversals, Transitivity Violations, and Decomposition Inconsistencies in Multi-Criteria Decision Analysis
by: Borda, Agustín, et al.
Published: (2025)
by: Borda, Agustín, et al.
Published: (2025)
Q3R: Quadratic Reweighted Rank Regularizer for Effective Low-Rank Training
by: Ghosh, Ipsita, et al.
Published: (2025)
by: Ghosh, Ipsita, et al.
Published: (2025)
A Deep Q-Network Based on Radial Basis Functions for Multi-Echelon Inventory Management
by: Cheng, Liqiang, et al.
Published: (2024)
by: Cheng, Liqiang, et al.
Published: (2024)
Q-Learning under Finite Model Uncertainty
by: Sester, Julian, et al.
Published: (2024)
by: Sester, Julian, et al.
Published: (2024)
Value of Information-based Deceptive Path Planning Under Adversarial Interventions
by: Suttle, Wesley A., et al.
Published: (2025)
by: Suttle, Wesley A., et al.
Published: (2025)
Multi-Year Maintenance Planning for Large-Scale Infrastructure Systems: A Novel Network Deep Q-Learning Approach
by: Fard, Amir, et al.
Published: (2025)
by: Fard, Amir, et al.
Published: (2025)
Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games
by: Jeong, Narim, et al.
Published: (2026)
by: Jeong, Narim, et al.
Published: (2026)
Parameter-Efficient Distributional RL via Normalizing Flows and a Geometry-Aware Cramér Surrogate
by: C., Simo Alami, et al.
Published: (2025)
by: C., Simo Alami, et al.
Published: (2025)
Is Scaling Learned Optimizers Worth It? Evaluating The Value of VeLO's 4000 TPU Months
by: Rezk, Fady, et al.
Published: (2023)
by: Rezk, Fady, et al.
Published: (2023)
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
by: Neufeld, Ariel, et al.
Published: (2022)
by: Neufeld, Ariel, et al.
Published: (2022)
Deflation-Free Optimal Scoring
by: Afroz, Sharmin, et al.
Published: (2026)
by: Afroz, Sharmin, et al.
Published: (2026)
The Cesàro Value Iteration
by: Mair, Jonas, et al.
Published: (2025)
by: Mair, Jonas, et al.
Published: (2025)
On the Error-Propagation of Inexact Hotelling's Deflation for Principal Component Analysis
by: Liao, Fangshuo, et al.
Published: (2023)
by: Liao, Fangshuo, et al.
Published: (2023)
Revisiting Inexact Fixed-Point Iterations for Min-Max Problems: Stochasticity and Structured Nonconvexity
by: Alacaoglu, Ahmet, et al.
Published: (2024)
by: Alacaoglu, Ahmet, et al.
Published: (2024)
Transformers Can Implement Preconditioned Richardson Iteration for In-Context Gaussian Kernel Regression
by: Yan, Mingsong, et al.
Published: (2026)
by: Yan, Mingsong, et al.
Published: (2026)
Strategizing against Q-learners: A Control-theoretical Approach
by: Arslantas, Yuksel, et al.
Published: (2024)
by: Arslantas, Yuksel, et al.
Published: (2024)
Column Generation for the Micro-Transit Zoning Problem
by: Hu, Hins, et al.
Published: (2026)
by: Hu, Hins, et al.
Published: (2026)
BONO-Bench: A Comprehensive Test Suite for Bi-objective Numerical Optimization with Traceable Pareto Sets
by: Schäpermeier, Lennart, et al.
Published: (2026)
by: Schäpermeier, Lennart, et al.
Published: (2026)
LLM4Branch: Large Language Model for Discovering Efficient Branching Policies of Integer Programs
by: Hou, Zhinan, et al.
Published: (2026)
by: Hou, Zhinan, et al.
Published: (2026)
Structural Segmentation of the Minimum Set Cover Problem: Exploiting Universe Decomposability for Metaheuristic Optimization
by: Hernández, Isidora, et al.
Published: (2026)
by: Hernández, Isidora, et al.
Published: (2026)
Cognitive Training for Language Models: Towards General Capabilities via Cross-Entropy Games
by: Hongler, Clément, et al.
Published: (2026)
by: Hongler, Clément, et al.
Published: (2026)
Integrated packing, placement, scheduling, and routing of personalized production: a pharmaceutical Industry 4.0 use-case with a planar transport system
by: Korladinov, Viktor Emil, et al.
Published: (2026)
by: Korladinov, Viktor Emil, et al.
Published: (2026)
Democratizing Large-Scale Re-Optimization with LLM-Guided Model Patches
by: Ye, Tinghan, et al.
Published: (2026)
by: Ye, Tinghan, et al.
Published: (2026)
Similar Items
-
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
by: Lee, Donghwan
Published: (2026) -
A Discrete-Time Switching System Analysis of Q-learning
by: Lee, Donghwan, et al.
Published: (2021) -
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
by: Lee, Donghwan
Published: (2023) -
Deflated Dynamics Value Iteration
by: Lee, Jongmin, et al.
Published: (2024) -
Lyapunov-Certified Direct Switching Theory for Q-Learning
by: Lee, Donghwan
Published: (2026)