Fitted Q-Iteration via Max-Plus-Linear Approximation
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Y., Kolarijani, M. A. S. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
From Optimization to Control: Quasi Policy Iteration
por: Kolarijani, Mohammad Amin Sharifi, et al.
Publicado: (2023)
por: Kolarijani, Mohammad Amin Sharifi, et al.
Publicado: (2023)
Local Linearity of LLMs Enables Activation Steering via Model-Based Linear Optimal Control
por: Skifstad, Julian, et al.
Publicado: (2026)
por: Skifstad, Julian, et al.
Publicado: (2026)
Rank-One Modified Value Iteration
por: Kolarijani, Arman Sharifi, et al.
Publicado: (2025)
por: Kolarijani, Arman Sharifi, et al.
Publicado: (2025)
Differentiable Optimization for Deep Learning-Enhanced DC Approximation of AC Optimal Power Flow
por: Rosemberg, Andrew, et al.
Publicado: (2025)
por: Rosemberg, Andrew, et al.
Publicado: (2025)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
por: Zhu, Feng, et al.
Publicado: (2025)
por: Zhu, Feng, et al.
Publicado: (2025)
Multi-Year Maintenance Planning for Large-Scale Infrastructure Systems: A Novel Network Deep Q-Learning Approach
por: Fard, Amir, et al.
Publicado: (2025)
por: Fard, Amir, et al.
Publicado: (2025)
Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems
por: Huang, Yilie, et al.
Publicado: (2024)
por: Huang, Yilie, et al.
Publicado: (2024)
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
por: Huang, Yilie, et al.
Publicado: (2025)
por: Huang, Yilie, et al.
Publicado: (2025)
DASA: Delay-Adaptive Multi-Agent Stochastic Approximation
por: Fabbro, Nicolò Dal, et al.
Publicado: (2024)
por: Fabbro, Nicolò Dal, et al.
Publicado: (2024)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
por: Adibi, Arman, et al.
Publicado: (2024)
por: Adibi, Arman, et al.
Publicado: (2024)
Analysis of Thompson Sampling for Controlling Unknown Linear Diffusion Processes
por: Faradonbeh, Mohamad Kazem Shirani, et al.
Publicado: (2022)
por: Faradonbeh, Mohamad Kazem Shirani, et al.
Publicado: (2022)
Universal Approximation Theorem for Deep Q-Learning via FBSDE System
por: Qi, Qian
Publicado: (2025)
por: Qi, Qian
Publicado: (2025)
Hereditary Geometric Meta-RL: Nonlocal Generalization via Task Symmetries
por: Nitschke, Paul, et al.
Publicado: (2026)
por: Nitschke, Paul, et al.
Publicado: (2026)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
por: Anand, Akhil S, et al.
Publicado: (2025)
por: Anand, Akhil S, et al.
Publicado: (2025)
Infinite-Horizon Reach-Avoid Zero-Sum Games via Deep Reinforcement Learning
por: Li, Jingqi, et al.
Publicado: (2022)
por: Li, Jingqi, et al.
Publicado: (2022)
Distributionally Robust Safety Verification of Neural Networks via Worst-Case CVaR
por: Kishida, Masako
Publicado: (2025)
por: Kishida, Masako
Publicado: (2025)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
por: Ibrahim, Sinan, et al.
Publicado: (2026)
por: Ibrahim, Sinan, et al.
Publicado: (2026)
Asynchronous Distributed Reinforcement Learning for LQR Control via Zeroth-Order Block Coordinate Descent
por: Jing, Gangshan, et al.
Publicado: (2021)
por: Jing, Gangshan, et al.
Publicado: (2021)
DiscoverDCP: A Data-Driven Approach for Construction of Disciplined Convex Programs via Symbolic Regression
por: Myhre, Sveinung
Publicado: (2025)
por: Myhre, Sveinung
Publicado: (2025)
Optimal Control Operator Perspective and a Neural Adaptive Spectral Method
por: Feng, Mingquan, et al.
Publicado: (2024)
por: Feng, Mingquan, et al.
Publicado: (2024)
Nonlinear Non-Gaussian Density Steering with Input and Noise Channel Mismatch: Sinkhorn with Memory for Solving the Control-affine Schrödinger Bridge Problem
por: Bondar, Georgiy A., et al.
Publicado: (2026)
por: Bondar, Georgiy A., et al.
Publicado: (2026)
Stochastic Learning of Computational Resource Usage as Graph Structured Multimarginal Schrödinger Bridge
por: Bondar, Georgiy A., et al.
Publicado: (2024)
por: Bondar, Georgiy A., et al.
Publicado: (2024)
Distributed Optimization via Energy Conservation Laws in Dilated Coordinates
por: Baranwal, Mayank, et al.
Publicado: (2024)
por: Baranwal, Mayank, et al.
Publicado: (2024)
A Riemannian Optimization Perspective of the Gauss-Newton Method for Feedforward Neural Networks
por: Cayci, Semih
Publicado: (2024)
por: Cayci, Semih
Publicado: (2024)
Stability of Primal-Dual Gradient Flow Dynamics for Multi-Block Convex Optimization Problems
por: Ozaslan, Ibrahim K., et al.
Publicado: (2024)
por: Ozaslan, Ibrahim K., et al.
Publicado: (2024)
Primitive Agentic First-Order Optimization
por: Sala, R.
Publicado: (2024)
por: Sala, R.
Publicado: (2024)
Differentiable Distributionally Robust Optimization Layers
por: Ma, Xutao, et al.
Publicado: (2024)
por: Ma, Xutao, et al.
Publicado: (2024)
Generative AI and Process Systems Engineering: The Next Frontier
por: Decardi-Nelson, Benjamin, et al.
Publicado: (2024)
por: Decardi-Nelson, Benjamin, et al.
Publicado: (2024)
PID Accelerated Temporal Difference Algorithms
por: Bedaywi, Mark, et al.
Publicado: (2024)
por: Bedaywi, Mark, et al.
Publicado: (2024)
Optimal Power Grid Operations with Foundation Models
por: Puech, Alban, et al.
Publicado: (2024)
por: Puech, Alban, et al.
Publicado: (2024)
Learning a local trading strategy: deep reinforcement learning for grid-scale renewable energy integration
por: Ju, Caleb, et al.
Publicado: (2024)
por: Ju, Caleb, et al.
Publicado: (2024)
On the Convergence of Overparameterized Problems: Inherent Properties of the Compositional Structure of Neural Networks
por: de Oliveira, Arthur Castello Branco, et al.
Publicado: (2025)
por: de Oliveira, Arthur Castello Branco, et al.
Publicado: (2025)
Action Dependency Graphs for Globally Optimal Coordinated Reinforcement Learning
por: Ding, Jianglin, et al.
Publicado: (2025)
por: Ding, Jianglin, et al.
Publicado: (2025)
Intersection of Reinforcement Learning and Bayesian Optimization for Intelligent Control of Industrial Processes: A Safe MPC-based DPG using Multi-Objective BO
por: Esfahani, Hossein Nejatbakhsh, et al.
Publicado: (2025)
por: Esfahani, Hossein Nejatbakhsh, et al.
Publicado: (2025)
WARP: A Benchmark for Primal-Dual Warm-Starting of Interior-Point Solvers
por: Suri, Dhruv, et al.
Publicado: (2026)
por: Suri, Dhruv, et al.
Publicado: (2026)
PGLearn -- An Open-Source Learning Toolkit for Optimal Power Flow
por: Klamkin, Michael, et al.
Publicado: (2025)
por: Klamkin, Michael, et al.
Publicado: (2025)
gridfm-datakit-v1: A Python Library for Scalable and Realistic Power Flow and Optimal Power Flow Data Generation
por: Puech, Alban, et al.
Publicado: (2025)
por: Puech, Alban, et al.
Publicado: (2025)
Optimizing Inventory Routing: A Decision-Focused Learning Approach using Neural Networks
por: Islam, MD Shafikul, et al.
Publicado: (2023)
por: Islam, MD Shafikul, et al.
Publicado: (2023)
Lyapunov Function Consistent Adaptive Network Signal Control with Back Pressure and Reinforcement Learning
por: Ma, Chaolun, et al.
Publicado: (2022)
por: Ma, Chaolun, et al.
Publicado: (2022)
From Automation to Autonomy in Smart Manufacturing: A Bayesian Optimization Framework for Modeling Multi-Objective Experimentation and Sequential Decision Making
por: Asru, Avijit Saha, et al.
Publicado: (2025)
por: Asru, Avijit Saha, et al.
Publicado: (2025)
Ejemplares similares
-
From Optimization to Control: Quasi Policy Iteration
por: Kolarijani, Mohammad Amin Sharifi, et al.
Publicado: (2023) -
Local Linearity of LLMs Enables Activation Steering via Model-Based Linear Optimal Control
por: Skifstad, Julian, et al.
Publicado: (2026) -
Rank-One Modified Value Iteration
por: Kolarijani, Arman Sharifi, et al.
Publicado: (2025) -
Differentiable Optimization for Deep Learning-Enhanced DC Approximation of AC Optimal Power Flow
por: Rosemberg, Andrew, et al.
Publicado: (2025) -
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
por: Zhu, Feng, et al.
Publicado: (2025)