Fill-and-Spill: Deep Reinforcement Learning Policy Gradient Methods for Reservoir Operation Decision and Control
Fuente:
arXiv
Saved in:
| Main Authors: | Tabas, Sadegh Sadeghi, Samadi, Vidya |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
by: Carmona, René, et al.
Published: (2019)
by: Carmona, René, et al.
Published: (2019)
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
Bilevel Learning with Inexact Stochastic Gradients
by: Salehi, Mohammad Sadegh, et al.
Published: (2024)
by: Salehi, Mohammad Sadegh, et al.
Published: (2024)
Elementary Analysis of Policy Gradient Methods
by: Liu, Jiacai, et al.
Published: (2024)
by: Liu, Jiacai, et al.
Published: (2024)
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
by: Xiao, Minheng, et al.
Published: (2024)
by: Xiao, Minheng, et al.
Published: (2024)
Deep Reinforcement Learning for Dynamic Order Picking in Warehouse Operations
by: Mahmoudinazlou, Sasan, et al.
Published: (2024)
by: Mahmoudinazlou, Sasan, et al.
Published: (2024)
Decision-Focused Learning with Directional Gradients
by: Huang, Michael, et al.
Published: (2024)
by: Huang, Michael, et al.
Published: (2024)
Natural Policy Gradient as Doubly Smoothed Policy Iteration: A Bellman-Operator Framework
by: Nanda, Phalguni, et al.
Published: (2026)
by: Nanda, Phalguni, et al.
Published: (2026)
Nonconvex Stochastic Bregman Proximal Gradient Method with Application to Deep Learning
by: Ding, Kuangyu, et al.
Published: (2023)
by: Ding, Kuangyu, et al.
Published: (2023)
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
Sequential Bayesian Optimal Experimental Design in Infinite Dimensions via Policy Gradient Reinforcement Learning
by: Shen, Kaichen, et al.
Published: (2026)
by: Shen, Kaichen, et al.
Published: (2026)
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
by: Basu, Debabrota, et al.
Published: (2025)
by: Basu, Debabrota, et al.
Published: (2025)
Analysis of On-policy Policy Gradient Methods under the Distribution Mismatch
by: Wang, Weizhen, et al.
Published: (2025)
by: Wang, Weizhen, et al.
Published: (2025)
An Efficient On-Policy Deep Learning Framework for Stochastic Optimal Control
by: Hua, Mengjian, et al.
Published: (2024)
by: Hua, Mengjian, et al.
Published: (2024)
Deterministic Policy Gradient for Reinforcement Learning with Continuous Time and State
by: Cheng, Ziheng, et al.
Published: (2025)
by: Cheng, Ziheng, et al.
Published: (2025)
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
by: Klein, Sara, et al.
Published: (2023)
by: Klein, Sara, et al.
Published: (2023)
Operator World Models for Reinforcement Learning
by: Novelli, Pietro, et al.
Published: (2024)
by: Novelli, Pietro, et al.
Published: (2024)
Model-Free Output Feedback Stabilization via Policy Gradient Methods
by: Zhang, Ankang, et al.
Published: (2026)
by: Zhang, Ankang, et al.
Published: (2026)
Structured Reinforcement Learning for Combinatorial Decision-Making
by: Hoppe, Heiko, et al.
Published: (2025)
by: Hoppe, Heiko, et al.
Published: (2025)
Policy Gradient Methods for Discrete Time Linear Quadratic Regulator With Random Parameters
by: Li, Deyue
Published: (2023)
by: Li, Deyue
Published: (2023)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
How Does the ReLU Activation Affect the Implicit Bias of Gradient Descent on High-dimensional Neural Network Regression?
by: Lai, Kuo-Wei, et al.
Published: (2026)
by: Lai, Kuo-Wei, et al.
Published: (2026)
Deep Operator Neural Network Model Predictive Control
by: de Jong, Thomas Oliver, et al.
Published: (2025)
by: de Jong, Thomas Oliver, et al.
Published: (2025)
An Adaptively Inexact Method for Bilevel Learning Using Primal-Dual Style Differentiation
by: Bogensperger, Lea, et al.
Published: (2024)
by: Bogensperger, Lea, et al.
Published: (2024)
On the Global Convergence of Risk-Averse Natural Policy Gradient Methods with Expected Conditional Risk Measures
by: Yu, Xian, et al.
Published: (2023)
by: Yu, Xian, et al.
Published: (2023)
Classical and Deep Reinforcement Learning Inventory Control Policies for Pharmaceutical Supply Chains with Perishability and Non-Stationarity
by: Stranieri, Francesco, et al.
Published: (2025)
by: Stranieri, Francesco, et al.
Published: (2025)
Quasi-Newton Compatible Actor-Critic for Deterministic Policies
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
Operator Models for Continuous-Time Offline Reinforcement Learning
by: Hoischen, Nicolas, et al.
Published: (2025)
by: Hoischen, Nicolas, et al.
Published: (2025)
Achieve Performatively Optimal Policy for Performative Reinforcement Learning
by: Chen, Ziyi, et al.
Published: (2025)
by: Chen, Ziyi, et al.
Published: (2025)
Recurrent Natural Policy Gradient for POMDPs
by: Cayci, Semih, et al.
Published: (2024)
by: Cayci, Semih, et al.
Published: (2024)
Non-stationary and Varying-discounting Markov Decision Processes for Reinforcement Learning
by: Chen, Zhizuo, et al.
Published: (2025)
by: Chen, Zhizuo, et al.
Published: (2025)
Online Reinforcement Learning in Markov Decision Process Using Linear Programming
by: Leon, Vincent, et al.
Published: (2023)
by: Leon, Vincent, et al.
Published: (2023)
Gradient Methods with Online Scaling
by: Gao, Wenzhi, et al.
Published: (2024)
by: Gao, Wenzhi, et al.
Published: (2024)
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
by: Mazumdar, Abhijit, et al.
Published: (2024)
by: Mazumdar, Abhijit, et al.
Published: (2024)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Non-Euclidean Gradient Descent Operates at the Edge of Stability
by: Islamov, Rustem, et al.
Published: (2026)
by: Islamov, Rustem, et al.
Published: (2026)
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
by: Han, Yinbin, et al.
Published: (2023)
by: Han, Yinbin, et al.
Published: (2023)
Deep Reinforcement Learning: A Convex Optimization Approach
by: Gattami, Ather
Published: (2024)
by: Gattami, Ather
Published: (2024)
Structure-Informed Deep Reinforcement Learning for Inventory Management
by: Maggiar, Alvaro, et al.
Published: (2025)
by: Maggiar, Alvaro, et al.
Published: (2025)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
by: Ganesh, Swetha, et al.
Published: (2024)
by: Ganesh, Swetha, et al.
Published: (2024)
Similar Items
-
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
by: Carmona, René, et al.
Published: (2019) -
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024) -
Bilevel Learning with Inexact Stochastic Gradients
by: Salehi, Mohammad Sadegh, et al.
Published: (2024) -
Elementary Analysis of Policy Gradient Methods
by: Liu, Jiacai, et al.
Published: (2024) -
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
by: Xiao, Minheng, et al.
Published: (2024)