Self-Organizing Dual-Buffer Adaptive Clustering Experience Replay (SODACER) for Safe Reinforcement Learning in Optimal Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Amirabadi, Roya Khalili, Farimani, Mohsen Jalaeian, Fard, Omid Solaymani |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient and Scalable Path-Planning Algorithms for Curvature Constrained Motion in the Hamilton-Jacobi Formulation
von: Parkinson, Christian, et al.
Veröffentlicht: (2023)
von: Parkinson, Christian, et al.
Veröffentlicht: (2023)
Risk-averse optimization under distributional uncertainty with Rockafellian relaxation
von: Antil, Harbir, et al.
Veröffentlicht: (2026)
von: Antil, Harbir, et al.
Veröffentlicht: (2026)
Rockafellian Relaxation for PDE-Constrained Optimization with Distributional Uncertainty
von: Antil, Harbir, et al.
Veröffentlicht: (2024)
von: Antil, Harbir, et al.
Veröffentlicht: (2024)
Optimal Control of Unbounded Functional Stochastic Evolution Systems in Hilbert Spaces: Second-Order Path-dependent HJB Equation
von: Tang, Shanjian, et al.
Veröffentlicht: (2024)
von: Tang, Shanjian, et al.
Veröffentlicht: (2024)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
von: Zheng, Yaowei, et al.
Veröffentlicht: (2026)
von: Zheng, Yaowei, et al.
Veröffentlicht: (2026)
Globally Optimal Inverse Kinematics as a Quadratic Program
von: Votroubek, Tomáš, et al.
Veröffentlicht: (2023)
von: Votroubek, Tomáš, et al.
Veröffentlicht: (2023)
Separable Approximations of Optimal Value Functions and Their Representation by Neural Networks
von: Sperl, Mario, et al.
Veröffentlicht: (2025)
von: Sperl, Mario, et al.
Veröffentlicht: (2025)
Distributed Online Optimization for Multi-Agent Optimal Transport
von: Krishnan, Vishaal, et al.
Veröffentlicht: (2018)
von: Krishnan, Vishaal, et al.
Veröffentlicht: (2018)
Optimal Control in Infinite Dimensional Spaces and Economic Modeling: State of the Art and Perspectives
von: Fabbri, Giorgio, et al.
Veröffentlicht: (2025)
von: Fabbri, Giorgio, et al.
Veröffentlicht: (2025)
Integrated Lander-Propulsion-GNC Framework for Autonomous Lunar Powered Descent
von: Aklan, Emre, et al.
Veröffentlicht: (2026)
von: Aklan, Emre, et al.
Veröffentlicht: (2026)
A Systematic Review and Taxonomy of Reinforcement Learning-Model Predictive Control Integration for Linear Systems
von: Farimani, Mohsen Jalaeian, et al.
Veröffentlicht: (2026)
von: Farimani, Mohsen Jalaeian, et al.
Veröffentlicht: (2026)
Lagrangian Relaxation for Continuous-Time Optimal Control of Coupled Hydrothermal Power Systems Including Storage Capacity and a Cascade of Hydropower Systems with Time Delays
von: Hammouda, Chiheb Ben, et al.
Veröffentlicht: (2023)
von: Hammouda, Chiheb Ben, et al.
Veröffentlicht: (2023)
Quantifying the Safety of Trajectories using Peak-Minimizing Control
von: Miller, Jared, et al.
Veröffentlicht: (2023)
von: Miller, Jared, et al.
Veröffentlicht: (2023)
Robust Ergodic Control of Jump-Diffusion Systems under Drift and Intensity Uncertainty
von: Azze, Abel, et al.
Veröffentlicht: (2026)
von: Azze, Abel, et al.
Veröffentlicht: (2026)
Stochastic Optimal Impulse Controls with Changing Running Costs
von: Cao, Yuchen, et al.
Veröffentlicht: (2025)
von: Cao, Yuchen, et al.
Veröffentlicht: (2025)
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
Optimality conditions for sparse optimal control of viscous Cahn-Hilliard systems with logarithmic potential
von: Colli, Pierluigi, et al.
Veröffentlicht: (2024)
von: Colli, Pierluigi, et al.
Veröffentlicht: (2024)
Second-order optimality conditions for the sparse optimal control of nonviscous Cahn-Hilliard systems
von: Colli, Pierluigi, et al.
Veröffentlicht: (2024)
von: Colli, Pierluigi, et al.
Veröffentlicht: (2024)
Engineering solutions for non-stationary gas pipeline reconstruction and emergency management
von: Aliyev, Ilgar
Veröffentlicht: (2025)
von: Aliyev, Ilgar
Veröffentlicht: (2025)
Optimality conditions in control problems with random state constraints in probabilistic or almost-sure form
von: Geiersbach, Caroline, et al.
Veröffentlicht: (2023)
von: Geiersbach, Caroline, et al.
Veröffentlicht: (2023)
A Scalable Method for Optimal Path Planning on Manifolds via a Hopf-Lax Type Formula
von: Huynh, Edward, et al.
Veröffentlicht: (2024)
von: Huynh, Edward, et al.
Veröffentlicht: (2024)
Adjoint-based calibration of nonlinear stochastic differential equations
von: Bartsch, Jan, et al.
Veröffentlicht: (2023)
von: Bartsch, Jan, et al.
Veröffentlicht: (2023)
Globally convergent homotopies for discrete-time optimal control
von: Esterhuizen, Willem, et al.
Veröffentlicht: (2023)
von: Esterhuizen, Willem, et al.
Veröffentlicht: (2023)
Energy Aware and Safe Path Planning for Unmanned Aircraft Systems
von: Gasche, Sebastian, et al.
Veröffentlicht: (2025)
von: Gasche, Sebastian, et al.
Veröffentlicht: (2025)
Machine Learning Algorithms for Improving Black Box Optimization Solvers
von: Kimiaei, Morteza, et al.
Veröffentlicht: (2025)
von: Kimiaei, Morteza, et al.
Veröffentlicht: (2025)
Convergence and turnpike properties of linear-quadratic mean field control problems with common noise
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2026)
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2026)
Analysis and Control of Input-Affine Dynamical Systems using Infinite-Dimensional Robust Counterparts
von: Miller, Jared, et al.
Veröffentlicht: (2021)
von: Miller, Jared, et al.
Veröffentlicht: (2021)
Monotone Causality in Opportunistically Stochastic Shortest Path Problems
von: Gaspard, Mallory E., et al.
Veröffentlicht: (2023)
von: Gaspard, Mallory E., et al.
Veröffentlicht: (2023)
Optimal power procurement for green cellular wireless networks under uncertainty and chance constraints
von: Rached, Nadhir Ben, et al.
Veröffentlicht: (2025)
von: Rached, Nadhir Ben, et al.
Veröffentlicht: (2025)
Technological schemes and control methods in the reconstruction of parallel gas pipeline systems under non-stationary conditions
von: Aliyev, Ilgar
Veröffentlicht: (2025)
von: Aliyev, Ilgar
Veröffentlicht: (2025)
Ergodicity and turnpike properties of linear-quadratic mean field control problems
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2025)
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2025)
Autonomous AI Agents for Real-Time Affordable Housing Site Selection: Multi-Objective Reinforcement Learning Under Regulatory Constraints
von: Imanov, Olaf Yunus Laitinen, et al.
Veröffentlicht: (2026)
von: Imanov, Olaf Yunus Laitinen, et al.
Veröffentlicht: (2026)
Second-Order $Λ$-Sets and Extensions to Non-Smooth, Hybrid, and Stochastic Optimal Control
von: Rashid, Mohammad H. M
Veröffentlicht: (2025)
von: Rashid, Mohammad H. M
Veröffentlicht: (2025)
The landscape of deterministic and stochastic optimal control problems: One-shot Optimization versus Dynamic Programming
von: Kim, Jihun, et al.
Veröffentlicht: (2024)
von: Kim, Jihun, et al.
Veröffentlicht: (2024)
An Inexact Trust-Region Method for Structured Nonsmooth Optimization with Application to Risk-Averse Stochastic Programming
von: Kouri, Drew P.
Veröffentlicht: (2026)
von: Kouri, Drew P.
Veröffentlicht: (2026)
Approximation of risk-averse optimal feedback control
von: Guth, Philipp A., et al.
Veröffentlicht: (2025)
von: Guth, Philipp A., et al.
Veröffentlicht: (2025)
A Dynamic-Growing Fuzzy-Neuro Controller, Application to a 3PSP Parallel Robot
von: Jalaeian-Farimani, Mohsen, et al.
Veröffentlicht: (2026)
von: Jalaeian-Farimani, Mohsen, et al.
Veröffentlicht: (2026)
Dual dynamic programming for stochastic programs over an infinite horizon
von: Ju, Caleb, et al.
Veröffentlicht: (2023)
von: Ju, Caleb, et al.
Veröffentlicht: (2023)
Tracking optimal feedback control under uncertain parameters
von: Guth, Philipp A., et al.
Veröffentlicht: (2024)
von: Guth, Philipp A., et al.
Veröffentlicht: (2024)
Non-Convex Global Optimization as an Optimal Stabilization Problem: Convergence Rates
von: Huang, Yuyang, et al.
Veröffentlicht: (2025)
von: Huang, Yuyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Efficient and Scalable Path-Planning Algorithms for Curvature Constrained Motion in the Hamilton-Jacobi Formulation
von: Parkinson, Christian, et al.
Veröffentlicht: (2023) -
Risk-averse optimization under distributional uncertainty with Rockafellian relaxation
von: Antil, Harbir, et al.
Veröffentlicht: (2026) -
Rockafellian Relaxation for PDE-Constrained Optimization with Distributional Uncertainty
von: Antil, Harbir, et al.
Veröffentlicht: (2024) -
Optimal Control of Unbounded Functional Stochastic Evolution Systems in Hilbert Spaces: Second-Order Path-dependent HJB Equation
von: Tang, Shanjian, et al.
Veröffentlicht: (2024) -
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
von: Zheng, Yaowei, et al.
Veröffentlicht: (2026)