Gespeichert in:
| Hauptverfasser: | Gros, Sebastien, Zanon, Mario |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2019
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/1906.04034 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Safe Reinforcement Learning Using NMPC and Policy Gradients: Part I - Stochastic case
von: Gros, Sebastien, et al.
Veröffentlicht: (2019)
von: Gros, Sebastien, et al.
Veröffentlicht: (2019)
Safe Reinforcement Learning Using Robust MPC
von: Zanon, Mario, et al.
Veröffentlicht: (2019)
von: Zanon, Mario, et al.
Veröffentlicht: (2019)
Quasi-Newton Compatible Actor-Critic for Deterministic Policies
von: Kordabad, Arash Bahari, et al.
Veröffentlicht: (2025)
von: Kordabad, Arash Bahari, et al.
Veröffentlicht: (2025)
Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression
von: Song, Bowen, et al.
Veröffentlicht: (2025)
von: Song, Bowen, et al.
Veröffentlicht: (2025)
Offline Guarded Safe Reinforcement Learning for Medical Treatment Optimization Strategies
von: Yan, Runze, et al.
Veröffentlicht: (2025)
von: Yan, Runze, et al.
Veröffentlicht: (2025)
Personalized Dynamic Pricing Policy for Electric Vehicles: Reinforcement learning approach
von: Bae, Sangjun, et al.
Veröffentlicht: (2024)
von: Bae, Sangjun, et al.
Veröffentlicht: (2024)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
Economic Linear Quadratic MPC With Non-Unique Optimal Solutions
von: Zanon, Mario
Veröffentlicht: (2025)
von: Zanon, Mario
Veröffentlicht: (2025)
Rethinking Strict Dissipativity for Economic MPC
von: Zanon, Mario
Veröffentlicht: (2026)
von: Zanon, Mario
Veröffentlicht: (2026)
Cost-Matching Model Predictive Control for Efficient Reinforcement Learning in Humanoid Locomotion
von: Cai, Wenqi, et al.
Veröffentlicht: (2026)
von: Cai, Wenqi, et al.
Veröffentlicht: (2026)
MPC4RL -- A Software Package for Reinforcement Learning based on Model Predictive Control
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2025)
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2025)
Uncertainty Propagation under Residual Disturbances: A Smart-Home Case Study
von: Pan, Guanru, et al.
Veröffentlicht: (2026)
von: Pan, Guanru, et al.
Veröffentlicht: (2026)
Off-Policy Reinforcement Learning with Anytime Safety Guarantees via Robust Safe Gradient Flow
von: Mestres, Pol, et al.
Veröffentlicht: (2025)
von: Mestres, Pol, et al.
Veröffentlicht: (2025)
Optimization of the Model Predictive Control Meta-Parameters Through Reinforcement Learning
von: Bøhn, Eivind, et al.
Veröffentlicht: (2021)
von: Bøhn, Eivind, et al.
Veröffentlicht: (2021)
NMPC-Augmented Visual Navigation and Safe Learning Control for Large-Scale Mobile Robots
von: Shahna, Mehdi Heydari, et al.
Veröffentlicht: (2026)
von: Shahna, Mehdi Heydari, et al.
Veröffentlicht: (2026)
Stabilization of Strictly Pre-Dissipative Receding Horizon Linear Quadratic Control by Terminal Costs
von: Zanon, Mario, et al.
Veröffentlicht: (2025)
von: Zanon, Mario, et al.
Veröffentlicht: (2025)
Convergent NMPC-based Reinforcement Learning Using Deep Expected Sarsa and Nonlinear Temporal Difference Learning
von: Salaje, Amine, et al.
Veröffentlicht: (2025)
von: Salaje, Amine, et al.
Veröffentlicht: (2025)
Learning disturbance models for offset-free reference tracking
von: Krupa, Pablo, et al.
Veröffentlicht: (2023)
von: Krupa, Pablo, et al.
Veröffentlicht: (2023)
Optimality Conditions for Model Predictive Control: Rethinking Predictive Model Design
von: Anand, Akhil S, et al.
Veröffentlicht: (2024)
von: Anand, Akhil S, et al.
Veröffentlicht: (2024)
Data-Driven Domestic Flexible Demand: Observations from experiments in cold climate
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2024)
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2024)
Feasible Policy Iteration for Safe Reinforcement Learning
von: Yang, Yujie, et al.
Veröffentlicht: (2023)
von: Yang, Yujie, et al.
Veröffentlicht: (2023)
Optimization Solution Functions as Deterministic Policies for Offline Reinforcement Learning
von: Khattar, Vanshaj, et al.
Veröffentlicht: (2024)
von: Khattar, Vanshaj, et al.
Veröffentlicht: (2024)
Solving Reach-Avoid-Stay Problems Using Deep Deterministic Policy Gradients
von: Chenevert, Gabriel, et al.
Veröffentlicht: (2024)
von: Chenevert, Gabriel, et al.
Veröffentlicht: (2024)
Probabilistic reachable sets of stochastic nonlinear systems with contextual uncertainties
von: Shen, Xun, et al.
Veröffentlicht: (2024)
von: Shen, Xun, et al.
Veröffentlicht: (2024)
Computationally efficient Gauss-Newton reinforcement learning for model predictive control
von: Brandner, Dean, et al.
Veröffentlicht: (2025)
von: Brandner, Dean, et al.
Veröffentlicht: (2025)
Optimization-based Coordination of Traffic Lights and Automated Vehicles at Intersections
von: Dabiri, Azita, et al.
Veröffentlicht: (2025)
von: Dabiri, Azita, et al.
Veröffentlicht: (2025)
Decentralized Real-Time Iterations for Distributed NMPC
von: Stomberg, Gösta, et al.
Veröffentlicht: (2024)
von: Stomberg, Gösta, et al.
Veröffentlicht: (2024)
Economic Model Predictive Control as a Solution to Markov Decision Processes
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2024)
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2024)
RTI-NMPC for Control of Autonomous Vehicles Using Implicit Discretization Methods
von: Wagner, Matheus, et al.
Veröffentlicht: (2024)
von: Wagner, Matheus, et al.
Veröffentlicht: (2024)
Optimization-based Heuristic for Vehicle Dynamic Coordination in Mixed Traffic Intersections
von: Faris, Muhammad, et al.
Veröffentlicht: (2024)
von: Faris, Muhammad, et al.
Veröffentlicht: (2024)
Active Learning MPC Objective Functions from Preferences
von: Hasnaouy, Hasna El, et al.
Veröffentlicht: (2026)
von: Hasnaouy, Hasna El, et al.
Veröffentlicht: (2026)
Learning the MPC objective function from human preferences
von: Krupa, Pablo, et al.
Veröffentlicht: (2025)
von: Krupa, Pablo, et al.
Veröffentlicht: (2025)
On Piecewise Quadratic Terminal Costs for MPC
von: Mulagaleti, Sampath Kumar, et al.
Veröffentlicht: (2026)
von: Mulagaleti, Sampath Kumar, et al.
Veröffentlicht: (2026)
Progressive Smoothing for Motion Planning in Real-Time NMPC
von: Reiter, Rudolf, et al.
Veröffentlicht: (2024)
von: Reiter, Rudolf, et al.
Veröffentlicht: (2024)
Hierarchical Policy-Gradient Reinforcement Learning for Multi-Agent Shepherding Control of Non-Cohesive Targets
von: Covone, Stefano, et al.
Veröffentlicht: (2025)
von: Covone, Stefano, et al.
Veröffentlicht: (2025)
Data-driven Nonlinear Model Reduction using Koopman Theory: Integrated Control Form and NMPC Case Study
von: Schulze, Jan C., et al.
Veröffentlicht: (2024)
von: Schulze, Jan C., et al.
Veröffentlicht: (2024)
DiffOP: Reinforcement Learning of Optimization-Based Control Policies via Implicit Policy Gradients
von: Bian, Yuexin, et al.
Veröffentlicht: (2024)
von: Bian, Yuexin, et al.
Veröffentlicht: (2024)
Anytime Safe Reinforcement Learning
von: Mestres, Pol, et al.
Veröffentlicht: (2025)
von: Mestres, Pol, et al.
Veröffentlicht: (2025)
Point-to-Cloud NMPC with Smooth Avoidance Constraints
von: Ferreira, Brener G., et al.
Veröffentlicht: (2026)
von: Ferreira, Brener G., et al.
Veröffentlicht: (2026)
NMPC and Deep Learning-Based Vibration Control of Satellite Beam Antenna Dynamics Using PZT Actuators and Sensors
von: Kalaycioglu, Sean, et al.
Veröffentlicht: (2024)
von: Kalaycioglu, Sean, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Safe Reinforcement Learning Using NMPC and Policy Gradients: Part I - Stochastic case
von: Gros, Sebastien, et al.
Veröffentlicht: (2019) -
Safe Reinforcement Learning Using Robust MPC
von: Zanon, Mario, et al.
Veröffentlicht: (2019) -
Quasi-Newton Compatible Actor-Critic for Deterministic Policies
von: Kordabad, Arash Bahari, et al.
Veröffentlicht: (2025) -
Sample-Efficient Model-Free Policy Gradient Methods for Stochastic LQR via Robust Linear Regression
von: Song, Bowen, et al.
Veröffentlicht: (2025) -
Offline Guarded Safe Reinforcement Learning for Medical Treatment Optimization Strategies
von: Yan, Runze, et al.
Veröffentlicht: (2025)