Safe Reinforcement Learning using Action Projection: Safeguard the Policy or the Environment?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Markgraf, Hannah, Sawant, Shambhuraj, Krasowski, Hanna, Schäfer, Lukas, Gros, Sebastien, Althoff, Matthias |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Provable Traffic Rule Compliance in Safe Reinforcement Learning on the Open Sea
von: Krasowski, Hanna, et al.
Veröffentlicht: (2024)
von: Krasowski, Hanna, et al.
Veröffentlicht: (2024)
Leveraging Analytic Gradients in Provably Safe Reinforcement Learning
von: Walter, Tim, et al.
Veröffentlicht: (2025)
von: Walter, Tim, et al.
Veröffentlicht: (2025)
Direct transfer of optimized controllers to similar systems using dimensionless MPC
von: Hromatko, Josip Kir, et al.
Veröffentlicht: (2025)
von: Hromatko, Josip Kir, et al.
Veröffentlicht: (2025)
All AI Models are Wrong, but Some are Optimal
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
CommonPower: A Framework for Safe Data-Driven Smart Grid Control
von: Eichelbeck, Michael, et al.
Veröffentlicht: (2024)
von: Eichelbeck, Michael, et al.
Veröffentlicht: (2024)
Excluding the Irrelevant: Focusing Reinforcement Learning through Continuous Action Masking
von: Stolz, Roland, et al.
Veröffentlicht: (2024)
von: Stolz, Roland, et al.
Veröffentlicht: (2024)
PyTupli: A Scalable Infrastructure for Collaborative Offline Reinforcement Learning Projects
von: Markgraf, Hannah, et al.
Veröffentlicht: (2025)
von: Markgraf, Hannah, et al.
Veröffentlicht: (2025)
Closing the Sim2Real Performance Gap in RL
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
Contingency-constrained economic dispatch with safe reinforcement learning
von: Eichelbeck, Michael, et al.
Veröffentlicht: (2022)
von: Eichelbeck, Michael, et al.
Veröffentlicht: (2022)
Falsification-driven reinforcement learning for maritime motion planning
von: Müller, Marlon, et al.
Veröffentlicht: (2025)
von: Müller, Marlon, et al.
Veröffentlicht: (2025)
Predictive Safety Shield for Dyna-Q Reinforcement Learning
von: Pin, Jin, et al.
Veröffentlicht: (2025)
von: Pin, Jin, et al.
Veröffentlicht: (2025)
Improving Stochastic Action-Constrained Reinforcement Learning via Truncated Distributions
von: Stolz, Roland, et al.
Veröffentlicht: (2025)
von: Stolz, Roland, et al.
Veröffentlicht: (2025)
Stepping Out of the Shadows: Reinforcement Learning in Shadow Mode
von: Gassert, Philipp, et al.
Veröffentlicht: (2024)
von: Gassert, Philipp, et al.
Veröffentlicht: (2024)
Data-Driven Predictive Control and MPC: Do we achieve optimality?
von: Anand, Akhil S, et al.
Veröffentlicht: (2024)
von: Anand, Akhil S, et al.
Veröffentlicht: (2024)
Economic Model Predictive Control as a Solution to Markov Decision Processes
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2024)
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2024)
Offline Guarded Safe Reinforcement Learning for Medical Treatment Optimization Strategies
von: Yan, Runze, et al.
Veröffentlicht: (2025)
von: Yan, Runze, et al.
Veröffentlicht: (2025)
Learning to Drive by Imitating Surrounding Vehicles
von: Sonmez, Yasin, et al.
Veröffentlicht: (2025)
von: Sonmez, Yasin, et al.
Veröffentlicht: (2025)
Training Verifiably Robust Agents Using Set-Based Reinforcement Learning
von: Wendl, Manuel, et al.
Veröffentlicht: (2024)
von: Wendl, Manuel, et al.
Veröffentlicht: (2024)
Progressive Safeguards for Safe and Model-Agnostic Reinforcement Learning
von: Omi, Nabil, et al.
Veröffentlicht: (2024)
von: Omi, Nabil, et al.
Veröffentlicht: (2024)
Out of the Shadows: Exploring a Latent Space for Neural Network Verification
von: Koller, Lukas, et al.
Veröffentlicht: (2025)
von: Koller, Lukas, et al.
Veröffentlicht: (2025)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
Synthesis of Model Predictive Control and Reinforcement Learning: Survey and Classification
von: Reiter, Rudolf, et al.
Veröffentlicht: (2025)
von: Reiter, Rudolf, et al.
Veröffentlicht: (2025)
On the Identifiability of Latent Action Policies
von: Lachapelle, Sébastien
Veröffentlicht: (2025)
von: Lachapelle, Sébastien
Veröffentlicht: (2025)
Policy Bifurcation in Safe Reinforcement Learning
von: Zou, Wenjun, et al.
Veröffentlicht: (2024)
von: Zou, Wenjun, et al.
Veröffentlicht: (2024)
On-Policy Policy Gradient Reinforcement Learning Without On-Policy Sampling
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2023)
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2023)
Using Offline Data to Speed Up Reinforcement Learning in Procedurally Generated Environments
von: Andres, Alain, et al.
Veröffentlicht: (2023)
von: Andres, Alain, et al.
Veröffentlicht: (2023)
Safe Continual Reinforcement Learning in Non-stationary Environments
von: Coursey, Austin, et al.
Veröffentlicht: (2026)
von: Coursey, Austin, et al.
Veröffentlicht: (2026)
Action Mapping for Reinforcement Learning in Continuous Environments with Constraints
von: Theile, Mirco, et al.
Veröffentlicht: (2024)
von: Theile, Mirco, et al.
Veröffentlicht: (2024)
Optimization of the Model Predictive Control Meta-Parameters Through Reinforcement Learning
von: Bøhn, Eivind, et al.
Veröffentlicht: (2021)
von: Bøhn, Eivind, et al.
Veröffentlicht: (2021)
Automata-Conditioned Cooperative Multi-Agent Reinforcement Learning
von: Yalcinkaya, Beyazit, et al.
Veröffentlicht: (2025)
von: Yalcinkaya, Beyazit, et al.
Veröffentlicht: (2025)
Off-Policy Primal-Dual Safe Reinforcement Learning
von: Wu, Zifan, et al.
Veröffentlicht: (2024)
von: Wu, Zifan, et al.
Veröffentlicht: (2024)
Extreme Value Policy Optimization for Safe Reinforcement Learning
von: Gao, Shiqing, et al.
Veröffentlicht: (2026)
von: Gao, Shiqing, et al.
Veröffentlicht: (2026)
Fully Automatic Neural Network Reduction for Formal Verification
von: Ladner, Tobias, et al.
Veröffentlicht: (2023)
von: Ladner, Tobias, et al.
Veröffentlicht: (2023)
Set-Based Training for Neural Network Verification
von: Koller, Lukas, et al.
Veröffentlicht: (2024)
von: Koller, Lukas, et al.
Veröffentlicht: (2024)
Feasible Policy Iteration for Safe Reinforcement Learning
von: Yang, Yujie, et al.
Veröffentlicht: (2023)
von: Yang, Yujie, et al.
Veröffentlicht: (2023)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
Intelligent Sailing Model for Open Sea Navigation
von: Krasowski, Hanna, et al.
Veröffentlicht: (2025)
von: Krasowski, Hanna, et al.
Veröffentlicht: (2025)
Multi-Horizon Representations with Hierarchical Forward Models for Reinforcement Learning
von: McInroe, Trevor, et al.
Veröffentlicht: (2022)
von: McInroe, Trevor, et al.
Veröffentlicht: (2022)
Towards Safe Reinforcement Learning Using NMPC and Policy Gradients: Part I - Stochastic case
von: Gros, Sebastien, et al.
Veröffentlicht: (2019)
von: Gros, Sebastien, et al.
Veröffentlicht: (2019)
Towards Safe Reinforcement Learning Using NMPC and Policy Gradients: Part II - Deterministic Case
von: Gros, Sebastien, et al.
Veröffentlicht: (2019)
von: Gros, Sebastien, et al.
Veröffentlicht: (2019)
Ähnliche Einträge
-
Provable Traffic Rule Compliance in Safe Reinforcement Learning on the Open Sea
von: Krasowski, Hanna, et al.
Veröffentlicht: (2024) -
Leveraging Analytic Gradients in Provably Safe Reinforcement Learning
von: Walter, Tim, et al.
Veröffentlicht: (2025) -
Direct transfer of optimized controllers to similar systems using dimensionless MPC
von: Hromatko, Josip Kir, et al.
Veröffentlicht: (2025) -
All AI Models are Wrong, but Some are Optimal
von: Anand, Akhil S, et al.
Veröffentlicht: (2025) -
CommonPower: A Framework for Safe Data-Driven Smart Grid Control
von: Eichelbeck, Michael, et al.
Veröffentlicht: (2024)