Probabilistic Shielding for Safe Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Court, Edwin Hamel-De le, Belardinelli, Francesco, Goodall, Alexander W. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ProSh: Probabilistic Shielding for Model-free Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2025)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2025)
Robust Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
Approximate Model-Based Shielding for Safe Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
Leveraging Approximate Model-based Shielding for Probabilistic Safety Guarantees in Continuous Environments
von: Goodall, Alexander W., et al.
Veröffentlicht: (2024)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2024)
Synthesis of Safety Specifications for Probabilistic Systems
von: Ohlmann, Gaspard, et al.
Veröffentlicht: (2025)
von: Ohlmann, Gaspard, et al.
Veröffentlicht: (2025)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
Robust Probabilistic Shielding for Safe Offline Reinforcement Learning
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
Adaptive GR(1) Specification Repair for Liveness-Preserving Shielding in Reinforcement Learning
von: Georgescu, Tiberiu-Andrei, et al.
Veröffentlicht: (2025)
von: Georgescu, Tiberiu-Andrei, et al.
Veröffentlicht: (2025)
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
von: Banerjee, Arko, et al.
Veröffentlicht: (2024)
von: Banerjee, Arko, et al.
Veröffentlicht: (2024)
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
Adaptive Shielding for Safe Reinforcement Learning under Hidden-Parameter Dynamics Shifts
von: Kwon, Minjae, et al.
Veröffentlicht: (2025)
von: Kwon, Minjae, et al.
Veröffentlicht: (2025)
Think Smart, Act SMARL! Analyzing Probabilistic Logic Shields for Multi-Agent Reinforcement Learning
von: Chatterji, Satchit, et al.
Veröffentlicht: (2024)
von: Chatterji, Satchit, et al.
Veröffentlicht: (2024)
Explainable Reinforcement Learning for Formula One Race Strategy
von: Thomas, Devin, et al.
Veröffentlicht: (2025)
von: Thomas, Devin, et al.
Veröffentlicht: (2025)
Expressive Temporal Specifications for Reward Monitoring
von: Adalat, Omar, et al.
Veröffentlicht: (2025)
von: Adalat, Omar, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning for Real-World Engine Control
von: Bedei, Julian, et al.
Veröffentlicht: (2025)
von: Bedei, Julian, et al.
Veröffentlicht: (2025)
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
von: Rajendran, Prajit T, et al.
Veröffentlicht: (2026)
von: Rajendran, Prajit T, et al.
Veröffentlicht: (2026)
Reinforcement Learning by Guided Safe Exploration
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
Measuring Goal-Directedness
von: MacDermott, Matt, et al.
Veröffentlicht: (2024)
von: MacDermott, Matt, et al.
Veröffentlicht: (2024)
Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation
von: Malik, Gitesh
Veröffentlicht: (2026)
von: Malik, Gitesh
Veröffentlicht: (2026)
Probabilistic Curriculum Learning for Goal-Based Reinforcement Learning
von: Salt, Llewyn, et al.
Veröffentlicht: (2025)
von: Salt, Llewyn, et al.
Veröffentlicht: (2025)
Online Optimization for Offline Safe Reinforcement Learning
von: Chemingui, Yassine, et al.
Veröffentlicht: (2025)
von: Chemingui, Yassine, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Learning in Probabilistic Reward Machines
von: Lin, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Lin, Xiaofeng, et al.
Veröffentlicht: (2024)
Probabilistic Constraint for Safety-Critical Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning
von: Zhao, Weiye, et al.
Veröffentlicht: (2024)
von: Zhao, Weiye, et al.
Veröffentlicht: (2024)
Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
von: Ji, Jiaming, et al.
Veröffentlicht: (2025)
von: Ji, Jiaming, et al.
Veröffentlicht: (2025)
Enhance Exploration in Safe Reinforcement Learning with Contrastive Representation Learning
von: Doan, Duc Kien, et al.
Veröffentlicht: (2025)
von: Doan, Duc Kien, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning with Learned Non-Markovian Safety Constraints
von: Low, Siow Meng, et al.
Veröffentlicht: (2024)
von: Low, Siow Meng, et al.
Veröffentlicht: (2024)
Sampling-Based Safe Reinforcement Learning
von: Vignola, Luca, et al.
Veröffentlicht: (2026)
von: Vignola, Luca, et al.
Veröffentlicht: (2026)
PNAct: Crafting Backdoor Attacks in Safe Reinforcement Learning
von: Guo, Weiran, et al.
Veröffentlicht: (2025)
von: Guo, Weiran, et al.
Veröffentlicht: (2025)
Skill-based Safe Reinforcement Learning with Risk Planning
von: Zhang, Hanping, et al.
Veröffentlicht: (2025)
von: Zhang, Hanping, et al.
Veröffentlicht: (2025)
Offline Safe Reinforcement Learning Using Trajectory Classification
von: Gong, Ze, et al.
Veröffentlicht: (2024)
von: Gong, Ze, et al.
Veröffentlicht: (2024)
A Survey of Constraint Formulations in Safe Reinforcement Learning
von: Wachi, Akifumi, et al.
Veröffentlicht: (2024)
von: Wachi, Akifumi, et al.
Veröffentlicht: (2024)
Spectral-Risk Safe Reinforcement Learning with Convergence Guarantees
von: Kim, Dohyeong, et al.
Veröffentlicht: (2024)
von: Kim, Dohyeong, et al.
Veröffentlicht: (2024)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
Identifiable Object-Centric Representation Learning via Probabilistic Slot Attention
von: Kori, Avinash, et al.
Veröffentlicht: (2024)
von: Kori, Avinash, et al.
Veröffentlicht: (2024)
GUARD: A Safe Reinforcement Learning Benchmark
von: Zhao, Weiye, et al.
Veröffentlicht: (2023)
von: Zhao, Weiye, et al.
Veröffentlicht: (2023)
Expressive Reward Synthesis with the Runtime Monitoring Language
von: Donnelly, Daniel, et al.
Veröffentlicht: (2025)
von: Donnelly, Daniel, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ProSh: Probabilistic Shielding for Model-free Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025) -
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2025) -
Robust Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026) -
Approximate Model-Based Shielding for Safe Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023) -
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)