ProSh: Probabilistic Shielding for Model-free Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Court, Edwin Hamel-De le, Ohlmann, Gaspard, Belardinelli, Francesco |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Probabilistic Shielding for Safe Reinforcement Learning
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
Synthesis of Safety Specifications for Probabilistic Systems
by: Ohlmann, Gaspard, et al.
Published: (2025)
by: Ohlmann, Gaspard, et al.
Published: (2025)
Robust Shielding for Safe Reinforcement Learning
by: Court, Edwin Hamel-De le, et al.
Published: (2026)
by: Court, Edwin Hamel-De le, et al.
Published: (2026)
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
by: Goodall, Alexander W., et al.
Published: (2025)
by: Goodall, Alexander W., et al.
Published: (2025)
Leveraging Approximate Model-based Shielding for Probabilistic Safety Guarantees in Continuous Environments
by: Goodall, Alexander W., et al.
Published: (2024)
by: Goodall, Alexander W., et al.
Published: (2024)
Approximate Model-Based Shielding for Safe Reinforcement Learning
by: Goodall, Alexander W., et al.
Published: (2023)
by: Goodall, Alexander W., et al.
Published: (2023)
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
by: Goodall, Alexander W., et al.
Published: (2026)
by: Goodall, Alexander W., et al.
Published: (2026)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
by: Anisimov, Maksim, et al.
Published: (2026)
by: Anisimov, Maksim, et al.
Published: (2026)
Robust Probabilistic Shielding for Safe Offline Reinforcement Learning
by: Galesloot, Maris F. L., et al.
Published: (2026)
by: Galesloot, Maris F. L., et al.
Published: (2026)
Think Smart, Act SMARL! Analyzing Probabilistic Logic Shields for Multi-Agent Reinforcement Learning
by: Chatterji, Satchit, et al.
Published: (2024)
by: Chatterji, Satchit, et al.
Published: (2024)
FreSh: Frequency Shifting for Accelerated Neural Representation Learning
by: Kania, Adam, et al.
Published: (2024)
by: Kania, Adam, et al.
Published: (2024)
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
by: Banerjee, Arko, et al.
Published: (2024)
by: Banerjee, Arko, et al.
Published: (2024)
Explainable Reinforcement Learning for Formula One Race Strategy
by: Thomas, Devin, et al.
Published: (2025)
by: Thomas, Devin, et al.
Published: (2025)
Expressive Temporal Specifications for Reward Monitoring
by: Adalat, Omar, et al.
Published: (2025)
by: Adalat, Omar, et al.
Published: (2025)
Measuring Goal-Directedness
by: MacDermott, Matt, et al.
Published: (2024)
by: MacDermott, Matt, et al.
Published: (2024)
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
by: Bethell, Daniel, et al.
Published: (2024)
by: Bethell, Daniel, et al.
Published: (2024)
Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation
by: Malik, Gitesh
Published: (2026)
by: Malik, Gitesh
Published: (2026)
Probabilistic Curriculum Learning for Goal-Based Reinforcement Learning
by: Salt, Llewyn, et al.
Published: (2025)
by: Salt, Llewyn, et al.
Published: (2025)
Efficient Reinforcement Learning in Probabilistic Reward Machines
by: Lin, Xiaofeng, et al.
Published: (2024)
by: Lin, Xiaofeng, et al.
Published: (2024)
Probabilistic Constraint for Safety-Critical Reinforcement Learning
by: Chen, Weiqin, et al.
Published: (2023)
by: Chen, Weiqin, et al.
Published: (2023)
Adaptive Shielding for Safe Reinforcement Learning under Hidden-Parameter Dynamics Shifts
by: Kwon, Minjae, et al.
Published: (2025)
by: Kwon, Minjae, et al.
Published: (2025)
TriShGAN: Enhancing Sparsity and Robustness in Multivariate Time Series Counterfactuals Explanation
by: Ma, Hongnan, et al.
Published: (2025)
by: Ma, Hongnan, et al.
Published: (2025)
State-free Reinforcement Learning
by: Chen, Mingyu, et al.
Published: (2024)
by: Chen, Mingyu, et al.
Published: (2024)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
by: Schnitzer, Yannik, et al.
Published: (2026)
by: Schnitzer, Yannik, et al.
Published: (2026)
Identifiable Object-Centric Representation Learning via Probabilistic Slot Attention
by: Kori, Avinash, et al.
Published: (2024)
by: Kori, Avinash, et al.
Published: (2024)
Expressive Reward Synthesis with the Runtime Monitoring Language
by: Donnelly, Daniel, et al.
Published: (2025)
by: Donnelly, Daniel, et al.
Published: (2025)
Scale-free Adversarial Reinforcement Learning
by: Chen, Mingyu, et al.
Published: (2024)
by: Chen, Mingyu, et al.
Published: (2024)
ShIOEnv: A Command Evaluation Environment for Grammar-Constrained Synthesis and Execution Behavior Modeling
by: Ragsdale, Jarrod, et al.
Published: (2025)
by: Ragsdale, Jarrod, et al.
Published: (2025)
Predictive Safety Shield for Dyna-Q Reinforcement Learning
by: Pin, Jin, et al.
Published: (2025)
by: Pin, Jin, et al.
Published: (2025)
Compositional Shielding and Reinforcement Learning for Multi-Agent Systems
by: Brorholt, Asger Horn, et al.
Published: (2024)
by: Brorholt, Asger Horn, et al.
Published: (2024)
EXPLAIN, AGREE, LEARN: Scaling Learning for Neural Probabilistic Logic
by: Verreet, Victor, et al.
Published: (2024)
by: Verreet, Victor, et al.
Published: (2024)
Real-World Reinforcement Learning of Active Perception Behaviors
by: Hu, Edward S., et al.
Published: (2025)
by: Hu, Edward S., et al.
Published: (2025)
Adaptive GR(1) Specification Repair for Liveness-Preserving Shielding in Reinforcement Learning
by: Georgescu, Tiberiu-Andrei, et al.
Published: (2025)
by: Georgescu, Tiberiu-Andrei, et al.
Published: (2025)
ProRL: Effective Reinforcement Learning for Proactive Recommendation via Rectified Policy Gradient Estimation
by: Hou, Hongru, et al.
Published: (2026)
by: Hou, Hongru, et al.
Published: (2026)
Reinforcement Learning for Control with Probabilistic Stability Guarantee: A Finite-Sample Approach
by: Han, Minghao, et al.
Published: (2026)
by: Han, Minghao, et al.
Published: (2026)
Dual-Criterion Curriculum Learning: Application to Temporal Data
by: Abel, Gaspard, et al.
Published: (2026)
by: Abel, Gaspard, et al.
Published: (2026)
Probabilistic Subgoal Representations for Hierarchical Reinforcement learning
by: Wang, Vivienne Huiling, et al.
Published: (2024)
by: Wang, Vivienne Huiling, et al.
Published: (2024)
Shortcomings of LLMs for Low-Resource Translation: Retrieval and Understanding are Both the Problem
by: Court, Sara, et al.
Published: (2024)
by: Court, Sara, et al.
Published: (2024)
A Probabilistic Model Behind Self-Supervised Learning
by: Bizeul, Alice, et al.
Published: (2024)
by: Bizeul, Alice, et al.
Published: (2024)
Probabilistic Abduction for Visual Abstract Reasoning via Learning Rules in Vector-symbolic Architectures
by: Hersche, Michael, et al.
Published: (2024)
by: Hersche, Michael, et al.
Published: (2024)
Similar Items
-
Probabilistic Shielding for Safe Reinforcement Learning
by: Court, Edwin Hamel-De le, et al.
Published: (2025) -
Synthesis of Safety Specifications for Probabilistic Systems
by: Ohlmann, Gaspard, et al.
Published: (2025) -
Robust Shielding for Safe Reinforcement Learning
by: Court, Edwin Hamel-De le, et al.
Published: (2026) -
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
by: Goodall, Alexander W., et al.
Published: (2025) -
Leveraging Approximate Model-based Shielding for Probabilistic Safety Guarantees in Continuous Environments
by: Goodall, Alexander W., et al.
Published: (2024)