Leveraging Approximate Model-based Shielding for Probabilistic Safety Guarantees in Continuous Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Goodall, Alexander W., Belardinelli, Francesco |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Approximate Model-Based Shielding for Safe Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
Probabilistic Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
ProSh: Probabilistic Shielding for Model-free Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2025)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2025)
Robust Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
Adaptive GR(1) Specification Repair for Liveness-Preserving Shielding in Reinforcement Learning
von: Georgescu, Tiberiu-Andrei, et al.
Veröffentlicht: (2025)
von: Georgescu, Tiberiu-Andrei, et al.
Veröffentlicht: (2025)
Rigorous Probabilistic Guarantees for Robust Counterfactual Explanations
von: Marzari, Luca, et al.
Veröffentlicht: (2024)
von: Marzari, Luca, et al.
Veröffentlicht: (2024)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
von: Anisimov, Maksim, et al.
Veröffentlicht: (2026)
Counterfactual Explanations with Probabilistic Guarantees on their Robustness to Model Change
von: Stępka, Ignacy, et al.
Veröffentlicht: (2024)
von: Stępka, Ignacy, et al.
Veröffentlicht: (2024)
Expressive Temporal Specifications for Reward Monitoring
von: Adalat, Omar, et al.
Veröffentlicht: (2025)
von: Adalat, Omar, et al.
Veröffentlicht: (2025)
Probabilistic Stability Guarantees for Feature Attributions
von: Jin, Helen, et al.
Veröffentlicht: (2025)
von: Jin, Helen, et al.
Veröffentlicht: (2025)
Robust Probabilistic Shielding for Safe Offline Reinforcement Learning
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
Shields to Guarantee Probabilistic Safety in MDPs
von: Heck, Linus, et al.
Veröffentlicht: (2026)
von: Heck, Linus, et al.
Veröffentlicht: (2026)
A Retention-Centric Framework for Continual Learning with Guaranteed Model Developmental Safety
von: Li, Gang, et al.
Veröffentlicht: (2024)
von: Li, Gang, et al.
Veröffentlicht: (2024)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
Measuring Goal-Directedness
von: MacDermott, Matt, et al.
Veröffentlicht: (2024)
von: MacDermott, Matt, et al.
Veröffentlicht: (2024)
Continuous Mixtures of Tractable Probabilistic Models
von: Correia, Alvaro H. C., et al.
Veröffentlicht: (2022)
von: Correia, Alvaro H. C., et al.
Veröffentlicht: (2022)
Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM
von: Kumarappan, Adarsh, et al.
Veröffentlicht: (2025)
von: Kumarappan, Adarsh, et al.
Veröffentlicht: (2025)
Enumerating Safe Regions in Deep Neural Networks with Provable Probabilistic Guarantees
von: Marzari, Luca, et al.
Veröffentlicht: (2023)
von: Marzari, Luca, et al.
Veröffentlicht: (2023)
Synthesis of Safety Specifications for Probabilistic Systems
von: Ohlmann, Gaspard, et al.
Veröffentlicht: (2025)
von: Ohlmann, Gaspard, et al.
Veröffentlicht: (2025)
On the Sparsifiability of Correlation Clustering: Approximation Guarantees under Edge Sampling
von: Shihab, Ibne Farabi, et al.
Veröffentlicht: (2026)
von: Shihab, Ibne Farabi, et al.
Veröffentlicht: (2026)
Reinforcement Learning for Control with Probabilistic Stability Guarantee: A Finite-Sample Approach
von: Han, Minghao, et al.
Veröffentlicht: (2026)
von: Han, Minghao, et al.
Veröffentlicht: (2026)
Circular Belief Propagation for Approximate Probabilistic Inference
von: Bouttier, Vincent, et al.
Veröffentlicht: (2024)
von: Bouttier, Vincent, et al.
Veröffentlicht: (2024)
Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation
von: Malik, Gitesh
Veröffentlicht: (2026)
von: Malik, Gitesh
Veröffentlicht: (2026)
Expressive Reward Synthesis with the Runtime Monitoring Language
von: Donnelly, Daniel, et al.
Veröffentlicht: (2025)
von: Donnelly, Daniel, et al.
Veröffentlicht: (2025)
Think Smart, Act SMARL! Analyzing Probabilistic Logic Shields for Multi-Agent Reinforcement Learning
von: Chatterji, Satchit, et al.
Veröffentlicht: (2024)
von: Chatterji, Satchit, et al.
Veröffentlicht: (2024)
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
Probabilistic Constraint for Safety-Critical Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
HardNet: Hard-Constrained Neural Networks with Universal Approximation Guarantees
von: Min, Youngjae, et al.
Veröffentlicht: (2024)
von: Min, Youngjae, et al.
Veröffentlicht: (2024)
Conformal Safety Shielding for Imperfect-Perception Agents
von: Scarbro, William, et al.
Veröffentlicht: (2025)
von: Scarbro, William, et al.
Veröffentlicht: (2025)
Neural Network Approximators for Marginal MAP in Probabilistic Circuits
von: Arya, Shivvrat, et al.
Veröffentlicht: (2024)
von: Arya, Shivvrat, et al.
Veröffentlicht: (2024)
Scaling Continuous Latent Variable Models as Probabilistic Integral Circuits
von: Gala, Gennaro, et al.
Veröffentlicht: (2024)
von: Gala, Gennaro, et al.
Veröffentlicht: (2024)
The Forgotten Shield: Safety Grafting in Parameter-Space for Medical MLLMs
von: Zhao, Jiale, et al.
Veröffentlicht: (2025)
von: Zhao, Jiale, et al.
Veröffentlicht: (2025)
Active Level Set Estimation for Continuous Search Space with Theoretical Guarantee
von: Ngo, Giang, et al.
Veröffentlicht: (2024)
von: Ngo, Giang, et al.
Veröffentlicht: (2024)
ChronosAD: Leveraging Time Series Foundation Models for Accurate Anomaly Detection
von: Khan, Uzair, et al.
Veröffentlicht: (2026)
von: Khan, Uzair, et al.
Veröffentlicht: (2026)
StyleShield: Exposing the Fragility of AIGC Detectors through Continuous Controllable Style Transfer
von: Zheng, Guantian
Veröffentlicht: (2026)
von: Zheng, Guantian
Veröffentlicht: (2026)
Counterfactual Probabilistic Diffusion with Expert Models
von: Mu, Wenhao, et al.
Veröffentlicht: (2025)
von: Mu, Wenhao, et al.
Veröffentlicht: (2025)
Predictive Safety Shield for Dyna-Q Reinforcement Learning
von: Pin, Jin, et al.
Veröffentlicht: (2025)
von: Pin, Jin, et al.
Veröffentlicht: (2025)
SDPM: Survival Diffusion Probabilistic Model for Continuous-Time Survival Analysis
von: Kirpichenko, Stanislav R., et al.
Veröffentlicht: (2026)
von: Kirpichenko, Stanislav R., et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Approximate Model-Based Shielding for Safe Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023) -
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026) -
Probabilistic Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025) -
ProSh: Probabilistic Shielding for Model-free Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025) -
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2025)