Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
Fuente:
arXiv
Saved in:
| Main Authors: | Bethell, Daniel, Gerasimou, Simos, Calinescu, Radu, Imrie, Calum |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Navigate Under Imperfect Perception: Conformalised Segmentation for Safe Reinforcement Learning
by: Bethell, Daniel, et al.
Published: (2025)
by: Bethell, Daniel, et al.
Published: (2025)
Safe But Not Sorry: Reducing Over-Conservatism in Safety Critics via Uncertainty-Aware Modulation
by: Bethell, Daniel, et al.
Published: (2025)
by: Bethell, Daniel, et al.
Published: (2025)
Robust Uncertainty Quantification Using Conformalised Monte Carlo Prediction
by: Bethell, Daniel, et al.
Published: (2023)
by: Bethell, Daniel, et al.
Published: (2023)
Robust Adversarial Quantification via Conflict-Aware Evidential Deep Learning
by: Barker, Charmaine, et al.
Published: (2025)
by: Barker, Charmaine, et al.
Published: (2025)
Conformal Safety Shielding for Imperfect-Perception Agents
by: Scarbro, William, et al.
Published: (2025)
by: Scarbro, William, et al.
Published: (2025)
Guided Uncertainty Learning Using a Post-Hoc Evidential Meta-Model
by: Barker, Charmaine, et al.
Published: (2025)
by: Barker, Charmaine, et al.
Published: (2025)
DeepKnowledge: Generalisation-Driven Deep Learning Testing
by: Missaoui, Sondess, et al.
Published: (2024)
by: Missaoui, Sondess, et al.
Published: (2024)
Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs
by: Vázquez, Gricel, et al.
Published: (2026)
by: Vázquez, Gricel, et al.
Published: (2026)
Probabilistic Shielding for Safe Reinforcement Learning
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
Adaptive Shielding for Safe Reinforcement Learning under Hidden-Parameter Dynamics Shifts
by: Kwon, Minjae, et al.
Published: (2025)
by: Kwon, Minjae, et al.
Published: (2025)
Assuring the Safety of Reinforcement Learning Components: AMLAS-RL
by: Imrie, Calum Corrie, et al.
Published: (2025)
by: Imrie, Calum Corrie, et al.
Published: (2025)
Robust Probabilistic Shielding for Safe Offline Reinforcement Learning
by: Galesloot, Maris F. L., et al.
Published: (2026)
by: Galesloot, Maris F. L., et al.
Published: (2026)
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
by: Banerjee, Arko, et al.
Published: (2024)
by: Banerjee, Arko, et al.
Published: (2024)
Robust Shielding for Safe Reinforcement Learning
by: Court, Edwin Hamel-De le, et al.
Published: (2026)
by: Court, Edwin Hamel-De le, et al.
Published: (2026)
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
by: Goodall, Alexander W., et al.
Published: (2026)
by: Goodall, Alexander W., et al.
Published: (2026)
Approximate Model-Based Shielding for Safe Reinforcement Learning
by: Goodall, Alexander W., et al.
Published: (2023)
by: Goodall, Alexander W., et al.
Published: (2023)
Accelerating Policy Synthesis in Large-Scale MDPs via Hierarchical Adaptive Refinement
by: Evangelidis, Alexandros, et al.
Published: (2025)
by: Evangelidis, Alexandros, et al.
Published: (2025)
Learning Fairer Representations with FairVIC
by: Barker, Charmaine, et al.
Published: (2024)
by: Barker, Charmaine, et al.
Published: (2024)
Is Prior-Free Black-Box Non-Stationary Reinforcement Learning Feasible?
by: Gerogiannis, Argyrios, et al.
Published: (2024)
by: Gerogiannis, Argyrios, et al.
Published: (2024)
SEBA: Sample-Efficient Black-Box Attacks on Visual Reinforcement Learning
by: Huang, Tairan, et al.
Published: (2025)
by: Huang, Tairan, et al.
Published: (2025)
Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning
by: Chemingui, Yassine, et al.
Published: (2024)
by: Chemingui, Yassine, et al.
Published: (2024)
Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning
by: Chen, Rufeng, et al.
Published: (2026)
by: Chen, Rufeng, et al.
Published: (2026)
Explaining the Behavior of Black-Box Prediction Algorithms with Causal Learning
by: Sani, Numair, et al.
Published: (2020)
by: Sani, Numair, et al.
Published: (2020)
Optimizing the Unknown: Black Box Bayesian Optimization with Energy-Based Model and Reinforcement Learning
by: Miao, Ruiyao, et al.
Published: (2025)
by: Miao, Ruiyao, et al.
Published: (2025)
Learning Surrogates for Offline Black-Box Optimization via Gradient Matching
by: Hoang, Minh, et al.
Published: (2025)
by: Hoang, Minh, et al.
Published: (2025)
Adaptive Primal-Dual Method for Safe Reinforcement Learning
by: Chen, Weiqin, et al.
Published: (2024)
by: Chen, Weiqin, et al.
Published: (2024)
Reinforced In-Context Black-Box Optimization
by: Song, Lei, et al.
Published: (2024)
by: Song, Lei, et al.
Published: (2024)
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
by: Rajendran, Prajit T, et al.
Published: (2026)
by: Rajendran, Prajit T, et al.
Published: (2026)
ProSh: Probabilistic Shielding for Model-free Reinforcement Learning
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
by: Court, Edwin Hamel-De le, et al.
Published: (2025)
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Reinforcement Learning by Guided Safe Exploration
by: Yang, Qisong, et al.
Published: (2023)
by: Yang, Qisong, et al.
Published: (2023)
Towards Fast Safe Online Reinforcement Learning via Policy Finetuning
by: Chen, Keru, et al.
Published: (2024)
by: Chen, Keru, et al.
Published: (2024)
BlackBoxToBlueprint: Extracting Interpretable Logic from Legacy Systems using Reinforcement Learning and Counterfactual Analysis
by: Rathore, Vidhi
Published: (2025)
by: Rathore, Vidhi
Published: (2025)
Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk
by: Ying, Chengyang, et al.
Published: (2022)
by: Ying, Chengyang, et al.
Published: (2022)
Self-Improving Safety Performance of Reinforcement Learning Based Driving with Black-Box Verification Algorithms
by: Dagdanov, Resul, et al.
Published: (2022)
by: Dagdanov, Resul, et al.
Published: (2022)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
by: Anisimov, Maksim, et al.
Published: (2026)
by: Anisimov, Maksim, et al.
Published: (2026)
Trust Regions for Explanations via Black-Box Probabilistic Certification
by: Dhurandhar, Amit, et al.
Published: (2024)
by: Dhurandhar, Amit, et al.
Published: (2024)
Adaptive Action Duration with Contextual Bandits for Deep Reinforcement Learning in Dynamic Environments
by: Verma, Abhishek, et al.
Published: (2025)
by: Verma, Abhishek, et al.
Published: (2025)
Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation
by: Malik, Gitesh
Published: (2026)
by: Malik, Gitesh
Published: (2026)
Black-Box Optimization From Small Offline Datasets via Meta Learning with Synthetic Tasks
by: Fadhel, Azza, et al.
Published: (2026)
by: Fadhel, Azza, et al.
Published: (2026)
Similar Items
-
Learning to Navigate Under Imperfect Perception: Conformalised Segmentation for Safe Reinforcement Learning
by: Bethell, Daniel, et al.
Published: (2025) -
Safe But Not Sorry: Reducing Over-Conservatism in Safety Critics via Uncertainty-Aware Modulation
by: Bethell, Daniel, et al.
Published: (2025) -
Robust Uncertainty Quantification Using Conformalised Monte Carlo Prediction
by: Bethell, Daniel, et al.
Published: (2023) -
Robust Adversarial Quantification via Conflict-Aware Evidential Deep Learning
by: Barker, Charmaine, et al.
Published: (2025) -
Conformal Safety Shielding for Imperfect-Perception Agents
by: Scarbro, William, et al.
Published: (2025)