Safety-Oriented Pruning and Interpretation of Reinforcement Learning Policies
Fuente:
arXiv
Saved in:
| Main Authors: | Gross, Dennis, Spieker, Helge |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Co-Activation Graph Analysis of Safety-Verified and Explainable Deep Reinforcement Learning Policies
by: Gross, Dennis, et al.
Published: (2025)
by: Gross, Dennis, et al.
Published: (2025)
Enhancing RL Safety with Counterfactual LLM Reasoning
by: Gross, Dennis, et al.
Published: (2024)
by: Gross, Dennis, et al.
Published: (2024)
Probabilistic Model Checking of Stochastic Reinforcement Learning Policies
by: Gross, Dennis, et al.
Published: (2024)
by: Gross, Dennis, et al.
Published: (2024)
Translating the Rashomon Effect to Sequential Decision-Making Tasks
by: Gross, Dennis, et al.
Published: (2025)
by: Gross, Dennis, et al.
Published: (2025)
Efficient Milling Quality Prediction with Explainable Machine Learning
by: Gross, Dennis, et al.
Published: (2024)
by: Gross, Dennis, et al.
Published: (2024)
Enhancing Manufacturing Quality Prediction Models through the Integration of Explainability Methods
by: Gross, Dennis, et al.
Published: (2024)
by: Gross, Dennis, et al.
Published: (2024)
Semi-supervised CAPP Transformer Learning via Pseudo-labeling
by: Gross, Dennis, et al.
Published: (2026)
by: Gross, Dennis, et al.
Published: (2026)
Turn-based Multi-Agent Reinforcement Learning Model Checking
by: Gross, Dennis
Published: (2025)
by: Gross, Dennis
Published: (2025)
Formally Verifying and Explaining Sepsis Treatment Policies with COOL-MC
by: Gross, Dennis
Published: (2026)
by: Gross, Dennis
Published: (2026)
COOL-MC: Verifying and Explaining RL Policies for Platelet Inventory Management
by: Gross, Dennis
Published: (2026)
by: Gross, Dennis
Published: (2026)
COOL-MC: Verifying and Explaining RL Policies for Multi-bridge Network Maintenance
by: Gross, Dennis
Published: (2026)
by: Gross, Dennis
Published: (2026)
Optimizing Interpretable Decision Tree Policies for Reinforcement Learning
by: Vos, Daniël, et al.
Published: (2024)
by: Vos, Daniël, et al.
Published: (2024)
Efficient Policy Evaluation with Safety Constraint for Reinforcement Learning
by: Chen, Claire, et al.
Published: (2024)
by: Chen, Claire, et al.
Published: (2024)
Understanding Annotator Safety Policy with Interpretability
by: Oesterling, Alex, et al.
Published: (2026)
by: Oesterling, Alex, et al.
Published: (2026)
Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning
by: Kohler, Hector, et al.
Published: (2024)
by: Kohler, Hector, et al.
Published: (2024)
Evaluating Interpretable Reinforcement Learning by Distilling Policies into Programs
by: Kohler, Hector, et al.
Published: (2025)
by: Kohler, Hector, et al.
Published: (2025)
Concurrent Learning of Policy and Unknown Safety Constraints in Reinforcement Learning
by: Yifru, Lunet, et al.
Published: (2024)
by: Yifru, Lunet, et al.
Published: (2024)
Latent Safety-Constrained Policy Approach for Safe Offline Reinforcement Learning
by: Koirala, Prajwal, et al.
Published: (2024)
by: Koirala, Prajwal, et al.
Published: (2024)
Towards Interpretable Reinforcement Learning with Constrained Normalizing Flow Policies
by: Rietz, Finn, et al.
Published: (2024)
by: Rietz, Finn, et al.
Published: (2024)
Reusable Test Suites for Reinforcement Learning
by: Betten, Jørn Eirik, et al.
Published: (2025)
by: Betten, Jørn Eirik, et al.
Published: (2025)
Verifying Memoryless Sequential Decision-making of Large Language Models
by: Gross, Dennis, et al.
Published: (2025)
by: Gross, Dennis, et al.
Published: (2025)
Bounded PCTL Model Checking of Large Language Model Outputs
by: Gross, Dennis, et al.
Published: (2025)
by: Gross, Dennis, et al.
Published: (2025)
From Explainability to Interpretability: Interpretable Policies in Reinforcement Learning Via Model Explanation
by: Li, Peilang, et al.
Published: (2025)
by: Li, Peilang, et al.
Published: (2025)
Three Pathways to Neurosymbolic Reinforcement Learning with Interpretable Model and Policy Networks
by: Graf, Peter, et al.
Published: (2024)
by: Graf, Peter, et al.
Published: (2024)
Prism: Policy Reuse via Interpretable Strategy Mapping in Reinforcement Learning
by: Pravetz, Thomas
Published: (2026)
by: Pravetz, Thomas
Published: (2026)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
by: Lanier, Michael, et al.
Published: (2024)
by: Lanier, Michael, et al.
Published: (2024)
Rashomon in the Streets: Explanation Ambiguity in Scene Understanding
by: Spieker, Helge, et al.
Published: (2025)
by: Spieker, Helge, et al.
Published: (2025)
Exploiting Hybrid Policy in Reinforcement Learning for Interpretable Temporal Logic Manipulation
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
Safety Representations for Safer Policy Learning
by: Mani, Kaustubh, et al.
Published: (2025)
by: Mani, Kaustubh, et al.
Published: (2025)
Interpretable Learning Dynamics in Unsupervised Reinforcement Learning
by: Pandey, Shashwat
Published: (2025)
by: Pandey, Shashwat
Published: (2025)
ObjectRL: An Object-Oriented Reinforcement Learning Codebase
by: Baykal, Gulcin, et al.
Published: (2025)
by: Baykal, Gulcin, et al.
Published: (2025)
Mechanistic Interpretability of Reinforcement Learning Agents
by: Trim, Tristan, et al.
Published: (2024)
by: Trim, Tristan, et al.
Published: (2024)
Interpreting Reinforcement Learning Agents with Susceptibilities
by: Elliott, Chris, et al.
Published: (2026)
by: Elliott, Chris, et al.
Published: (2026)
Pruning Cannot Hurt Robustness: Certified Trade-offs in Reinforcement Learning
by: Pedley, James, et al.
Published: (2025)
by: Pedley, James, et al.
Published: (2025)
The Impact of Quantization and Pruning on Deep Reinforcement Learning Models
by: Lu, Heng, et al.
Published: (2024)
by: Lu, Heng, et al.
Published: (2024)
On-Policy Policy Gradient Reinforcement Learning Without On-Policy Sampling
by: Corrado, Nicholas E., et al.
Published: (2023)
by: Corrado, Nicholas E., et al.
Published: (2023)
Policy Improvement Reinforcement Learning
by: Wang, Huaiyang, et al.
Published: (2026)
by: Wang, Huaiyang, et al.
Published: (2026)
Safety Optimized Reinforcement Learning via Multi-Objective Policy Optimization
by: Honari, Homayoun, et al.
Published: (2024)
by: Honari, Homayoun, et al.
Published: (2024)
Online Training and Pruning of Deep Reinforcement Learning Networks
by: Guenter, Valentin Frank Ingmar, et al.
Published: (2025)
by: Guenter, Valentin Frank Ingmar, et al.
Published: (2025)
Enhancing Safety in Reinforcement Learning with Human Feedback via Rectified Policy Optimization
by: Peng, Xiyue, et al.
Published: (2024)
by: Peng, Xiyue, et al.
Published: (2024)
Similar Items
-
Co-Activation Graph Analysis of Safety-Verified and Explainable Deep Reinforcement Learning Policies
by: Gross, Dennis, et al.
Published: (2025) -
Enhancing RL Safety with Counterfactual LLM Reasoning
by: Gross, Dennis, et al.
Published: (2024) -
Probabilistic Model Checking of Stochastic Reinforcement Learning Policies
by: Gross, Dennis, et al.
Published: (2024) -
Translating the Rashomon Effect to Sequential Decision-Making Tasks
by: Gross, Dennis, et al.
Published: (2025) -
Efficient Milling Quality Prediction with Explainable Machine Learning
by: Gross, Dennis, et al.
Published: (2024)