Evaluating Interpretable Reinforcement Learning by Distilling Policies into Programs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kohler, Hector, Delfosse, Quentin, Radji, Waris, Akrour, Riad, Preux, Philippe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning
von: Kohler, Hector, et al.
Veröffentlicht: (2024)
von: Kohler, Hector, et al.
Veröffentlicht: (2024)
Limits of Actor-Critic Algorithms for Decision Tree Policies Learning in IBMDPs
von: Kohler, Hector, et al.
Veröffentlicht: (2023)
von: Kohler, Hector, et al.
Veröffentlicht: (2023)
Towards a Research Community in Interpretable Reinforcement Learning: the InterpPol Workshop
von: Kohler, Hector, et al.
Veröffentlicht: (2024)
von: Kohler, Hector, et al.
Veröffentlicht: (2024)
Breiman meets Bellman: Non-Greedy Decision Trees with MDPs
von: Kohler, Hector, et al.
Veröffentlicht: (2023)
von: Kohler, Hector, et al.
Veröffentlicht: (2023)
When (and How) to Trust the Expert: Diagnosing Query-Time Expert-Guided Reinforcement Learning
von: Berthelot, Yann, et al.
Veröffentlicht: (2026)
von: Berthelot, Yann, et al.
Veröffentlicht: (2026)
PB$^2$: Preference Space Exploration via Population-Based Methods in Preference-Based Reinforcement Learning
von: Driss, Brahim, et al.
Veröffentlicht: (2025)
von: Driss, Brahim, et al.
Veröffentlicht: (2025)
Intrinsic-Energy Joint Embedding Predictive Architectures Induce Quasimetric Spaces
von: Kobanda, Anthony, et al.
Veröffentlicht: (2026)
von: Kobanda, Anthony, et al.
Veröffentlicht: (2026)
StaQ it! Growing neural networks for Policy Mirror Descent
von: Shilova, Alena, et al.
Veröffentlicht: (2025)
von: Shilova, Alena, et al.
Veröffentlicht: (2025)
Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAX
von: Radji, Waris, et al.
Veröffentlicht: (2025)
von: Radji, Waris, et al.
Veröffentlicht: (2025)
HackAtari: Atari Learning Environments for Robust and Continual Reinforcement Learning
von: Delfosse, Quentin, et al.
Veröffentlicht: (2024)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2024)
GRAIL: Autonomous Concept Grounding for Neuro-Symbolic Reinforcement Learning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2026)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2026)
IDEQ: an improved diffusion model for the TSP
von: Basson, Mickael, et al.
Veröffentlicht: (2024)
von: Basson, Mickael, et al.
Veröffentlicht: (2024)
BlendRL: A Framework for Merging Symbolic and Neural Policy Learning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning via Object-Centric Attention
von: Blüml, Jannis, et al.
Veröffentlicht: (2025)
von: Blüml, Jannis, et al.
Veröffentlicht: (2025)
Boosting deep Reinforcement Learning using pretraining with Logical Options
von: Ye, Zihan, et al.
Veröffentlicht: (2026)
von: Ye, Zihan, et al.
Veröffentlicht: (2026)
OCAtari: Object-Centric Atari 2600 Reinforcement Learning Environments
von: Delfosse, Quentin, et al.
Veröffentlicht: (2023)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2023)
Distilling Reinforcement Learning Policies for Interpretable Robot Locomotion: Gradient Boosting Machines and Symbolic Regression
von: Acero, Fernando, et al.
Veröffentlicht: (2024)
von: Acero, Fernando, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning Agents are not even close to Human Intelligence
von: Delfosse, Quentin, et al.
Veröffentlicht: (2025)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2025)
Interpretable end-to-end Neurosymbolic Reinforcement Learning agents
von: Grandien, Nils, et al.
Veröffentlicht: (2024)
von: Grandien, Nils, et al.
Veröffentlicht: (2024)
Pix2Code: Learning to Compose Neural Visual Concepts as Programs
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
How Hard is it to Confuse a World Model?
von: Radji, Waris, et al.
Veröffentlicht: (2025)
von: Radji, Waris, et al.
Veröffentlicht: (2025)
The Confusing Instance Principle for Online Linear Quadratic Control
von: Radji, Waris, et al.
Veröffentlicht: (2025)
von: Radji, Waris, et al.
Veröffentlicht: (2025)
How Ensembles of Distilled Policies Improve Generalisation in Reinforcement Learning
von: Weltevrede, Max, et al.
Veröffentlicht: (2025)
von: Weltevrede, Max, et al.
Veröffentlicht: (2025)
Interpretable Policy Distillation for Power Grid Topology Control
von: Dmitruka, Aleksandra, et al.
Veröffentlicht: (2026)
von: Dmitruka, Aleksandra, et al.
Veröffentlicht: (2026)
Towards Interpretable Reinforcement Learning with Constrained Normalizing Flow Policies
von: Rietz, Finn, et al.
Veröffentlicht: (2024)
von: Rietz, Finn, et al.
Veröffentlicht: (2024)
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
von: Xu, Charles, et al.
Veröffentlicht: (2024)
von: Xu, Charles, et al.
Veröffentlicht: (2024)
Prism: Policy Reuse via Interpretable Strategy Mapping in Reinforcement Learning
von: Pravetz, Thomas
Veröffentlicht: (2026)
von: Pravetz, Thomas
Veröffentlicht: (2026)
Three Pathways to Neurosymbolic Reinforcement Learning with Interpretable Model and Policy Networks
von: Graf, Peter, et al.
Veröffentlicht: (2024)
von: Graf, Peter, et al.
Veröffentlicht: (2024)
IPD: Boosting Sequential Policy with Imaginary Planning Distillation in Offline Reinforcement Learning
von: Qin, Yihao, et al.
Veröffentlicht: (2026)
von: Qin, Yihao, et al.
Veröffentlicht: (2026)
Augmented Bayesian Policy Search
von: Kallel, Mahdi, et al.
Veröffentlicht: (2024)
von: Kallel, Mahdi, et al.
Veröffentlicht: (2024)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
von: Lanier, Michael, et al.
Veröffentlicht: (2024)
von: Lanier, Michael, et al.
Veröffentlicht: (2024)
Evaluation-Time Policy Switching for Offline Reinforcement Learning
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2025)
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2025)
From Explainability to Interpretability: Interpretable Policies in Reinforcement Learning Via Model Explanation
von: Li, Peilang, et al.
Veröffentlicht: (2025)
von: Li, Peilang, et al.
Veröffentlicht: (2025)
Exploiting Hybrid Policy in Reinforcement Learning for Interpretable Temporal Logic Manipulation
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
Offline Goal-Conditioned Reinforcement Learning with Projective Quasimetric Planning
von: Kobanda, Anthony, et al.
Veröffentlicht: (2025)
von: Kobanda, Anthony, et al.
Veröffentlicht: (2025)
Offline Policy Evaluation for Reinforcement Learning with Adaptively Collected Data
von: Madhow, Sunil, et al.
Veröffentlicht: (2023)
von: Madhow, Sunil, et al.
Veröffentlicht: (2023)
Dataset Distillation for Offline Reinforcement Learning
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
Reinforcement Learning via Self-Distillation
von: Hübotter, Jonas, et al.
Veröffentlicht: (2026)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2026)
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
von: Li, Lanpei, et al.
Veröffentlicht: (2024)
von: Li, Lanpei, et al.
Veröffentlicht: (2024)
Better Decisions through the Right Causal World Model
von: Dillies, Elisabeth, et al.
Veröffentlicht: (2025)
von: Dillies, Elisabeth, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning
von: Kohler, Hector, et al.
Veröffentlicht: (2024) -
Limits of Actor-Critic Algorithms for Decision Tree Policies Learning in IBMDPs
von: Kohler, Hector, et al.
Veröffentlicht: (2023) -
Towards a Research Community in Interpretable Reinforcement Learning: the InterpPol Workshop
von: Kohler, Hector, et al.
Veröffentlicht: (2024) -
Breiman meets Bellman: Non-Greedy Decision Trees with MDPs
von: Kohler, Hector, et al.
Veröffentlicht: (2023) -
When (and How) to Trust the Expert: Diagnosing Query-Time Expert-Guided Reinforcement Learning
von: Berthelot, Yann, et al.
Veröffentlicht: (2026)