'Explaining RL Decisions with Trajectories': A Reproducibility Study
Fuente:
arXiv
Saved in:
| Main Authors: | Sadek, Karim Abdel, Nulli, Matteo, Velja, Joan, Vincenti, Jort |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dynamic Vocabulary Pruning in Early-Exit LLMs
by: Vincenti, Jort, et al.
Published: (2024)
by: Vincenti, Jort, et al.
Published: (2024)
Explaining RL Decisions with Trajectories
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2023)
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2023)
Studying Cross-cluster Modularity in Neural Networks
by: Golechha, Satvik, et al.
Published: (2025)
by: Golechha, Satvik, et al.
Published: (2025)
Learning the Preferences of a Learning Agent
by: Sadek, Karim Abdel, et al.
Published: (2026)
by: Sadek, Karim Abdel, et al.
Published: (2026)
In-Context Learning Improves Compositional Understanding of Vision-Language Models
by: Nulli, Matteo, et al.
Published: (2024)
by: Nulli, Matteo, et al.
Published: (2024)
When can we trust untrusted monitoring? A safety case sketch across collusion strategies
by: Gardner-Challis, Nelson, et al.
Published: (2026)
by: Gardner-Challis, Nelson, et al.
Published: (2026)
Explaining Decisions of Agents in Mixed-Motive Games
by: Orner, Maayan, et al.
Published: (2024)
by: Orner, Maayan, et al.
Published: (2024)
Semantic segmentation with coarse annotations
by: de Jong, Jort, et al.
Published: (2025)
by: de Jong, Jort, et al.
Published: (2025)
The Effective Horizon Explains Deep RL Performance in Stochastic Environments
by: Laidlaw, Cassidy, et al.
Published: (2023)
by: Laidlaw, Cassidy, et al.
Published: (2023)
Explaining Learned Reward Functions with Counterfactual Trajectories
by: Wehner, Jan, et al.
Published: (2024)
by: Wehner, Jan, et al.
Published: (2024)
A Uniform Language to Explain Decision Trees
by: Arenas, Marcelo, et al.
Published: (2023)
by: Arenas, Marcelo, et al.
Published: (2023)
Explaining Hitori Puzzles: Neurosymbolic Proof Staging for Sequential Decisions
by: Pacheco, Maria Leonor, et al.
Published: (2025)
by: Pacheco, Maria Leonor, et al.
Published: (2025)
ACTER: Diverse and Actionable Counterfactual Sequences for Explaining and Diagnosing RL Policies
by: Gajcin, Jasmina, et al.
Published: (2024)
by: Gajcin, Jasmina, et al.
Published: (2024)
COOL-MC: Verifying and Explaining RL Policies for Platelet Inventory Management
by: Gross, Dennis
Published: (2026)
by: Gross, Dennis
Published: (2026)
A Comprehensive Survey of Belief Rule Base (BRB) Hybrid Expert system: Bridging Decision Science and Professional Services
by: Derrick, Karim
Published: (2024)
by: Derrick, Karim
Published: (2024)
COOL-MC: Verifying and Explaining RL Policies for Multi-bridge Network Maintenance
by: Gross, Dennis
Published: (2026)
by: Gross, Dennis
Published: (2026)
XAI-LAW: A Logic Programming Tool for Modeling, Explaining, and Learning Legal Decisions
by: Dovier, Agostino, et al.
Published: (2026)
by: Dovier, Agostino, et al.
Published: (2026)
Explaining Control Policies through Predicate Decision Diagrams
by: Chakraborty, Debraj, et al.
Published: (2025)
by: Chakraborty, Debraj, et al.
Published: (2025)
AI Copilots for Reproducibility in Science: A Case Study
by: Bibal, Adrien, et al.
Published: (2025)
by: Bibal, Adrien, et al.
Published: (2025)
Explaining Decisions in ML Models: a Parameterized Complexity Analysis (Part I)
by: Ordyniak, Sebastian, et al.
Published: (2025)
by: Ordyniak, Sebastian, et al.
Published: (2025)
Stitching Sub-Trajectories with Conditional Diffusion Model for Goal-Conditioned Offline RL
by: Kim, Sungyoon, et al.
Published: (2024)
by: Kim, Sungyoon, et al.
Published: (2024)
Explaining Decisions in ML Models: a Parameterized Complexity Analysis
by: Ordyniak, Sebastian, et al.
Published: (2024)
by: Ordyniak, Sebastian, et al.
Published: (2024)
Explaining and Improving Information Complementarities in Multi-Agent Decision-making
by: Guo, Ziyang, et al.
Published: (2025)
by: Guo, Ziyang, et al.
Published: (2025)
Online Finetuning Decision Transformers with Pure RL Gradients
by: Luo, Junkai, et al.
Published: (2026)
by: Luo, Junkai, et al.
Published: (2026)
Dynamic Mixture-of-Experts for Visual Autoregressive Model
by: Vincenti, Jort, et al.
Published: (2025)
by: Vincenti, Jort, et al.
Published: (2025)
BET: Explaining Deep Reinforcement Learning through The Error-Prone Decisions
by: Liu, Xiao, et al.
Published: (2024)
by: Liu, Xiao, et al.
Published: (2024)
On Explaining Proxy Discrimination and Unfairness in Individual Decisions Made by AI Systems
by: Sonna, Belona, et al.
Published: (2025)
by: Sonna, Belona, et al.
Published: (2025)
Meeting SLOs, Slashing Hours: Automated Enterprise LLM Optimization with OptiKIT
by: Santavas, Nicholas, et al.
Published: (2026)
by: Santavas, Nicholas, et al.
Published: (2026)
Formally Explaining Decision Tree Models with Answer Set Programming
by: Takemura, Akihiro, et al.
Published: (2026)
by: Takemura, Akihiro, et al.
Published: (2026)
Crafting Desirable Climate Trajectories with RL Explored Socio-Environmental Simulations
by: Rudd-Jones, James, et al.
Published: (2024)
by: Rudd-Jones, James, et al.
Published: (2024)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
by: Schmied, Thomas, et al.
Published: (2024)
by: Schmied, Thomas, et al.
Published: (2024)
Teaching LLMs to Think Mathematically: A Critical Study of Decision-Making via Optimization
by: Abdel-Rahman, Mohammad J., et al.
Published: (2025)
by: Abdel-Rahman, Mohammad J., et al.
Published: (2025)
Towards Self-Evolving Benchmarks: Synthesizing Agent Trajectories via Test-Time Exploration under Validate-by-Reproduce Paradigm
by: Guo, Dadi, et al.
Published: (2025)
by: Guo, Dadi, et al.
Published: (2025)
A Systematic Reproducibility Study of BSARec for Sequential Recommendation
by: Hutter, Jan, et al.
Published: (2025)
by: Hutter, Jan, et al.
Published: (2025)
Using Reinforcement Learning to Train Large Language Models to Explain Human Decisions
by: Zhu, Jian-Qiao, et al.
Published: (2025)
by: Zhu, Jian-Qiao, et al.
Published: (2025)
TraceGraph: Shared Decision Landscapes for Diagnosing and Improving Agent Trajectories
by: Nian, Junjie, et al.
Published: (2026)
by: Nian, Junjie, et al.
Published: (2026)
Mind Your Entropy: From Maximum Entropy to Trajectory Entropy-Constrained RL
by: Zhan, Guojian, et al.
Published: (2025)
by: Zhan, Guojian, et al.
Published: (2025)
Explaining Large Language Models Decisions Using Shapley Values
by: Mohammadi, Behnam
Published: (2024)
by: Mohammadi, Behnam
Published: (2024)
Think Twice, Act Once: A Co-Evolution Framework of LLM and RL for Large-Scale Decision Making
by: Wan, Xu, et al.
Published: (2025)
by: Wan, Xu, et al.
Published: (2025)
Med-CAM: Minimal Evidence for Explaining Medical Decision Making
by: Suhail, Pirzada, et al.
Published: (2026)
by: Suhail, Pirzada, et al.
Published: (2026)
Similar Items
-
Dynamic Vocabulary Pruning in Early-Exit LLMs
by: Vincenti, Jort, et al.
Published: (2024) -
Explaining RL Decisions with Trajectories
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2023) -
Studying Cross-cluster Modularity in Neural Networks
by: Golechha, Satvik, et al.
Published: (2025) -
Learning the Preferences of a Learning Agent
by: Sadek, Karim Abdel, et al.
Published: (2026) -
In-Context Learning Improves Compositional Understanding of Vision-Language Models
by: Nulli, Matteo, et al.
Published: (2024)