SAFE-RL: Saliency-Aware Counterfactual Explainer for Deep Reinforcement Learning Policies
Fuente:
arXiv
Guardado en:
| Autores principales: | Samadi, Amir, Koufos, Konstantinos, Debattista, Kurt, Dianati, Mehrdad |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Good Data Is All Imitation Learning Needs
por: Samadi, Amir, et al.
Publicado: (2024)
por: Samadi, Amir, et al.
Publicado: (2024)
Taming Transformers for Realistic Lidar Point Cloud Generation
por: Haghighi, Hamed, et al.
Publicado: (2024)
por: Haghighi, Hamed, et al.
Publicado: (2024)
Run-time Monitoring of 3D Object Detection in Automated Driving Systems Using Early Layer Neural Activation Patterns
por: Yatbaz, Hakan Yekta, et al.
Publicado: (2024)
por: Yatbaz, Hakan Yekta, et al.
Publicado: (2024)
CoDy: Counterfactual Explainers for Dynamic Graphs
por: Qu, Zhan, et al.
Publicado: (2024)
por: Qu, Zhan, et al.
Publicado: (2024)
UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning Models
por: Kang, Hyunju, et al.
Publicado: (2026)
por: Kang, Hyunju, et al.
Publicado: (2026)
A Unified Generative Framework for Realistic Lidar Simulation in Autonomous Driving Systems
por: Haghighi, Hamed, et al.
Publicado: (2023)
por: Haghighi, Hamed, et al.
Publicado: (2023)
ACTER: Diverse and Actionable Counterfactual Sequences for Explaining and Diagnosing RL Policies
por: Gajcin, Jasmina, et al.
Publicado: (2024)
por: Gajcin, Jasmina, et al.
Publicado: (2024)
Fast Explanations via Policy Gradient-Optimized Explainer
por: Pan, Deng, et al.
Publicado: (2024)
por: Pan, Deng, et al.
Publicado: (2024)
Budgeting Counterfactual for Offline RL
por: Liu, Yao, et al.
Publicado: (2023)
por: Liu, Yao, et al.
Publicado: (2023)
M-CELS: Counterfactual Explanation for Multivariate Time Series Data Guided by Learned Saliency Maps
por: Li, Peiyu, et al.
Publicado: (2024)
por: Li, Peiyu, et al.
Publicado: (2024)
LPPG-RL: Lexicographically Projected Policy Gradient Reinforcement Learning with Subproblem Exploration
por: Qiu, Ruiyu, et al.
Publicado: (2025)
por: Qiu, Ruiyu, et al.
Publicado: (2025)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
por: Belaire, Roman, et al.
Publicado: (2024)
por: Belaire, Roman, et al.
Publicado: (2024)
CARE-RL: Capability-Aware Reinforcement Learning for Mitigating Cross-Domain Conflicts
por: Zhang, Rui, et al.
Publicado: (2026)
por: Zhang, Rui, et al.
Publicado: (2026)
SCOPE-RL: A Python Library for Offline Reinforcement Learning and Off-Policy Evaluation
por: Kiyohara, Haruka, et al.
Publicado: (2023)
por: Kiyohara, Haruka, et al.
Publicado: (2023)
AstRL: Analog and Mixed-Signal Circuit Synthesis with Deep Reinforcement Learning
por: Guo, Felicia B., et al.
Publicado: (2026)
por: Guo, Felicia B., et al.
Publicado: (2026)
Counteractive RL: Rethinking Core Principles for Efficient and Scalable Deep Reinforcement Learning
por: Korkmaz, Ezgi
Publicado: (2026)
por: Korkmaz, Ezgi
Publicado: (2026)
SPEC-RL: Accelerating On-Policy Reinforcement Learning with Speculative Rollouts
por: Liu, Bingshuai, et al.
Publicado: (2025)
por: Liu, Bingshuai, et al.
Publicado: (2025)
SAFE-SVD: Sensitivity-Aware Fidelity-Enforcing SVD for Physics Foundation Models
por: Hong, Chengjie, et al.
Publicado: (2026)
por: Hong, Chengjie, et al.
Publicado: (2026)
Counterfactual Explanations for Continuous Action Reinforcement Learning
por: Dong, Shuyang, et al.
Publicado: (2025)
por: Dong, Shuyang, et al.
Publicado: (2025)
RL$^3$: Boosting Meta Reinforcement Learning via RL inside RL$^2$
por: Bhatia, Abhinav, et al.
Publicado: (2023)
por: Bhatia, Abhinav, et al.
Publicado: (2023)
On-Policy RL Meets Off-Policy Experts: Harmonizing Supervised Fine-Tuning and Reinforcement Learning via Dynamic Weighting
por: Zhang, Wenhao, et al.
Publicado: (2025)
por: Zhang, Wenhao, et al.
Publicado: (2025)
ProRL: Effective Reinforcement Learning for Proactive Recommendation via Rectified Policy Gradient Estimation
por: Hou, Hongru, et al.
Publicado: (2026)
por: Hou, Hongru, et al.
Publicado: (2026)
Deep RL With Information Constrained Policies: Generalization in Continuous Control
por: Malloy, Tailia, et al.
Publicado: (2020)
por: Malloy, Tailia, et al.
Publicado: (2020)
Using Part-based Representations for Explainable Deep Reinforcement Learning
por: Kirtas, Manos, et al.
Publicado: (2024)
por: Kirtas, Manos, et al.
Publicado: (2024)
RL-Based Method for Benchmarking the Adversarial Resilience and Robustness of Deep Reinforcement Learning Policies
por: Behzadan, Vahid, et al.
Publicado: (2019)
por: Behzadan, Vahid, et al.
Publicado: (2019)
Faster Synchronous On-Policy RL via Straggler-Aware Group Sizing
por: Khan, Azal Ahmad, et al.
Publicado: (2026)
por: Khan, Azal Ahmad, et al.
Publicado: (2026)
Evaluating Explainability in Machine Learning Predictions through Explainer-Agnostic Metrics
por: Munoz, Cristian, et al.
Publicado: (2023)
por: Munoz, Cristian, et al.
Publicado: (2023)
Knowledge Transfer in Deep Reinforcement Learning via an RL-Specific GAN-Based Correspondence Function
por: Ruman, Marko, et al.
Publicado: (2022)
por: Ruman, Marko, et al.
Publicado: (2022)
A Study of Plasticity Loss in On-Policy Deep Reinforcement Learning
por: Juliani, Arthur, et al.
Publicado: (2024)
por: Juliani, Arthur, et al.
Publicado: (2024)
DeepStock: Reinforcement Learning with Policy Regularizations for Inventory Management
por: Xie, Yaqi, et al.
Publicado: (2026)
por: Xie, Yaqi, et al.
Publicado: (2026)
Learning by Doing: An Online Causal Reinforcement Learning Framework with Causal-Aware Policy
por: Cai, Ruichu, et al.
Publicado: (2024)
por: Cai, Ruichu, et al.
Publicado: (2024)
Counterfactual Explanations for Deep Learning-Based Traffic Forecasting
por: Wang, Rushan, et al.
Publicado: (2024)
por: Wang, Rushan, et al.
Publicado: (2024)
Dissecting Deep RL with High Update Ratios: Combatting Value Divergence
por: Hussing, Marcel, et al.
Publicado: (2024)
por: Hussing, Marcel, et al.
Publicado: (2024)
Do No Harm: A Counterfactual Approach to Safe Reinforcement Learning
por: Vaskov, Sean, et al.
Publicado: (2024)
por: Vaskov, Sean, et al.
Publicado: (2024)
Null Counterfactual Factor Interactions for Goal-Conditioned Reinforcement Learning
por: Chuck, Caleb, et al.
Publicado: (2025)
por: Chuck, Caleb, et al.
Publicado: (2025)
Policy Learning for Off-Dynamics RL with Deficient Support
por: Van, Linh Le Pham, et al.
Publicado: (2024)
por: Van, Linh Le Pham, et al.
Publicado: (2024)
Saliency-Aware Regularized Quantization Calibration for Large Language Models
por: Zhao, Yanlong, et al.
Publicado: (2026)
por: Zhao, Yanlong, et al.
Publicado: (2026)
Spatial-Aware Deep Reinforcement Learning for the Traveling Officer Problem
por: Strauß, Niklas, et al.
Publicado: (2024)
por: Strauß, Niklas, et al.
Publicado: (2024)
Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies
por: Lee, Haanvid, et al.
Publicado: (2024)
por: Lee, Haanvid, et al.
Publicado: (2024)
Critic-Driven Voronoi-Quantization for Distilling Deep RL Policies to Explainable Models
por: Deproost, Senne, et al.
Publicado: (2026)
por: Deproost, Senne, et al.
Publicado: (2026)
Ejemplares similares
-
Good Data Is All Imitation Learning Needs
por: Samadi, Amir, et al.
Publicado: (2024) -
Taming Transformers for Realistic Lidar Point Cloud Generation
por: Haghighi, Hamed, et al.
Publicado: (2024) -
Run-time Monitoring of 3D Object Detection in Automated Driving Systems Using Early Layer Neural Activation Patterns
por: Yatbaz, Hakan Yekta, et al.
Publicado: (2024) -
CoDy: Counterfactual Explainers for Dynamic Graphs
por: Qu, Zhan, et al.
Publicado: (2024) -
UNR-Explainer: Counterfactual Explanations for Unsupervised Node Representation Learning Models
por: Kang, Hyunju, et al.
Publicado: (2026)