Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autor principal: | Olaz, Xabier |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Understanding and Diagnosing Deep Reinforcement Learning
por: Korkmaz, Ezgi
Publicado: (2024)
por: Korkmaz, Ezgi
Publicado: (2024)
A New Paradigm in Tuning Learned Indexes: A Reinforcement Learning Enhanced Approach
por: Wang, Taiyi, et al.
Publicado: (2025)
por: Wang, Taiyi, et al.
Publicado: (2025)
Understanding and Improving Hyperbolic Deep Reinforcement Learning
por: Klein, Timo, et al.
Publicado: (2025)
por: Klein, Timo, et al.
Publicado: (2025)
A Study of Plasticity Loss in On-Policy Deep Reinforcement Learning
por: Juliani, Arthur, et al.
Publicado: (2024)
por: Juliani, Arthur, et al.
Publicado: (2024)
Deep Reinforcement Learning for Power Grid Multi-Stage Cascading Failure Mitigation
por: Meng, Bo, et al.
Publicado: (2025)
por: Meng, Bo, et al.
Publicado: (2025)
Learning from Failures in Multi-Attempt Reinforcement Learning
por: Chung, Stephen, et al.
Publicado: (2025)
por: Chung, Stephen, et al.
Publicado: (2025)
Policy-Based Deep Reinforcement Learning Hyperheuristics for Job-Shop Scheduling Problems
por: Lassoued, Sofiene, et al.
Publicado: (2026)
por: Lassoued, Sofiene, et al.
Publicado: (2026)
DeepStock: Reinforcement Learning with Policy Regularizations for Inventory Management
por: Xie, Yaqi, et al.
Publicado: (2026)
por: Xie, Yaqi, et al.
Publicado: (2026)
Topological Deep Learning: A Review of an Emerging Paradigm
por: Zia, Ali, et al.
Publicado: (2023)
por: Zia, Ali, et al.
Publicado: (2023)
Internalizing Outcome Supervision into Process Supervision: A New Paradigm for Reinforcement Learning for Reasoning
por: Ding, Fei, et al.
Publicado: (2026)
por: Ding, Fei, et al.
Publicado: (2026)
TRIMMER: A New Paradigm for Video Summarization through Self-Supervised Reinforcement Learning
por: Mishra, Pritam, et al.
Publicado: (2026)
por: Mishra, Pritam, et al.
Publicado: (2026)
Learning in Context, Guided by Choice: A Reward-Free Paradigm for Reinforcement Learning with Transformers
por: Dong, Juncheng, et al.
Publicado: (2026)
por: Dong, Juncheng, et al.
Publicado: (2026)
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
por: Mayor, Walter, et al.
Publicado: (2025)
por: Mayor, Walter, et al.
Publicado: (2025)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
por: Anisimov, Maksim, et al.
Publicado: (2026)
por: Anisimov, Maksim, et al.
Publicado: (2026)
Efficient Deep Reinforcement Learning with Predictive Processing Proximal Policy Optimization
por: Küçükoğlu, Burcu, et al.
Publicado: (2022)
por: Küçükoğlu, Burcu, et al.
Publicado: (2022)
Optimal Policy Sparsification and Low Rank Decomposition for Deep Reinforcement Learning
por: Goddla, Vikram
Publicado: (2024)
por: Goddla, Vikram
Publicado: (2024)
Deep Reinforcement Learning for Inventory Networks: Toward Reliable Policy Optimization
por: Alvo, Matias, et al.
Publicado: (2023)
por: Alvo, Matias, et al.
Publicado: (2023)
A Model of Understanding in Deep Learning Systems
por: Freeborn, David Peter Wallis
Publicado: (2026)
por: Freeborn, David Peter Wallis
Publicado: (2026)
Policy-Based Reinforcement Learning with Action Masking for Dynamic Job Shop Scheduling under Uncertainty: Handling Random Arrivals and Machine Failures
por: Lassoued, Sofiene, et al.
Publicado: (2026)
por: Lassoued, Sofiene, et al.
Publicado: (2026)
Pessimistic Auxiliary Policy for Offline Reinforcement Learning
por: Zhang, Fan, et al.
Publicado: (2026)
por: Zhang, Fan, et al.
Publicado: (2026)
Textual Explanations and Their Evaluations for Reinforcement Learning Policy
por: Terra, Ahmad, et al.
Publicado: (2026)
por: Terra, Ahmad, et al.
Publicado: (2026)
Composing Reinforcement Learning Policies, with Formal Guarantees
por: Delgrange, Florent, et al.
Publicado: (2024)
por: Delgrange, Florent, et al.
Publicado: (2024)
SAFE-RL: Saliency-Aware Counterfactual Explainer for Deep Reinforcement Learning Policies
por: Samadi, Amir, et al.
Publicado: (2024)
por: Samadi, Amir, et al.
Publicado: (2024)
Relative Importance Sampling for off-Policy Actor-Critic in Deep Reinforcement Learning
por: Humayoo, Mahammad, et al.
Publicado: (2018)
por: Humayoo, Mahammad, et al.
Publicado: (2018)
Improving Deep Reinforcement Learning by Reducing the Chain Effect of Value and Policy Churn
por: Tang, Hongyao, et al.
Publicado: (2024)
por: Tang, Hongyao, et al.
Publicado: (2024)
Solving Deep Reinforcement Learning Tasks with Evolution Strategies and Linear Policy Networks
por: Wong, Annie, et al.
Publicado: (2024)
por: Wong, Annie, et al.
Publicado: (2024)
The Definitive Guide to Policy Gradients in Deep Reinforcement Learning: Theory, Algorithms and Implementations
por: Lehmann, Matthias
Publicado: (2024)
por: Lehmann, Matthias
Publicado: (2024)
Salience-Invariant Consistent Policy Learning for Generalization in Visual Reinforcement Learning
por: Sun, Jingbo, et al.
Publicado: (2025)
por: Sun, Jingbo, et al.
Publicado: (2025)
Learning from Less: Guiding Deep Reinforcement Learning with Differentiable Symbolic Planning
por: Ye, Zihan, et al.
Publicado: (2025)
por: Ye, Zihan, et al.
Publicado: (2025)
Co-Activation Graph Analysis of Safety-Verified and Explainable Deep Reinforcement Learning Policies
por: Gross, Dennis, et al.
Publicado: (2025)
por: Gross, Dennis, et al.
Publicado: (2025)
Rank-1 Approximation of Inverse Fisher for Natural Policy Gradients in Deep Reinforcement Learning
por: Huo, Yingxiao, et al.
Publicado: (2026)
por: Huo, Yingxiao, et al.
Publicado: (2026)
Explaining Decentralized Multi-Agent Reinforcement Learning Policies
por: Boggess, Kayla, et al.
Publicado: (2025)
por: Boggess, Kayla, et al.
Publicado: (2025)
Policy Expansion for Bridging Offline-to-Online Reinforcement Learning
por: Zhang, Haichao, et al.
Publicado: (2023)
por: Zhang, Haichao, et al.
Publicado: (2023)
On Generating Explanations for Reinforcement Learning Policies: An Empirical Study
por: Yuasa, Mikihisa, et al.
Publicado: (2023)
por: Yuasa, Mikihisa, et al.
Publicado: (2023)
Probabilistic Model Checking of Stochastic Reinforcement Learning Policies
por: Gross, Dennis, et al.
Publicado: (2024)
por: Gross, Dennis, et al.
Publicado: (2024)
Zero-Knowledge Federated Learning: A New Trustworthy and Privacy-Preserving Distributed Learning Paradigm
por: Wang, Taotao, et al.
Publicado: (2025)
por: Wang, Taotao, et al.
Publicado: (2025)
Learning Robot Soccer from Egocentric Vision with Deep Reinforcement Learning
por: Tirumala, Dhruva, et al.
Publicado: (2024)
por: Tirumala, Dhruva, et al.
Publicado: (2024)
An Invitation to Deep Reinforcement Learning
por: Jaeger, Bernhard, et al.
Publicado: (2023)
por: Jaeger, Bernhard, et al.
Publicado: (2023)
Sample-Efficient Neurosymbolic Deep Reinforcement Learning
por: Veronese, Celeste, et al.
Publicado: (2026)
por: Veronese, Celeste, et al.
Publicado: (2026)
A New Error Temporal Difference Algorithm for Deep Reinforcement Learning in Microgrid Optimization
por: Yao, Fulong, et al.
Publicado: (2025)
por: Yao, Fulong, et al.
Publicado: (2025)
Ejemplares similares
-
Understanding and Diagnosing Deep Reinforcement Learning
por: Korkmaz, Ezgi
Publicado: (2024) -
A New Paradigm in Tuning Learned Indexes: A Reinforcement Learning Enhanced Approach
por: Wang, Taiyi, et al.
Publicado: (2025) -
Understanding and Improving Hyperbolic Deep Reinforcement Learning
por: Klein, Timo, et al.
Publicado: (2025) -
A Study of Plasticity Loss in On-Policy Deep Reinforcement Learning
por: Juliani, Arthur, et al.
Publicado: (2024) -
Deep Reinforcement Learning for Power Grid Multi-Stage Cascading Failure Mitigation
por: Meng, Bo, et al.
Publicado: (2025)