Interpretable Policy Distillation for Power Grid Topology Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Dmitruka, Aleksandra, Freivalds, Karlis |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi-Objective Reinforcement Learning for Power Grid Topology Control
por: Lautenbacher, Thomas, et al.
Publicado: (2025)
por: Lautenbacher, Thomas, et al.
Publicado: (2025)
Evaluating Interpretable Reinforcement Learning by Distilling Policies into Programs
por: Kohler, Hector, et al.
Publicado: (2025)
por: Kohler, Hector, et al.
Publicado: (2025)
Towards Efficient Multi-Objective Optimisation for Real-World Power Grid Topology Control
por: Manyari, Yassine El, et al.
Publicado: (2025)
por: Manyari, Yassine El, et al.
Publicado: (2025)
Proximal Policy Distillation
por: Spigler, Giacomo
Publicado: (2024)
por: Spigler, Giacomo
Publicado: (2024)
Graph Neural Networks for Transmission Grid Topology Control: Busbar Information Asymmetry and Heterogeneous Representations
por: de Jong, Matthijs, et al.
Publicado: (2025)
por: de Jong, Matthijs, et al.
Publicado: (2025)
RL2Grid: Benchmarking Reinforcement Learning in Power Grid Operations
por: Marchesini, Enrico, et al.
Publicado: (2025)
por: Marchesini, Enrico, et al.
Publicado: (2025)
Imitation Learning for Intra-Day Power Grid Operation through Topology Actions
por: de Jong, Matthijs, et al.
Publicado: (2024)
por: de Jong, Matthijs, et al.
Publicado: (2024)
Optimizing Power Grid Topologies with Reinforcement Learning: A Survey of Methods and Challenges
por: van der Sar, Erica, et al.
Publicado: (2025)
por: van der Sar, Erica, et al.
Publicado: (2025)
Extreme Region Policy Distillation
por: Chen, Changyu, et al.
Publicado: (2026)
por: Chen, Changyu, et al.
Publicado: (2026)
Distilling Reinforcement Learning Policies for Interpretable Robot Locomotion: Gradient Boosting Machines and Symbolic Regression
por: Acero, Fernando, et al.
Publicado: (2024)
por: Acero, Fernando, et al.
Publicado: (2024)
PACED: Distillation and On-Policy Self-Distillation at the Frontier of Student Competence
por: Xu, Yuanda, et al.
Publicado: (2026)
por: Xu, Yuanda, et al.
Publicado: (2026)
$\boldsymbol{f}$-OPD: Stabilizing Long-Horizon On-Policy Distillation with Freshness-Aware Control
por: Chen, Xianwei, et al.
Publicado: (2026)
por: Chen, Xianwei, et al.
Publicado: (2026)
SymLight: Exploring Interpretable and Deployable Symbolic Policies for Traffic Signal Control
por: Liao, Xiao-Cheng, et al.
Publicado: (2025)
por: Liao, Xiao-Cheng, et al.
Publicado: (2025)
HDPO: Hybrid Distillation Policy Optimization via Privileged Self-Distillation
por: Ding, Ken
Publicado: (2026)
por: Ding, Ken
Publicado: (2026)
TIP: Token Importance in On-Policy Distillation
por: Xu, Yuanda, et al.
Publicado: (2026)
por: Xu, Yuanda, et al.
Publicado: (2026)
Online Policy Distillation with Decision-Attention
por: Yu, Xinqiang, et al.
Publicado: (2024)
por: Yu, Xinqiang, et al.
Publicado: (2024)
KL for a KL: On-Policy Distillation with Control Variate Baseline
por: Oh, Minjae, et al.
Publicado: (2026)
por: Oh, Minjae, et al.
Publicado: (2026)
UPath: Universal Planner Across Topological Heterogeneity For Grid-Based Pathfinding
por: Ananikian, Aleksandr, et al.
Publicado: (2026)
por: Ananikian, Aleksandr, et al.
Publicado: (2026)
OPD+: Rethinking the Advantage Design for On-Policy Distillation
por: Zhao, Hanyang, et al.
Publicado: (2026)
por: Zhao, Hanyang, et al.
Publicado: (2026)
Trust-Region Behavior Blending for On-Policy Distillation
por: Plyusov, Daniil, et al.
Publicado: (2026)
por: Plyusov, Daniil, et al.
Publicado: (2026)
Understanding Annotator Safety Policy with Interpretability
por: Oesterling, Alex, et al.
Publicado: (2026)
por: Oesterling, Alex, et al.
Publicado: (2026)
Learning to Reason: Temporal Saliency Distillation for Interpretable Knowledge Transfer
por: Dehigahawattage, Nilushika Udayangani Hewa, et al.
Publicado: (2026)
por: Dehigahawattage, Nilushika Udayangani Hewa, et al.
Publicado: (2026)
Power Plays: Unleashing Machine Learning Magic in Smart Grids
por: Rashid, Abdur, et al.
Publicado: (2024)
por: Rashid, Abdur, et al.
Publicado: (2024)
SafePowerGraph: Safety-aware Evaluation of Graph Neural Networks for Transmission Power Grids
por: Ghamizi, Salah, et al.
Publicado: (2024)
por: Ghamizi, Salah, et al.
Publicado: (2024)
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
por: Li, Lanpei, et al.
Publicado: (2024)
por: Li, Lanpei, et al.
Publicado: (2024)
The Illusion of Certainty: Decoupling Capability and Calibration in On-Policy Distillation
por: Zhang, Jiaxin, et al.
Publicado: (2026)
por: Zhang, Jiaxin, et al.
Publicado: (2026)
ADWIN: Adaptive Windows for Horizon-Aware On-Policy Distillation
por: Liang, Kun, et al.
Publicado: (2026)
por: Liang, Kun, et al.
Publicado: (2026)
Stable On-Policy Distillation through Adaptive Target Reformulation
por: Jang, Ijun, et al.
Publicado: (2026)
por: Jang, Ijun, et al.
Publicado: (2026)
Graph Embedding Dynamic Feature-based Supervised Contrastive Learning of Transient Stability for Changing Power Grid Topologies
por: Lv, Zijian, et al.
Publicado: (2023)
por: Lv, Zijian, et al.
Publicado: (2023)
Interpreting and Controlling LLM Reasoning through Integrated Policy Gradient
por: Li, Changming, et al.
Publicado: (2026)
por: Li, Changming, et al.
Publicado: (2026)
Explainable RL Policies by Distilling to Locally-Specialized Linear Policies with Voronoi State Partitioning
por: Deproost, Senne, et al.
Publicado: (2025)
por: Deproost, Senne, et al.
Publicado: (2025)
Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation
por: Malik, Gitesh
Publicado: (2026)
por: Malik, Gitesh
Publicado: (2026)
Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level
por: Jia, Nan, et al.
Publicado: (2026)
por: Jia, Nan, et al.
Publicado: (2026)
Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why
por: Armandpour, Mohammadreza, et al.
Publicado: (2026)
por: Armandpour, Mohammadreza, et al.
Publicado: (2026)
How Ensembles of Distilled Policies Improve Generalisation in Reinforcement Learning
por: Weltevrede, Max, et al.
Publicado: (2025)
por: Weltevrede, Max, et al.
Publicado: (2025)
Toward Dynamic Stability Assessment of Power Grid Topologies using Graph Neural Networks
por: Nauck, Christian, et al.
Publicado: (2022)
por: Nauck, Christian, et al.
Publicado: (2022)
CINDI: Conditional Imputation and Noisy Data Integrity with Flows in Power Grid Data
por: Baumgartner, David, et al.
Publicado: (2026)
por: Baumgartner, David, et al.
Publicado: (2026)
Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning
por: Kohler, Hector, et al.
Publicado: (2024)
por: Kohler, Hector, et al.
Publicado: (2024)
Pragmatic Policy Development via Interpretable Behavior Cloning
por: Matsson, Anton, et al.
Publicado: (2025)
por: Matsson, Anton, et al.
Publicado: (2025)
Variational Distillation of Diffusion Policies into Mixture of Experts
por: Zhou, Hongyi, et al.
Publicado: (2024)
por: Zhou, Hongyi, et al.
Publicado: (2024)
Ejemplares similares
-
Multi-Objective Reinforcement Learning for Power Grid Topology Control
por: Lautenbacher, Thomas, et al.
Publicado: (2025) -
Evaluating Interpretable Reinforcement Learning by Distilling Policies into Programs
por: Kohler, Hector, et al.
Publicado: (2025) -
Towards Efficient Multi-Objective Optimisation for Real-World Power Grid Topology Control
por: Manyari, Yassine El, et al.
Publicado: (2025) -
Proximal Policy Distillation
por: Spigler, Giacomo
Publicado: (2024) -
Graph Neural Networks for Transmission Grid Topology Control: Busbar Information Asymmetry and Heterogeneous Representations
por: de Jong, Matthijs, et al.
Publicado: (2025)