Data-Driven Policy Mapping for Safe RL-based Energy Management Systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zangato, Theo, Osmani, Aomar, Alizadeh, Pegah |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Meta-RL with Shared Representations Enables Fast Adaptation in Energy Systems
von: Zangato, Théo, et al.
Veröffentlicht: (2026)
von: Zangato, Théo, et al.
Veröffentlicht: (2026)
General Machine Learning: Theory for Learning Under Variable Regimes
von: Osmani, Aomar
Veröffentlicht: (2026)
von: Osmani, Aomar
Veröffentlicht: (2026)
SMGI: A Structural Theory of General Artificial Intelligence
von: Osmani, Aomar
Veröffentlicht: (2026)
von: Osmani, Aomar
Veröffentlicht: (2026)
On the Necessity of Metalearning: Learning Suitable Parameterizations for Learning Processes
von: Hamidi, Massinissa, et al.
Veröffentlicht: (2023)
von: Hamidi, Massinissa, et al.
Veröffentlicht: (2023)
From Ontology Conformance to Admissible Reconfiguration: A RoSO/SMGI Adequacy Argument for Robotic Service Governance
von: Osmani, Aomar
Veröffentlicht: (2026)
von: Osmani, Aomar
Veröffentlicht: (2026)
Reducing Class Bias In Data-Balanced Datasets Through Hardness-Based Resampling
von: Pukowski, Pawel, et al.
Veröffentlicht: (2025)
von: Pukowski, Pawel, et al.
Veröffentlicht: (2025)
Directional Optimism for Safe Linear Bandits
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2023)
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2023)
SPoRt -- Safe Policy Ratio: Certified Training and Deployment of Task Policies in Model-Free RL
von: Cloete, Jacques, et al.
Veröffentlicht: (2025)
von: Cloete, Jacques, et al.
Veröffentlicht: (2025)
Selecting Offline Reinforcement Learning Algorithms for Stochastic Network Control
von: Helson, Nicolas, et al.
Veröffentlicht: (2026)
von: Helson, Nicolas, et al.
Veröffentlicht: (2026)
Safe Online Convex Optimization with Multi-Point Feedback
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2024)
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2024)
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026)
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026)
APC-RL: Exceeding Data-Driven Behavior Priors with Adaptive Policy Composition
von: Rietz, Finn, et al.
Veröffentlicht: (2026)
von: Rietz, Finn, et al.
Veröffentlicht: (2026)
On the Design of Safe Continual RL Methods for Control of Nonlinear Systems
von: Coursey, Austin, et al.
Veröffentlicht: (2025)
von: Coursey, Austin, et al.
Veröffentlicht: (2025)
Data-and-Semantic Dual-Driven Spectrum Map Construction for 6G Spectrum Management
von: Liu, Jiayu, et al.
Veröffentlicht: (2025)
von: Liu, Jiayu, et al.
Veröffentlicht: (2025)
COOL-MC: Verifying and Explaining RL Policies for Platelet Inventory Management
von: Gross, Dennis
Veröffentlicht: (2026)
von: Gross, Dennis
Veröffentlicht: (2026)
Critic-Driven Voronoi-Quantization for Distilling Deep RL Policies to Explainable Models
von: Deproost, Senne, et al.
Veröffentlicht: (2026)
von: Deproost, Senne, et al.
Veröffentlicht: (2026)
SCOPE-RL: Stable and Quantitative Control of Policy Entropy in RL Post-Training
von: Wang, Chen, et al.
Veröffentlicht: (2025)
von: Wang, Chen, et al.
Veröffentlicht: (2025)
Fat-to-Thin Policy Optimization: Offline RL with Sparse Policies
von: Zhu, Lingwei, et al.
Veröffentlicht: (2025)
von: Zhu, Lingwei, et al.
Veröffentlicht: (2025)
MNIST-Fraction: Enhancing Math Education with AI-Driven Fraction Detection and Analysis
von: Ahadian, Pegah, et al.
Veröffentlicht: (2024)
von: Ahadian, Pegah, et al.
Veröffentlicht: (2024)
SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data
von: Romeo, Carlo, et al.
Veröffentlicht: (2026)
von: Romeo, Carlo, et al.
Veröffentlicht: (2026)
Green AI: A Preliminary Empirical Study on Energy Consumption in DL Models Across Different Runtime Infrastructures
von: Alizadeh, Negar, et al.
Veröffentlicht: (2024)
von: Alizadeh, Negar, et al.
Veröffentlicht: (2024)
Policy Bifurcation in Safe Reinforcement Learning
von: Zou, Wenjun, et al.
Veröffentlicht: (2024)
von: Zou, Wenjun, et al.
Veröffentlicht: (2024)
D5RL: Diverse Datasets for Data-Driven Deep Reinforcement Learning
von: Rafailov, Rafael, et al.
Veröffentlicht: (2024)
von: Rafailov, Rafael, et al.
Veröffentlicht: (2024)
Residual Off-Policy RL for Finetuning Behavior Cloning Policies
von: Ankile, Lars, et al.
Veröffentlicht: (2025)
von: Ankile, Lars, et al.
Veröffentlicht: (2025)
Data-Driven Permissible Safe Control with Barrier Certificates
von: Mazouz, Rayan, et al.
Veröffentlicht: (2024)
von: Mazouz, Rayan, et al.
Veröffentlicht: (2024)
Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
von: Mark, Max Sobol, et al.
Veröffentlicht: (2024)
von: Mark, Max Sobol, et al.
Veröffentlicht: (2024)
Robust Policy Expansion for Offline-to-Online RL under Diverse Data Corruption
von: He, Longxiang, et al.
Veröffentlicht: (2025)
von: He, Longxiang, et al.
Veröffentlicht: (2025)
Integrating DeepRL with Robust Low-Level Control in Robotic Manipulators for Non-Repetitive Reaching Tasks
von: Shahna, Mehdi Heydari, et al.
Veröffentlicht: (2024)
von: Shahna, Mehdi Heydari, et al.
Veröffentlicht: (2024)
On-Policy RL with Optimal Reward Baseline
von: Hao, Yaru, et al.
Veröffentlicht: (2025)
von: Hao, Yaru, et al.
Veröffentlicht: (2025)
Partial Policy Gradients for RL in LLMs
von: Mathur, Puneet, et al.
Veröffentlicht: (2026)
von: Mathur, Puneet, et al.
Veröffentlicht: (2026)
Decision-Point Guided Safe Policy Improvement
von: Sharma, Abhishek, et al.
Veröffentlicht: (2024)
von: Sharma, Abhishek, et al.
Veröffentlicht: (2024)
SaVeR: Optimal Data Collection Strategy for Safe Policy Evaluation in Tabular MDP
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
ShuttleEnv: An Interactive Data-Driven RL Environment for Badminton Strategy Modeling
von: Li, Ang, et al.
Veröffentlicht: (2026)
von: Li, Ang, et al.
Veröffentlicht: (2026)
SafeAR: Safe Algorithmic Recourse by Risk-Aware Policies
von: Wu, Haochen, et al.
Veröffentlicht: (2023)
von: Wu, Haochen, et al.
Veröffentlicht: (2023)
Approximate Next Policy Sampling: Replacing Conservative Target Policy Updates in Deep RL
von: Sandhu, Dillon, et al.
Veröffentlicht: (2026)
von: Sandhu, Dillon, et al.
Veröffentlicht: (2026)
Data-Driven Energy Estimation for Virtual Servers Using Combined System Metrics and Machine Learning
von: Sangha, Amandip
Veröffentlicht: (2025)
von: Sangha, Amandip
Veröffentlicht: (2025)
Malware Detection in IOT Systems Using Machine Learning Techniques
von: Mehrban, Ali, et al.
Veröffentlicht: (2023)
von: Mehrban, Ali, et al.
Veröffentlicht: (2023)
Cross-Species Transfer Learning for Electrophysiology-to-Transcriptomics Mapping in Cortical GABAergic Interneurons
von: Schwider, Theo, et al.
Veröffentlicht: (2026)
von: Schwider, Theo, et al.
Veröffentlicht: (2026)
Fuz-RL: A Fuzzy-Guided Robust Framework for Safe Reinforcement Learning under Uncertainty
von: Wan, Xu, et al.
Veröffentlicht: (2026)
von: Wan, Xu, et al.
Veröffentlicht: (2026)
Learning-Zone Energy: Online Data Selection for Efficient RL Post-Training
von: Cui, Peng, et al.
Veröffentlicht: (2026)
von: Cui, Peng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Meta-RL with Shared Representations Enables Fast Adaptation in Energy Systems
von: Zangato, Théo, et al.
Veröffentlicht: (2026) -
General Machine Learning: Theory for Learning Under Variable Regimes
von: Osmani, Aomar
Veröffentlicht: (2026) -
SMGI: A Structural Theory of General Artificial Intelligence
von: Osmani, Aomar
Veröffentlicht: (2026) -
On the Necessity of Metalearning: Learning Suitable Parameterizations for Learning Processes
von: Hamidi, Massinissa, et al.
Veröffentlicht: (2023) -
From Ontology Conformance to Admissible Reconfiguration: A RoSO/SMGI Adequacy Argument for Robotic Service Governance
von: Osmani, Aomar
Veröffentlicht: (2026)