Generalized Policy Learning for Smart Grids: FL TRPO Approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yunxiang, Cuadrado, Nicolas Mauricio, Horváth, Samuel, Takáč, Martin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generalising Battery Control in Net-Zero Buildings via Personalised Federated RL
von: Avila, Nicolas M Cuadrado, et al.
Veröffentlicht: (2024)
von: Avila, Nicolas M Cuadrado, et al.
Veröffentlicht: (2024)
What Scalable Second-Order Information Knows for Pruning at Initialization
von: Navarrete, Ivo Gollini, et al.
Veröffentlicht: (2025)
von: Navarrete, Ivo Gollini, et al.
Veröffentlicht: (2025)
Enhancing Policy Gradient with the Polyak Step-Size Adaption
von: Li, Yunxiang, et al.
Veröffentlicht: (2024)
von: Li, Yunxiang, et al.
Veröffentlicht: (2024)
FRESCO: Federated Reinforcement Energy System for Cooperative Optimization
von: Cuadrado, Nicolas Mauricio, et al.
Veröffentlicht: (2024)
von: Cuadrado, Nicolas Mauricio, et al.
Veröffentlicht: (2024)
Collaborative and Efficient Personalization with Mixtures of Adaptors
von: Almansoori, Abdulla Jasem, et al.
Veröffentlicht: (2024)
von: Almansoori, Abdulla Jasem, et al.
Veröffentlicht: (2024)
Federated Learning Can Find Friends That Are Advantageous
von: Tupitsa, Nazarii, et al.
Veröffentlicht: (2024)
von: Tupitsa, Nazarii, et al.
Veröffentlicht: (2024)
Knowledge Distillation from Large Language Models for Household Energy Modeling
von: Takrouri, Mohannad, et al.
Veröffentlicht: (2025)
von: Takrouri, Mohannad, et al.
Veröffentlicht: (2025)
FRUGAL: Memory-Efficient Optimization by Reducing State Overhead for Scalable Training
von: Zmushko, Philip, et al.
Veröffentlicht: (2024)
von: Zmushko, Philip, et al.
Veröffentlicht: (2024)
Byzantine-Robust Optimization under $(L_0, L_1)$-Smoothness
von: Bolatov, Arman, et al.
Veröffentlicht: (2026)
von: Bolatov, Arman, et al.
Veröffentlicht: (2026)
FedPeWS: Personalized Warmup via Subnetworks for Enhanced Heterogeneous Federated Learning
von: Tastan, Nurbek, et al.
Veröffentlicht: (2024)
von: Tastan, Nurbek, et al.
Veröffentlicht: (2024)
PaDPaF: Partial Disentanglement with Partially-Federated GANs
von: Almansoori, Abdulla Jasem, et al.
Veröffentlicht: (2022)
von: Almansoori, Abdulla Jasem, et al.
Veröffentlicht: (2022)
Escaping Local Optima in the Waddington Landscape: A Two-Stage TRPO-PPO Approach for Single-Cell Perturbation Analysis
von: Boabang, Francis, et al.
Veröffentlicht: (2025)
von: Boabang, Francis, et al.
Veröffentlicht: (2025)
LoFT: Low-Rank Adaptation That Behaves Like Full Fine-Tuning
von: Tastan, Nurbek, et al.
Veröffentlicht: (2025)
von: Tastan, Nurbek, et al.
Veröffentlicht: (2025)
SB-TRPO: Towards Safe Reinforcement Learning with Hard Constraints
von: Wagner, Dominik, et al.
Veröffentlicht: (2025)
von: Wagner, Dominik, et al.
Veröffentlicht: (2025)
Remove that Square Root: A New Efficient Scale-Invariant Version of AdaGrad
von: Choudhury, Sayantan, et al.
Veröffentlicht: (2024)
von: Choudhury, Sayantan, et al.
Veröffentlicht: (2024)
Tractable Probabilistic Models for Investment Planning
von: A., Nicolas M. Cuadrado, et al.
Veröffentlicht: (2025)
von: A., Nicolas M. Cuadrado, et al.
Veröffentlicht: (2025)
Revisiting LocalSGD and SCAFFOLD: Improved Rates and Missing Analysis
von: Luo, Ruichen, et al.
Veröffentlicht: (2025)
von: Luo, Ruichen, et al.
Veröffentlicht: (2025)
Methods with Local Steps and Random Reshuffling for Generally Smooth Non-Convex Federated Optimization
von: Demidovich, Yury, et al.
Veröffentlicht: (2024)
von: Demidovich, Yury, et al.
Veröffentlicht: (2024)
Faster Than SVD, Smarter Than SGD: The OPLoRA Alternating Update
von: Almansoori, Abdulla Jasem, et al.
Veröffentlicht: (2025)
von: Almansoori, Abdulla Jasem, et al.
Veröffentlicht: (2025)
Beyond SGD, Without SVD: Proximal Subspace Iteration LoRA with Diagonal Fractional K-FAC
von: Almansoori, Abdulla Jasem, et al.
Veröffentlicht: (2026)
von: Almansoori, Abdulla Jasem, et al.
Veröffentlicht: (2026)
Simple Stepsize for Quasi-Newton Methods with Global Convergence Guarantees
von: Agafonov, Artem, et al.
Veröffentlicht: (2025)
von: Agafonov, Artem, et al.
Veröffentlicht: (2025)
LionMuon: Alternating Spectral and Sign Descent for Efficient Training
von: Bolatov, Arman, et al.
Veröffentlicht: (2026)
von: Bolatov, Arman, et al.
Veröffentlicht: (2026)
Methods for Convex $(L_0,L_1)$-Smooth Optimization: Clipping, Acceleration, and Adaptivity
von: Gorbunov, Eduard, et al.
Veröffentlicht: (2024)
von: Gorbunov, Eduard, et al.
Veröffentlicht: (2024)
Preconditioned Norms: A Unified Framework for Steepest Descent, Quasi-Newton and Adaptive Methods
von: Veprikov, Andrey, et al.
Veröffentlicht: (2025)
von: Veprikov, Andrey, et al.
Veröffentlicht: (2025)
Zero-Shot Off-Policy Learning
von: Asadulaev, Arip, et al.
Veröffentlicht: (2026)
von: Asadulaev, Arip, et al.
Veröffentlicht: (2026)
SusFL: Energy-Aware Federated Learning-based Monitoring for Sustainable Smart Farms
von: Chen, Dian, et al.
Veröffentlicht: (2024)
von: Chen, Dian, et al.
Veröffentlicht: (2024)
Clipping Improves Adam-Norm and AdaGrad-Norm when the Noise Is Heavy-Tailed
von: Chezhegov, Savelii, et al.
Veröffentlicht: (2024)
von: Chezhegov, Savelii, et al.
Veröffentlicht: (2024)
Who to Trust? Aggregating Client Predictions in Federated Distillation
von: Kovalchuk, Viktor, et al.
Veröffentlicht: (2025)
von: Kovalchuk, Viktor, et al.
Veröffentlicht: (2025)
Random-reshuffled SARAH does not need a full gradient computations
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2021)
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2021)
Electrical Load Forecasting in Smart Grid: A Personalized Federated Learning Approach
von: Rahman, Ratun, et al.
Veröffentlicht: (2024)
von: Rahman, Ratun, et al.
Veröffentlicht: (2024)
Latent Reasoning in TRMs is Secretly a Policy Improvement Operator
von: Asadulaev, Arip, et al.
Veröffentlicht: (2025)
von: Asadulaev, Arip, et al.
Veröffentlicht: (2025)
Efficient Conformal Prediction under Data Heterogeneity
von: Plassier, Vincent, et al.
Veröffentlicht: (2023)
von: Plassier, Vincent, et al.
Veröffentlicht: (2023)
Gradient Clipping Beyond Vector Norms: A Spectral Approach for Matrix-Valued Parameters
von: Yukhimchuk, Alexander, et al.
Veröffentlicht: (2026)
von: Yukhimchuk, Alexander, et al.
Veröffentlicht: (2026)
ProDER: A Continual Learning Approach for Fault Prediction in Evolving Smart Grids
von: Efatinasab, Emad, et al.
Veröffentlicht: (2025)
von: Efatinasab, Emad, et al.
Veröffentlicht: (2025)
Cybersecurity Assessment of Smart Grid Exposure Using a Machine Learning Based Approach
von: Jeje, Mofe O.
Veröffentlicht: (2025)
von: Jeje, Mofe O.
Veröffentlicht: (2025)
Expert or not? assessing data quality in offline reinforcement learning
von: Asadulaev, Arip, et al.
Veröffentlicht: (2025)
von: Asadulaev, Arip, et al.
Veröffentlicht: (2025)
HybridFL: A Federated Learning Approach for Financial Crime Detection
von: Khan, Afsana, et al.
Veröffentlicht: (2026)
von: Khan, Afsana, et al.
Veröffentlicht: (2026)
From Risk to Uncertainty: Generating Predictive Uncertainty Measures via Bayesian Estimation
von: Kotelevskii, Nikita, et al.
Veröffentlicht: (2024)
von: Kotelevskii, Nikita, et al.
Veröffentlicht: (2024)
GeFL: Model-Agnostic Federated Learning with Generative Models
von: Kang, Honggu, et al.
Veröffentlicht: (2024)
von: Kang, Honggu, et al.
Veröffentlicht: (2024)
Few-Shot Load Forecasting Under Data Scarcity in Smart Grids: A Meta-Learning Approach
von: Tsoumplekas, Georgios, et al.
Veröffentlicht: (2024)
von: Tsoumplekas, Georgios, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Generalising Battery Control in Net-Zero Buildings via Personalised Federated RL
von: Avila, Nicolas M Cuadrado, et al.
Veröffentlicht: (2024) -
What Scalable Second-Order Information Knows for Pruning at Initialization
von: Navarrete, Ivo Gollini, et al.
Veröffentlicht: (2025) -
Enhancing Policy Gradient with the Polyak Step-Size Adaption
von: Li, Yunxiang, et al.
Veröffentlicht: (2024) -
FRESCO: Federated Reinforcement Energy System for Cooperative Optimization
von: Cuadrado, Nicolas Mauricio, et al.
Veröffentlicht: (2024) -
Collaborative and Efficient Personalization with Mixtures of Adaptors
von: Almansoori, Abdulla Jasem, et al.
Veröffentlicht: (2024)