Convergence Guarantees for Federated SARSA with Local Training and Heterogeneous Agents
Fuente:
arXiv
Salvato in:
| Autori principali: | Mangold, Paul, Berthier, Eloïse, Moulines, Eric |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On Global Convergence Rates for Federated Softmax Policy Gradient under Heterogeneous Environments
di: Labbi, Safwan, et al.
Pubblicazione: (2025)
di: Labbi, Safwan, et al.
Pubblicazione: (2025)
Federated UCBVI: Communication-Efficient Federated Regret Minimization with Heterogeneous Agents
di: Labbi, Safwan, et al.
Pubblicazione: (2024)
di: Labbi, Safwan, et al.
Pubblicazione: (2024)
Beyond Softmax and Entropy: Convergence Rates of Policy Gradients with f-SoftArgmax Parameterization & Coupled Regularization
di: Labbi, Safwan, et al.
Pubblicazione: (2026)
di: Labbi, Safwan, et al.
Pubblicazione: (2026)
Refined Analysis of Federated Averaging and Federated Richardson-Romberg
di: Mangold, Paul, et al.
Pubblicazione: (2024)
di: Mangold, Paul, et al.
Pubblicazione: (2024)
SCAFFLSA: Taming Heterogeneity in Federated Linear Stochastic Approximation and TD Learning
di: Mangold, Paul, et al.
Pubblicazione: (2024)
di: Mangold, Paul, et al.
Pubblicazione: (2024)
Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation
di: Levin, Ilya, et al.
Pubblicazione: (2026)
di: Levin, Ilya, et al.
Pubblicazione: (2026)
Semi-Gradient SARSA Routing with Theoretical Guarantee on Traffic Stability and Weight Convergence
di: Wu, Yidan, et al.
Pubblicazione: (2025)
di: Wu, Yidan, et al.
Pubblicazione: (2025)
Refined Analysis of Entropy-Regularized Actor-Critic
di: Labbi, Safwan, et al.
Pubblicazione: (2026)
di: Labbi, Safwan, et al.
Pubblicazione: (2026)
Scaffold with Stochastic Gradients: New Analysis with Linear Speed-Up
di: Mangold, Paul, et al.
Pubblicazione: (2025)
di: Mangold, Paul, et al.
Pubblicazione: (2025)
Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of Gaussians
di: Huix, Tom, et al.
Pubblicazione: (2024)
di: Huix, Tom, et al.
Pubblicazione: (2024)
Federated Dynamical Low-Rank Training with Global Loss Convergence Guarantees
di: Schotthöfer, Steffen, et al.
Pubblicazione: (2024)
di: Schotthöfer, Steffen, et al.
Pubblicazione: (2024)
Implicit Q-Learning and SARSA: Liberating Policy Control from Step-Size Calibration
di: Kim, Hwanwoo, et al.
Pubblicazione: (2026)
di: Kim, Hwanwoo, et al.
Pubblicazione: (2026)
No-Regret Gaussian Process Optimization of Time-Varying Functions
di: Mauduit, Eliabelle, et al.
Pubblicazione: (2025)
di: Mauduit, Eliabelle, et al.
Pubblicazione: (2025)
Balanced Training of Energy-Based Models with Adaptive Flow Sampling
di: Grenioux, Louis, et al.
Pubblicazione: (2023)
di: Grenioux, Louis, et al.
Pubblicazione: (2023)
Joint Channel Selection using FedDRL in V2X
di: Mancini, Lorenzo, et al.
Pubblicazione: (2024)
di: Mancini, Lorenzo, et al.
Pubblicazione: (2024)
On Double Descent in Reinforcement Learning with LSTD and Random Features
di: Brellmann, David, et al.
Pubblicazione: (2023)
di: Brellmann, David, et al.
Pubblicazione: (2023)
Queuing dynamics of asynchronous Federated Learning
di: Leconte, Louis, et al.
Pubblicazione: (2024)
di: Leconte, Louis, et al.
Pubblicazione: (2024)
Fairness-aware Federated Minimax Optimization with Convergence Guarantee
di: Dunda, Gerry Windiarto Mohamad, et al.
Pubblicazione: (2023)
di: Dunda, Gerry Windiarto Mohamad, et al.
Pubblicazione: (2023)
Adaptive Guidance for Local Training in Heterogeneous Federated Learning
di: Zhang, Jianqing, et al.
Pubblicazione: (2024)
di: Zhang, Jianqing, et al.
Pubblicazione: (2024)
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games
di: Ocello, Antonio, et al.
Pubblicazione: (2025)
di: Ocello, Antonio, et al.
Pubblicazione: (2025)
A Generalized Meta Federated Learning Framework with Theoretical Convergence Guarantees
di: Jamali, Mohammad Vahid, et al.
Pubblicazione: (2025)
di: Jamali, Mohammad Vahid, et al.
Pubblicazione: (2025)
State-Separated SARSA: A Practical Sequential Decision-Making Algorithm with Recovering Rewards
di: Tanimoto, Yuto, et al.
Pubblicazione: (2024)
di: Tanimoto, Yuto, et al.
Pubblicazione: (2024)
torchsom: The Reference PyTorch Library for Self-Organizing Maps
di: Berthier, Louis, et al.
Pubblicazione: (2025)
di: Berthier, Louis, et al.
Pubblicazione: (2025)
Concept-Based Mechanistic Interpretability Using Structured Knowledge Graphs
di: Chorna, Sofiia, et al.
Pubblicazione: (2025)
di: Chorna, Sofiia, et al.
Pubblicazione: (2025)
Loss-Guided Auxiliary Agents for Overcoming Mode Collapse in GFlowNets
di: Malek, Idriss, et al.
Pubblicazione: (2025)
di: Malek, Idriss, et al.
Pubblicazione: (2025)
Dimensionality Reduction for Robust Federated Learning: A Theoretical Analysis and Convergence Guarantee
di: Zuo, Shiyuan, et al.
Pubblicazione: (2026)
di: Zuo, Shiyuan, et al.
Pubblicazione: (2026)
Convergence Analysis of Sequential Federated Learning on Heterogeneous Data
di: Li, Yipeng, et al.
Pubblicazione: (2023)
di: Li, Yipeng, et al.
Pubblicazione: (2023)
Optimizing Asynchronous Federated Learning: A Delicate Trade-Off Between Model-Parameter Staleness and Update Frequency
di: Alahyane, Abdelkrim, et al.
Pubblicazione: (2025)
di: Alahyane, Abdelkrim, et al.
Pubblicazione: (2025)
Learning with Locally Private Examples by Inverse Weierstrass Private Stochastic Gradient Descent
di: Dufraiche, Jean, et al.
Pubblicazione: (2026)
di: Dufraiche, Jean, et al.
Pubblicazione: (2026)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
di: Ganesh, Swetha, et al.
Pubblicazione: (2024)
di: Ganesh, Swetha, et al.
Pubblicazione: (2024)
Learning Over-Relaxation Policies for ADMM with Convergence Guarantees
di: Lin, Junan, et al.
Pubblicazione: (2026)
di: Lin, Junan, et al.
Pubblicazione: (2026)
Theoretical Convergence Guarantees for Variational Autoencoders
di: Surendran, Sobihan, et al.
Pubblicazione: (2024)
di: Surendran, Sobihan, et al.
Pubblicazione: (2024)
Tight Analysis of Decentralized SGD: A Markov Chain Perspective
di: Versini, Lucas, et al.
Pubblicazione: (2026)
di: Versini, Lucas, et al.
Pubblicazione: (2026)
Efficient Conformal Prediction under Data Heterogeneity
di: Plassier, Vincent, et al.
Pubblicazione: (2023)
di: Plassier, Vincent, et al.
Pubblicazione: (2023)
Briding Diffusion Posterior Sampling and Monte Carlo methods: a survey
di: Janati, Yazid, et al.
Pubblicazione: (2025)
di: Janati, Yazid, et al.
Pubblicazione: (2025)
On Sampling with Approximate Transport Maps
di: Grenioux, Louis, et al.
Pubblicazione: (2023)
di: Grenioux, Louis, et al.
Pubblicazione: (2023)
Communication Efficient Federated Learning with Linear Convergence on Heterogeneous Data
di: Liu, Jie, et al.
Pubblicazione: (2025)
di: Liu, Jie, et al.
Pubblicazione: (2025)
Decentralized Sporadic Federated Learning: A Unified Algorithmic Framework with Convergence Guarantees
di: Zehtabi, Shahryar, et al.
Pubblicazione: (2024)
di: Zehtabi, Shahryar, et al.
Pubblicazione: (2024)
Convergence Analysis of Split Federated Learning on Heterogeneous Data
di: Han, Pengchao, et al.
Pubblicazione: (2024)
di: Han, Pengchao, et al.
Pubblicazione: (2024)
Segmenting Action-Value Functions Over Time-Scales in SARSA via TD($Δ$)
di: Humayoo, Mahammad
Pubblicazione: (2024)
di: Humayoo, Mahammad
Pubblicazione: (2024)
Documenti analoghi
-
On Global Convergence Rates for Federated Softmax Policy Gradient under Heterogeneous Environments
di: Labbi, Safwan, et al.
Pubblicazione: (2025) -
Federated UCBVI: Communication-Efficient Federated Regret Minimization with Heterogeneous Agents
di: Labbi, Safwan, et al.
Pubblicazione: (2024) -
Beyond Softmax and Entropy: Convergence Rates of Policy Gradients with f-SoftArgmax Parameterization & Coupled Regularization
di: Labbi, Safwan, et al.
Pubblicazione: (2026) -
Refined Analysis of Federated Averaging and Federated Richardson-Romberg
di: Mangold, Paul, et al.
Pubblicazione: (2024) -
SCAFFLSA: Taming Heterogeneity in Federated Linear Stochastic Approximation and TD Learning
di: Mangold, Paul, et al.
Pubblicazione: (2024)