Convergence Guarantees for Federated SARSA with Local Training and Heterogeneous Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Mangold, Paul, Berthier, Eloïse, Moulines, Eric |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Global Convergence Rates for Federated Softmax Policy Gradient under Heterogeneous Environments
by: Labbi, Safwan, et al.
Published: (2025)
by: Labbi, Safwan, et al.
Published: (2025)
Federated UCBVI: Communication-Efficient Federated Regret Minimization with Heterogeneous Agents
by: Labbi, Safwan, et al.
Published: (2024)
by: Labbi, Safwan, et al.
Published: (2024)
Beyond Softmax and Entropy: Convergence Rates of Policy Gradients with f-SoftArgmax Parameterization & Coupled Regularization
by: Labbi, Safwan, et al.
Published: (2026)
by: Labbi, Safwan, et al.
Published: (2026)
Refined Analysis of Federated Averaging and Federated Richardson-Romberg
by: Mangold, Paul, et al.
Published: (2024)
by: Mangold, Paul, et al.
Published: (2024)
SCAFFLSA: Taming Heterogeneity in Federated Linear Stochastic Approximation and TD Learning
by: Mangold, Paul, et al.
Published: (2024)
by: Mangold, Paul, et al.
Published: (2024)
Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation
by: Levin, Ilya, et al.
Published: (2026)
by: Levin, Ilya, et al.
Published: (2026)
Semi-Gradient SARSA Routing with Theoretical Guarantee on Traffic Stability and Weight Convergence
by: Wu, Yidan, et al.
Published: (2025)
by: Wu, Yidan, et al.
Published: (2025)
Refined Analysis of Entropy-Regularized Actor-Critic
by: Labbi, Safwan, et al.
Published: (2026)
by: Labbi, Safwan, et al.
Published: (2026)
Scaffold with Stochastic Gradients: New Analysis with Linear Speed-Up
by: Mangold, Paul, et al.
Published: (2025)
by: Mangold, Paul, et al.
Published: (2025)
Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of Gaussians
by: Huix, Tom, et al.
Published: (2024)
by: Huix, Tom, et al.
Published: (2024)
Federated Dynamical Low-Rank Training with Global Loss Convergence Guarantees
by: Schotthöfer, Steffen, et al.
Published: (2024)
by: Schotthöfer, Steffen, et al.
Published: (2024)
Implicit Q-Learning and SARSA: Liberating Policy Control from Step-Size Calibration
by: Kim, Hwanwoo, et al.
Published: (2026)
by: Kim, Hwanwoo, et al.
Published: (2026)
No-Regret Gaussian Process Optimization of Time-Varying Functions
by: Mauduit, Eliabelle, et al.
Published: (2025)
by: Mauduit, Eliabelle, et al.
Published: (2025)
Balanced Training of Energy-Based Models with Adaptive Flow Sampling
by: Grenioux, Louis, et al.
Published: (2023)
by: Grenioux, Louis, et al.
Published: (2023)
Joint Channel Selection using FedDRL in V2X
by: Mancini, Lorenzo, et al.
Published: (2024)
by: Mancini, Lorenzo, et al.
Published: (2024)
On Double Descent in Reinforcement Learning with LSTD and Random Features
by: Brellmann, David, et al.
Published: (2023)
by: Brellmann, David, et al.
Published: (2023)
Queuing dynamics of asynchronous Federated Learning
by: Leconte, Louis, et al.
Published: (2024)
by: Leconte, Louis, et al.
Published: (2024)
Fairness-aware Federated Minimax Optimization with Convergence Guarantee
by: Dunda, Gerry Windiarto Mohamad, et al.
Published: (2023)
by: Dunda, Gerry Windiarto Mohamad, et al.
Published: (2023)
Adaptive Guidance for Local Training in Heterogeneous Federated Learning
by: Zhang, Jianqing, et al.
Published: (2024)
by: Zhang, Jianqing, et al.
Published: (2024)
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games
by: Ocello, Antonio, et al.
Published: (2025)
by: Ocello, Antonio, et al.
Published: (2025)
A Generalized Meta Federated Learning Framework with Theoretical Convergence Guarantees
by: Jamali, Mohammad Vahid, et al.
Published: (2025)
by: Jamali, Mohammad Vahid, et al.
Published: (2025)
State-Separated SARSA: A Practical Sequential Decision-Making Algorithm with Recovering Rewards
by: Tanimoto, Yuto, et al.
Published: (2024)
by: Tanimoto, Yuto, et al.
Published: (2024)
torchsom: The Reference PyTorch Library for Self-Organizing Maps
by: Berthier, Louis, et al.
Published: (2025)
by: Berthier, Louis, et al.
Published: (2025)
Concept-Based Mechanistic Interpretability Using Structured Knowledge Graphs
by: Chorna, Sofiia, et al.
Published: (2025)
by: Chorna, Sofiia, et al.
Published: (2025)
Loss-Guided Auxiliary Agents for Overcoming Mode Collapse in GFlowNets
by: Malek, Idriss, et al.
Published: (2025)
by: Malek, Idriss, et al.
Published: (2025)
Dimensionality Reduction for Robust Federated Learning: A Theoretical Analysis and Convergence Guarantee
by: Zuo, Shiyuan, et al.
Published: (2026)
by: Zuo, Shiyuan, et al.
Published: (2026)
Convergence Analysis of Sequential Federated Learning on Heterogeneous Data
by: Li, Yipeng, et al.
Published: (2023)
by: Li, Yipeng, et al.
Published: (2023)
Optimizing Asynchronous Federated Learning: A Delicate Trade-Off Between Model-Parameter Staleness and Update Frequency
by: Alahyane, Abdelkrim, et al.
Published: (2025)
by: Alahyane, Abdelkrim, et al.
Published: (2025)
Learning with Locally Private Examples by Inverse Weierstrass Private Stochastic Gradient Descent
by: Dufraiche, Jean, et al.
Published: (2026)
by: Dufraiche, Jean, et al.
Published: (2026)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
by: Ganesh, Swetha, et al.
Published: (2024)
by: Ganesh, Swetha, et al.
Published: (2024)
Learning Over-Relaxation Policies for ADMM with Convergence Guarantees
by: Lin, Junan, et al.
Published: (2026)
by: Lin, Junan, et al.
Published: (2026)
Theoretical Convergence Guarantees for Variational Autoencoders
by: Surendran, Sobihan, et al.
Published: (2024)
by: Surendran, Sobihan, et al.
Published: (2024)
Tight Analysis of Decentralized SGD: A Markov Chain Perspective
by: Versini, Lucas, et al.
Published: (2026)
by: Versini, Lucas, et al.
Published: (2026)
Efficient Conformal Prediction under Data Heterogeneity
by: Plassier, Vincent, et al.
Published: (2023)
by: Plassier, Vincent, et al.
Published: (2023)
Briding Diffusion Posterior Sampling and Monte Carlo methods: a survey
by: Janati, Yazid, et al.
Published: (2025)
by: Janati, Yazid, et al.
Published: (2025)
On Sampling with Approximate Transport Maps
by: Grenioux, Louis, et al.
Published: (2023)
by: Grenioux, Louis, et al.
Published: (2023)
Communication Efficient Federated Learning with Linear Convergence on Heterogeneous Data
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
Decentralized Sporadic Federated Learning: A Unified Algorithmic Framework with Convergence Guarantees
by: Zehtabi, Shahryar, et al.
Published: (2024)
by: Zehtabi, Shahryar, et al.
Published: (2024)
Convergence Analysis of Split Federated Learning on Heterogeneous Data
by: Han, Pengchao, et al.
Published: (2024)
by: Han, Pengchao, et al.
Published: (2024)
Segmenting Action-Value Functions Over Time-Scales in SARSA via TD($Δ$)
by: Humayoo, Mahammad
Published: (2024)
by: Humayoo, Mahammad
Published: (2024)
Similar Items
-
On Global Convergence Rates for Federated Softmax Policy Gradient under Heterogeneous Environments
by: Labbi, Safwan, et al.
Published: (2025) -
Federated UCBVI: Communication-Efficient Federated Regret Minimization with Heterogeneous Agents
by: Labbi, Safwan, et al.
Published: (2024) -
Beyond Softmax and Entropy: Convergence Rates of Policy Gradients with f-SoftArgmax Parameterization & Coupled Regularization
by: Labbi, Safwan, et al.
Published: (2026) -
Refined Analysis of Federated Averaging and Federated Richardson-Romberg
by: Mangold, Paul, et al.
Published: (2024) -
SCAFFLSA: Taming Heterogeneity in Federated Linear Stochastic Approximation and TD Learning
by: Mangold, Paul, et al.
Published: (2024)