Collaborative Value Function Estimation Under Model Mismatch: A Federated Temporal Difference Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Beikmohammadi, Ali, Khirirat, Sarit, Richtárik, Peter, Magnússon, Sindri |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Parallel Momentum Methods Under Biased Gradient Estimations
by: Beikmohammadi, Ali, et al.
Published: (2024)
by: Beikmohammadi, Ali, et al.
Published: (2024)
Compressed Federated Reinforcement Learning with a Generative Model
by: Beikmohammadi, Ali, et al.
Published: (2024)
by: Beikmohammadi, Ali, et al.
Published: (2024)
On the Convergence of Federated Learning Algorithms without Data Similarity
by: Beikmohammadi, Ali, et al.
Published: (2024)
by: Beikmohammadi, Ali, et al.
Published: (2024)
Human-Inspired Framework to Accelerate Reinforcement Learning
by: Beikmohammadi, Ali, et al.
Published: (2023)
by: Beikmohammadi, Ali, et al.
Published: (2023)
Smoothed Normalization for Efficient Distributed Private Optimization
by: Shulgin, Egor, et al.
Published: (2025)
by: Shulgin, Egor, et al.
Published: (2025)
A Cost-Sensitive Transformer Model for Prognostics Under Highly Imbalanced Industrial Data
by: Beikmohammadi, Ali, et al.
Published: (2024)
by: Beikmohammadi, Ali, et al.
Published: (2024)
First Provable Guarantees for Practical Private FL: Beyond Restrictive Assumptions
by: Shulgin, Egor, et al.
Published: (2025)
by: Shulgin, Egor, et al.
Published: (2025)
Better LMO-based Momentum Methods with Second-Order Information
by: Khirirat, Sarit, et al.
Published: (2025)
by: Khirirat, Sarit, et al.
Published: (2025)
PriPG-RL: Privileged Planner-Guided Reinforcement Learning for Partially Observable Systems with Anytime-Feasible MPC
by: Amiri, Mohsen, et al.
Published: (2026)
by: Amiri, Mohsen, et al.
Published: (2026)
Error Feedback under $(L_0,L_1)$-Smoothness: Normalization and Momentum
by: Khirirat, Sarit, et al.
Published: (2024)
by: Khirirat, Sarit, et al.
Published: (2024)
Reinforcement Learning in Switching Non-Stationary Markov Decision Processes: Algorithms and Convergence Analysis
by: Amiri, Mohsen, et al.
Published: (2025)
by: Amiri, Mohsen, et al.
Published: (2025)
Improved Convergence in Parameter-Agnostic Error Feedback through Momentum
by: Sadiev, Abdurakhmon, et al.
Published: (2025)
by: Sadiev, Abdurakhmon, et al.
Published: (2025)
Challenger-Based Combinatorial Bandits for Subcarrier Selection in OFDM Systems
by: Amiri, Mohsen, et al.
Published: (2025)
by: Amiri, Mohsen, et al.
Published: (2025)
Asynchronous Distributed Optimization with Delay-free Parameters
by: Wu, Xuyang, et al.
Published: (2023)
by: Wu, Xuyang, et al.
Published: (2023)
MARBLE: Multi-Armed Restless Bandits in Latent Markovian Environment
by: Amiri, Mohsen, et al.
Published: (2025)
by: Amiri, Mohsen, et al.
Published: (2025)
Adaptive Budgeted Multi-Armed Bandits for IoT with Dynamic Resource Constraints
by: Vaishnav, Shubham, et al.
Published: (2025)
by: Vaishnav, Shubham, et al.
Published: (2025)
Dynamic and Distributed Routing in IoT Networks based on Multi-Objective Q-Learning
by: Vaishnav, Shubham, et al.
Published: (2025)
by: Vaishnav, Shubham, et al.
Published: (2025)
SCANIA Component X Dataset: A Real-World Multivariate Time Series Dataset for Predictive Maintenance
by: Kharazian, Zahra, et al.
Published: (2024)
by: Kharazian, Zahra, et al.
Published: (2024)
MARINA-P: Superior Performance in Non-smooth Federated Optimization with Adaptive Stepsizes
by: Sokolov, Igor, et al.
Published: (2024)
by: Sokolov, Igor, et al.
Published: (2024)
Communication-Efficient Gluon in Federated Learning
by: Qian, Xun, et al.
Published: (2026)
by: Qian, Xun, et al.
Published: (2026)
Convergence Analysis of the PAGE Stochastic Algorithm for Weakly Convex Finite-Sum Optimization
by: Condat, Laurent, et al.
Published: (2025)
by: Condat, Laurent, et al.
Published: (2025)
Constrained Reinforcement Learning Under Model Mismatch
by: Sun, Zhongchang, et al.
Published: (2024)
by: Sun, Zhongchang, et al.
Published: (2024)
Symmetric Pruning of Large Language Models
by: Yi, Kai, et al.
Published: (2025)
by: Yi, Kai, et al.
Published: (2025)
Automatic Fused Multimodal Deep Learning for Plant Identification
by: Lapkovskis, Alfreds, et al.
Published: (2024)
by: Lapkovskis, Alfreds, et al.
Published: (2024)
FedP3: Federated Personalized and Privacy-friendly Network Pruning under Model Heterogeneity
by: Yi, Kai, et al.
Published: (2024)
by: Yi, Kai, et al.
Published: (2024)
An Analysis of Action-Value Temporal-Difference Methods That Learn State Values
by: Daley, Brett, et al.
Published: (2025)
by: Daley, Brett, et al.
Published: (2025)
Cohort Squeeze: Beyond a Single Communication Round per Cohort in Cross-Device Federated Learning
by: Yi, Kai, et al.
Published: (2024)
by: Yi, Kai, et al.
Published: (2024)
A Data Mining-Based Dynamical Anomaly Detection Method for Integrating with an Advance Metering System
by: Maitra, Sarit
Published: (2024)
by: Maitra, Sarit
Published: (2024)
Improving the Worst-Case Bidirectional Communication Complexity for Nonconvex Distributed Optimization under Function Similarity
by: Gruntkowska, Kaja, et al.
Published: (2024)
by: Gruntkowska, Kaja, et al.
Published: (2024)
Non-Euclidean Broximal Point Method: A Blueprint for Geometry-Aware Optimization
by: Gruntkowska, Kaja, et al.
Published: (2025)
by: Gruntkowska, Kaja, et al.
Published: (2025)
Improving RCT-Based CATE Estimation Under Covariate Mismatch via Calibrated Alignment
by: Asiaee, Amir, et al.
Published: (2026)
by: Asiaee, Amir, et al.
Published: (2026)
Understanding Uncertainty-based Active Learning Under Model Mismatch
by: Rahmati, Amir Hossein, et al.
Published: (2024)
by: Rahmati, Amir Hossein, et al.
Published: (2024)
A Computation and Communication Efficient Method for Distributed Nonconvex Problems in the Partial Participation Setting
by: Tyurin, Alexander, et al.
Published: (2022)
by: Tyurin, Alexander, et al.
Published: (2022)
Sparse-ProxSkip: Accelerated Sparse-to-Sparse Training in Federated Learning
by: Meinhardt, Georg, et al.
Published: (2024)
by: Meinhardt, Georg, et al.
Published: (2024)
Benchmarking Dynamic SLO Compliance in Distributed Computing Continuum Systems
by: Lapkovskis, Alfreds, et al.
Published: (2025)
by: Lapkovskis, Alfreds, et al.
Published: (2025)
Shadowheart SGD: Distributed Asynchronous SGD with Optimal Time Complexity Under Arbitrary Computation and Communication Heterogeneity
by: Tyurin, Alexander, et al.
Published: (2024)
by: Tyurin, Alexander, et al.
Published: (2024)
Reward Models Are Secretly Value Functions: Temporally Coherent Reward Modeling
by: Nikulkov, Alex
Published: (2026)
by: Nikulkov, Alex
Published: (2026)
Federated Temporal Difference Learning with Linear Function Approximation under Environmental Heterogeneity
by: Wang, Han, et al.
Published: (2023)
by: Wang, Han, et al.
Published: (2023)
On the Convergence of DP-SGD with Adaptive Clipping
by: Shulgin, Egor, et al.
Published: (2024)
by: Shulgin, Egor, et al.
Published: (2024)
Thanos: A Block-wise Pruning Algorithm for Efficient Large Language Model Compression
by: Ilin, Ivan, et al.
Published: (2025)
by: Ilin, Ivan, et al.
Published: (2025)
Similar Items
-
Parallel Momentum Methods Under Biased Gradient Estimations
by: Beikmohammadi, Ali, et al.
Published: (2024) -
Compressed Federated Reinforcement Learning with a Generative Model
by: Beikmohammadi, Ali, et al.
Published: (2024) -
On the Convergence of Federated Learning Algorithms without Data Similarity
by: Beikmohammadi, Ali, et al.
Published: (2024) -
Human-Inspired Framework to Accelerate Reinforcement Learning
by: Beikmohammadi, Ali, et al.
Published: (2023) -
Smoothed Normalization for Efficient Distributed Private Optimization
by: Shulgin, Egor, et al.
Published: (2025)