FedRDMA: Communication-Efficient Cross-Silo Federated LLM via Chunked RDMA Transmission
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Zeling, Cai, Dongqi, Zhang, Yiran, Xu, Mengwei, Wang, Shangguang, Zhou, Ao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reimagining RDMA Through the Lens of ML
by: Warraich, Ertza, et al.
Published: (2025)
by: Warraich, Ertza, et al.
Published: (2025)
Varuna: Enabling Failure-Type Aware RDMA Failover
by: Wang, Xiaoyang, et al.
Published: (2026)
by: Wang, Xiaoyang, et al.
Published: (2026)
Closing the HPC-Cloud Convergence Gap: Multi-Tenant Slingshot RDMA for Kubernetes
by: Friese, Philipp A., et al.
Published: (2025)
by: Friese, Philipp A., et al.
Published: (2025)
OptiNIC: A Resilient and Tail-Optimal RDMA NIC for Distributed ML Workloads
by: Warraich, Ertza, et al.
Published: (2025)
by: Warraich, Ertza, et al.
Published: (2025)
Palladium: A DPU-enabled Multi-Tenant Serverless Cloud over Zero-copy Multi-node RDMA Fabrics
by: Qi, Shixiong, et al.
Published: (2025)
by: Qi, Shixiong, et al.
Published: (2025)
Handling of Memory Page Faults during Virtual-Address RDMA
by: Psistakis, Antonis
Published: (2025)
by: Psistakis, Antonis
Published: (2025)
Diffusion Models on the Edge: Challenges, Optimizations, and Applications
by: Zheng, Dongqi
Published: (2025)
by: Zheng, Dongqi
Published: (2025)
FAST: An Efficient Scheduler for All-to-All GPU Communication
by: Lei, Yiran, et al.
Published: (2025)
by: Lei, Yiran, et al.
Published: (2025)
FedSkipTwin: Digital-Twin-Guided Client Skipping for Communication-Efficient Federated Learning
by: Commey, Daniel, et al.
Published: (2025)
by: Commey, Daniel, et al.
Published: (2025)
fabric-lib: RDMA Point-to-Point Communication for LLM Systems
by: Licker, Nandor, et al.
Published: (2025)
by: Licker, Nandor, et al.
Published: (2025)
FedMFS: Federated Multimodal Fusion Learning with Selective Modality Communication
by: Yuan, Liangqi, et al.
Published: (2023)
by: Yuan, Liangqi, et al.
Published: (2023)
JANUS: Resilient and Adaptive Data Transmission for Enabling Timely and Efficient Cross-Facility Scientific Workflows
by: Esaulov, Vladislav, et al.
Published: (2025)
by: Esaulov, Vladislav, et al.
Published: (2025)
Towards Efficient and Scalable Distributed Vector Search with RDMA
by: Zhi, Xiangyu, et al.
Published: (2025)
by: Zhi, Xiangyu, et al.
Published: (2025)
Reliable Image Transmission in CPS-based Pub/Sub
by: Flores, Everson, et al.
Published: (2025)
by: Flores, Everson, et al.
Published: (2025)
Efficient All-to-All Collective Communication Schedules for Direct-Connect Topologies
by: Basu, Prithwish, et al.
Published: (2023)
by: Basu, Prithwish, et al.
Published: (2023)
Revisiting Bruck: Phase-Efficient All-to-All Communication in Reconfigurable Networks
by: Juerss, Anton, et al.
Published: (2026)
by: Juerss, Anton, et al.
Published: (2026)
LOAM: Low-latency Communication, Caching, and Computation Placement in Data-Intensive Computing Networks
by: Zhang, Jinkun, et al.
Published: (2024)
by: Zhang, Jinkun, et al.
Published: (2024)
Towards Practical Overlay Networks for Decentralized Federated Learning
by: Hua, Yifan, et al.
Published: (2024)
by: Hua, Yifan, et al.
Published: (2024)
GORGO: Maximizing KV-Cache Reuse While Minimizing Network Latency in Cross-Region LLM Load Balancing
by: Toniolo, Alessio Ricci, et al.
Published: (2026)
by: Toniolo, Alessio Ricci, et al.
Published: (2026)
Federated Learning and Evolutionary Game Model for Fog Federation Formation
by: Yasser, Zyad, et al.
Published: (2024)
by: Yasser, Zyad, et al.
Published: (2024)
Multi-stage Flow Scheduling for LLM Serving
by: Sun, Yijun, et al.
Published: (2026)
by: Sun, Yijun, et al.
Published: (2026)
FedCod: An Efficient Communication Protocol for Cross-Silo Federated Learning with Coding
by: Yan, Peishen, et al.
Published: (2024)
by: Yan, Peishen, et al.
Published: (2024)
A Task Decomposition and Planning Framework for Efficient LLM Inference in AI-Enabled WiFi-Offload Networks
by: Han, Mingqi, et al.
Published: (2026)
by: Han, Mingqi, et al.
Published: (2026)
FedAQ: Communication-Efficient Federated Edge Learning via Joint Uplink and Downlink Adaptive Quantization
by: Qu, Linping, et al.
Published: (2024)
by: Qu, Linping, et al.
Published: (2024)
Towards Integrated Energy-Communication-Transportation Hub: A Base-Station-Centric Design in 5G and Beyond
by: Shen, Linfeng, et al.
Published: (2025)
by: Shen, Linfeng, et al.
Published: (2025)
ALock: Asymmetric Lock Primitive for RDMA Systems
by: Baran, Amanda, et al.
Published: (2024)
by: Baran, Amanda, et al.
Published: (2024)
RailS: Load Balancing for All-to-All Communication in Distributed Mixture-of-Experts Training
by: Xu, Heng, et al.
Published: (2025)
by: Xu, Heng, et al.
Published: (2025)
ZKP-FedEval: Verifiable and Privacy-Preserving Federated Evaluation using Zero-Knowledge Proofs
by: Commey, Daniel, et al.
Published: (2025)
by: Commey, Daniel, et al.
Published: (2025)
Toward Edge General Intelligence with Multiple-Large Language Model (Multi-LLM): Architecture, Trust, and Orchestration
by: Luo, Haoxiang, et al.
Published: (2025)
by: Luo, Haoxiang, et al.
Published: (2025)
PerLLM: Personalized Inference Scheduling with Edge-Cloud Collaboration for Diverse LLM Services
by: Yang, Zheming, et al.
Published: (2024)
by: Yang, Zheming, et al.
Published: (2024)
Surviving the Edge: Federated Learning under Networking and Resource Constraints
by: Mwanje, Mike, et al.
Published: (2026)
by: Mwanje, Mike, et al.
Published: (2026)
Enabling Scalability in Asynchronous and Bidirectional Communication in LPWAN
by: Rahman, Mahbubur
Published: (2025)
by: Rahman, Mahbubur
Published: (2025)
Recursive Offloading for LLM Serving in Multi-tier Networks
by: Wu, Zhiyuan, et al.
Published: (2025)
by: Wu, Zhiyuan, et al.
Published: (2025)
KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving
by: Liu, Zedong, et al.
Published: (2026)
by: Liu, Zedong, et al.
Published: (2026)
Hierarchical Online-Scheduling for Energy-Efficient Split Inference with Progressive Transmission
by: Tang, Zengzipeng, et al.
Published: (2026)
by: Tang, Zengzipeng, et al.
Published: (2026)
Real-Time Scheduling for 802.1Qbv Time-Sensitive Networking (TSN): A Systematic Review and Experimental Study
by: Xue, Chuanyu, et al.
Published: (2023)
by: Xue, Chuanyu, et al.
Published: (2023)
A Mathematical Theory of Hyper-simplex Fractal Network for Blockchain: Part I
by: Yang, Kaiwen, et al.
Published: (2024)
by: Yang, Kaiwen, et al.
Published: (2024)
RDMA-Based Algorithms for Sparse Matrix Multiplication on GPUs
by: Brock, Benjamin, et al.
Published: (2023)
by: Brock, Benjamin, et al.
Published: (2023)
A Survey on Resource Management in Joint Communication and Computing-Embedded SAGIN
by: Chen, Qian, et al.
Published: (2024)
by: Chen, Qian, et al.
Published: (2024)
Efficient Data Management for IPFS dApps
by: Estrada-Galiñanes, Vero, et al.
Published: (2024)
by: Estrada-Galiñanes, Vero, et al.
Published: (2024)
Similar Items
-
Reimagining RDMA Through the Lens of ML
by: Warraich, Ertza, et al.
Published: (2025) -
Varuna: Enabling Failure-Type Aware RDMA Failover
by: Wang, Xiaoyang, et al.
Published: (2026) -
Closing the HPC-Cloud Convergence Gap: Multi-Tenant Slingshot RDMA for Kubernetes
by: Friese, Philipp A., et al.
Published: (2025) -
OptiNIC: A Resilient and Tail-Optimal RDMA NIC for Distributed ML Workloads
by: Warraich, Ertza, et al.
Published: (2025) -
Palladium: A DPU-enabled Multi-Tenant Serverless Cloud over Zero-copy Multi-node RDMA Fabrics
by: Qi, Shixiong, et al.
Published: (2025)