DynamiQ: Accelerating Gradient Synchronization using Compressed Multi-hop All-reduce
Fuente:
arXiv
Saved in:
| Main Authors: | Han, Wenchen, Vargaftik, Shay, Mitzenmacher, Michael, Basat, Ran Ben |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Combinatorial Client-Master Multiagent Deep Reinforcement Learning for Task Offloading in Mobile Edge Computing
by: Gebrekidan, Tesfay Zemuy, et al.
Published: (2024)
by: Gebrekidan, Tesfay Zemuy, et al.
Published: (2024)
Asynchronous Federated Reinforcement Learning with Policy Gradient Updates: Algorithm Design and Convergence Analysis
by: Lan, Guangchen, et al.
Published: (2024)
by: Lan, Guangchen, et al.
Published: (2024)
HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference
by: Zhang, Zeyu, et al.
Published: (2025)
by: Zhang, Zeyu, et al.
Published: (2025)
OptiReduce: Resilient and Tail-Optimal AllReduce for Distributed Deep Learning in the Cloud
by: Warraich, Ertza, et al.
Published: (2023)
by: Warraich, Ertza, et al.
Published: (2023)
Naeural AI OS -- Decentralized ubiquitous computing MLOps execution engine
by: Bleotiu, Cristian, et al.
Published: (2023)
by: Bleotiu, Cristian, et al.
Published: (2023)
OptiNIC: A Resilient and Tail-Optimal RDMA NIC for Distributed ML Workloads
by: Warraich, Ertza, et al.
Published: (2025)
by: Warraich, Ertza, et al.
Published: (2025)
Reimagining RDMA Through the Lens of ML
by: Warraich, Ertza, et al.
Published: (2025)
by: Warraich, Ertza, et al.
Published: (2025)
EPARA: Parallelizing Categorized AI Inference in Edge Clouds
by: Wang, Yubo, et al.
Published: (2025)
by: Wang, Yubo, et al.
Published: (2025)
Beyond Throughput and Compression Ratios: Towards High End-to-end Utility of Gradient Compression
by: Han, Wenchen, et al.
Published: (2024)
by: Han, Wenchen, et al.
Published: (2024)
A Framework for testing Federated Learning algorithms using an edge-like environment
by: Schwanck, Felipe Machado, et al.
Published: (2024)
by: Schwanck, Felipe Machado, et al.
Published: (2024)
BitSov: A Composable Bitcoin-Native Architecture for Sovereign Internet Infrastructure
by: Larsen, Oliver Aleksander, et al.
Published: (2026)
by: Larsen, Oliver Aleksander, et al.
Published: (2026)
Deterministic and Reliable Software-Defined Vehicles: key building blocks, challenges, and vision
by: Teixeira, Pedro Veloso, et al.
Published: (2024)
by: Teixeira, Pedro Veloso, et al.
Published: (2024)
A Survey on Heterogeneous Computing Using SmartNICs and Emerging Data Processing Units
by: Tibbetts, Nathan, et al.
Published: (2025)
by: Tibbetts, Nathan, et al.
Published: (2025)
Konnektor: Connection Protocol for Ensuring Peer Uniqueness in Decentralized P2P Networks
by: Ozkan, Onur
Published: (2024)
by: Ozkan, Onur
Published: (2024)
Nezha: Deployable and High-Performance Consensus Using Synchronized Clocks
by: Geng, Jinkun, et al.
Published: (2022)
by: Geng, Jinkun, et al.
Published: (2022)
Joint Task Offloading and Routing in Wireless Multi-hop Networks Using Biased Backpressure Algorithm
by: Zhao, Zhongyuan, et al.
Published: (2024)
by: Zhao, Zhongyuan, et al.
Published: (2024)
ABACUS: A FinOps Service for Cloud Cost Optimization
by: Deochake, Saurabh
Published: (2024)
by: Deochake, Saurabh
Published: (2024)
Towards Message Brokers for Generative AI: Survey, Challenges, and Opportunities
by: Saleh, Alaa, et al.
Published: (2023)
by: Saleh, Alaa, et al.
Published: (2023)
Towards Policy-Enabled Multi-Hop Routing for Cross-Chain Message Delivery
by: Rezaei, Amin, et al.
Published: (2026)
by: Rezaei, Amin, et al.
Published: (2026)
Quantize Once, Train Fast: Allreduce-Compatible Compression with Provable Guarantees
by: Xin, Jihao, et al.
Published: (2023)
by: Xin, Jihao, et al.
Published: (2023)
Security Analysis of Bitcoin's V2 Transport Protocol: Exploiting Design Implications for Sustained Eclipse and Downgrade Attacks
by: Ndolo, Charmaine, et al.
Published: (2026)
by: Ndolo, Charmaine, et al.
Published: (2026)
Self-Adaptive Probabilistic Skyline Query Processing in Distributed Edge Computing via Deep Reinforcement Learning
by: Lai, Chuan-Chi
Published: (2026)
by: Lai, Chuan-Chi
Published: (2026)
How Can We Train Deep Learning Models Across Clouds and Continents? An Experimental Study
by: Erben, Alexander, et al.
Published: (2023)
by: Erben, Alexander, et al.
Published: (2023)
Federated Learning for Traffic Flow Prediction with Synthetic Data Augmentation
by: Orozco, Fermin, et al.
Published: (2024)
by: Orozco, Fermin, et al.
Published: (2024)
Relay-Based Synchronization of Replicated Data Types in Opportunistic Networks
by: Guidec, Frédéric, et al.
Published: (2026)
by: Guidec, Frédéric, et al.
Published: (2026)
FAST: An Efficient Scheduler for All-to-All GPU Communication
by: Lei, Yiran, et al.
Published: (2025)
by: Lei, Yiran, et al.
Published: (2025)
Prediction based computation offloading and resource allocation for multi-access ISAC enabled IoT system
by: Le, Duc-Thuan
Published: (2024)
by: Le, Duc-Thuan
Published: (2024)
Efficient All-to-All Collective Communication Schedules for Direct-Connect Topologies
by: Basu, Prithwish, et al.
Published: (2023)
by: Basu, Prithwish, et al.
Published: (2023)
Revisiting Bruck: Phase-Efficient All-to-All Communication in Reconfigurable Networks
by: Juerss, Anton, et al.
Published: (2026)
by: Juerss, Anton, et al.
Published: (2026)
RailS: Load Balancing for All-to-All Communication in Distributed Mixture-of-Experts Training
by: Xu, Heng, et al.
Published: (2025)
by: Xu, Heng, et al.
Published: (2025)
Dynamic Hierarchical Birkhoff-von Neumann Decomposition for All-to-All GPU Communication
by: Wu, Yen-Chieh, et al.
Published: (2026)
by: Wu, Yen-Chieh, et al.
Published: (2026)
A Full Compression Pipeline for Green Federated Learning in Communication-Constrained Environments
by: Colybes, Elouan, et al.
Published: (2026)
by: Colybes, Elouan, et al.
Published: (2026)
Aethon: A Reference-Based Replication Primitive for Constant-Time Instantiation of Stateful AI Agents
by: Rao, Swanand, et al.
Published: (2026)
by: Rao, Swanand, et al.
Published: (2026)
AllReduce Scheduling with Hierarchical Deep Reinforcement Learning
by: Wei, Yufan, et al.
Published: (2025)
by: Wei, Yufan, et al.
Published: (2025)
Short-circuiting Rings for Low-Latency AllReduce
by: Hammer, Sarah-Michelle, et al.
Published: (2025)
by: Hammer, Sarah-Michelle, et al.
Published: (2025)
Trivance: Latency-Optimal AllReduce by Shortcutting Multiport Networks
by: Juerss, Anton, et al.
Published: (2026)
by: Juerss, Anton, et al.
Published: (2026)
Q-adaptive: A Multi-Agent Reinforcement Learning Based Routing on Dragonfly Network
by: Kang, Yao, et al.
Published: (2024)
by: Kang, Yao, et al.
Published: (2024)
PeerSync: Accelerating Containerized Service Delivery at the Network Edge
by: Deng, Yinuo, et al.
Published: (2025)
by: Deng, Yinuo, et al.
Published: (2025)
Accelerating the Operation of Complex Workflows through Standard Data Interfaces
by: Paul, Taylor, et al.
Published: (2024)
by: Paul, Taylor, et al.
Published: (2024)
CommonSense: Efficient Set Intersection (SetX) Protocol Based on Compressed Sensing
by: Meng, Jingfan, et al.
Published: (2025)
by: Meng, Jingfan, et al.
Published: (2025)
Similar Items
-
Combinatorial Client-Master Multiagent Deep Reinforcement Learning for Task Offloading in Mobile Edge Computing
by: Gebrekidan, Tesfay Zemuy, et al.
Published: (2024) -
Asynchronous Federated Reinforcement Learning with Policy Gradient Updates: Algorithm Design and Convergence Analysis
by: Lan, Guangchen, et al.
Published: (2024) -
HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference
by: Zhang, Zeyu, et al.
Published: (2025) -
OptiReduce: Resilient and Tail-Optimal AllReduce for Distributed Deep Learning in the Cloud
by: Warraich, Ertza, et al.
Published: (2023) -
Naeural AI OS -- Decentralized ubiquitous computing MLOps execution engine
by: Bleotiu, Cristian, et al.
Published: (2023)