THC: Accelerating Distributed Deep Learning Using Tensor Homomorphic Compression
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Minghao, Basat, Ran Ben, Vargaftik, Shay, Lao, ChonLam, Xu, Kevin, Mitzenmacher, Michael, Yu, Minlan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Throughput and Compression Ratios: Towards High End-to-end Utility of Gradient Compression
di: Han, Wenchen, et al.
Pubblicazione: (2024)
di: Han, Wenchen, et al.
Pubblicazione: (2024)
DynamiQ: Accelerating Gradient Synchronization using Compressed Multi-hop All-reduce
di: Han, Wenchen, et al.
Pubblicazione: (2026)
di: Han, Wenchen, et al.
Pubblicazione: (2026)
Optimal and Near-Optimal Adaptive Vector Quantization
di: Ben-Basat, Ran, et al.
Pubblicazione: (2024)
di: Ben-Basat, Ran, et al.
Pubblicazione: (2024)
Quantizing With Randomized Hadamard Transforms: Efficient Heuristic Now Proven
di: Ben-Basat, Ran, et al.
Pubblicazione: (2026)
di: Ben-Basat, Ran, et al.
Pubblicazione: (2026)
Counter Pools: Counter Representation for Efficient Stream Processing
di: Basat, Ran Ben, et al.
Pubblicazione: (2025)
di: Basat, Ran Ben, et al.
Pubblicazione: (2025)
EdgeSight: Enabling Modeless and Cost-Efficient Inference at the Edge
di: Lao, ChonLam, et al.
Pubblicazione: (2024)
di: Lao, ChonLam, et al.
Pubblicazione: (2024)
FRANCIS: Fast Reaction Algorithms for Network Coordination In Switches
di: Han, Wenchen, et al.
Pubblicazione: (2022)
di: Han, Wenchen, et al.
Pubblicazione: (2022)
HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference
di: Zhang, Zeyu, et al.
Pubblicazione: (2025)
di: Zhang, Zeyu, et al.
Pubblicazione: (2025)
A Note on TurboQuant and the Earlier DRIVE/EDEN Line of Work
di: Ben-Basat, Ran, et al.
Pubblicazione: (2026)
di: Ben-Basat, Ran, et al.
Pubblicazione: (2026)
Sketch Disaggregation Across Time and Space
di: Langlet, Jonatan, et al.
Pubblicazione: (2025)
di: Langlet, Jonatan, et al.
Pubblicazione: (2025)
Towards Easy and Realistic Network Infrastructure Testing for Large-scale Machine Learning
di: Yoo, Jinsun, et al.
Pubblicazione: (2025)
di: Yoo, Jinsun, et al.
Pubblicazione: (2025)
An Extensible Software Transport Layer for GPU Networking
di: Zhou, Yang, et al.
Pubblicazione: (2025)
di: Zhou, Yang, et al.
Pubblicazione: (2025)
Accelerating Distributed Deep Learning using Lossless Homomorphic Compression
di: Li, Haoyu, et al.
Pubblicazione: (2024)
di: Li, Haoyu, et al.
Pubblicazione: (2024)
SQUID: Faster Analytics via Sampled Quantile Estimation
di: Ben-Basat, Ran, et al.
Pubblicazione: (2022)
di: Ben-Basat, Ran, et al.
Pubblicazione: (2022)
OptiReduce: Resilient and Tail-Optimal AllReduce for Distributed Deep Learning in the Cloud
di: Warraich, Ertza, et al.
Pubblicazione: (2023)
di: Warraich, Ertza, et al.
Pubblicazione: (2023)
OptiNIC: A Resilient and Tail-Optimal RDMA NIC for Distributed ML Workloads
di: Warraich, Ertza, et al.
Pubblicazione: (2025)
di: Warraich, Ertza, et al.
Pubblicazione: (2025)
Reimagining RDMA Through the Lens of ML
di: Warraich, Ertza, et al.
Pubblicazione: (2025)
di: Warraich, Ertza, et al.
Pubblicazione: (2025)
Teal: Learning-Accelerated Optimization of WAN Traffic Engineering
di: Xu, Zhiying, et al.
Pubblicazione: (2022)
di: Xu, Zhiying, et al.
Pubblicazione: (2022)
OLAF: Programmable Data Plane Acceleration for Asynchronous Distributed Reinforcement Learning
di: Krishna, Nehal Baganal, et al.
Pubblicazione: (2025)
di: Krishna, Nehal Baganal, et al.
Pubblicazione: (2025)
Cora: Accelerating Stateful Network Applications with SmartNICs
di: Xi, Shaoke, et al.
Pubblicazione: (2024)
di: Xi, Shaoke, et al.
Pubblicazione: (2024)
OffRAC: Offloading Through Remote Accelerator Calls
di: Yang, Ziyi, et al.
Pubblicazione: (2025)
di: Yang, Ziyi, et al.
Pubblicazione: (2025)
Adaptive Compression of Massive MIMO Channel State Information with Deep Learning
di: Mismar, Faris B., et al.
Pubblicazione: (2024)
di: Mismar, Faris B., et al.
Pubblicazione: (2024)
An LLM-based Agentic Framework for Accessible Network Control
di: Lin, Samuel, et al.
Pubblicazione: (2025)
di: Lin, Samuel, et al.
Pubblicazione: (2025)
Arcalis: Accelerating Remote Procedure Calls Using a Lightweight Near-Cache Solution
di: Umeike, Johnson, et al.
Pubblicazione: (2026)
di: Umeike, Johnson, et al.
Pubblicazione: (2026)
Learning to Compress and Transmit: Adaptive Rate Control for Semantic Communications over LEO Satellite-to-Ground Links
di: Luo, Jiangtao, et al.
Pubblicazione: (2026)
di: Luo, Jiangtao, et al.
Pubblicazione: (2026)
Memory-efficient Sketch Acceleration for Handling Large Network Flows on FPGAs
di: Han, Zhaoyang, et al.
Pubblicazione: (2025)
di: Han, Zhaoyang, et al.
Pubblicazione: (2025)
The Carrier Pigeon Internet Protocol: An Algorithmic (and Lighthearted) Perspective
di: Bentert, Matthias, et al.
Pubblicazione: (2026)
di: Bentert, Matthias, et al.
Pubblicazione: (2026)
Distributed Resource Allocation and Application Deployment in Mesh Edge Networks
di: Bernard, Antoine, et al.
Pubblicazione: (2025)
di: Bernard, Antoine, et al.
Pubblicazione: (2025)
State-Aware IoT Scheduling Using Deep Q-Networks and Edge-Based Coordination
di: He, Qingyuan, et al.
Pubblicazione: (2025)
di: He, Qingyuan, et al.
Pubblicazione: (2025)
NetFlowGen: Leveraging Generative Pre-training for Network Traffic Dynamics
di: Zhou, Jiawei, et al.
Pubblicazione: (2024)
di: Zhou, Jiawei, et al.
Pubblicazione: (2024)
Selfish Carrier Monitoring in WIFI Using Distributed Sniffers
di: Sinthuja, U, et al.
Pubblicazione: (2024)
di: Sinthuja, U, et al.
Pubblicazione: (2024)
Scalability Assurance in SFC provisioning via Distributed Design for Deep Reinforcement Learning
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
Distributed Split Computing Using Diffusive Metrics for UAV Swarms
di: Sarı, Talip Tolga, et al.
Pubblicazione: (2025)
di: Sarı, Talip Tolga, et al.
Pubblicazione: (2025)
Diagnosing and Repairing Distributed Routing Configurations Using Selective Symbolic Simulation
di: Yang, Rulan, et al.
Pubblicazione: (2024)
di: Yang, Rulan, et al.
Pubblicazione: (2024)
Lossy Compression of Cellular Network KPIs
di: Pimpinella, Andrea, et al.
Pubblicazione: (2026)
di: Pimpinella, Andrea, et al.
Pubblicazione: (2026)
Distributed Accountability in Democracy: Using MANETs and DTNs in the Face of Acts of Questionable Legality
di: Schmidheiser, Mathew, et al.
Pubblicazione: (2025)
di: Schmidheiser, Mathew, et al.
Pubblicazione: (2025)
Toward Communication-Efficient Space Data Centers: Bottlenecks, Architectures, and New Paradigms
di: Sun, Minghao, et al.
Pubblicazione: (2026)
di: Sun, Minghao, et al.
Pubblicazione: (2026)
Cooperative Jamming for Physical Layer Security Enhancement Using Deep Reinforcement Learning
di: Hoseini, Sayed Amir, et al.
Pubblicazione: (2024)
di: Hoseini, Sayed Amir, et al.
Pubblicazione: (2024)
Hybrid Centralized-Distributed Resource Allocation Based on Deep Reinforcement Learning for Cooperative D2D Communications
di: Yu, Yang, et al.
Pubblicazione: (2024)
di: Yu, Yang, et al.
Pubblicazione: (2024)
Joint Scheduling and Resource Allocation in mmWave IAB Networks Using Deep RL
di: Abbasalizadeh, Maryam, et al.
Pubblicazione: (2025)
di: Abbasalizadeh, Maryam, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Beyond Throughput and Compression Ratios: Towards High End-to-end Utility of Gradient Compression
di: Han, Wenchen, et al.
Pubblicazione: (2024) -
DynamiQ: Accelerating Gradient Synchronization using Compressed Multi-hop All-reduce
di: Han, Wenchen, et al.
Pubblicazione: (2026) -
Optimal and Near-Optimal Adaptive Vector Quantization
di: Ben-Basat, Ran, et al.
Pubblicazione: (2024) -
Quantizing With Randomized Hadamard Transforms: Efficient Heuristic Now Proven
di: Ben-Basat, Ran, et al.
Pubblicazione: (2026) -
Counter Pools: Counter Representation for Efficient Stream Processing
di: Basat, Ran Ben, et al.
Pubblicazione: (2025)