A Note on TurboQuant and the Earlier DRIVE/EDEN Line of Work
Fuente:
arXiv
Saved in:
| Main Authors: | Ben-Basat, Ran, Ben-Itzhak, Yaniv, Mendelson, Gal, Mitzenmacher, Michael, Portnoy, Amit, Vargaftik, Shay |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimal and Near-Optimal Adaptive Vector Quantization
by: Ben-Basat, Ran, et al.
Published: (2024)
by: Ben-Basat, Ran, et al.
Published: (2024)
Quantizing With Randomized Hadamard Transforms: Efficient Heuristic Now Proven
by: Ben-Basat, Ran, et al.
Published: (2026)
by: Ben-Basat, Ran, et al.
Published: (2026)
Beyond Throughput and Compression Ratios: Towards High End-to-end Utility of Gradient Compression
by: Han, Wenchen, et al.
Published: (2024)
by: Han, Wenchen, et al.
Published: (2024)
DynamiQ: Accelerating Gradient Synchronization using Compressed Multi-hop All-reduce
by: Han, Wenchen, et al.
Published: (2026)
by: Han, Wenchen, et al.
Published: (2026)
Counter Pools: Counter Representation for Efficient Stream Processing
by: Basat, Ran Ben, et al.
Published: (2025)
by: Basat, Ran Ben, et al.
Published: (2025)
THC: Accelerating Distributed Deep Learning Using Tensor Homomorphic Compression
by: Li, Minghao, et al.
Published: (2023)
by: Li, Minghao, et al.
Published: (2023)
Sketch Disaggregation Across Time and Space
by: Langlet, Jonatan, et al.
Published: (2025)
by: Langlet, Jonatan, et al.
Published: (2025)
FRANCIS: Fast Reaction Algorithms for Network Coordination In Switches
by: Han, Wenchen, et al.
Published: (2022)
by: Han, Wenchen, et al.
Published: (2022)
SQUID: Faster Analytics via Sampled Quantile Estimation
by: Ben-Basat, Ran, et al.
Published: (2022)
by: Ben-Basat, Ran, et al.
Published: (2022)
Design Conductor 2.0: An agent builds a TurboQuant inference accelerator in 80 hours
by: The Verkor Team, et al.
Published: (2026)
by: The Verkor Team, et al.
Published: (2026)
OptiNIC: A Resilient and Tail-Optimal RDMA NIC for Distributed ML Workloads
by: Warraich, Ertza, et al.
Published: (2025)
by: Warraich, Ertza, et al.
Published: (2025)
Reimagining RDMA Through the Lens of ML
by: Warraich, Ertza, et al.
Published: (2025)
by: Warraich, Ertza, et al.
Published: (2025)
SareQuant: Towards a quantum-based communication network
by: Sanz, Ane, et al.
Published: (2024)
by: Sanz, Ane, et al.
Published: (2024)
OptiReduce: Resilient and Tail-Optimal AllReduce for Distributed Deep Learning in the Cloud
by: Warraich, Ertza, et al.
Published: (2023)
by: Warraich, Ertza, et al.
Published: (2023)
Evaluating Small Language Models for Front-Door Routing: A Harmonized Benchmark and Synthetic-Traffic Experiment
by: Johnson, Warren, et al.
Published: (2026)
by: Johnson, Warren, et al.
Published: (2026)
HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference
by: Zhang, Zeyu, et al.
Published: (2025)
by: Zhang, Zeyu, et al.
Published: (2025)
Semantic Caching for Improving Web Affordability
by: Akbar, Hafsa, et al.
Published: (2025)
by: Akbar, Hafsa, et al.
Published: (2025)
Representing LLMs in Prompt Semantic Task Space
by: Kashani, Idan, et al.
Published: (2025)
by: Kashani, Idan, et al.
Published: (2025)
Optimizing Energy and Latency in 6G Smart Cities with Edge CyberTwins
by: Abouaomar, Amine, et al.
Published: (2025)
by: Abouaomar, Amine, et al.
Published: (2025)
Langues en danger et multilinguisme num{é}rique
by: Henda, Mokhtar Ben
Published: (2024)
by: Henda, Mokhtar Ben
Published: (2024)
Technical Report: Performance Comparison of Service Mesh Frameworks: the MTLS Test Case
by: Barr, Anat Bremler, et al.
Published: (2024)
by: Barr, Anat Bremler, et al.
Published: (2024)
The Carrier Pigeon Internet Protocol: An Algorithmic (and Lighthearted) Perspective
by: Bentert, Matthias, et al.
Published: (2026)
by: Bentert, Matthias, et al.
Published: (2026)
Distributed Resource Allocation and Application Deployment in Mesh Edge Networks
by: Bernard, Antoine, et al.
Published: (2025)
by: Bernard, Antoine, et al.
Published: (2025)
Intelligent Channel Allocation for IEEE 802.11be Multi-Link Operation: When MAB Meets LLM
by: Lian, Shumin, et al.
Published: (2025)
by: Lian, Shumin, et al.
Published: (2025)
Advantages of Global Entanglement-Distillation Policies in Quantum Repeater Chains
by: Yakar, Iftach, et al.
Published: (2025)
by: Yakar, Iftach, et al.
Published: (2025)
Trust, But Verify, Operator-Reported Geolocation
by: Izhikevich, Katherine, et al.
Published: (2024)
by: Izhikevich, Katherine, et al.
Published: (2024)
An Algorithmic Approach to Line Construction in Existing Transit Networks
by: Deng, Zezhi, et al.
Published: (2024)
by: Deng, Zezhi, et al.
Published: (2024)
Programmable Packet Scheduling with Dynamic Reordering at Line Rate
by: Wang, Zekun, et al.
Published: (2026)
by: Wang, Zekun, et al.
Published: (2026)
Communication Traffic Characteristics Reveal an IoT Devices Identity
by: Chowdhury, Rajarshi Roy, et al.
Published: (2024)
by: Chowdhury, Rajarshi Roy, et al.
Published: (2024)
Empirical Line-of-Sight Probability Modeling for UAVs in Random Urban Layouts
by: Saboor, Abdul, et al.
Published: (2025)
by: Saboor, Abdul, et al.
Published: (2025)
SpliDT: Partitioned Decision Trees for Scalable Stateful Inference at Line Rate
by: Parvez, Murayyiam, et al.
Published: (2025)
by: Parvez, Murayyiam, et al.
Published: (2025)
FlexiNS: A SmartNIC-Centric, Line-Rate and Flexible Network Stack
by: Chen, Xuzheng, et al.
Published: (2025)
by: Chen, Xuzheng, et al.
Published: (2025)
Traffic-Aware Domain Partitioning and Load-Balanced Inter-Domain Routing for LEO Satellite Networks
by: Zhou, Chen, et al.
Published: (2026)
by: Zhou, Chen, et al.
Published: (2026)
Task Scheduling in Space-Air-Ground Uniformly Integrated Networks with Ripple Effects
by: Huang, Chuan, et al.
Published: (2025)
by: Huang, Chuan, et al.
Published: (2025)
Joint Semantic Coding and Routing for Multi-Hop Semantic Transmission in LEO Satellite Networks
by: Zeng, Hong, et al.
Published: (2026)
by: Zeng, Hong, et al.
Published: (2026)
IRR-Based AS Type of Relationship Inference
by: Zulan, Amit, et al.
Published: (2025)
by: Zulan, Amit, et al.
Published: (2025)
Notes on Degeneracy and Robustness
by: Dey, Indrakshi, et al.
Published: (2025)
by: Dey, Indrakshi, et al.
Published: (2025)
Hercules: Heterogeneous Requirements Congestion Control Protocol
by: Rozen-Schiff, Neta, et al.
Published: (2024)
by: Rozen-Schiff, Neta, et al.
Published: (2024)
5G Network Automation Using Local Large Language Models and Retrieval-Augmented Generation
by: Majlesara, Ahmadreza, et al.
Published: (2025)
by: Majlesara, Ahmadreza, et al.
Published: (2025)
AIvailable: A Software-Defined Architecture for LLM-as-a-Service on Heterogeneous and Legacy GPUs
by: Antunes, Pedro, et al.
Published: (2025)
by: Antunes, Pedro, et al.
Published: (2025)
Similar Items
-
Optimal and Near-Optimal Adaptive Vector Quantization
by: Ben-Basat, Ran, et al.
Published: (2024) -
Quantizing With Randomized Hadamard Transforms: Efficient Heuristic Now Proven
by: Ben-Basat, Ran, et al.
Published: (2026) -
Beyond Throughput and Compression Ratios: Towards High End-to-end Utility of Gradient Compression
by: Han, Wenchen, et al.
Published: (2024) -
DynamiQ: Accelerating Gradient Synchronization using Compressed Multi-hop All-reduce
by: Han, Wenchen, et al.
Published: (2026) -
Counter Pools: Counter Representation for Efficient Stream Processing
by: Basat, Ran Ben, et al.
Published: (2025)