Saved in:
| Main Author: | D'Alberto, Paolo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.08114 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FibQuant: Universal Vector Quantization for Random-Access KV-Cache Compression
by: Lee, Namyoon, et al.
Published: (2026)
by: Lee, Namyoon, et al.
Published: (2026)
RateQuant: Optimal Mixed-Precision KV Cache Quantization via Rate-Distortion Theory
by: Zuo, Fei, et al.
Published: (2026)
by: Zuo, Fei, et al.
Published: (2026)
TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate
by: Zandieh, Amir, et al.
Published: (2025)
by: Zandieh, Amir, et al.
Published: (2025)
Scorio.jl: A Julia package for ranking stochastic responses
by: Hariri, Mohsen, et al.
Published: (2026)
by: Hariri, Mohsen, et al.
Published: (2026)
A Note on TurboQuant and the Earlier DRIVE/EDEN Line of Work
by: Ben-Basat, Ran, et al.
Published: (2026)
by: Ben-Basat, Ran, et al.
Published: (2026)
TurboSAT: Gradient-Guided Boolean Satisfiability Accelerated on GPU-CPU Hybrid System
by: Dai, Steve, et al.
Published: (2025)
by: Dai, Steve, et al.
Published: (2025)
Revisiting RaBitQ and TurboQuant: A Symmetric Comparison of Methods, Theory, and Experiments
by: Gao, Jianyang, et al.
Published: (2026)
by: Gao, Jianyang, et al.
Published: (2026)
Statistical Complexity of Quantum Learning
by: Banchi, Leonardo, et al.
Published: (2023)
by: Banchi, Leonardo, et al.
Published: (2023)
Spectral Toolkit of Algorithms for Graphs: Technical Report (2)
by: Macgregor, Peter, et al.
Published: (2024)
by: Macgregor, Peter, et al.
Published: (2024)
Bottlenecked Transformers: Periodic KV Cache Consolidation for Generalised Reasoning
by: Oomerjee, Adnan, et al.
Published: (2025)
by: Oomerjee, Adnan, et al.
Published: (2025)
Zero-Truncated Poisson Regression for Sparse Multiway Count Data Corrupted by False Zeros
by: López, Oscar, et al.
Published: (2022)
by: López, Oscar, et al.
Published: (2022)
SQuat: Subspace-orthogonal KV Cache Quantization
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
NUBO: A Transparent Python Package for Bayesian Optimization
by: Diessner, Mike, et al.
Published: (2023)
by: Diessner, Mike, et al.
Published: (2023)
scikit-fda: A Python Package for Functional Data Analysis
by: Ramos-Carreño, Carlos, et al.
Published: (2022)
by: Ramos-Carreño, Carlos, et al.
Published: (2022)
evomap: A Toolbox for Dynamic Mapping in Python
by: Matthe, Maximilian
Published: (2025)
by: Matthe, Maximilian
Published: (2025)
Rethinking KV Cache Eviction via a Unified Information-Theoretic Objective
by: Yang, Jiaming, et al.
Published: (2026)
by: Yang, Jiaming, et al.
Published: (2026)
eOptShrinkQ: Near-Lossless KV Cache Compression Through Optimal Spectral Denoising and Quantization
by: Su, Pei-Chun
Published: (2026)
by: Su, Pei-Chun
Published: (2026)
Virtual Parameter Sharpening: Dynamic Low-Rank Perturbations for Inference-Time Reasoning Enhancement
by: Kublashvili, Saba
Published: (2025)
by: Kublashvili, Saba
Published: (2025)
A Fast and Scalable Pathwise-Solver for Group Lasso and Elastic Net Penalized Regression via Block-Coordinate Descent
by: Yang, James, et al.
Published: (2024)
by: Yang, James, et al.
Published: (2024)
Self-Scaled Broyden Family of Quasi-Newton Methods in JAX
by: Bioli, Ivan, et al.
Published: (2026)
by: Bioli, Ivan, et al.
Published: (2026)
Differentiable Parameter Optimization for DAEs with State-Dependent Events
by: Matei, Ion, et al.
Published: (2026)
by: Matei, Ion, et al.
Published: (2026)
Dimensional Peeking for Low-Variance Gradients in Zeroth-Order Discrete Optimization via Simulation
by: Andelfinger, Philipp, et al.
Published: (2026)
by: Andelfinger, Philipp, et al.
Published: (2026)
auto-fpt: Automating Free Probability Theory Calculations for Machine Learning Theory
by: Subramonian, Arjun, et al.
Published: (2025)
by: Subramonian, Arjun, et al.
Published: (2025)
MissMecha: An All-in-One Python Package for Studying Missing Data Mechanisms
by: Zhou, Youran, et al.
Published: (2025)
by: Zhou, Youran, et al.
Published: (2025)
Jaya R Package -- A Parameter-Free Solution for Advanced Single and Multi-Objective Optimization
by: Bokde, Neeraj Dhanraj
Published: (2024)
by: Bokde, Neeraj Dhanraj
Published: (2024)
A method of using RSVD in residual calculation of LowBit GEMM
by: Gu, Hongyaoxing
Published: (2024)
by: Gu, Hongyaoxing
Published: (2024)
CausalVerse: Benchmarking Causal Representation Learning with Configurable High-Fidelity Simulations
by: Chen, Guangyi, et al.
Published: (2025)
by: Chen, Guangyi, et al.
Published: (2025)
EarlyStopping: Implicit Regularization for Iterative Learning Procedures in Python
by: Ziebell, Eric, et al.
Published: (2025)
by: Ziebell, Eric, et al.
Published: (2025)
Sparser, Better, Faster, Stronger: Sparsity Detection for Efficient Automatic Differentiation
by: Hill, Adrian, et al.
Published: (2025)
by: Hill, Adrian, et al.
Published: (2025)
Reproducibility, energy efficiency and performance of pseudorandom number generators in machine learning: a comparative study of python, numpy, tensorflow, and pytorch implementations
by: Antunes, Benjamin, et al.
Published: (2024)
by: Antunes, Benjamin, et al.
Published: (2024)
Cooper: A Library for Constrained Optimization in Deep Learning
by: Gallego-Posada, Jose, et al.
Published: (2025)
by: Gallego-Posada, Jose, et al.
Published: (2025)
NNTile: a machine learning framework capable of training extremely large GPT language models on a single node
by: Mikhalev, Aleksandr, et al.
Published: (2025)
by: Mikhalev, Aleksandr, et al.
Published: (2025)
PlasmoData.jl -- A Julia Framework for Modeling and Analyzing Complex Data as Graphs
by: Cole, David L, et al.
Published: (2024)
by: Cole, David L, et al.
Published: (2024)
TorchDA: A Python package for performing data assimilation with deep learning forward and transformation functions
by: Cheng, Sibo, et al.
Published: (2024)
by: Cheng, Sibo, et al.
Published: (2024)
BlackJAX: Composable Bayesian inference in JAX
by: Cabezas, Alberto, et al.
Published: (2024)
by: Cabezas, Alberto, et al.
Published: (2024)
MinGRU-Based Encoder for Turbo Autoencoder Frameworks
by: Fritschek, Rick, et al.
Published: (2025)
by: Fritschek, Rick, et al.
Published: (2025)
Rigorous dynamical mean field theory for stochastic gradient descent methods
by: Gerbelot, Cedric, et al.
Published: (2022)
by: Gerbelot, Cedric, et al.
Published: (2022)
Adaptation of XAI to Auto-tuning for Numerical Libraries
by: Aoki, Shota, et al.
Published: (2024)
by: Aoki, Shota, et al.
Published: (2024)
Polynomial Context-Truncation Sensitivity in Autoregressive Language Models: Sequential Wyner-Ziv Bounds for KV Cache Compression
by: Kim, Munsik
Published: (2026)
by: Kim, Munsik
Published: (2026)
Linearized Optimal Transport pyLOT Library: A Toolkit for Machine Learning on Point Clouds
by: Linwu, Jun, et al.
Published: (2025)
by: Linwu, Jun, et al.
Published: (2025)
Similar Items
-
FibQuant: Universal Vector Quantization for Random-Access KV-Cache Compression
by: Lee, Namyoon, et al.
Published: (2026) -
RateQuant: Optimal Mixed-Precision KV Cache Quantization via Rate-Distortion Theory
by: Zuo, Fei, et al.
Published: (2026) -
TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate
by: Zandieh, Amir, et al.
Published: (2025) -
Scorio.jl: A Julia package for ranking stochastic responses
by: Hariri, Mohsen, et al.
Published: (2026) -
A Note on TurboQuant and the Earlier DRIVE/EDEN Line of Work
by: Ben-Basat, Ran, et al.
Published: (2026)