List-Level Distribution Coupling with Applications to Speculative Decoding and Lossy Compression
Fuente:
arXiv
Saved in:
| Main Authors: | Rowan, Joseph, Phan, Buu, Khisti, Ashish |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Channel Simulation and Distributed Compression with Ensemble Rejection Sampling
by: Phan, Buu, et al.
Published: (2025)
by: Phan, Buu, et al.
Published: (2025)
Importance Matching Lemma for Lossy Compression with Side Information
by: Phan, Buu, et al.
Published: (2024)
by: Phan, Buu, et al.
Published: (2024)
On Self-Adaptive Perception Loss Function for Sequential Lossy Compression
by: Salehkalaibar, Sadaf, et al.
Published: (2025)
by: Salehkalaibar, Sadaf, et al.
Published: (2025)
Cross-Tokenizer Likelihood Scoring Algorithms for Language Model Distillation
by: Phan, Buu, et al.
Published: (2025)
by: Phan, Buu, et al.
Published: (2025)
One-Shot Broadcast Joint Source-Channel Coding with Codebook Diversity
by: Rowan, Joseph, et al.
Published: (2026)
by: Rowan, Joseph, et al.
Published: (2026)
Multi-Marginal Couplings for Metropolis-Hastings
by: Phan, Buu, et al.
Published: (2026)
by: Phan, Buu, et al.
Published: (2026)
Random Cycle Coding: Lossless Compression of Cluster Assignments via Bits-Back Coding
by: Severo, Daniel, et al.
Published: (2024)
by: Severo, Daniel, et al.
Published: (2024)
Minimum Entropy Coupling with Bottleneck
by: Ebrahimi, M. Reza, et al.
Published: (2024)
by: Ebrahimi, M. Reza, et al.
Published: (2024)
Black-Box Detection of LLM-Generated Text Using Generalized Jensen-Shannon Divergence
by: Chen, Shuangyi, et al.
Published: (2025)
by: Chen, Shuangyi, et al.
Published: (2025)
Rate-Distortion-Perception Tradeoff Based on the Conditional-Distribution Perception Measure
by: Salehkalaibar, Sadaf, et al.
Published: (2024)
by: Salehkalaibar, Sadaf, et al.
Published: (2024)
Speculative Speculative Decoding
by: Kumar, Tanishq, et al.
Published: (2026)
by: Kumar, Tanishq, et al.
Published: (2026)
Multi-Draft Speculative Sampling: Canonical Decomposition and Theoretical Limits
by: Khisti, Ashish, et al.
Published: (2024)
by: Khisti, Ashish, et al.
Published: (2024)
Delayed Attention Training Improves Length Generalization in Transformer--RNN Hybrids
by: Phan, Buu, et al.
Published: (2025)
by: Phan, Buu, et al.
Published: (2025)
Cross-Domain Lossy Compression via Constrained Minimum Entropy Coupling
by: Nguyen, Nam, et al.
Published: (2026)
by: Nguyen, Nam, et al.
Published: (2026)
Exact Byte-Level Probabilities from Tokenized Language Models for FIM-Tasks and Model Ensembles
by: Phan, Buu, et al.
Published: (2024)
by: Phan, Buu, et al.
Published: (2024)
Decoding Speculative Decoding
by: Yan, Minghao, et al.
Published: (2024)
by: Yan, Minghao, et al.
Published: (2024)
Generative Decompression: Optimal Lossy Decoding Against Distribution Mismatch
by: Khosravirad, Saeed R., et al.
Published: (2026)
by: Khosravirad, Saeed R., et al.
Published: (2026)
Understanding and Mitigating Tokenization Bias in Language Models
by: Phan, Buu, et al.
Published: (2024)
by: Phan, Buu, et al.
Published: (2024)
Beat the long tail: Distribution-Aware Speculative Decoding for RL Training
by: Shao, Zelei, et al.
Published: (2025)
by: Shao, Zelei, et al.
Published: (2025)
Coupling without Communication and Drafter-Invariant Speculative Decoding
by: Daliri, Majid, et al.
Published: (2024)
by: Daliri, Majid, et al.
Published: (2024)
Foundation Model for Lossy Compression of Spatiotemporal Scientific Data
by: Li, Xiao, et al.
Published: (2024)
by: Li, Xiao, et al.
Published: (2024)
On the Generalization of Stochastic Gradient Descent with Momentum
by: Ramezani-Kebrya, Ali, et al.
Published: (2018)
by: Ramezani-Kebrya, Ali, et al.
Published: (2018)
Exploration-Exploitation Tradeoff in Universal Lossy Compression
by: Weinberger, Nir, et al.
Published: (2025)
by: Weinberger, Nir, et al.
Published: (2025)
Speculative Safety-Aware Decoding
by: Wang, Xuekang, et al.
Published: (2025)
by: Wang, Xuekang, et al.
Published: (2025)
Speculative Decoding Across Languages
by: Paudel, Nirajan, et al.
Published: (2026)
by: Paudel, Nirajan, et al.
Published: (2026)
Enhancing Lossy Compression Through Cross-Field Information for Scientific Applications
by: Liu, Youyuan, et al.
Published: (2024)
by: Liu, Youyuan, et al.
Published: (2024)
ML-SpecQD: Multi-Level Speculative Decoding with Quantized Drafts
by: Georganas, Evangelos, et al.
Published: (2025)
by: Georganas, Evangelos, et al.
Published: (2025)
Online Speculative Decoding
by: Liu, Xiaoxuan, et al.
Published: (2023)
by: Liu, Xiaoxuan, et al.
Published: (2023)
SuffixDecoding: Extreme Speculative Decoding for Emerging AI Applications
by: Oliaro, Gabriele, et al.
Published: (2024)
by: Oliaro, Gabriele, et al.
Published: (2024)
Accelerating Communication in Deep Learning Recommendation Model Training with Dual-Level Adaptive Lossy Compression
by: Feng, Hao, et al.
Published: (2024)
by: Feng, Hao, et al.
Published: (2024)
Fast Inference via Hierarchical Speculative Decoding
by: Mohri, Clara, et al.
Published: (2025)
by: Mohri, Clara, et al.
Published: (2025)
Robust Federated Finetuning of Foundation Models via Alternating Minimization of LoRA
by: Chen, Shuangyi, et al.
Published: (2024)
by: Chen, Shuangyi, et al.
Published: (2024)
Scaling Speculative Decoding with Lookahead Reasoning
by: Fu, Yichao, et al.
Published: (2025)
by: Fu, Yichao, et al.
Published: (2025)
Steering Pretrained Drafters during Speculative Decoding
by: Berdoz, Frédéric, et al.
Published: (2025)
by: Berdoz, Frédéric, et al.
Published: (2025)
Robust Federated Finetuning of LLMs via Alternating Optimization of LoRA
by: Chen, Shuangyi, et al.
Published: (2025)
by: Chen, Shuangyi, et al.
Published: (2025)
Polybasic Speculative Decoding Through a Theoretical Perspective
by: Wang, Ruilin, et al.
Published: (2025)
by: Wang, Ruilin, et al.
Published: (2025)
FastEagle: Cascaded Drafting for Accelerating Speculative Decoding
by: Huang, Haiduo, et al.
Published: (2025)
by: Huang, Haiduo, et al.
Published: (2025)
SPIRe: Boosting LLM Inference Throughput with Speculative Decoding
by: Neelam, Sanjit, et al.
Published: (2025)
by: Neelam, Sanjit, et al.
Published: (2025)
Accelerating Time Series Foundation Models with Speculative Decoding
by: Subbaraman, Pranav, et al.
Published: (2025)
by: Subbaraman, Pranav, et al.
Published: (2025)
Mixture of Attentions For Speculative Decoding
by: Zimmer, Matthieu, et al.
Published: (2024)
by: Zimmer, Matthieu, et al.
Published: (2024)
Similar Items
-
Channel Simulation and Distributed Compression with Ensemble Rejection Sampling
by: Phan, Buu, et al.
Published: (2025) -
Importance Matching Lemma for Lossy Compression with Side Information
by: Phan, Buu, et al.
Published: (2024) -
On Self-Adaptive Perception Loss Function for Sequential Lossy Compression
by: Salehkalaibar, Sadaf, et al.
Published: (2025) -
Cross-Tokenizer Likelihood Scoring Algorithms for Language Model Distillation
by: Phan, Buu, et al.
Published: (2025) -
One-Shot Broadcast Joint Source-Channel Coding with Codebook Diversity
by: Rowan, Joseph, et al.
Published: (2026)