CARVQ: Corrective Adaptor with Group Residual Vector Quantization for LLM Embedding Compression
Fuente:
arXiv
Saved in:
| Main Authors: | Gou, Dayin, Byun, Sanghyun, Malpeddi, Nilesh, De Micheli, Gabrielle, Vaste, Prathamesh, Song, Jacob, Chung, Woo Seong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MultiDepth: Multi-Sample Priors for Refining Monocular Metric Depth Estimations in Indoor Scenes
by: Byun, Sanghyun, et al.
Published: (2024)
by: Byun, Sanghyun, et al.
Published: (2024)
APCE: Adaptive Progressive Context Expansion for Long Context Processing
by: Lee, Baisub, et al.
Published: (2025)
by: Lee, Baisub, et al.
Published: (2025)
OneNet: A Channel-Wise 1D Convolutional U-Net
by: Byun, Sanghyun, et al.
Published: (2024)
by: Byun, Sanghyun, et al.
Published: (2024)
Unifying Vision-Language Latents for Zero-label Image Caption Enhancement
by: Byun, Sanghyun, et al.
Published: (2025)
by: Byun, Sanghyun, et al.
Published: (2025)
3-Model Speculative Decoding
by: Byun, Sanghyun, et al.
Published: (2025)
by: Byun, Sanghyun, et al.
Published: (2025)
Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression
by: Zhang, Xi, et al.
Published: (2025)
by: Zhang, Xi, et al.
Published: (2025)
ProGIC: Progressive and Lightweight Generative Image Compression with Residual Vector Quantization
by: Cao, Hao, et al.
Published: (2026)
by: Cao, Hao, et al.
Published: (2026)
RQ-MoE: Residual Quantization via Mixture of Experts for Efficient Input-Dependent Vector Compression
by: Zhong, Zhengjia, et al.
Published: (2026)
by: Zhong, Zhengjia, et al.
Published: (2026)
Leech Lattice Vector Quantization for Efficient LLM Compression
by: van der Ouderaa, Tycho F. A., et al.
Published: (2026)
by: van der Ouderaa, Tycho F. A., et al.
Published: (2026)
Search-Adaptor: Embedding Customization for Information Retrieval
by: Yoon, Jinsung, et al.
Published: (2023)
by: Yoon, Jinsung, et al.
Published: (2023)
Variable Bitrate Residual Vector Quantization for Audio Coding
by: Chae, Yunkee, et al.
Published: (2024)
by: Chae, Yunkee, et al.
Published: (2024)
Matryoshka-Adaptor: Unsupervised and Supervised Tuning for Smaller Embedding Dimensions
by: Yoon, Jinsung, et al.
Published: (2024)
by: Yoon, Jinsung, et al.
Published: (2024)
HE-LRM: Efficient Private Embedding Lookups for Neural Inference Using Fully Homomorphic Encryption
by: Garimella, Karthik, et al.
Published: (2025)
by: Garimella, Karthik, et al.
Published: (2025)
HAS-VQ: Hessian-Adaptive Sparse Vector Quantization for High-Fidelity LLM Compression
by: Khasia, Vladimer
Published: (2026)
by: Khasia, Vladimer
Published: (2026)
Robust Residual Finite Scalar Quantization for Neural Compression
by: Zhu, Xiaoxu, et al.
Published: (2025)
by: Zhu, Xiaoxu, et al.
Published: (2025)
4bit-Quantization in Vector-Embedding for RAG
by: Jeong, Taehee
Published: (2025)
by: Jeong, Taehee
Published: (2025)
Balance of Number of Embedding and their Dimensions in Vector Quantization
by: Chen, Hang, et al.
Published: (2024)
by: Chen, Hang, et al.
Published: (2024)
Residual Vector Quantization For Communication-Efficient Multi-Agent Perception
by: Shenkut, Dereje, et al.
Published: (2025)
by: Shenkut, Dereje, et al.
Published: (2025)
Efficient Generative Modeling with Residual Vector Quantization-Based Tokens
by: Kim, Jaehyeon, et al.
Published: (2024)
by: Kim, Jaehyeon, et al.
Published: (2024)
Residual Governance
by: Hecht, Gabrielle
Published: (2023)
by: Hecht, Gabrielle
Published: (2023)
FaTRQ: Tiered Residual Quantization for LLM Vector Search in Far-Memory-Aware ANNS Systems
by: Zhang, Tianqi, et al.
Published: (2026)
by: Zhang, Tianqi, et al.
Published: (2026)
Mamba-Adaptor: State Space Model Adaptor for Visual Recognition
by: Xie, Fei, et al.
Published: (2025)
by: Xie, Fei, et al.
Published: (2025)
Vector-Quantized Soft Label Compression for Dataset Distillation
by: Abbasi, Ali, et al.
Published: (2026)
by: Abbasi, Ali, et al.
Published: (2026)
LitE-SQL: A Lightweight and Efficient Text-to-SQL Framework with Vector-based Schema Linking and Execution-Guided Self-Correction
by: Piao, Shengmin, et al.
Published: (2025)
by: Piao, Shengmin, et al.
Published: (2025)
QET: Enhancing Quantized LLM Parameters and KV cache Compression through Element Substitution and Residual Clustering
by: Wang, Yanshu, et al.
Published: (2024)
by: Wang, Yanshu, et al.
Published: (2024)
The Impact of Teacher Social‐Emotional Competence Training on Students' Social‐Emotional Competence: A Theoretical Framework Based on the KAB Model and Empirical Analysis Using SSES Data
by: Dayin Li, et al.
Published: (2025)
by: Dayin Li, et al.
Published: (2025)
Low-Rank Adapters Meet Neural Architecture Search for LLM Compression
by: Muñoz, J. Pablo, et al.
Published: (2025)
by: Muñoz, J. Pablo, et al.
Published: (2025)
Quantized Embedding Vectors for Controllable Diffusion Language Models
by: Kang, Cheng, et al.
Published: (2024)
by: Kang, Cheng, et al.
Published: (2024)
GranQ: Efficient Channel-wise Quantization via Vectorized Pre-Scaling for Zero-Shot QAT
by: Hong, Inpyo, et al.
Published: (2025)
by: Hong, Inpyo, et al.
Published: (2025)
CommVQ: Commutative Vector Quantization for KV Cache Compression
by: Li, Junyan, et al.
Published: (2025)
by: Li, Junyan, et al.
Published: (2025)
CRVQ: Channel-Relaxed Vector Quantization for Extreme Compression of LLMs
by: Xu, Yuzhuang, et al.
Published: (2024)
by: Xu, Yuzhuang, et al.
Published: (2024)
ESC: Efficient Speech Coding with Cross-Scale Residual Vector Quantized Transformers
by: Gu, Yuzhe, et al.
Published: (2024)
by: Gu, Yuzhe, et al.
Published: (2024)
Making Pose Representations More Expressive and Disentangled via Residual Vector Quantization
by: Jeong, Sukhyun, et al.
Published: (2025)
by: Jeong, Sukhyun, et al.
Published: (2025)
Rethinking Residual Errors in Compensation-based LLM Quantization
by: Li, Shuaiting, et al.
Published: (2026)
by: Li, Shuaiting, et al.
Published: (2026)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
by: Kim, Bumsoo, et al.
Published: (2024)
by: Kim, Bumsoo, et al.
Published: (2024)
SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories
by: Yu, Zhuoyun, et al.
Published: (2026)
by: Yu, Zhuoyun, et al.
Published: (2026)
RAVE: Residual Vector Embedding for CLIP-Guided Backlit Image Enhancement
by: Gaintseva, Tatiana, et al.
Published: (2024)
by: Gaintseva, Tatiana, et al.
Published: (2024)
Hierarchical Vector-Quantized Latents for Perceptual Low-Resolution Video Compression
by: Kotthapalli, Manikanta, et al.
Published: (2025)
by: Kotthapalli, Manikanta, et al.
Published: (2025)
RAVQ-HoloNet: Rate-Adaptive Vector-Quantized Hologram Compression
by: Rafiei, Shima, et al.
Published: (2025)
by: Rafiei, Shima, et al.
Published: (2025)
Differentiable Vector Quantization for Rate-Distortion Optimization of Generative Image Compression
by: Jiang, Shiyin, et al.
Published: (2026)
by: Jiang, Shiyin, et al.
Published: (2026)
Similar Items
-
MultiDepth: Multi-Sample Priors for Refining Monocular Metric Depth Estimations in Indoor Scenes
by: Byun, Sanghyun, et al.
Published: (2024) -
APCE: Adaptive Progressive Context Expansion for Long Context Processing
by: Lee, Baisub, et al.
Published: (2025) -
OneNet: A Channel-Wise 1D Convolutional U-Net
by: Byun, Sanghyun, et al.
Published: (2024) -
Unifying Vision-Language Latents for Zero-label Image Caption Enhancement
by: Byun, Sanghyun, et al.
Published: (2025) -
3-Model Speculative Decoding
by: Byun, Sanghyun, et al.
Published: (2025)