Reclaiming Residual Knowledge: A Novel Paradigm to Low-Bit Quantization
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Róisín, Drimbarean, Alexandru, McDermott, James, O'Riordan, Colm |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Interpreting Global Perturbation Robustness of Image Models using Axiomatic Spectral Importance Decomposition
by: Luo, Róisín, et al.
Published: (2024)
by: Luo, Róisín, et al.
Published: (2024)
Sampling Matters in Explanations: Towards Trustworthy Attribution Analysis Building Block in Visual Models through Maximizing Explanation Certainty
by: Luo, Róisín, et al.
Published: (2025)
by: Luo, Róisín, et al.
Published: (2025)
Optimization-Induced Dynamics of Lipschitz Continuity in Neural Networks
by: Luo, Róisín, et al.
Published: (2025)
by: Luo, Róisín, et al.
Published: (2025)
Higher-Order Singular-Value Derivatives of Rectangular Real Matrices
by: Luo, Róisín, et al.
Published: (2025)
by: Luo, Róisín, et al.
Published: (2025)
MARR: Module-Adaptive Residual Reconstruction for Low-Bit Post-Training Quantization
by: Su, Le, et al.
Published: (2026)
by: Su, Le, et al.
Published: (2026)
Q$^2$: Quantization-Aware Gradient Balancing and Attention Alignment for Low-Bit Quantization
by: Wang, Zhaoyang, et al.
Published: (2025)
by: Wang, Zhaoyang, et al.
Published: (2025)
Breaking Modality Heterogeneity in Low-Bit Quantization for Large Vision-Language Models
by: Zhong, Yi, et al.
Published: (2026)
by: Zhong, Yi, et al.
Published: (2026)
Video Anomaly Detection via Spatio-Temporal Pseudo-Anomaly Generation : A Unified Approach
by: Rai, Ayush K., et al.
Published: (2023)
by: Rai, Ayush K., et al.
Published: (2023)
Progressive Fine-to-Coarse Reconstruction for Accurate Low-Bit Post-Training Quantization in Vision Transformers
by: Ding, Rui, et al.
Published: (2024)
by: Ding, Rui, et al.
Published: (2024)
Quantization Variation: A New Perspective on Training Transformers with Low-Bit Precision
by: Huang, Xijie, et al.
Published: (2023)
by: Huang, Xijie, et al.
Published: (2023)
Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models
by: Du, Jinyang, et al.
Published: (2026)
by: Du, Jinyang, et al.
Published: (2026)
When Bits Break Recourse: Counterfactual-Faithful Quantization
by: Yahyati, Chaymae, et al.
Published: (2026)
by: Yahyati, Chaymae, et al.
Published: (2026)
QuEPT: Quantized Elastic Precision Transformers with One-Shot Calibration for Multi-Bit Switching
by: Xu, Ke, et al.
Published: (2026)
by: Xu, Ke, et al.
Published: (2026)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
by: Choi, Kanghyun, et al.
Published: (2024)
by: Choi, Kanghyun, et al.
Published: (2024)
Embedding Compression for Efficient Re-Identification
by: McDermott, Luke
Published: (2024)
by: McDermott, Luke
Published: (2024)
LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models
by: Bhatnagar, Shubhang, et al.
Published: (2025)
by: Bhatnagar, Shubhang, et al.
Published: (2025)
ReSpinQuant: Efficient Layer-Wise LLM Quantization via Subspace Residual Rotation Approximation
by: Kim, Suyoung, et al.
Published: (2026)
by: Kim, Suyoung, et al.
Published: (2026)
Multinex: Lightweight Low-light Image Enhancement via Multi-prior Retinex
by: Brateanu, Alexandru, et al.
Published: (2026)
by: Brateanu, Alexandru, et al.
Published: (2026)
Reference-Guided Diffusion Inpainting For Multimodal Counterfactual Generation
by: Buburuzan, Alexandru
Published: (2025)
by: Buburuzan, Alexandru
Published: (2025)
MOGO: Residual Quantized Hierarchical Causal Transformer for High-Quality and Real-Time 3D Human Motion Generation
by: Fu, Dongjie, et al.
Published: (2025)
by: Fu, Dongjie, et al.
Published: (2025)
VQ-Style: Disentangling Style and Content in Motion with Residual Quantized Representations
by: Zargarbashi, Fatemeh, et al.
Published: (2026)
by: Zargarbashi, Fatemeh, et al.
Published: (2026)
Free-Mask: A Novel Paradigm of Integration Between the Segmentation Diffusion Model and Image Editing
by: Gao, Bo, et al.
Published: (2024)
by: Gao, Bo, et al.
Published: (2024)
BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook
by: Gu, Hao, et al.
Published: (2025)
by: Gu, Hao, et al.
Published: (2025)
Nearly Lossless Adaptive Bit Switching
by: Huang, Haiduo, et al.
Published: (2025)
by: Huang, Haiduo, et al.
Published: (2025)
Shedding the Bits: Pushing the Boundaries of Quantization with Minifloats on FPGAs
by: Aggarwal, Shivam, et al.
Published: (2023)
by: Aggarwal, Shivam, et al.
Published: (2023)
Lightweight Embedded FPGA Deployment of Learned Image Compression with Knowledge Distillation and Hybrid Quantization
by: Mazouz, Alaa, et al.
Published: (2025)
by: Mazouz, Alaa, et al.
Published: (2025)
LLM-FP4: 4-Bit Floating-Point Quantized Transformers
by: Liu, Shih-yang, et al.
Published: (2023)
by: Liu, Shih-yang, et al.
Published: (2023)
Scalable Image Tokenization with Index Backpropagation Quantization
by: Shi, Fengyuan, et al.
Published: (2024)
by: Shi, Fengyuan, et al.
Published: (2024)
Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision Language Models
by: Zou, Zhengtao, et al.
Published: (2025)
by: Zou, Zhengtao, et al.
Published: (2025)
Continuous Expert Assembly: Instance-Conditioned Low-Rank Residuals for All-in-One Image Restoration
by: He, Haisen, et al.
Published: (2026)
by: He, Haisen, et al.
Published: (2026)
FADRM: Fast and Accurate Data Residual Matching for Dataset Distillation
by: Cui, Jiacheng, et al.
Published: (2025)
by: Cui, Jiacheng, et al.
Published: (2025)
Self-Supervised Quantization-Aware Knowledge Distillation
by: Zhao, Kaiqi, et al.
Published: (2024)
by: Zhao, Kaiqi, et al.
Published: (2024)
TransFace++: Rethinking the Face Recognition Paradigm with a Focus on Accuracy, Efficiency, and Security
by: Dan, Jun, et al.
Published: (2023)
by: Dan, Jun, et al.
Published: (2023)
ResLink: A Novel Deep Learning Architecture for Brain Tumor Classification with Area Attention and Residual Connections
by: Arya, Sumedha, et al.
Published: (2025)
by: Arya, Sumedha, et al.
Published: (2025)
An Analysis of Multi-Task Architectures for the Hierarchic Multi-Label Problem of Vehicle Model and Make Classification
by: Manole, Alexandru, et al.
Published: (2026)
by: Manole, Alexandru, et al.
Published: (2026)
Denoising Diffusion Probabilistic Model for Point Cloud Compression at Low Bit-Rates
by: Spadaro, Gabriele, et al.
Published: (2025)
by: Spadaro, Gabriele, et al.
Published: (2025)
Residual SODAP: Residual Self-Organizing Domain-Adaptive Prompting with Structural Knowledge Preservation for Continual Learning
by: Oh, Gyutae, et al.
Published: (2026)
by: Oh, Gyutae, et al.
Published: (2026)
On the Robustness of Diffusion-Based Image Compression to Bit-Flip Errors
by: Vaisman, Amit, et al.
Published: (2026)
by: Vaisman, Amit, et al.
Published: (2026)
BitDance: Scaling Autoregressive Generative Models with Binary Tokens
by: Ai, Yuang, et al.
Published: (2026)
by: Ai, Yuang, et al.
Published: (2026)
BitMark: Watermarking Bitwise Autoregressive Image Generative Models
by: Kerner, Louis, et al.
Published: (2025)
by: Kerner, Louis, et al.
Published: (2025)
Similar Items
-
Interpreting Global Perturbation Robustness of Image Models using Axiomatic Spectral Importance Decomposition
by: Luo, Róisín, et al.
Published: (2024) -
Sampling Matters in Explanations: Towards Trustworthy Attribution Analysis Building Block in Visual Models through Maximizing Explanation Certainty
by: Luo, Róisín, et al.
Published: (2025) -
Optimization-Induced Dynamics of Lipschitz Continuity in Neural Networks
by: Luo, Róisín, et al.
Published: (2025) -
Higher-Order Singular-Value Derivatives of Rectangular Real Matrices
by: Luo, Róisín, et al.
Published: (2025) -
MARR: Module-Adaptive Residual Reconstruction for Low-Bit Post-Training Quantization
by: Su, Le, et al.
Published: (2026)