CoDeQ: End-to-End Joint Model Compression with Dead-Zone Quantizer for High-Sparsity and Low-Precision Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Wenshøj, Jonathan, Chen, Tong, Pepin, Bob, Selvan, Raghavendra |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Oscillations Make Neural Networks Robust to Quantization
by: Wenshøj, Jonathan, et al.
Published: (2025)
by: Wenshøj, Jonathan, et al.
Published: (2025)
Algorithmic Simplification of Neural Networks with Mosaic-of-Motifs
by: Bakhtiarifard, Pedram, et al.
Published: (2026)
by: Bakhtiarifard, Pedram, et al.
Published: (2026)
When Can Memorization Improve Fairness?
by: Pepin, Bob, et al.
Published: (2024)
by: Pepin, Bob, et al.
Published: (2024)
Is Adversarial Training with Compressed Datasets Effective?
by: Chen, Tong, et al.
Published: (2024)
by: Chen, Tong, et al.
Published: (2024)
Activation Compression of Graph Neural Networks using Block-wise Quantization with Improved Variance Minimization
by: Eliassen, Sebastian, et al.
Published: (2023)
by: Eliassen, Sebastian, et al.
Published: (2023)
Characterizing Learning in Deep Neural Networks using Tractable Algorithmic Complexity Analysis
by: Bakhtiarifard, Pedram, et al.
Published: (2026)
by: Bakhtiarifard, Pedram, et al.
Published: (2026)
FairQuant: Fairness-Aware Mixed-Precision Quantization for Medical Image Classification
by: Woergaard, Thomas, et al.
Published: (2026)
by: Woergaard, Thomas, et al.
Published: (2026)
PePR: Performance Per Resource Unit as a Metric to Promote Small-Scale Deep Learning in Medical Image Analysis
by: Selvan, Raghavendra, et al.
Published: (2024)
by: Selvan, Raghavendra, et al.
Published: (2024)
A Discrepancy-Based Perspective on Dataset Condensation
by: Chen, Tong, et al.
Published: (2025)
by: Chen, Tong, et al.
Published: (2025)
An Empirical Study of the Influence of Adversarial Fine-Tuning on Compressed Neural Networks
by: Thorsteinsson, Hallgrimur, et al.
Published: (2024)
by: Thorsteinsson, Hallgrimur, et al.
Published: (2024)
UniPhyNet: A Unified Network For Multimodal Physiological Raw Signal Classification
by: Qiu, Renxiang, et al.
Published: (2025)
by: Qiu, Renxiang, et al.
Published: (2025)
MiCo: End-to-End Mixed Precision Neural Network Co-Exploration Framework for Edge AI
by: Jiang, Zijun, et al.
Published: (2025)
by: Jiang, Zijun, et al.
Published: (2025)
Extreme Model Compression with Structured Sparsity at Low Precision
by: Liu, Dan, et al.
Published: (2025)
by: Liu, Dan, et al.
Published: (2025)
End-to-End Compression for Tabular Foundation Models
by: Zabërgja, Guri, et al.
Published: (2026)
by: Zabërgja, Guri, et al.
Published: (2026)
BMRS: Bayesian Model Reduction for Structured Pruning
by: Wright, Dustin, et al.
Published: (2024)
by: Wright, Dustin, et al.
Published: (2024)
End-to-End Heterogeneous Graph Neural Networks for Traffic Assignment
by: Liu, Tong, et al.
Published: (2023)
by: Liu, Tong, et al.
Published: (2023)
EC-NAS: Energy Consumption Aware Tabular Benchmarks for Neural Architecture Search
by: Bakhtiarifard, Pedram, et al.
Published: (2022)
by: Bakhtiarifard, Pedram, et al.
Published: (2022)
Compression Scaling Laws:Unifying Sparsity and Quantization
by: Frantar, Elias, et al.
Published: (2025)
by: Frantar, Elias, et al.
Published: (2025)
SutureBot: A Precision Framework & Benchmark For Autonomous End-to-End Suturing
by: Haworth, Jesse, et al.
Published: (2025)
by: Haworth, Jesse, et al.
Published: (2025)
End-to-End On-Device Quantization-Aware Training for LLMs at Inference Cost
by: Tan, Qitao, et al.
Published: (2025)
by: Tan, Qitao, et al.
Published: (2025)
PQuantML: A Tool for End-to-End Hardware-aware Model Compression
by: Niemi, Roope, et al.
Published: (2026)
by: Niemi, Roope, et al.
Published: (2026)
ResQ: Mixed-Precision Quantization of Large Language Models with Low-Rank Residuals
by: Saxena, Utkarsh, et al.
Published: (2024)
by: Saxena, Utkarsh, et al.
Published: (2024)
Janus-Q: End-to-End Event-Driven Trading via Hierarchical-Gated Reward Modeling
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
FineQ: Software-Hardware Co-Design for Low-Bit Fine-Grained Mixed-Precision Quantization of LLMs
by: Xie, Xilong, et al.
Published: (2025)
by: Xie, Xilong, et al.
Published: (2025)
E2E-GRec: An End-to-End Joint Training Framework for Graph Neural Networks and Recommender Systems
by: Xue, Rui, et al.
Published: (2025)
by: Xue, Rui, et al.
Published: (2025)
SLiM: One-shot Quantization and Sparsity with Low-rank Approximation for LLM Weight Compression
by: Mozaffari, Mohammad, et al.
Published: (2024)
by: Mozaffari, Mohammad, et al.
Published: (2024)
Operating critical machine learning models in resource constrained regimes
by: Selvan, Raghavendra, et al.
Published: (2023)
by: Selvan, Raghavendra, et al.
Published: (2023)
Ultra-Efficient Decoding for End-to-End Neural Compression and Reconstruction
by: Rogers, Ethan G., et al.
Published: (2025)
by: Rogers, Ethan G., et al.
Published: (2025)
Huff-LLM: End-to-End Lossless Compression for Efficient LLM Inference
by: Yubeaton, Patrick, et al.
Published: (2025)
by: Yubeaton, Patrick, et al.
Published: (2025)
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
by: Maes, Lucas, et al.
Published: (2026)
by: Maes, Lucas, et al.
Published: (2026)
Loss Function Considering Dead Zone for Neural Networks
by: Inami, Koki, et al.
Published: (2024)
by: Inami, Koki, et al.
Published: (2024)
PoTAcc: A Pipeline for End-to-End Acceleration of Power-of-Two Quantized DNNs
by: Saha, Rappy, et al.
Published: (2026)
by: Saha, Rappy, et al.
Published: (2026)
Efficiency is Not Enough: A Critical Perspective of Environmentally Sustainable AI
by: Wright, Dustin, et al.
Published: (2023)
by: Wright, Dustin, et al.
Published: (2023)
Dead End at the Shelves
by: Perrault, Anna H.
Published: (1972)
by: Perrault, Anna H.
Published: (1972)
Beyond Throughput and Compression Ratios: Towards High End-to-end Utility of Gradient Compression
by: Han, Wenchen, et al.
Published: (2024)
by: Han, Wenchen, et al.
Published: (2024)
Towards End-to-End Network Intent Management with Large Language Models
by: Dinh, Lam, et al.
Published: (2025)
by: Dinh, Lam, et al.
Published: (2025)
An End-to-End Differentiable, Graph Neural Network-Embedded Pore Network Model for Permeability Prediction
by: Zhao, Qingqi, et al.
Published: (2025)
by: Zhao, Qingqi, et al.
Published: (2025)
ImaginationPolicy: Towards Generalizable, Precise and Reliable End-to-End Policy for Robotic Manipulation
by: Lu, Dekun, et al.
Published: (2025)
by: Lu, Dekun, et al.
Published: (2025)
The Harmonic Indel Distance
by: Pepin, Bob
Published: (2020)
by: Pepin, Bob
Published: (2020)
OMG-HD: A High-Resolution AI Weather Model for End-to-End Forecasts from Observations
by: Zhao, Pengcheng, et al.
Published: (2024)
by: Zhao, Pengcheng, et al.
Published: (2024)
Similar Items
-
Oscillations Make Neural Networks Robust to Quantization
by: Wenshøj, Jonathan, et al.
Published: (2025) -
Algorithmic Simplification of Neural Networks with Mosaic-of-Motifs
by: Bakhtiarifard, Pedram, et al.
Published: (2026) -
When Can Memorization Improve Fairness?
by: Pepin, Bob, et al.
Published: (2024) -
Is Adversarial Training with Compressed Datasets Effective?
by: Chen, Tong, et al.
Published: (2024) -
Activation Compression of Graph Neural Networks using Block-wise Quantization with Improved Variance Minimization
by: Eliassen, Sebastian, et al.
Published: (2023)