Adaptive Distribution-aware Quantization for Mixed-Precision Neural Networks
Fuente:
arXiv
Salvato in:
| Autori principali: | Jia, Shaohang, Huang, Zhiyong, Yu, Zhi, Hou, Mingyang, Miao, Shuai, Yang, Han |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Euclidean Prototypes: Spectral Disentanglement and Geodesic Matching for Few-Shot Medical Image Segmentation
di: Jia, Penghao, et al.
Pubblicazione: (2026)
di: Jia, Penghao, et al.
Pubblicazione: (2026)
Precision Neural Network Quantization via Learnable Adaptive Modules
di: Zhou, Wenqiang, et al.
Pubblicazione: (2025)
di: Zhou, Wenqiang, et al.
Pubblicazione: (2025)
Quantization Meets OOD: Generalizable Quantization-aware Training from a Flatness Perspective
di: Jiang, Jiacheng, et al.
Pubblicazione: (2025)
di: Jiang, Jiacheng, et al.
Pubblicazione: (2025)
Mix-QSAM: Mixed-Precision Quantization of the Segment Anything Model
di: Ranjan, Navin, et al.
Pubblicazione: (2025)
di: Ranjan, Navin, et al.
Pubblicazione: (2025)
MPTQ-ViT: Mixed-Precision Post-Training Quantization for Vision Transformer
di: Tai, Yu-Shan, et al.
Pubblicazione: (2024)
di: Tai, Yu-Shan, et al.
Pubblicazione: (2024)
Mix-QViT: Mixed-Precision Vision Transformer Quantization Driven by Layer Importance and Quantization Sensitivity
di: Ranjan, Navin, et al.
Pubblicazione: (2025)
di: Ranjan, Navin, et al.
Pubblicazione: (2025)
MPQ-DM: Mixed Precision Quantization for Extremely Low Bit Diffusion Models
di: Feng, Weilun, et al.
Pubblicazione: (2024)
di: Feng, Weilun, et al.
Pubblicazione: (2024)
Efficient and Effective Methods for Mixed Precision Neural Network Quantization for Faster, Energy-efficient Inference
di: Bablani, Deepika, et al.
Pubblicazione: (2023)
di: Bablani, Deepika, et al.
Pubblicazione: (2023)
MetaMix: Meta-state Precision Searcher for Mixed-precision Activation Quantization
di: Kim, Han-Byul, et al.
Pubblicazione: (2023)
di: Kim, Han-Byul, et al.
Pubblicazione: (2023)
Dynamics-aware Adversarial Attack of Adaptive Neural Networks
di: Tao, An, et al.
Pubblicazione: (2022)
di: Tao, An, et al.
Pubblicazione: (2022)
Q-SNNs: Quantized Spiking Neural Networks
di: Wei, Wenjie, et al.
Pubblicazione: (2024)
di: Wei, Wenjie, et al.
Pubblicazione: (2024)
MPQ-DMv2: Flexible Residual Mixed Precision Quantization for Low-Bit Diffusion Models with Temporal Distillation
di: Feng, Weilun, et al.
Pubblicazione: (2025)
di: Feng, Weilun, et al.
Pubblicazione: (2025)
TreeQ: Pushing the Quantization Boundary of Diffusion Transformer via Tree-Structured Mixed-Precision Search
di: Yang, Kaicheng, et al.
Pubblicazione: (2025)
di: Yang, Kaicheng, et al.
Pubblicazione: (2025)
QP-SNN: Quantized and Pruned Spiking Neural Networks
di: Wei, Wenjie, et al.
Pubblicazione: (2025)
di: Wei, Wenjie, et al.
Pubblicazione: (2025)
Memory-Free and Parallel Computation for Quantized Spiking Neural Networks
di: Zhang, Dehao, et al.
Pubblicazione: (2025)
di: Zhang, Dehao, et al.
Pubblicazione: (2025)
EDGE: Unknown-aware Multi-label Learning by Energy Distribution Gap Expansion
di: Sun, Yuchen, et al.
Pubblicazione: (2024)
di: Sun, Yuchen, et al.
Pubblicazione: (2024)
MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts
di: Tao, Wei, et al.
Pubblicazione: (2025)
di: Tao, Wei, et al.
Pubblicazione: (2025)
MPQ-Diff: Mixed Precision Quantization for Diffusion Models
di: Maruzzelli, Rocco Manz, et al.
Pubblicazione: (2024)
di: Maruzzelli, Rocco Manz, et al.
Pubblicazione: (2024)
Towards Accurate Binary Spiking Neural Networks: Learning with Adaptive Gradient Modulation Mechanism
di: Liang, Yu, et al.
Pubblicazione: (2025)
di: Liang, Yu, et al.
Pubblicazione: (2025)
MixA-Q: Revisiting Activation Sparsity for Vision Transformers from a Mixed-Precision Quantization Perspective
di: Wang, Weitian, et al.
Pubblicazione: (2025)
di: Wang, Weitian, et al.
Pubblicazione: (2025)
MixDQ: Memory-Efficient Few-Step Text-to-Image Diffusion Models with Metric-Decoupled Mixed Precision Quantization
di: Zhao, Tianchen, et al.
Pubblicazione: (2024)
di: Zhao, Tianchen, et al.
Pubblicazione: (2024)
Data-Free Quantization via Mixed-Precision Compensation without Fine-Tuning
di: Chen, Jun, et al.
Pubblicazione: (2023)
di: Chen, Jun, et al.
Pubblicazione: (2023)
LampQ: Towards Accurate Layer-wise Mixed Precision Quantization for Vision Transformers
di: Kim, Minjun, et al.
Pubblicazione: (2025)
di: Kim, Minjun, et al.
Pubblicazione: (2025)
6Bit-Diffusion: Inference-Time Mixed-Precision Quantization for Video Diffusion Models
di: Su, Rundong, et al.
Pubblicazione: (2026)
di: Su, Rundong, et al.
Pubblicazione: (2026)
Efficient and Robust Quantization-aware Training via Adaptive Coreset Selection
di: Huang, Xijie, et al.
Pubblicazione: (2023)
di: Huang, Xijie, et al.
Pubblicazione: (2023)
Learning from Loss Landscape: Generalizable Mixed-Precision Quantization via Adaptive Sharpness-Aware Gradient Aligning
di: Ma, Lianbo, et al.
Pubblicazione: (2025)
di: Ma, Lianbo, et al.
Pubblicazione: (2025)
Dual Precision Quantization for Efficient and Accurate Deep Neural Networks Inference
di: Gafni, Tomer, et al.
Pubblicazione: (2025)
di: Gafni, Tomer, et al.
Pubblicazione: (2025)
LRP-QViT: Mixed-Precision Vision Transformer Quantization via Layer-wise Relevance Propagation
di: Ranjan, Navin, et al.
Pubblicazione: (2024)
di: Ranjan, Navin, et al.
Pubblicazione: (2024)
Query Quantized Neural SLAM
di: Jiang, Sijia, et al.
Pubblicazione: (2024)
di: Jiang, Sijia, et al.
Pubblicazione: (2024)
Fisher-aware Quantization for DETR Detectors with Critical-category Objectives
di: Yang, Huanrui, et al.
Pubblicazione: (2024)
di: Yang, Huanrui, et al.
Pubblicazione: (2024)
DynaQuant: Dynamic Mixed-Precision Quantization for Learned Image Compression
di: Bao, Youneng, et al.
Pubblicazione: (2025)
di: Bao, Youneng, et al.
Pubblicazione: (2025)
Test-Time Model Adaptation for Quantized Neural Networks
di: Deng, Zeshuai, et al.
Pubblicazione: (2025)
di: Deng, Zeshuai, et al.
Pubblicazione: (2025)
High-Precision Fabric Defect Detection via Adaptive Shape Convolutions and Large Kernel Spatial Modeling
di: Wang, Shuai, et al.
Pubblicazione: (2025)
di: Wang, Shuai, et al.
Pubblicazione: (2025)
Value-Driven Mixed-Precision Quantization for Patch-Based Inference on Microcontrollers
di: Tao, Wei, et al.
Pubblicazione: (2024)
di: Tao, Wei, et al.
Pubblicazione: (2024)
QMix: Quality-aware Learning with Mixed Noise for Robust Retinal Disease Diagnosis
di: Hou, Junlin, et al.
Pubblicazione: (2024)
di: Hou, Junlin, et al.
Pubblicazione: (2024)
Real-Time Spacecraft Pose Estimation Using Mixed-Precision Quantized Neural Network on COTS Reconfigurable MPSoC
di: Posso, Julien, et al.
Pubblicazione: (2024)
di: Posso, Julien, et al.
Pubblicazione: (2024)
Frequency-aware Event Cloud Network
di: Ren, Hongwei, et al.
Pubblicazione: (2024)
di: Ren, Hongwei, et al.
Pubblicazione: (2024)
DC-PCN: Point Cloud Completion Network with Dual-Codebook Guided Quantization
di: Wu, Qiuxia, et al.
Pubblicazione: (2025)
di: Wu, Qiuxia, et al.
Pubblicazione: (2025)
AdaLog: Post-Training Quantization for Vision Transformers with Adaptive Logarithm Quantizer
di: Wu, Zhuguanyu, et al.
Pubblicazione: (2024)
di: Wu, Zhuguanyu, et al.
Pubblicazione: (2024)
FLIQS: One-Shot Mixed-Precision Floating-Point and Integer Quantization Search
di: Dotzel, Jordan, et al.
Pubblicazione: (2023)
di: Dotzel, Jordan, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Beyond Euclidean Prototypes: Spectral Disentanglement and Geodesic Matching for Few-Shot Medical Image Segmentation
di: Jia, Penghao, et al.
Pubblicazione: (2026) -
Precision Neural Network Quantization via Learnable Adaptive Modules
di: Zhou, Wenqiang, et al.
Pubblicazione: (2025) -
Quantization Meets OOD: Generalizable Quantization-aware Training from a Flatness Perspective
di: Jiang, Jiacheng, et al.
Pubblicazione: (2025) -
Mix-QSAM: Mixed-Precision Quantization of the Segment Anything Model
di: Ranjan, Navin, et al.
Pubblicazione: (2025) -
MPTQ-ViT: Mixed-Precision Post-Training Quantization for Vision Transformer
di: Tai, Yu-Shan, et al.
Pubblicazione: (2024)