Trainable Fixed-Point Quantization for Deep Learning Acceleration on FPGAs
Fuente:
arXiv
Saved in:
| Main Authors: | Dai, Dingyi, Zhang, Yichi, Zhang, Jiahao, Hu, Zhanqiu, Cai, Yaohui, Sun, Qi, Zhang, Zhiru |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Accelerating Deep Learning with Fixed Time Budget
by: Khan, Muhammad Asif, et al.
Published: (2024)
by: Khan, Muhammad Asif, et al.
Published: (2024)
FLIQS: One-Shot Mixed-Precision Floating-Point and Integer Quantization Search
by: Dotzel, Jordan, et al.
Published: (2023)
by: Dotzel, Jordan, et al.
Published: (2023)
PointFix: Learning to Fix Domain Bias for Robust Online Stereo Adaptation
by: Kim, Kwonyoung, et al.
Published: (2022)
by: Kim, Kwonyoung, et al.
Published: (2022)
Accelerating Non-Maximum Suppression: A Graph Theory Perspective
by: Si, King-Siong, et al.
Published: (2024)
by: Si, King-Siong, et al.
Published: (2024)
A Trainable Feature Extractor Module for Deep Neural Networks and Scanpath Classification
by: Fuhl, Wolfgang
Published: (2024)
by: Fuhl, Wolfgang
Published: (2024)
DQA: An Efficient Method for Deep Quantization of Deep Neural Network Activations
by: Hu, Wenhao, et al.
Published: (2024)
by: Hu, Wenhao, et al.
Published: (2024)
SpargeAttention2: Trainable Sparse Attention via Hybrid Top-k+Top-p Masking and Distillation Fine-Tuning
by: Zhang, Jintao, et al.
Published: (2026)
by: Zhang, Jintao, et al.
Published: (2026)
Shedding the Bits: Pushing the Boundaries of Quantization with Minifloats on FPGAs
by: Aggarwal, Shivam, et al.
Published: (2023)
by: Aggarwal, Shivam, et al.
Published: (2023)
Learning from Students: Applying t-Distributions to Explore Accurate and Efficient Formats for LLMs
by: Dotzel, Jordan, et al.
Published: (2024)
by: Dotzel, Jordan, et al.
Published: (2024)
Towards Hierarchical Rectified Flow
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
Hierarchical Rectified Flow Matching with Mini-Batch Couplings
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
Physics Inspired Criterion for Pruning-Quantization Joint Learning
by: Xie, Weiying, et al.
Published: (2023)
by: Xie, Weiying, et al.
Published: (2023)
Modulated Diffusion: Accelerating Generative Modeling with Modulated Quantization
by: Gao, Weizhi, et al.
Published: (2025)
by: Gao, Weizhi, et al.
Published: (2025)
Quantized and Interpretable Learning Scheme for Deep Neural Networks in Classification Task
by: Maleki, Alireza, et al.
Published: (2024)
by: Maleki, Alireza, et al.
Published: (2024)
A Data-Free Analytical Quantization Scheme for Deep Learning Models
by: Luqman, Ahmed, et al.
Published: (2024)
by: Luqman, Ahmed, et al.
Published: (2024)
BRIDLE: Generalized Self-supervised Learning with Quantization
by: Nguyen, Hoang M., et al.
Published: (2025)
by: Nguyen, Hoang M., et al.
Published: (2025)
Multi-Task Model Merging via Adaptive Weight Disentanglement
by: Xiong, Feng, et al.
Published: (2024)
by: Xiong, Feng, et al.
Published: (2024)
Scaling Laws for Black box Adversarial Attacks
by: Liu, Chuan, et al.
Published: (2024)
by: Liu, Chuan, et al.
Published: (2024)
COMQ: A Backpropagation-Free Algorithm for Post-Training Quantization
by: Zhang, Aozhong, et al.
Published: (2024)
by: Zhang, Aozhong, et al.
Published: (2024)
TrAct: Making First-layer Pre-Activations Trainable
by: Petersen, Felix, et al.
Published: (2024)
by: Petersen, Felix, et al.
Published: (2024)
Fixed Point Diffusion Models
by: Bai, Xingjian, et al.
Published: (2024)
by: Bai, Xingjian, et al.
Published: (2024)
MoQE: Improve Quantization Model performance via Mixture of Quantization Experts
by: Zhang, Jinhao, et al.
Published: (2025)
by: Zhang, Jinhao, et al.
Published: (2025)
Beyond Fixed Formulas: Data-Driven Linear Predictor for Efficient Diffusion Models
by: Shen, Zhirong, et al.
Published: (2026)
by: Shen, Zhirong, et al.
Published: (2026)
Deep Regression Representation Learning with Topology
by: Zhang, Shihao, et al.
Published: (2024)
by: Zhang, Shihao, et al.
Published: (2024)
Scaling Up Quantization-Aware Neural Architecture Search for Efficient Deep Learning on the Edge
by: Lu, Yao, et al.
Published: (2024)
by: Lu, Yao, et al.
Published: (2024)
Rethinking Data Input for Point Cloud Upsampling
by: Zhang, Tongxu
Published: (2024)
by: Zhang, Tongxu
Published: (2024)
Revisiting Transformation Invariant Geometric Deep Learning: An Initial Representation Perspective
by: Zhang, Ziwei, et al.
Published: (2021)
by: Zhang, Ziwei, et al.
Published: (2021)
Systematic Characterization of Minimal Deep Learning Architectures: A Unified Analysis of Convergence, Pruning, and Quantization
by: Zheng, Ziwei, et al.
Published: (2026)
by: Zheng, Ziwei, et al.
Published: (2026)
Token Pruning for Caching Better: 9 Times Acceleration on Stable Diffusion for Free
by: Zhang, Evelyn, et al.
Published: (2024)
by: Zhang, Evelyn, et al.
Published: (2024)
Quantize-then-Rectify: Efficient VQ-VAE Training
by: Zhang, Borui, et al.
Published: (2025)
by: Zhang, Borui, et al.
Published: (2025)
TimePoint: Accelerated Time Series Alignment via Self-Supervised Keypoint and Descriptor Learning
by: Weber, Ron Shapira, et al.
Published: (2025)
by: Weber, Ron Shapira, et al.
Published: (2025)
CovMatch: Cross-Covariance Guided Multimodal Dataset Distillation with Trainable Text Encoder
by: Lee, Yongmin, et al.
Published: (2025)
by: Lee, Yongmin, et al.
Published: (2025)
Preventing Local Pitfalls in Vector Quantization via Optimal Transport
by: Zhang, Borui, et al.
Published: (2024)
by: Zhang, Borui, et al.
Published: (2024)
EfficientViT-SAM: Accelerated Segment Anything Model Without Accuracy Loss
by: Zhang, Zhuoyang, et al.
Published: (2024)
by: Zhang, Zhuoyang, et al.
Published: (2024)
Learning Topological Representations for Deep Image Understanding
by: Hu, Xiaoling
Published: (2024)
by: Hu, Xiaoling
Published: (2024)
Enhancing Post-Training Quantization via Future Activation Awareness
by: Lv, Zheqi, et al.
Published: (2026)
by: Lv, Zheqi, et al.
Published: (2026)
Sparse Forcing: Native Trainable Sparse Attention for Real-time Autoregressive Diffusion Video Generation
by: Xu, Boxun, et al.
Published: (2026)
by: Xu, Boxun, et al.
Published: (2026)
Deep Learning-based Point Cloud Registration for Augmented Reality-guided Surgery
by: Weber, Maximilian, et al.
Published: (2024)
by: Weber, Maximilian, et al.
Published: (2024)
Differential-informed Sample Selection Accelerates Multimodal Contrastive Learning
by: Zhao, Zihua, et al.
Published: (2025)
by: Zhao, Zihua, et al.
Published: (2025)
Vector Quantization Prompting for Continual Learning
by: Jiao, Li, et al.
Published: (2024)
by: Jiao, Li, et al.
Published: (2024)
Similar Items
-
Accelerating Deep Learning with Fixed Time Budget
by: Khan, Muhammad Asif, et al.
Published: (2024) -
FLIQS: One-Shot Mixed-Precision Floating-Point and Integer Quantization Search
by: Dotzel, Jordan, et al.
Published: (2023) -
PointFix: Learning to Fix Domain Bias for Robust Online Stereo Adaptation
by: Kim, Kwonyoung, et al.
Published: (2022) -
Accelerating Non-Maximum Suppression: A Graph Theory Perspective
by: Si, King-Siong, et al.
Published: (2024) -
A Trainable Feature Extractor Module for Deep Neural Networks and Scanpath Classification
by: Fuhl, Wolfgang
Published: (2024)