FraQAT: Quantization Aware Training with Fractional bits
Fuente:
arXiv
Saved in:
| Main Authors: | Morreale, Luca, Ramos, Alberto Gil C. P., Chadwick, Malcolm, Noroozi, Mehid, Chavhan, Ruchika, Mehrotra, Abhinav, Bhattacharya, Sourav |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Guidance Free Image Editing via Explicit Conditioning
by: Noroozi, Mehdi, et al.
Published: (2025)
by: Noroozi, Mehdi, et al.
Published: (2025)
Upcycling Text-to-Image Diffusion Models for Multi-Task Capabilities
by: Chavhan, Ruchika, et al.
Published: (2025)
by: Chavhan, Ruchika, et al.
Published: (2025)
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention
by: Becker, Philipp, et al.
Published: (2025)
by: Becker, Philipp, et al.
Published: (2025)
RFDM: Residual Flow Diffusion Model for Efficient Causal Video Editing
by: Salehi, Mohammadreza, et al.
Published: (2026)
by: Salehi, Mohammadreza, et al.
Published: (2026)
NanoFLUX: Distillation-Driven Compression of Large Text-to-Image Generation Models for Mobile Devices
by: Chavhan, Ruchika, et al.
Published: (2026)
by: Chavhan, Ruchika, et al.
Published: (2026)
Efficient High-Resolution Image Editing with Hallucination-Aware Loss and Adaptive Tiling
by: Kwon, Young D., et al.
Published: (2025)
by: Kwon, Young D., et al.
Published: (2025)
CR-QAT: Curriculum Relational Quantization-Aware Training for Open-Vocabulary Object Detection
by: Park, Jinyeong, et al.
Published: (2026)
by: Park, Jinyeong, et al.
Published: (2026)
ConceptPrune: Concept Editing in Diffusion Models via Skilled Neuron Pruning
by: Chavhan, Ruchika, et al.
Published: (2024)
by: Chavhan, Ruchika, et al.
Published: (2024)
DL-QAT: Weight-Decomposed Low-Rank Quantization-Aware Training for Large Language Models
by: Ke, Wenjin, et al.
Published: (2025)
by: Ke, Wenjin, et al.
Published: (2025)
Fast Sampling Through The Reuse Of Attention Maps In Diffusion Models
by: Hunter, Rosco, et al.
Published: (2023)
by: Hunter, Rosco, et al.
Published: (2023)
AdaLoRA-QAT: Adaptive Low-Rank and Quantization-Aware Segmentation
by: Deb, Prantik, et al.
Published: (2026)
by: Deb, Prantik, et al.
Published: (2026)
EfQAT: An Efficient Framework for Quantization-Aware Training
by: Ashkboos, Saleh, et al.
Published: (2024)
by: Ashkboos, Saleh, et al.
Published: (2024)
Memorized Images in Diffusion Models share a Subspace that can be Located and Deleted
by: Chavhan, Ruchika, et al.
Published: (2024)
by: Chavhan, Ruchika, et al.
Published: (2024)
AdaQAT: Adaptive Bit-Width Quantization-Aware Training
by: Gernigon, Cédric, et al.
Published: (2024)
by: Gernigon, Cédric, et al.
Published: (2024)
Attn-QAT: 4-Bit Attention With Quantization-Aware Training
by: Zhang, Peiyuan, et al.
Published: (2026)
by: Zhang, Peiyuan, et al.
Published: (2026)
EfficientQAT: Efficient Quantization-Aware Training for Large Language Models
by: Chen, Mengzhao, et al.
Published: (2024)
by: Chen, Mengzhao, et al.
Published: (2024)
StableQAT: Stable Quantization-Aware Training at Ultra-Low Bitwidths
by: Chen, Tianyi, et al.
Published: (2026)
by: Chen, Tianyi, et al.
Published: (2026)
MF-QAT: Multi-Format Quantization-Aware Training for Elastic Inference
by: Xu, Zifei, et al.
Published: (2026)
by: Xu, Zifei, et al.
Published: (2026)
Topology-Constrained Quantized nnUNet for Efficient and Anatomically Accurate 3D Tooth Segmentation
by: Prasad, Paarth, et al.
Published: (2026)
by: Prasad, Paarth, et al.
Published: (2026)
DPVO-QAT++: Heterogeneous QAT and CUDA Kernel Fusion for High-Performance Deep Patch Visual Odometry
by: Liao, Cheng
Published: (2025)
by: Liao, Cheng
Published: (2025)
GranQ: Efficient Channel-wise Quantization via Vectorized Pre-Scaling for Zero-Shot QAT
by: Hong, Inpyo, et al.
Published: (2025)
by: Hong, Inpyo, et al.
Published: (2025)
Power-of-Two Quantization-Aware-Training (PoT-QAT) in Large Language Models (LLMs)
by: Elgenedy, Mahmoud
Published: (2026)
by: Elgenedy, Mahmoud
Published: (2026)
1-Bit FQT: Pushing the Limit of Fully Quantized Training to 1-bit
by: Gao, Chang, et al.
Published: (2024)
by: Gao, Chang, et al.
Published: (2024)
DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers
by: Sharify, Sayeh, et al.
Published: (2026)
by: Sharify, Sayeh, et al.
Published: (2026)
2DQuant: Low-bit Post-Training Quantization for Image Super-Resolution
by: Liu, Kai, et al.
Published: (2024)
by: Liu, Kai, et al.
Published: (2024)
Scheduling Weight Transitions for Quantization-Aware Training
by: Lee, Junghyup, et al.
Published: (2024)
by: Lee, Junghyup, et al.
Published: (2024)
Unifying Block-wise PTQ and Distillation-based QAT for Progressive Quantization toward 2-bit Instruction-Tuned LLMs
by: Lee, Jung Hyun, et al.
Published: (2025)
by: Lee, Jung Hyun, et al.
Published: (2025)
GPLQ: A General, Practical, and Lightning QAT Method for Vision Transformers
by: Liang, Guang, et al.
Published: (2025)
by: Liang, Guang, et al.
Published: (2025)
Efficient Multi-bit Quantization Network Training via Weight Bias Correction and Bit-wise Coreset Sampling
by: Kim, Jinhee, et al.
Published: (2025)
by: Kim, Jinhee, et al.
Published: (2025)
$γ$-Quant: Towards Learnable Quantization for Low-bit Pattern Recognition
by: Fatima, Mishal, et al.
Published: (2025)
by: Fatima, Mishal, et al.
Published: (2025)
BitsFusion: 1.99 bits Weight Quantization of Diffusion Model
by: Sui, Yang, et al.
Published: (2024)
by: Sui, Yang, et al.
Published: (2024)
QGen: On the Ability to Generalize in Quantization Aware Training
by: AskariHemmat, MohammadHossein, et al.
Published: (2024)
by: AskariHemmat, MohammadHossein, et al.
Published: (2024)
Neural Semantic Surface Maps
by: Morreale, Luca, et al.
Published: (2023)
by: Morreale, Luca, et al.
Published: (2023)
SLA2: Sparse-Linear Attention with Learnable Routing and QAT
by: Zhang, Jintao, et al.
Published: (2026)
by: Zhang, Jintao, et al.
Published: (2026)
A Simple Low-bit Quantization Framework for Video Snapshot Compressive Imaging
by: Cao, Miao, et al.
Published: (2024)
by: Cao, Miao, et al.
Published: (2024)
Sensitivity-Aware Post-Training Quantization for Deep Neural Networks
by: Zheng, Zekang, et al.
Published: (2025)
by: Zheng, Zekang, et al.
Published: (2025)
Stabilizing Quantization-Aware Training by Implicit-Regularization on Hessian Matrix
by: Pang, Junbiao, et al.
Published: (2025)
by: Pang, Junbiao, et al.
Published: (2025)
In-Distribution Consistency Regularization Improves the Generalization of Quantization-Aware Training
by: Pang, Junbiao, et al.
Published: (2024)
by: Pang, Junbiao, et al.
Published: (2024)
Outlier-Aware Post-Training Quantization for Image Super-Resolution
by: Wang, Hailing, et al.
Published: (2025)
by: Wang, Hailing, et al.
Published: (2025)
StatQAT: Statistical Quantizer Optimization for Deep Networks
by: Aktukmak, Mehmet, et al.
Published: (2026)
by: Aktukmak, Mehmet, et al.
Published: (2026)
Similar Items
-
Guidance Free Image Editing via Explicit Conditioning
by: Noroozi, Mehdi, et al.
Published: (2025) -
Upcycling Text-to-Image Diffusion Models for Multi-Task Capabilities
by: Chavhan, Ruchika, et al.
Published: (2025) -
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention
by: Becker, Philipp, et al.
Published: (2025) -
RFDM: Residual Flow Diffusion Model for Efficient Causal Video Editing
by: Salehi, Mohammadreza, et al.
Published: (2026) -
NanoFLUX: Distillation-Driven Compression of Large Text-to-Image Generation Models for Mobile Devices
by: Chavhan, Ruchika, et al.
Published: (2026)