DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers
Fuente:
arXiv
Salvato in:
| Autori principali: | Sharify, Sayeh, Salmani, Mahsa, Mostafa, Hesham |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MF-QAT: Multi-Format Quantization-Aware Training for Elastic Inference
di: Xu, Zifei, et al.
Pubblicazione: (2026)
di: Xu, Zifei, et al.
Pubblicazione: (2026)
ConvRot: Rotation-Based Plug-and-Play 4-bit Quantization for Diffusion Transformers
di: Huang, Feice, et al.
Pubblicazione: (2025)
di: Huang, Feice, et al.
Pubblicazione: (2025)
Q-Drift: Quantization-Aware Drift Correction for Diffusion Model Sampling
di: Ryu, Sooyoung, et al.
Pubblicazione: (2026)
di: Ryu, Sooyoung, et al.
Pubblicazione: (2026)
TerDiT: Ternary Diffusion Models with Transformers
di: Lu, Xudong, et al.
Pubblicazione: (2024)
di: Lu, Xudong, et al.
Pubblicazione: (2024)
Qua$^2$SeDiMo: Quantifiable Quantization Sensitivity of Diffusion Models
di: Mills, Keith G., et al.
Pubblicazione: (2024)
di: Mills, Keith G., et al.
Pubblicazione: (2024)
ParetoQ: Improving Scaling Laws in Extremely Low-bit LLM Quantization
di: Liu, Zechun, et al.
Pubblicazione: (2025)
di: Liu, Zechun, et al.
Pubblicazione: (2025)
Instance-Aware Group Quantization for Vision Transformers
di: Moon, Jaehyeon, et al.
Pubblicazione: (2024)
di: Moon, Jaehyeon, et al.
Pubblicazione: (2024)
Exploring Magnitude Preservation and Rotation Modulation in Diffusion Transformers
di: Bill, Eric Tillman, et al.
Pubblicazione: (2025)
di: Bill, Eric Tillman, et al.
Pubblicazione: (2025)
1-Bit FQT: Pushing the Limit of Fully Quantized Training to 1-bit
di: Gao, Chang, et al.
Pubblicazione: (2024)
di: Gao, Chang, et al.
Pubblicazione: (2024)
Restructuring Vector Quantization with the Rotation Trick
di: Fifty, Christopher, et al.
Pubblicazione: (2024)
di: Fifty, Christopher, et al.
Pubblicazione: (2024)
DiGAN: Diffusion-Guided Attention Network for Early Alzheimer's Disease Detection
di: Rahman, Maxx Richard, et al.
Pubblicazione: (2026)
di: Rahman, Maxx Richard, et al.
Pubblicazione: (2026)
TanDiT: Tangent-Plane Diffusion Transformer for High-Quality 360° Panorama Generation
di: Çapuk, Hakan, et al.
Pubblicazione: (2025)
di: Çapuk, Hakan, et al.
Pubblicazione: (2025)
DGQ: Distribution-Aware Group Quantization for Text-to-Image Diffusion Models
di: Ryu, Hyogon, et al.
Pubblicazione: (2025)
di: Ryu, Hyogon, et al.
Pubblicazione: (2025)
Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers
di: Chen, Lei, et al.
Pubblicazione: (2024)
di: Chen, Lei, et al.
Pubblicazione: (2024)
ResQ: Mixed-Precision Quantization of Large Language Models with Low-Rank Residuals
di: Saxena, Utkarsh, et al.
Pubblicazione: (2024)
di: Saxena, Utkarsh, et al.
Pubblicazione: (2024)
FlexiDiT: Your Diffusion Transformer Can Easily Generate High-Quality Samples with Less Compute
di: Anagnostidis, Sotiris, et al.
Pubblicazione: (2025)
di: Anagnostidis, Sotiris, et al.
Pubblicazione: (2025)
Quantized and Interpretable Learning Scheme for Deep Neural Networks in Classification Task
di: Maleki, Alireza, et al.
Pubblicazione: (2024)
di: Maleki, Alireza, et al.
Pubblicazione: (2024)
Q-SENN: Quantized Self-Explaining Neural Networks
di: Norrenbrock, Thomas, et al.
Pubblicazione: (2023)
di: Norrenbrock, Thomas, et al.
Pubblicazione: (2023)
ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation
di: Zhao, Tianchen, et al.
Pubblicazione: (2024)
di: Zhao, Tianchen, et al.
Pubblicazione: (2024)
Efficient Multi-bit Quantization Network Training via Weight Bias Correction and Bit-wise Coreset Sampling
di: Kim, Jinhee, et al.
Pubblicazione: (2025)
di: Kim, Jinhee, et al.
Pubblicazione: (2025)
PTQ4DiT: Post-training Quantization for Diffusion Transformers
di: Wu, Junyi, et al.
Pubblicazione: (2024)
di: Wu, Junyi, et al.
Pubblicazione: (2024)
LRQ-DiT: Log-Rotation Post-Training Quantization of Diffusion Transformers for Image and Video Generation
di: Yang, Lianwei, et al.
Pubblicazione: (2025)
di: Yang, Lianwei, et al.
Pubblicazione: (2025)
Fast-DiM: Towards Fast Diffusion Morphs
di: Blasingame, Zander W., et al.
Pubblicazione: (2023)
di: Blasingame, Zander W., et al.
Pubblicazione: (2023)
BrainRotViT: Transformer-ResNet Hybrid for Explainable Modeling of Brain Aging from 3D sMRI
di: Jalal, Wasif, et al.
Pubblicazione: (2025)
di: Jalal, Wasif, et al.
Pubblicazione: (2025)
BayesQ: Uncertainty-Guided Bayesian Quantization
di: Lamaakal, Ismail, et al.
Pubblicazione: (2025)
di: Lamaakal, Ismail, et al.
Pubblicazione: (2025)
SDQ-LLM: Sigma-Delta Quantization for 1-bit LLMs of any size
di: Xia, Junhao, et al.
Pubblicazione: (2025)
di: Xia, Junhao, et al.
Pubblicazione: (2025)
ReDiF: Reinforced Distillation for Few Step Diffusion
di: Tighkhorshid, Amirhossein, et al.
Pubblicazione: (2025)
di: Tighkhorshid, Amirhossein, et al.
Pubblicazione: (2025)
Q-DiT4SR: Exploration of Detail-Preserving Diffusion Transformer Quantization for Real-World Image Super-Resolution
di: Zhang, Xun, et al.
Pubblicazione: (2026)
di: Zhang, Xun, et al.
Pubblicazione: (2026)
FIMA-Q: Post-Training Quantization for Vision Transformers by Fisher Information Matrix Approximation
di: Wu, Zhuguanyu, et al.
Pubblicazione: (2025)
di: Wu, Zhuguanyu, et al.
Pubblicazione: (2025)
QGen: On the Ability to Generalize in Quantization Aware Training
di: AskariHemmat, MohammadHossein, et al.
Pubblicazione: (2024)
di: AskariHemmat, MohammadHossein, et al.
Pubblicazione: (2024)
Data-Augmented Quantization-Aware Knowledge Distillation
di: Kur, Justin, et al.
Pubblicazione: (2025)
di: Kur, Justin, et al.
Pubblicazione: (2025)
Progressive Compression with Universally Quantized Diffusion Models
di: Yang, Yibo, et al.
Pubblicazione: (2024)
di: Yang, Yibo, et al.
Pubblicazione: (2024)
Shiva-DiT: Residual-Based Differentiable Top-$k$ Selection for Efficient Diffusion Transformers
di: Zhang, Jiaji, et al.
Pubblicazione: (2026)
di: Zhang, Jiaji, et al.
Pubblicazione: (2026)
GalaxyDiT: Efficient Video Generation with Guidance Alignment and Adaptive Proxy in Diffusion Transformers
di: Song, Zhiye, et al.
Pubblicazione: (2025)
di: Song, Zhiye, et al.
Pubblicazione: (2025)
Sparse-vDiT: Unleashing the Power of Sparse Attention to Accelerate Video Diffusion Transformers
di: Chen, Pengtao, et al.
Pubblicazione: (2025)
di: Chen, Pengtao, et al.
Pubblicazione: (2025)
Sharpness-Aware Data Generation for Zero-shot Quantization
di: Hoang-Anh, Dung, et al.
Pubblicazione: (2025)
di: Hoang-Anh, Dung, et al.
Pubblicazione: (2025)
DiC: Rethinking Conv3x3 Designs in Diffusion Models
di: Tian, Yuchuan, et al.
Pubblicazione: (2024)
di: Tian, Yuchuan, et al.
Pubblicazione: (2024)
DiCTI: Diffusion-based Clothing Designer via Text-guided Input
di: Lampe, Ajda, et al.
Pubblicazione: (2024)
di: Lampe, Ajda, et al.
Pubblicazione: (2024)
Eff-GRot: Efficient and Generalizable Rotation Estimation with Transformers
di: Mathioulakis, Fanis, et al.
Pubblicazione: (2025)
di: Mathioulakis, Fanis, et al.
Pubblicazione: (2025)
DiWA: Diffusion Policy Adaptation with World Models
di: Chandra, Akshay L, et al.
Pubblicazione: (2025)
di: Chandra, Akshay L, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MF-QAT: Multi-Format Quantization-Aware Training for Elastic Inference
di: Xu, Zifei, et al.
Pubblicazione: (2026) -
ConvRot: Rotation-Based Plug-and-Play 4-bit Quantization for Diffusion Transformers
di: Huang, Feice, et al.
Pubblicazione: (2025) -
Q-Drift: Quantization-Aware Drift Correction for Diffusion Model Sampling
di: Ryu, Sooyoung, et al.
Pubblicazione: (2026) -
TerDiT: Ternary Diffusion Models with Transformers
di: Lu, Xudong, et al.
Pubblicazione: (2024) -
Qua$^2$SeDiMo: Quantifiable Quantization Sensitivity of Diffusion Models
di: Mills, Keith G., et al.
Pubblicazione: (2024)