VQ4DiT: Efficient Post-Training Vector Quantization for Diffusion Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Deng, Juncan, Li, Shuaiting, Wang, Zeyu, Gu, Hong, Xu, Kedong, Huang, Kejie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ViM-VQ: Efficient Post-Training Vector Quantization for Visual Mamba
by: Deng, Juncan, et al.
Published: (2025)
by: Deng, Juncan, et al.
Published: (2025)
Efficiency Meets Fidelity: A Novel Quantization Framework for Stable Diffusion
by: Li, Shuaiting, et al.
Published: (2024)
by: Li, Shuaiting, et al.
Published: (2024)
SSVQ: Unleashing the Potential of Vector Quantization with Sign-Splitting
by: Li, Shuaiting, et al.
Published: (2025)
by: Li, Shuaiting, et al.
Published: (2025)
VQ4ALL: Efficient Neural Network Representation via a Universal Codebook
by: Deng, Juncan, et al.
Published: (2024)
by: Deng, Juncan, et al.
Published: (2024)
Transformer-Based Vector Font Classification Using Different Font Formats: TrueType versus PostScript
by: Fujioka, Takumu, et al.
Published: (2025)
by: Fujioka, Takumu, et al.
Published: (2025)
S-HR-VQVAE: Sequential Hierarchical Residual Learning Vector Quantized Variational Autoencoder for Video Prediction
by: Adiban, Mohammad, et al.
Published: (2023)
by: Adiban, Mohammad, et al.
Published: (2023)
Efficient Temporally-Aware DeepFake Detection using H.264 Motion Vectors
by: Grönquist, Peter, et al.
Published: (2023)
by: Grönquist, Peter, et al.
Published: (2023)
Exploring Diffusion with Test-Time Training on Efficient Image Restoration
by: Lu, Rongchang, et al.
Published: (2025)
by: Lu, Rongchang, et al.
Published: (2025)
Efficient Diffusion Training through Parallelization with Truncated Karhunen-Loève Expansion
by: Ren, Yumeng, et al.
Published: (2025)
by: Ren, Yumeng, et al.
Published: (2025)
Boundary-Protection W8A8 HiFloat8 Quantization for Large-Scale Text-to-Video Diffusion Transformers
by: Zhao, Yiming
Published: (2026)
by: Zhao, Yiming
Published: (2026)
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
by: Yasuno, Takato
Published: (2026)
by: Yasuno, Takato
Published: (2026)
Quantized FCA: Efficient Zero-Shot Texture Anomaly Detection
by: Ardelean, Andrei-Timotei, et al.
Published: (2025)
by: Ardelean, Andrei-Timotei, et al.
Published: (2025)
Global-Local Similarity for Efficient Fine-Grained Image Recognition with Vision Transformers
by: Rios, Edwin Arkel, et al.
Published: (2024)
by: Rios, Edwin Arkel, et al.
Published: (2024)
A Storage-Efficient Feature for 3D Concrete Defect Segmentation to Replace Normal Vector
by: Hua, Linxin, et al.
Published: (2025)
by: Hua, Linxin, et al.
Published: (2025)
NumeriKontrol: Adding Numeric Control to Diffusion Transformers for Instruction-based Image Editing
by: Xu, Zhenyu, et al.
Published: (2025)
by: Xu, Zhenyu, et al.
Published: (2025)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
by: Shahin, Nada, et al.
Published: (2025)
by: Shahin, Nada, et al.
Published: (2025)
CFFormer: Cross CNN-Transformer Channel Attention and Spatial Feature Fusion for Improved Segmentation of Heterogeneous Medical Images
by: Li, Jiaxuan, et al.
Published: (2025)
by: Li, Jiaxuan, et al.
Published: (2025)
AttEntropy: On the Generalization Ability of Supervised Semantic Segmentation Transformers to New Objects in New Domains
by: Lis, Krzysztof, et al.
Published: (2022)
by: Lis, Krzysztof, et al.
Published: (2022)
HySparK: Hybrid Sparse Masking for Large Scale Medical Image Pre-Training
by: Tang, Fenghe, et al.
Published: (2024)
by: Tang, Fenghe, et al.
Published: (2024)
MVQ:Towards Efficient DNN Compression and Acceleration with Masked Vector Quantization
by: Li, Shuaiting, et al.
Published: (2024)
by: Li, Shuaiting, et al.
Published: (2024)
Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
by: Chen, Pei-Chi, et al.
Published: (2025)
by: Chen, Pei-Chi, et al.
Published: (2025)
DCT-HistoTransformer: Efficient Lightweight Vision Transformer with DCT Integration for histopathological image analysis
by: Ranjbar, Mahtab, et al.
Published: (2024)
by: Ranjbar, Mahtab, et al.
Published: (2024)
Outline-Guided Object Inpainting with Diffusion Models
by: Pobitzer, Markus, et al.
Published: (2024)
by: Pobitzer, Markus, et al.
Published: (2024)
DeltaVLM: Interactive Remote Sensing Image Change Analysis via Instruction-guided Difference Perception
by: Deng, Pei, et al.
Published: (2025)
by: Deng, Pei, et al.
Published: (2025)
Efficient Diffusion Model for Image Restoration by Residual Shifting
by: Yue, Zongsheng, et al.
Published: (2024)
by: Yue, Zongsheng, et al.
Published: (2024)
Rethinking Video Deblurring with Wavelet-Aware Dynamic Transformer and Diffusion Model
by: Rao, Chen, et al.
Published: (2024)
by: Rao, Chen, et al.
Published: (2024)
MoDE: Mixture of Diffusion Experts for Any Occluded Face Recognition
by: Fan, Qiannan, et al.
Published: (2025)
by: Fan, Qiannan, et al.
Published: (2025)
LeDiFlow: Learned Distribution-guided Flow Matching to Accelerate Image Generation
by: Zwick, Pascal, et al.
Published: (2025)
by: Zwick, Pascal, et al.
Published: (2025)
POC-SLT: Partial Object Completion with SDF Latent Transformers
by: Zakeri, Faezeh, et al.
Published: (2024)
by: Zakeri, Faezeh, et al.
Published: (2024)
A Resource-Efficient Hybrid CNN-LSTM network for image-based bean leaf disease classification
by: Rhee, Hye Jin, et al.
Published: (2026)
by: Rhee, Hye Jin, et al.
Published: (2026)
Low-Cost Tree Crown Dieback Estimation Using Deep Learning-Based Segmentation
by: Allen, M. J., et al.
Published: (2024)
by: Allen, M. J., et al.
Published: (2024)
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
by: Lu, Wanglong, et al.
Published: (2024)
by: Lu, Wanglong, et al.
Published: (2024)
Automated MRI Tumor Segmentation using hybrid U-Net with Transformer and Efficient Attention
by: Ali, Syed Haider, et al.
Published: (2025)
by: Ali, Syed Haider, et al.
Published: (2025)
FlowIBR: Leveraging Pre-Training for Efficient Neural Image-Based Rendering of Dynamic Scenes
by: Büsching, Marcel, et al.
Published: (2023)
by: Büsching, Marcel, et al.
Published: (2023)
PyCAT4: A Hierarchical Vision Transformer-based Framework for 3D Human Pose Estimation
by: Yang, Zongyou, et al.
Published: (2025)
by: Yang, Zongyou, et al.
Published: (2025)
ControlHair: Physically-based Video Diffusion for Controllable Dynamic Hair Rendering
by: Lin, Weikai, et al.
Published: (2025)
by: Lin, Weikai, et al.
Published: (2025)
Pan-Arctic Permafrost Landform and Human-built Infrastructure Feature Detection with Vision Transformers and Location Embeddings
by: Perera, Amal S., et al.
Published: (2025)
by: Perera, Amal S., et al.
Published: (2025)
Circuit Mechanisms for Spatial Relation Generation in Diffusion Transformers
by: Wang, Binxu, et al.
Published: (2026)
by: Wang, Binxu, et al.
Published: (2026)
A Large-Scale Study on the Accuracy vs Cost Trade-offs of Training and Evaluation Settings in Fine-Grained Image Recognition
by: Rios, Edwin Arkel, et al.
Published: (2026)
by: Rios, Edwin Arkel, et al.
Published: (2026)
Towards Accurate and Efficient Waste Image Classification: A Hybrid Deep Learning and Machine Learning Approach
by: Nguyen, Ngoc-Bao-Quang, et al.
Published: (2025)
by: Nguyen, Ngoc-Bao-Quang, et al.
Published: (2025)
Similar Items
-
ViM-VQ: Efficient Post-Training Vector Quantization for Visual Mamba
by: Deng, Juncan, et al.
Published: (2025) -
Efficiency Meets Fidelity: A Novel Quantization Framework for Stable Diffusion
by: Li, Shuaiting, et al.
Published: (2024) -
SSVQ: Unleashing the Potential of Vector Quantization with Sign-Splitting
by: Li, Shuaiting, et al.
Published: (2025) -
VQ4ALL: Efficient Neural Network Representation via a Universal Codebook
by: Deng, Juncan, et al.
Published: (2024) -
Transformer-Based Vector Font Classification Using Different Font Formats: TrueType versus PostScript
by: Fujioka, Takumu, et al.
Published: (2025)