Hardware-Friendly Static Quantization Method for Video Diffusion Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yi, Sanghyun, Liu, Qingfeng, El-Khamy, Mostafa |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation
von: Liu, Qingfeng, et al.
Veröffentlicht: (2024)
von: Liu, Qingfeng, et al.
Veröffentlicht: (2024)
Open Vocabulary Panoptic Segmentation With Retrieval Augmentation
von: Sadeq, Nafis, et al.
Veröffentlicht: (2026)
von: Sadeq, Nafis, et al.
Veröffentlicht: (2026)
Model Already Knows the Best Noise: Bayesian Active Noise Selection via Attention in Video Diffusion Model
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2025)
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2025)
QVD: Post-training Quantization for Video Diffusion Models
von: Tian, Shilong, et al.
Veröffentlicht: (2024)
von: Tian, Shilong, et al.
Veröffentlicht: (2024)
HQ-DiT: Efficient Diffusion Transformer with FP4 Hybrid Quantization
von: Liu, Wenxuan, et al.
Veröffentlicht: (2024)
von: Liu, Wenxuan, et al.
Veröffentlicht: (2024)
CLQ: Cross-Layer Guided Orthogonal-based Quantization for Diffusion Transformers
von: Liu, Kai, et al.
Veröffentlicht: (2025)
von: Liu, Kai, et al.
Veröffentlicht: (2025)
HadaNorm: Diffusion Transformer Quantization through Mean-Centered Transformations
von: Federici, Marco, et al.
Veröffentlicht: (2025)
von: Federici, Marco, et al.
Veröffentlicht: (2025)
FashionFlow: Leveraging Diffusion Models for Dynamic Fashion Video Synthesis from Static Imagery
von: Islam, Tasin, et al.
Veröffentlicht: (2023)
von: Islam, Tasin, et al.
Veröffentlicht: (2023)
SANA-Video: Efficient Video Generation with Block Linear Diffusion Transformer
von: Chen, Junsong, et al.
Veröffentlicht: (2025)
von: Chen, Junsong, et al.
Veröffentlicht: (2025)
Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers
von: Chen, Lei, et al.
Veröffentlicht: (2024)
von: Chen, Lei, et al.
Veröffentlicht: (2024)
EdgeDiT: Hardware-Aware Diffusion Transformers for Efficient On-Device Image Generation
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
von: Kodavanti, Sravanth, et al.
Veröffentlicht: (2026)
Quasar-ViT: Hardware-Oriented Quantization-Aware Architecture Search for Vision Transformers
von: Li, Zhengang, et al.
Veröffentlicht: (2024)
von: Li, Zhengang, et al.
Veröffentlicht: (2024)
Safety Alignment Backfires: Preventing the Re-emergence of Suppressed Concepts in Fine-tuned Text-to-Image Diffusion Models
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
VMonarch: Efficient Video Diffusion Transformers with Structured Attention
von: Liang, Cheng, et al.
Veröffentlicht: (2026)
von: Liang, Cheng, et al.
Veröffentlicht: (2026)
Boosting Camera Motion Control for Video Diffusion Transformers
von: Cheong, Soon Yau, et al.
Veröffentlicht: (2024)
von: Cheong, Soon Yau, et al.
Veröffentlicht: (2024)
Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models
von: Du, Jinyang, et al.
Veröffentlicht: (2026)
von: Du, Jinyang, et al.
Veröffentlicht: (2026)
EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion
von: Wang, Haotian, et al.
Veröffentlicht: (2024)
von: Wang, Haotian, et al.
Veröffentlicht: (2024)
FPQVAR: Floating Point Quantization for Visual Autoregressive Model with FPGA Hardware Co-design
von: Wei, Renjie, et al.
Veröffentlicht: (2025)
von: Wei, Renjie, et al.
Veröffentlicht: (2025)
Post-Training Quantization for Video Matting
von: Zhu, Tianrui, et al.
Veröffentlicht: (2025)
von: Zhu, Tianrui, et al.
Veröffentlicht: (2025)
MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization
von: Yu, JiangYong, et al.
Veröffentlicht: (2025)
von: Yu, JiangYong, et al.
Veröffentlicht: (2025)
SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer
von: Zhao, Yuyang, et al.
Veröffentlicht: (2026)
von: Zhao, Yuyang, et al.
Veröffentlicht: (2026)
Safeguard Text-to-Image Diffusion Models with Human Feedback Inversion
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
Enabling Versatile Controls for Video Diffusion Models
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
Regularizing Differentiable Architecture Search with Smooth Activation
von: Zhou, Yanlin, et al.
Veröffentlicht: (2025)
von: Zhou, Yanlin, et al.
Veröffentlicht: (2025)
DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching
von: Zou, Chang, et al.
Veröffentlicht: (2026)
von: Zou, Chang, et al.
Veröffentlicht: (2026)
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and $\mathcal{O}(T)$ Complexity
von: Zou, Shihao, et al.
Veröffentlicht: (2025)
von: Zou, Shihao, et al.
Veröffentlicht: (2025)
BWCache: Accelerating Video Diffusion Transformers through Block-Wise Caching
von: Cui, Hanshuai, et al.
Veröffentlicht: (2025)
von: Cui, Hanshuai, et al.
Veröffentlicht: (2025)
Mixed Non-linear Quantization for Vision Transformers
von: Kim, Gihwan, et al.
Veröffentlicht: (2024)
von: Kim, Gihwan, et al.
Veröffentlicht: (2024)
QNCD: Quantization Noise Correction for Diffusion Models
von: Chu, Huanpeng, et al.
Veröffentlicht: (2024)
von: Chu, Huanpeng, et al.
Veröffentlicht: (2024)
DilateQuant: Accurate and Efficient Diffusion Quantization via Weight Dilation
von: Liu, Xuewen, et al.
Veröffentlicht: (2024)
von: Liu, Xuewen, et al.
Veröffentlicht: (2024)
DFQ-ViT: Data-Free Quantization for Vision Transformers without Fine-tuning
von: Tong, Yujia, et al.
Veröffentlicht: (2025)
von: Tong, Yujia, et al.
Veröffentlicht: (2025)
LuxDiT: Lighting Estimation with Video Diffusion Transformer
von: Liang, Ruofan, et al.
Veröffentlicht: (2025)
von: Liang, Ruofan, et al.
Veröffentlicht: (2025)
Efficient Personalization of Quantized Diffusion Model without Backpropagation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Amortized-Precision Quantization for Early-Exit Vision Transformers
von: Fang, Rui, et al.
Veröffentlicht: (2026)
von: Fang, Rui, et al.
Veröffentlicht: (2026)
Memory-Efficient Fine-Tuning for Quantized Diffusion Model
von: Ryu, Hyogon, et al.
Veröffentlicht: (2024)
von: Ryu, Hyogon, et al.
Veröffentlicht: (2024)
Video Motion Transfer with Diffusion Transformers
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
SegQuant: A Semantics-Aware and Generalizable Quantization Framework for Diffusion Models
von: Zhang, Jiaji, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaji, et al.
Veröffentlicht: (2025)
GVD: Guiding Video Diffusion Model for Scalable Video Distillation
von: Li, Kunyang, et al.
Veröffentlicht: (2025)
von: Li, Kunyang, et al.
Veröffentlicht: (2025)
Re-Attentional Controllable Video Diffusion Editing
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2024)
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2024)
AI-based System for Transforming text and sound to Educational Videos
von: ElAlami, M. E., et al.
Veröffentlicht: (2026)
von: ElAlami, M. E., et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation
von: Liu, Qingfeng, et al.
Veröffentlicht: (2024) -
Open Vocabulary Panoptic Segmentation With Retrieval Augmentation
von: Sadeq, Nafis, et al.
Veröffentlicht: (2026) -
Model Already Knows the Best Noise: Bayesian Active Noise Selection via Attention in Video Diffusion Model
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2025) -
QVD: Post-training Quantization for Video Diffusion Models
von: Tian, Shilong, et al.
Veröffentlicht: (2024) -
HQ-DiT: Efficient Diffusion Transformer with FP4 Hybrid Quantization
von: Liu, Wenxuan, et al.
Veröffentlicht: (2024)