S$^2$Q-VDiT: Accurate Quantized Video Diffusion Transformer with Salient Data and Sparse Token Distillation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Feng, Weilun, Qin, Haotong, Yang, Chuanguang, Li, Xiangqi, Yang, Han, Li, Yuqi, An, Zhulin, Huang, Libo, Magno, Michele, Xu, Yongjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
QuantSparse: Comprehensively Compressing Video Diffusion Transformer with Model Quantization and Attention Sparsification
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
Quantized Visual Geometry Grounded Transformer
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
MPQ-DMv2: Flexible Residual Mixed Precision Quantization for Low-Bit Diffusion Models with Temporal Distillation
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
MPQ-DM: Mixed Precision Quantization for Extremely Low Bit Diffusion Models
von: Feng, Weilun, et al.
Veröffentlicht: (2024)
von: Feng, Weilun, et al.
Veröffentlicht: (2024)
WorldCache: Accelerating World Models for Free via Heterogeneous Token Caching
von: Feng, Weilun, et al.
Veröffentlicht: (2026)
von: Feng, Weilun, et al.
Veröffentlicht: (2026)
Relational Diffusion Distillation for Efficient Image Generation
von: Feng, Weilun, et al.
Veröffentlicht: (2024)
von: Feng, Weilun, et al.
Veröffentlicht: (2024)
Echo-Forcing: A Scene Memory Framework for Interactive Long Video Generation
von: Wu, Mingqiang, et al.
Veröffentlicht: (2026)
von: Wu, Mingqiang, et al.
Veröffentlicht: (2026)
Teacher-Guided Student Self-Knowledge Distillation Using Diffusion Model
von: Wang, Yu, et al.
Veröffentlicht: (2026)
von: Wang, Yu, et al.
Veröffentlicht: (2026)
Online Policy Distillation with Decision-Attention
von: Yu, Xinqiang, et al.
Veröffentlicht: (2024)
von: Yu, Xinqiang, et al.
Veröffentlicht: (2024)
Multi-party Collaborative Attention Control for Image Customization
von: Yang, Han, et al.
Veröffentlicht: (2025)
von: Yang, Han, et al.
Veröffentlicht: (2025)
SRKD: Towards Efficient 3D Point Cloud Segmentation via Structure- and Relation-aware Knowledge Distillation
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
Multi-Teacher Knowledge Distillation with Reinforcement Learning for Visual Recognition
von: Yang, Chuanguang, et al.
Veröffentlicht: (2025)
von: Yang, Chuanguang, et al.
Veröffentlicht: (2025)
Q-SAM2: Accurate Quantization for Segment Anything Model 2
von: Farronato, Nicola, et al.
Veröffentlicht: (2025)
von: Farronato, Nicola, et al.
Veröffentlicht: (2025)
CLIP-KD: An Empirical Study of CLIP Model Distillation
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
Frequency-Aligned Knowledge Distillation for Lightweight Spatiotemporal Forecasting
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration
von: Chen, Yujie, et al.
Veröffentlicht: (2025)
von: Chen, Yujie, et al.
Veröffentlicht: (2025)
Exemplar-Free Class Incremental Learning via Incremental Representation
von: Huang, Libo, et al.
Veröffentlicht: (2024)
von: Huang, Libo, et al.
Veröffentlicht: (2024)
Efficient Continual Learning through Frequency Decomposition and Integration
von: Liu, Ruiqi, et al.
Veröffentlicht: (2025)
von: Liu, Ruiqi, et al.
Veröffentlicht: (2025)
Post-Training Quantization for Video Matting
von: Zhu, Tianrui, et al.
Veröffentlicht: (2025)
von: Zhu, Tianrui, et al.
Veröffentlicht: (2025)
Accurate LoRA-Finetuning Quantization of LLMs via Information Retention
von: Qin, Haotong, et al.
Veröffentlicht: (2024)
von: Qin, Haotong, et al.
Veröffentlicht: (2024)
Parameterized Prompt for Incremental Object Detection
von: An, Zijia, et al.
Veröffentlicht: (2025)
von: An, Zijia, et al.
Veröffentlicht: (2025)
DVD-Quant: Data-free Video Diffusion Transformers Quantization
von: Li, Zhiteng, et al.
Veröffentlicht: (2025)
von: Li, Zhiteng, et al.
Veröffentlicht: (2025)
First-Order Error Matters: Accurate Compensation for Quantized Large Language Models
von: Zheng, Xingyu, et al.
Veröffentlicht: (2025)
von: Zheng, Xingyu, et al.
Veröffentlicht: (2025)
PrePrompt: Predictive prompting for class incremental learning
von: Huang, Libo, et al.
Veröffentlicht: (2025)
von: Huang, Libo, et al.
Veröffentlicht: (2025)
TreeQ: Pushing the Quantization Boundary of Diffusion Transformer via Tree-Structured Mixed-Precision Search
von: Yang, Kaicheng, et al.
Veröffentlicht: (2025)
von: Yang, Kaicheng, et al.
Veröffentlicht: (2025)
ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
MultiAnimate: Pose-Guided Image Animation Made Extensible
von: Hu, Yingcheng, et al.
Veröffentlicht: (2026)
von: Hu, Yingcheng, et al.
Veröffentlicht: (2026)
AMMKD: Adaptive Multimodal Multi-teacher Distillation for Lightweight Vision-Language Models
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
von: Li, Yuqi, et al.
Veröffentlicht: (2025)
PassionSR: Post-Training Quantization with Adaptive Scale in One-Step Diffusion based Image Super-Resolution
von: Zhu, Libo, et al.
Veröffentlicht: (2024)
von: Zhu, Libo, et al.
Veröffentlicht: (2024)
BiVM: Accurate Binarized Neural Network for Efficient Video Matting
von: Qin, Haotong, et al.
Veröffentlicht: (2025)
von: Qin, Haotong, et al.
Veröffentlicht: (2025)
Q-DiT4SR: Exploration of Detail-Preserving Diffusion Transformer Quantization for Real-World Image Super-Resolution
von: Zhang, Xun, et al.
Veröffentlicht: (2026)
von: Zhang, Xun, et al.
Veröffentlicht: (2026)
Event-Priori-Based Vision-Language Model for Efficient Visual Understanding
von: Qin, Haotong, et al.
Veröffentlicht: (2025)
von: Qin, Haotong, et al.
Veröffentlicht: (2025)
QArtSR: Quantization via Reverse-Module and Timestep-Retraining in One-Step Diffusion based Image Super-Resolution
von: Zhu, Libo, et al.
Veröffentlicht: (2025)
von: Zhu, Libo, et al.
Veröffentlicht: (2025)
Low-redundancy Distillation for Continual Learning
von: Liu, RuiQi, et al.
Veröffentlicht: (2023)
von: Liu, RuiQi, et al.
Veröffentlicht: (2023)
ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification
von: He, Yefei, et al.
Veröffentlicht: (2024)
von: He, Yefei, et al.
Veröffentlicht: (2024)
Q&C: When Quantization Meets Cache in Efficient Image Generation
von: Ding, Xin, et al.
Veröffentlicht: (2025)
von: Ding, Xin, et al.
Veröffentlicht: (2025)
Distilling Time Series Foundation Models for Efficient Forecasting
von: Li, Yuqi, et al.
Veröffentlicht: (2026)
von: Li, Yuqi, et al.
Veröffentlicht: (2026)
Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers
von: Chen, Lei, et al.
Veröffentlicht: (2024)
von: Chen, Lei, et al.
Veröffentlicht: (2024)
QuantFace: Efficient Quantization for Face Restoration
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers
von: Feng, Weilun, et al.
Veröffentlicht: (2025) -
QuantSparse: Comprehensively Compressing Video Diffusion Transformer with Model Quantization and Attention Sparsification
von: Feng, Weilun, et al.
Veröffentlicht: (2025) -
Quantized Visual Geometry Grounded Transformer
von: Feng, Weilun, et al.
Veröffentlicht: (2025) -
MPQ-DMv2: Flexible Residual Mixed Precision Quantization for Low-Bit Diffusion Models with Temporal Distillation
von: Feng, Weilun, et al.
Veröffentlicht: (2025) -
MPQ-DM: Mixed Precision Quantization for Extremely Low Bit Diffusion Models
von: Feng, Weilun, et al.
Veröffentlicht: (2024)