Timestep-Aware SVDQuant-GPTQ for W4A4 Quantization of Wan2.2-I2V
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Junhao, Yao, Dezhong, Jin, Hai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models
von: Du, Jinyang, et al.
Veröffentlicht: (2026)
von: Du, Jinyang, et al.
Veröffentlicht: (2026)
Fine-Tuning Open Video Generators for Cinematic Scene Synthesis: A Small-Data Pipeline with LoRA and Wan2.1 I2V
von: Akarsu, Meftun, et al.
Veröffentlicht: (2025)
von: Akarsu, Meftun, et al.
Veröffentlicht: (2025)
Q$^2$: Quantization-Aware Gradient Balancing and Attention Alignment for Low-Bit Quantization
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2025)
Timestep-Aware Correction for Quantized Diffusion Models
von: Yao, Yuzhe, et al.
Veröffentlicht: (2024)
von: Yao, Yuzhe, et al.
Veröffentlicht: (2024)
TARO: Timestep-Adaptive Representation Alignment with Onset-Aware Conditioning for Synchronized Video-to-Audio Synthesis
von: Ton, Tri, et al.
Veröffentlicht: (2025)
von: Ton, Tri, et al.
Veröffentlicht: (2025)
P4Q: Learning to Prompt for Quantization in Visual-language Models
von: Sun, Huixin, et al.
Veröffentlicht: (2024)
von: Sun, Huixin, et al.
Veröffentlicht: (2024)
Explainable Synthetic Image Detection through Diffusion Timestep Ensembling
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
HybridStitch: Pixel and Timestep Level Model Stitching for Diffusion Acceleration
von: Sun, Desen, et al.
Veröffentlicht: (2026)
von: Sun, Desen, et al.
Veröffentlicht: (2026)
Fine-Grained Post-Training Quantization for Large Vision Language Models with Quantization-Aware Integrated Gradients
von: Xiang, Ziwei, et al.
Veröffentlicht: (2026)
von: Xiang, Ziwei, et al.
Veröffentlicht: (2026)
Q-SAM2: Accurate Quantization for Segment Anything Model 2
von: Farronato, Nicola, et al.
Veröffentlicht: (2025)
von: Farronato, Nicola, et al.
Veröffentlicht: (2025)
DuQuant++: Fine-grained Rotation Enhances Microscaling FP4 Quantization
von: Lin, Haokun, et al.
Veröffentlicht: (2026)
von: Lin, Haokun, et al.
Veröffentlicht: (2026)
Is it safe to cross? Interpretable Risk Assessment with GPT-4V for Safety-Aware Street Crossing
von: Hwang, Hochul, et al.
Veröffentlicht: (2024)
von: Hwang, Hochul, et al.
Veröffentlicht: (2024)
HQ-DiT: Efficient Diffusion Transformer with FP4 Hybrid Quantization
von: Liu, Wenxuan, et al.
Veröffentlicht: (2024)
von: Liu, Wenxuan, et al.
Veröffentlicht: (2024)
SegQuant: A Semantics-Aware and Generalizable Quantization Framework for Diffusion Models
von: Zhang, Jiaji, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaji, et al.
Veröffentlicht: (2025)
Accelerating Diffusion-based Video Editing via Heterogeneous Caching: Beyond Full Computing at Sampled Denoising Timestep
von: Liu, Tianyi, et al.
Veröffentlicht: (2026)
von: Liu, Tianyi, et al.
Veröffentlicht: (2026)
Tail-Aware HiFloat4: W4A4 Post-Training Quantization for Wan2.2
von: Feng, Zhanfeng, et al.
Veröffentlicht: (2026)
von: Feng, Zhanfeng, et al.
Veröffentlicht: (2026)
The Disappearance of Timestep Embedding in Modern Time-Dependent Neural Networks
von: Kim, Bum Jun, et al.
Veröffentlicht: (2024)
von: Kim, Bum Jun, et al.
Veröffentlicht: (2024)
Anti-I2V: Safeguarding your photos from malicious image-to-video generation
von: Vu, Duc, et al.
Veröffentlicht: (2026)
von: Vu, Duc, et al.
Veröffentlicht: (2026)
PTQ4ARVG: Post-Training Quantization for AutoRegressive Visual Generation Models
von: Liu, Xuewen, et al.
Veröffentlicht: (2026)
von: Liu, Xuewen, et al.
Veröffentlicht: (2026)
Characterizing Motion Encoding in Video Diffusion Timesteps
von: Baherwani, Vatsal, et al.
Veröffentlicht: (2025)
von: Baherwani, Vatsal, et al.
Veröffentlicht: (2025)
Wan-S2V: Audio-Driven Cinematic Video Generation
von: Gao, Xin, et al.
Veröffentlicht: (2025)
von: Gao, Xin, et al.
Veröffentlicht: (2025)
SVDQuant: Absorbing Outliers by Low-Rank Components for 4-Bit Diffusion Models
von: Li, Muyang, et al.
Veröffentlicht: (2024)
von: Li, Muyang, et al.
Veröffentlicht: (2024)
Modality-Aware and Anatomical Vector-Quantized Autoencoding for Multimodal Brain MRI
von: Li, Mingjie, et al.
Veröffentlicht: (2026)
von: Li, Mingjie, et al.
Veröffentlicht: (2026)
T2I-VeRW: Part-level Fine-grained Perception for Text-to-Image Vehicle Retrieval
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
Event Stream-based Visual Object Tracking: HDETrack V2 and A High-Definition Benchmark
von: Wang, Shiao, et al.
Veröffentlicht: (2025)
von: Wang, Shiao, et al.
Veröffentlicht: (2025)
T2S-GPT: Dynamic Vector Quantization for Autoregressive Sign Language Production from Text
von: Yin, Aoxiong, et al.
Veröffentlicht: (2024)
von: Yin, Aoxiong, et al.
Veröffentlicht: (2024)
Echo4DIR: 4D Implicit Heart Reconstruction from 2D Echocardiography Videos
von: Liu, Yanan, et al.
Veröffentlicht: (2026)
von: Liu, Yanan, et al.
Veröffentlicht: (2026)
Interruption-Aware Cooperative Perception for V2X Communication-Aided Autonomous Driving
von: Ren, Shunli, et al.
Veröffentlicht: (2023)
von: Ren, Shunli, et al.
Veröffentlicht: (2023)
QAPruner: Quantization-Aware Vision Token Pruning for Multimodal Large Language Models
von: Wang, Xinhao, et al.
Veröffentlicht: (2026)
von: Wang, Xinhao, et al.
Veröffentlicht: (2026)
Semantic-Consistent Bidirectional Contrastive Hashing for Noisy Multi-Label Cross-Modal Retrieval
von: Peng, Likang, et al.
Veröffentlicht: (2025)
von: Peng, Likang, et al.
Veröffentlicht: (2025)
The Dawn of KAN in Image-to-Image (I2I) Translation: Integrating Kolmogorov-Arnold Networks with GANs for Unpaired I2I Translation
von: Mahara, Arpan, et al.
Veröffentlicht: (2024)
von: Mahara, Arpan, et al.
Veröffentlicht: (2024)
RSwinV2-MD: An Enhanced Residual SwinV2 Transformer for Monkeypox Detection from Skin Images
von: Iqbal, Rashid, et al.
Veröffentlicht: (2026)
von: Iqbal, Rashid, et al.
Veröffentlicht: (2026)
Temporal Aware Pruning for Efficient Diffusion-based Video Generation
von: Li, Sheng, et al.
Veröffentlicht: (2026)
von: Li, Sheng, et al.
Veröffentlicht: (2026)
SDQ-LLM: Sigma-Delta Quantization for 1-bit LLMs of any size
von: Xia, Junhao, et al.
Veröffentlicht: (2025)
von: Xia, Junhao, et al.
Veröffentlicht: (2025)
GAT-NeRF: Geometry-Aware-Transformer Enhanced Neural Radiance Fields for High-Fidelity 4D Facial Avatars
von: Chang, Zhe, et al.
Veröffentlicht: (2026)
von: Chang, Zhe, et al.
Veröffentlicht: (2026)
Layer- and Timestep-Adaptive Differentiable Token Compression Ratios for Efficient Diffusion Transformers
von: You, Haoran, et al.
Veröffentlicht: (2024)
von: You, Haoran, et al.
Veröffentlicht: (2024)
Q-Sched: Pushing the Boundaries of Few-Step Diffusion Models with Quantization-Aware Scheduling
von: Frumkin, Natalia, et al.
Veröffentlicht: (2025)
von: Frumkin, Natalia, et al.
Veröffentlicht: (2025)
Efficient Quantization-Aware Training on Segment Anything Model in Medical Images and Its Deployment
von: Lu, Haisheng, et al.
Veröffentlicht: (2024)
von: Lu, Haisheng, et al.
Veröffentlicht: (2024)
4D-GSW: Kinematic-Aware Spatio-Temporal Consistent Watermarking for 4D Gaussian Splatting
von: Zhou, Sifan, et al.
Veröffentlicht: (2026)
von: Zhou, Sifan, et al.
Veröffentlicht: (2026)
LoCoT2V-Bench: Benchmarking Long-Form and Complex Text-to-Video Generation
von: Zheng, Xiangqing, et al.
Veröffentlicht: (2025)
von: Zheng, Xiangqing, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models
von: Du, Jinyang, et al.
Veröffentlicht: (2026) -
Fine-Tuning Open Video Generators for Cinematic Scene Synthesis: A Small-Data Pipeline with LoRA and Wan2.1 I2V
von: Akarsu, Meftun, et al.
Veröffentlicht: (2025) -
Q$^2$: Quantization-Aware Gradient Balancing and Attention Alignment for Low-Bit Quantization
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2025) -
Timestep-Aware Correction for Quantized Diffusion Models
von: Yao, Yuzhe, et al.
Veröffentlicht: (2024) -
TARO: Timestep-Adaptive Representation Alignment with Onset-Aware Conditioning for Synchronized Video-to-Audio Synthesis
von: Ton, Tri, et al.
Veröffentlicht: (2025)