DraftAttention: Fast Video Diffusion via Low-Resolution Attention Guidance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shen, Xuan, Han, Chenxia, Zhou, Yufa, Xie, Yanyue, Gong, Yifan, Wang, Quanyi, Wang, Yiwei, Wang, Yanzhi, Zhao, Pu, Gu, Jiuxiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
Efficient Reasoning with Hidden Thinking
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers
von: Shen, Xuan, et al.
Veröffentlicht: (2024)
von: Shen, Xuan, et al.
Veröffentlicht: (2024)
QuartDepth: Post-Training Quantization for Real-Time Depth Estimation on the Edge
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
HybridFlow: Infusing Continuity into Masked Codebook for Extreme Low-Bitrate Image Compression
von: Lu, Lei, et al.
Veröffentlicht: (2024)
von: Lu, Lei, et al.
Veröffentlicht: (2024)
Fast and Memory-Efficient Video Diffusion Using Streamlined Inference
von: Zhan, Zheng, et al.
Veröffentlicht: (2024)
von: Zhan, Zheng, et al.
Veröffentlicht: (2024)
Squat: Quant Small Language Models on the Edge
von: Shen, Xuan, et al.
Veröffentlicht: (2024)
von: Shen, Xuan, et al.
Veröffentlicht: (2024)
OmniMem: Scalable and Adaptive Memory Retrieval for Long Video Generation
von: Zhao, Lin, et al.
Veröffentlicht: (2026)
von: Zhao, Lin, et al.
Veröffentlicht: (2026)
Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention
von: Lv, Chengtao, et al.
Veröffentlicht: (2026)
von: Lv, Chengtao, et al.
Veröffentlicht: (2026)
Differentially Private Attention Computation
von: Gao, Yeqi, et al.
Veröffentlicht: (2023)
von: Gao, Yeqi, et al.
Veröffentlicht: (2023)
Collaborative Compression for Large-Scale MoE Deployment on Edge
von: Chen, Yixiao, et al.
Veröffentlicht: (2025)
von: Chen, Yixiao, et al.
Veröffentlicht: (2025)
Numerical Pruning for Efficient Autoregressive Models
von: Shen, Xuan, et al.
Veröffentlicht: (2024)
von: Shen, Xuan, et al.
Veröffentlicht: (2024)
Video Super-Resolution Transformer with Masked Inter&Intra-Frame Attention
von: Zhou, Xingyu, et al.
Veröffentlicht: (2024)
von: Zhou, Xingyu, et al.
Veröffentlicht: (2024)
FastFace: Tuning Identity Preservation in Distilled Diffusion via Guidance and Attention
von: Karpukhin, Sergey, et al.
Veröffentlicht: (2025)
von: Karpukhin, Sergey, et al.
Veröffentlicht: (2025)
Higher-order Linear Attention
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
SALAD: Achieve High-Sparsity Attention via Efficient Linear Attention Tuning for Video Diffusion Transformer
von: Fang, Tongcheng, et al.
Veröffentlicht: (2026)
von: Fang, Tongcheng, et al.
Veröffentlicht: (2026)
GRA: Detecting Oriented Objects through Group-wise Rotating and Attention
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models
von: Chen, Dar-Yen, et al.
Veröffentlicht: (2025)
von: Chen, Dar-Yen, et al.
Veröffentlicht: (2025)
FastAttention: Extend FlashAttention2 to NPUs and Low-resource GPUs
von: Lin, Haoran, et al.
Veröffentlicht: (2024)
von: Lin, Haoran, et al.
Veröffentlicht: (2024)
Rethinking Token Reduction for State Space Models
von: Zhan, Zheng, et al.
Veröffentlicht: (2024)
von: Zhan, Zheng, et al.
Veröffentlicht: (2024)
Blockwise SFT for Diffusion Language Models: Reconciling Bidirectional Attention and Autoregressive Decoding
von: Sun, Bowen, et al.
Veröffentlicht: (2025)
von: Sun, Bowen, et al.
Veröffentlicht: (2025)
Understanding and Improving Training-free Loss-based Diffusion Guidance
von: Shen, Yifei, et al.
Veröffentlicht: (2024)
von: Shen, Yifei, et al.
Veröffentlicht: (2024)
Cavia: Camera-controllable Multi-view Video Diffusion with View-Integrated Attention
von: Xu, Dejia, et al.
Veröffentlicht: (2024)
von: Xu, Dejia, et al.
Veröffentlicht: (2024)
Attention-Constrained Inference for Robust Decoder-Only Text-to-Speech
von: Wang, Hankun, et al.
Veröffentlicht: (2024)
von: Wang, Hankun, et al.
Veröffentlicht: (2024)
Understanding Attention Mechanism in Video Diffusion Models
von: Liu, Bingyan, et al.
Veröffentlicht: (2025)
von: Liu, Bingyan, et al.
Veröffentlicht: (2025)
Re-Attentional Controllable Video Diffusion Editing
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2024)
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2024)
METAL: A Multi-Agent Framework for Chart Generation with Test-Time Scaling
von: Li, Bingxuan, et al.
Veröffentlicht: (2025)
von: Li, Bingxuan, et al.
Veröffentlicht: (2025)
STDAN: Deformable Attention Network for Space-Time Video Super-Resolution
von: Wang, Hai, et al.
Veröffentlicht: (2022)
von: Wang, Hai, et al.
Veröffentlicht: (2022)
DAMRO: Dive into the Attention Mechanism of LVLM to Reduce Object Hallucination
von: Gong, Xuan, et al.
Veröffentlicht: (2024)
von: Gong, Xuan, et al.
Veröffentlicht: (2024)
Attention Beats Linear for Fast Implicit Neural Representation Generation
von: Zhang, Shuyi, et al.
Veröffentlicht: (2024)
von: Zhang, Shuyi, et al.
Veröffentlicht: (2024)
Beyond Linear Approximations: A Novel Pruning Approach for Attention Matrix
von: Liang, Yingyu, et al.
Veröffentlicht: (2024)
von: Liang, Yingyu, et al.
Veröffentlicht: (2024)
MoE-Pruner: Pruning Mixture-of-Experts Large Language Model using the Hints from Its Router
von: Xie, Yanyue, et al.
Veröffentlicht: (2024)
von: Xie, Yanyue, et al.
Veröffentlicht: (2024)
SRFLOW‐Attention: Super Resolution with Multi‐Head Attention for Effective Use of Low‐Resolution Information
von: Shinya Ohtani, et al.
Veröffentlicht: (2024)
von: Shinya Ohtani, et al.
Veröffentlicht: (2024)
Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance
von: Ahn, Donghoon, et al.
Veröffentlicht: (2024)
von: Ahn, Donghoon, et al.
Veröffentlicht: (2024)
ChronoTailor: Harnessing Attention Guidance for Fine-Grained Video Virtual Try-On
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
Fast Cross-Operator Optimization of Attention Dataflow
von: Chang, Haodong, et al.
Veröffentlicht: (2026)
von: Chang, Haodong, et al.
Veröffentlicht: (2026)
HDCompression: Hybrid-Diffusion Image Compression for Ultra-Low Bitrates
von: Lu, Lei, et al.
Veröffentlicht: (2025)
von: Lu, Lei, et al.
Veröffentlicht: (2025)
Efficient Diffusion Transformer with Step-wise Dynamic Attention Mediators
von: Pu, Yifan, et al.
Veröffentlicht: (2024)
von: Pu, Yifan, et al.
Veröffentlicht: (2024)
RAPTOR: Real-Time High-Resolution UAV Video Prediction with Efficient Video Attention
von: Chen, Zhan, et al.
Veröffentlicht: (2025)
von: Chen, Zhan, et al.
Veröffentlicht: (2025)
VMonarch: Efficient Video Diffusion Transformers with Structured Attention
von: Liang, Cheng, et al.
Veröffentlicht: (2026)
von: Liang, Cheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge
von: Shen, Xuan, et al.
Veröffentlicht: (2025) -
Efficient Reasoning with Hidden Thinking
von: Shen, Xuan, et al.
Veröffentlicht: (2025) -
LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers
von: Shen, Xuan, et al.
Veröffentlicht: (2024) -
QuartDepth: Post-Training Quantization for Real-Time Depth Estimation on the Edge
von: Shen, Xuan, et al.
Veröffentlicht: (2025) -
HybridFlow: Infusing Continuity into Masked Codebook for Extreme Low-Bitrate Image Compression
von: Lu, Lei, et al.
Veröffentlicht: (2024)