Accelerating Frequency Domain Diffusion Models with Error-Feedback Event-Driven Caching
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Dong, Wang, Haisheng, Yu, Yanxuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FastCache: Fast Caching for Diffusion Transformer Through Learnable Linear Approximation
von: Liu, Dong, et al.
Veröffentlicht: (2025)
von: Liu, Dong, et al.
Veröffentlicht: (2025)
LLMEasyQuant: Scalable Quantization for Parallel and Distributed LLM Inference
von: Liu, Dong, et al.
Veröffentlicht: (2024)
von: Liu, Dong, et al.
Veröffentlicht: (2024)
MT2ST: Adaptive Multi-Task to Single-Task Learning
von: Liu, Dong, et al.
Veröffentlicht: (2024)
von: Liu, Dong, et al.
Veröffentlicht: (2024)
Efficient Graph Optimization via Distance-Aware Graph Representation Learning
von: Liu, Dong, et al.
Veröffentlicht: (2024)
von: Liu, Dong, et al.
Veröffentlicht: (2024)
FreqCa: Accelerating Diffusion Models via Frequency-Aware Caching
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
Window-Diffusion: Accelerating Diffusion Language Model Inference with Windowed Token Pruning and Caching
von: Zuo, Fengrui, et al.
Veröffentlicht: (2026)
von: Zuo, Fengrui, et al.
Veröffentlicht: (2026)
dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching
von: Ma, Xinyin, et al.
Veröffentlicht: (2024)
von: Ma, Xinyin, et al.
Veröffentlicht: (2024)
AB-Cache: Training-Free Acceleration of Diffusion Models via Adams-Bashforth Cached Feature Reuse
von: Yu, Zichao, et al.
Veröffentlicht: (2025)
von: Yu, Zichao, et al.
Veröffentlicht: (2025)
Token Caching for Diffusion Transformer Acceleration
von: Lou, Jinming, et al.
Veröffentlicht: (2024)
von: Lou, Jinming, et al.
Veröffentlicht: (2024)
SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2026)
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2026)
SmoothCache: A Universal Inference Acceleration Technique for Diffusion Transformers
von: Liu, Joseph, et al.
Veröffentlicht: (2024)
von: Liu, Joseph, et al.
Veröffentlicht: (2024)
High-Frequency Space Diffusion Models for Accelerated MRI
von: Cao, Chentao, et al.
Veröffentlicht: (2022)
von: Cao, Chentao, et al.
Veröffentlicht: (2022)
Accelerated Distributed Optimization with Compression and Error Feedback
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
MKA: Memory-Keyed Attention for Efficient Long-Context Reasoning
von: Liu, Dong, et al.
Veröffentlicht: (2026)
von: Liu, Dong, et al.
Veröffentlicht: (2026)
Adaptive Hybrid Caching for Efficient Text-to-Video Diffusion Model Acceleration
von: Wei, Yuanxin, et al.
Veröffentlicht: (2025)
von: Wei, Yuanxin, et al.
Veröffentlicht: (2025)
Accelerating Diffusion Transformers with Token-wise Feature Caching
von: Zou, Chang, et al.
Veröffentlicht: (2024)
von: Zou, Chang, et al.
Veröffentlicht: (2024)
ReasonCache: Accelerating Large Reasoning Model Serving through KV Cache Sharing
von: Chen, Kaiwen, et al.
Veröffentlicht: (2025)
von: Chen, Kaiwen, et al.
Veröffentlicht: (2025)
Relational Feature Caching for Accelerating Diffusion Transformers
von: Son, Byunggwan, et al.
Veröffentlicht: (2026)
von: Son, Byunggwan, et al.
Veröffentlicht: (2026)
FEB-Cache: Frequency-Guided Exposure Bias Reduction for Enhancing Diffusion Transformer Caching
von: Zou, Zhen, et al.
Veröffentlicht: (2025)
von: Zou, Zhen, et al.
Veröffentlicht: (2025)
SpeCa: Accelerating Diffusion Transformers with Speculative Feature Caching
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
AttnCache: Accelerating Self-Attention Inference for LLM Prefill via Attention Cache
von: Song, Dinghong, et al.
Veröffentlicht: (2025)
von: Song, Dinghong, et al.
Veröffentlicht: (2025)
RelayCaching: Accelerating LLM Collaboration via Decoding KV Cache Reuse
von: Geng, Yingsheng, et al.
Veröffentlicht: (2026)
von: Geng, Yingsheng, et al.
Veröffentlicht: (2026)
Time Series Diffusion in the Frequency Domain
von: Crabbé, Jonathan, et al.
Veröffentlicht: (2024)
von: Crabbé, Jonathan, et al.
Veröffentlicht: (2024)
FlexCache: Flexible Approximate Cache System for Video Diffusion
von: Sun, Desen, et al.
Veröffentlicht: (2024)
von: Sun, Desen, et al.
Veröffentlicht: (2024)
TinyServe: Query-Aware Cache Selection for Efficient LLM Serving
von: Liu, Dong, et al.
Veröffentlicht: (2025)
von: Liu, Dong, et al.
Veröffentlicht: (2025)
SPA-Cache: Singular Proxies for Adaptive Caching in Diffusion Language Models
von: Sun, Wenhao, et al.
Veröffentlicht: (2026)
von: Sun, Wenhao, et al.
Veröffentlicht: (2026)
Rethinking Token-wise Feature Caching: Accelerating Diffusion Transformers with Dual Feature Caching
von: Zou, Chang, et al.
Veröffentlicht: (2024)
von: Zou, Chang, et al.
Veröffentlicht: (2024)
Accelerating LLM Inference Throughput via Asynchronous KV Cache Prefetching
von: Dong, Yanhao, et al.
Veröffentlicht: (2025)
von: Dong, Yanhao, et al.
Veröffentlicht: (2025)
CacheClip: Accelerating RAG with Effective KV Cache Reuse
von: Yang, Bin, et al.
Veröffentlicht: (2025)
von: Yang, Bin, et al.
Veröffentlicht: (2025)
CXL-SpecKV: A Disaggregated FPGA Speculative KV-Cache for Datacenter LLM Serving
von: Liu, Dong, et al.
Veröffentlicht: (2025)
von: Liu, Dong, et al.
Veröffentlicht: (2025)
LAPA: Log-Domain Prediction-Driven Dynamic Sparsity Accelerator for Transformer Model
von: Wang, Huizheng, et al.
Veröffentlicht: (2025)
von: Wang, Huizheng, et al.
Veröffentlicht: (2025)
InvarDiff: Cross-Scale Invariance Caching for Accelerated Diffusion Models
von: Wu, Zihao
Veröffentlicht: (2025)
von: Wu, Zihao
Veröffentlicht: (2025)
Event-Driven Digital-Time-Domain Inference Architectures for Tsetlin Machines
von: Lan, Tian, et al.
Veröffentlicht: (2025)
von: Lan, Tian, et al.
Veröffentlicht: (2025)
Circuit Representation Learning with Masked Gate Modeling and Verilog-AIG Alignment
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
Accelerating Parallel Sampling of Diffusion Models
von: Tang, Zhiwei, et al.
Veröffentlicht: (2024)
von: Tang, Zhiwei, et al.
Veröffentlicht: (2024)
On Diffusion Models for Multi-Agent Partial Observability: Shared Attractors, Error Bounds, and Composite Flow
von: Wang, Tonghan, et al.
Veröffentlicht: (2024)
von: Wang, Tonghan, et al.
Veröffentlicht: (2024)
Token Pruning for Caching Better: 9 Times Acceleration on Stable Diffusion for Free
von: Zhang, Evelyn, et al.
Veröffentlicht: (2024)
von: Zhang, Evelyn, et al.
Veröffentlicht: (2024)
IC-Cache: Efficient Large Language Model Serving via In-context Caching
von: Yu, Yifan, et al.
Veröffentlicht: (2025)
von: Yu, Yifan, et al.
Veröffentlicht: (2025)
NI Sampling: Accelerating Discrete Diffusion Sampling by Token Order Optimization
von: Liu, Enshu, et al.
Veröffentlicht: (2026)
von: Liu, Enshu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
FastCache: Fast Caching for Diffusion Transformer Through Learnable Linear Approximation
von: Liu, Dong, et al.
Veröffentlicht: (2025) -
LLMEasyQuant: Scalable Quantization for Parallel and Distributed LLM Inference
von: Liu, Dong, et al.
Veröffentlicht: (2024) -
MT2ST: Adaptive Multi-Task to Single-Task Learning
von: Liu, Dong, et al.
Veröffentlicht: (2024) -
Efficient Graph Optimization via Distance-Aware Graph Representation Learning
von: Liu, Dong, et al.
Veröffentlicht: (2024) -
FreqCa: Accelerating Diffusion Models via Frequency-Aware Caching
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)