AB-Cache: Training-Free Acceleration of Diffusion Models via Adams-Bashforth Cached Feature Reuse

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Yu, Zichao, Zou, Zhen, Shao, Guojiang, Zhang, Chengwei, Xu, Shengze, Huang, Jie, Zhao, Feng, Cun, Xiaodong, Zhang, Wenyi
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917985749303296
author Yu, Zichao
Zou, Zhen
Shao, Guojiang
Zhang, Chengwei
Xu, Shengze
Huang, Jie
Zhao, Feng
Cun, Xiaodong
Zhang, Wenyi
author_facet Yu, Zichao
Zou, Zhen
Shao, Guojiang
Zhang, Chengwei
Xu, Shengze
Huang, Jie
Zhao, Feng
Cun, Xiaodong
Zhang, Wenyi
contents Diffusion models have demonstrated remarkable success in generative tasks, yet their iterative denoising process results in slow inference, limiting their practicality. While existing acceleration methods exploit the well-known U-shaped similarity pattern between adjacent steps through caching mechanisms, they lack theoretical foundation and rely on simplistic computation reuse, often leading to performance degradation. In this work, we provide a theoretical understanding by analyzing the denoising process through the second-order Adams-Bashforth method, revealing a linear relationship between the outputs of consecutive steps. This analysis explains why the outputs of adjacent steps exhibit a U-shaped pattern. Furthermore, extending Adams-Bashforth method to higher order, we propose a novel caching-based acceleration approach for diffusion models, instead of directly reusing cached results, with a truncation error bound of only \(O(h^k)\) where $h$ is the step size. Extensive validation across diverse image and video diffusion models (including HunyuanVideo and FLUX.1-dev) with various schedulers demonstrates our method's effectiveness in achieving nearly $3\times$ speedup while maintaining original performance levels, offering a practical real-time solution without compromising generation quality.
format Preprint
id arxiv_https___arxiv_org_abs_2504_10540
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle AB-Cache: Training-Free Acceleration of Diffusion Models via Adams-Bashforth Cached Feature Reuse
Yu, Zichao
Zou, Zhen
Shao, Guojiang
Zhang, Chengwei
Xu, Shengze
Huang, Jie
Zhao, Feng
Cun, Xiaodong
Zhang, Wenyi
Machine Learning
Artificial Intelligence
Diffusion models have demonstrated remarkable success in generative tasks, yet their iterative denoising process results in slow inference, limiting their practicality. While existing acceleration methods exploit the well-known U-shaped similarity pattern between adjacent steps through caching mechanisms, they lack theoretical foundation and rely on simplistic computation reuse, often leading to performance degradation. In this work, we provide a theoretical understanding by analyzing the denoising process through the second-order Adams-Bashforth method, revealing a linear relationship between the outputs of consecutive steps. This analysis explains why the outputs of adjacent steps exhibit a U-shaped pattern. Furthermore, extending Adams-Bashforth method to higher order, we propose a novel caching-based acceleration approach for diffusion models, instead of directly reusing cached results, with a truncation error bound of only \(O(h^k)\) where $h$ is the step size. Extensive validation across diverse image and video diffusion models (including HunyuanVideo and FLUX.1-dev) with various schedulers demonstrates our method's effectiveness in achieving nearly $3\times$ speedup while maintaining original performance levels, offering a practical real-time solution without compromising generation quality.
title AB-Cache: Training-Free Acceleration of Diffusion Models via Adams-Bashforth Cached Feature Reuse
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2504.10540