Saved in:
Bibliographic Details
Main Authors: Kim, Youngjoong, Kim, Duhoe, Kim, Woosung, Park, Jaesik
Format: Preprint
Published: 2026
Subjects:
Online Access:https://arxiv.org/abs/2601.22679
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866908799732809728
author Kim, Youngjoong
Kim, Duhoe
Kim, Woosung
Park, Jaesik
author_facet Kim, Youngjoong
Kim, Duhoe
Kim, Woosung
Park, Jaesik
contents Consistency models have been proposed for fast generative modeling, achieving results competitive with diffusion and flow models. However, these methods exhibit inherent instability and limited reproducibility when training from scratch, motivating subsequent work to explain and stabilize these issues. While these efforts have provided valuable insights, the explanations remain fragmented, and the theoretical relationships remain unclear. In this work, we provide a theoretical examination of consistency models by analyzing them from a flow map-based perspective. This joint analysis clarifies how training stability and convergence behavior can give rise to degenerate solutions. Building on these insights, we revisit self-distillation as a practical remedy for certain forms of suboptimal convergence and reformulate it to avoid excessive gradient norms for stable optimization. We further demonstrate that our strategy extends beyond image generation to diffusion-based policy learning, without reliance on a pretrained diffusion model for initialization, thereby illustrating its broader applicability.
format Preprint
id arxiv_https___arxiv_org_abs_2601_22679
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Stabilizing Consistency Training: A Flow Map Analysis and Self-Distillation
Kim, Youngjoong
Kim, Duhoe
Kim, Woosung
Park, Jaesik
Machine Learning
Computer Vision and Pattern Recognition
Consistency models have been proposed for fast generative modeling, achieving results competitive with diffusion and flow models. However, these methods exhibit inherent instability and limited reproducibility when training from scratch, motivating subsequent work to explain and stabilize these issues. While these efforts have provided valuable insights, the explanations remain fragmented, and the theoretical relationships remain unclear. In this work, we provide a theoretical examination of consistency models by analyzing them from a flow map-based perspective. This joint analysis clarifies how training stability and convergence behavior can give rise to degenerate solutions. Building on these insights, we revisit self-distillation as a practical remedy for certain forms of suboptimal convergence and reformulate it to avoid excessive gradient norms for stable optimization. We further demonstrate that our strategy extends beyond image generation to diffusion-based policy learning, without reliance on a pretrained diffusion model for initialization, thereby illustrating its broader applicability.
title Stabilizing Consistency Training: A Flow Map Analysis and Self-Distillation
topic Machine Learning
Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2601.22679