Value Gradient Guidance for Flow Matching Alignment

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Liu, Zhen, Xiao, Tim Z., Domingo-Enrich, Carles, Liu, Weiyang, Zhang, Dinghuai
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914363909079040
author Liu, Zhen
Xiao, Tim Z.
Domingo-Enrich, Carles
Liu, Weiyang
Zhang, Dinghuai
author_facet Liu, Zhen
Xiao, Tim Z.
Domingo-Enrich, Carles
Liu, Weiyang
Zhang, Dinghuai
contents While methods exist for aligning flow matching models--a popular and effective class of generative models--with human preferences, existing approaches fail to achieve both adaptation efficiency and probabilistically sound prior preservation. In this work, we leverage the theory of optimal control and propose VGG-Flow, a gradient-matching-based method for finetuning pretrained flow matching models. The key idea behind this algorithm is that the optimal difference between the finetuned velocity field and the pretrained one should be matched with the gradient field of a value function. This method not only incorporates first-order information from the reward model but also benefits from heuristic initialization of the value function to enable fast adaptation. Empirically, we show on a popular text-to-image flow matching model, Stable Diffusion 3, that our method can finetune flow matching models under limited computational budgets while achieving effective and prior-preserving alignment.
format Preprint
id arxiv_https___arxiv_org_abs_2512_05116
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Value Gradient Guidance for Flow Matching Alignment
Liu, Zhen
Xiao, Tim Z.
Domingo-Enrich, Carles
Liu, Weiyang
Zhang, Dinghuai
Machine Learning
Computer Vision and Pattern Recognition
While methods exist for aligning flow matching models--a popular and effective class of generative models--with human preferences, existing approaches fail to achieve both adaptation efficiency and probabilistically sound prior preservation. In this work, we leverage the theory of optimal control and propose VGG-Flow, a gradient-matching-based method for finetuning pretrained flow matching models. The key idea behind this algorithm is that the optimal difference between the finetuned velocity field and the pretrained one should be matched with the gradient field of a value function. This method not only incorporates first-order information from the reward model but also benefits from heuristic initialization of the value function to enable fast adaptation. Empirically, we show on a popular text-to-image flow matching model, Stable Diffusion 3, that our method can finetune flow matching models under limited computational budgets while achieving effective and prior-preserving alignment.
title Value Gradient Guidance for Flow Matching Alignment
topic Machine Learning
Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2512.05116