DLM-SWAI: Steering Diffusion Language Models Before They Unmask

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: An, Hyeseon, Han, Yo-Sub
Natura: Preprint
Pubblicazione: 2026
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866914613101068288
author An, Hyeseon
Han, Yo-Sub
author_facet An, Hyeseon
Han, Yo-Sub
contents Steering language model generation toward desired textual properties is essential for practical deployment, and inference-time methods are particularly appealing because they enable controllable generation without retraining. Recent work has also highlighted diffusion language models as an emerging generation paradigm with distinct decoding properties. However, most existing steering approaches either rely on auxiliary models or are designed for autoregressive next-token decoding, making them difficult to apply to diffusion language models DLMs, which generate text through iterative denoising of partially masked sequences. Therefore, we propose DLM-SWAI, a simple training-free steering method that biases the token distribution at each denoising step using pre-computed token-level style scores. Experiments on style and safety control tasks show that DLM-SWAI effectively steers diffusion language models while preserving generation quality and requiring minimal computational overhead. Ablations further reveal a controllable trade-off between steering strength and fluency, and our analysis links class-wise steerability to the strength of token-level attribute cues.
format Preprint
id arxiv_https___arxiv_org_abs_2605_29626
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle DLM-SWAI: Steering Diffusion Language Models Before They Unmask
An, Hyeseon
Han, Yo-Sub
Computation and Language
Artificial Intelligence
Steering language model generation toward desired textual properties is essential for practical deployment, and inference-time methods are particularly appealing because they enable controllable generation without retraining. Recent work has also highlighted diffusion language models as an emerging generation paradigm with distinct decoding properties. However, most existing steering approaches either rely on auxiliary models or are designed for autoregressive next-token decoding, making them difficult to apply to diffusion language models DLMs, which generate text through iterative denoising of partially masked sequences. Therefore, we propose DLM-SWAI, a simple training-free steering method that biases the token distribution at each denoising step using pre-computed token-level style scores. Experiments on style and safety control tasks show that DLM-SWAI effectively steers diffusion language models while preserving generation quality and requiring minimal computational overhead. Ablations further reveal a controllable trade-off between steering strength and fluency, and our analysis links class-wise steerability to the strength of token-level attribute cues.
title DLM-SWAI: Steering Diffusion Language Models Before They Unmask
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2605.29626