ConsistentChat: Building Skeleton-Guided Consistent Multi-Turn Dialogues for Large Language Models from Scratch

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Chen, Jiawei, Guan, Xinyan, Yuan, Qianhao, Mo, Guozhao, Zhou, Weixiang, Lu, Yaojie, Lin, Hongyu, He, Ben, Sun, Le, Han, Xianpei
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866912603925643264
author Chen, Jiawei
Guan, Xinyan
Yuan, Qianhao
Mo, Guozhao
Zhou, Weixiang
Lu, Yaojie
Lin, Hongyu
He, Ben
Sun, Le
Han, Xianpei
author_facet Chen, Jiawei
Guan, Xinyan
Yuan, Qianhao
Mo, Guozhao
Zhou, Weixiang
Lu, Yaojie
Lin, Hongyu
He, Ben
Sun, Le
Han, Xianpei
contents Current instruction data synthesis methods primarily focus on single-turn instructions and often neglect cross-turn coherence, resulting in context drift and reduced task completion rates in extended conversations. To address this limitation, we propose Skeleton-Guided Multi-Turn Dialogue Generation, a framework that constrains multi-turn instruction synthesis by explicitly modeling human conversational intent. It operates in two stages: (1) Intent Modeling, which captures the global structure of human dialogues by assigning each conversation to one of nine well-defined intent trajectories, ensuring a coherent and goal-oriented information flow; and (2) Skeleton Generation, which constructs a structurally grounded sequence of user queries aligned with the modeled intent, thereby serving as a scaffold that constrains and guides the downstream instruction synthesis process. Based on this process, we construct ConsistentChat, a multi-turn instruction dataset with approximately 15,000 multi-turn conversations and 224,392 utterances. Experiments on the Light, Topdial, and MT-Eval benchmarks show that models fine-tuned on ConsistentChat achieve a 20-30% improvement in chat consistency and up to a 15% increase in task success rate, significantly outperforming models trained on existing single-turn and multi-turn instruction datasets.
format Preprint
id arxiv_https___arxiv_org_abs_2506_03558
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle ConsistentChat: Building Skeleton-Guided Consistent Multi-Turn Dialogues for Large Language Models from Scratch
Chen, Jiawei
Guan, Xinyan
Yuan, Qianhao
Mo, Guozhao
Zhou, Weixiang
Lu, Yaojie
Lin, Hongyu
He, Ben
Sun, Le
Han, Xianpei
Computation and Language
Current instruction data synthesis methods primarily focus on single-turn instructions and often neglect cross-turn coherence, resulting in context drift and reduced task completion rates in extended conversations. To address this limitation, we propose Skeleton-Guided Multi-Turn Dialogue Generation, a framework that constrains multi-turn instruction synthesis by explicitly modeling human conversational intent. It operates in two stages: (1) Intent Modeling, which captures the global structure of human dialogues by assigning each conversation to one of nine well-defined intent trajectories, ensuring a coherent and goal-oriented information flow; and (2) Skeleton Generation, which constructs a structurally grounded sequence of user queries aligned with the modeled intent, thereby serving as a scaffold that constrains and guides the downstream instruction synthesis process. Based on this process, we construct ConsistentChat, a multi-turn instruction dataset with approximately 15,000 multi-turn conversations and 224,392 utterances. Experiments on the Light, Topdial, and MT-Eval benchmarks show that models fine-tuned on ConsistentChat achieve a 20-30% improvement in chat consistency and up to a 15% increase in task success rate, significantly outperforming models trained on existing single-turn and multi-turn instruction datasets.
title ConsistentChat: Building Skeleton-Guided Consistent Multi-Turn Dialogues for Large Language Models from Scratch
topic Computation and Language
url https://arxiv.org/abs/2506.03558