Instruction Tuning with and without Context: Behavioral Shifts and Downstream Impact

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Lee, Hyunji, Yoon, Seunghyun, Won, Yunjae, Oh, Hanseok, Kim, Geewook, Bui, Trung, Dernoncourt, Franck, Stengel-Eskin, Elias, Bansal, Mohit, Seo, Minjoon
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914240677281792
author Lee, Hyunji
Yoon, Seunghyun
Won, Yunjae
Oh, Hanseok
Kim, Geewook
Bui, Trung
Dernoncourt, Franck
Stengel-Eskin, Elias
Bansal, Mohit
Seo, Minjoon
author_facet Lee, Hyunji
Yoon, Seunghyun
Won, Yunjae
Oh, Hanseok
Kim, Geewook
Bui, Trung
Dernoncourt, Franck
Stengel-Eskin, Elias
Bansal, Mohit
Seo, Minjoon
contents Instruction tuning is a widely used approach to improve the instruction-following ability of large language models (LLMs). Instruction-tuning datasets typically include a mixture of context-augmented and context-free examples, yet prior work has largely combined these data types without examining their distinct effects. In this paper, we investigate how training LLMs with or without context affects model behavior and downstream performance. First, in the text domain, we show that LLMs trained with context attend more strongly to the provided knowledge, achieving better grounding. We also observe that context-augmented training shifts how LLMs use knowledge: models store and leverage less on parametric knowledge and instead depend more on the provided context. Second, we observe that using LLM trained with context-augmented data as the backbone for vision-language models reduces hallucination and improves grounding in the visual domain. Finally, we explore practical strategies for real-world deployments where context availability varies. We show that maintaining separate context-augmented and context-free models and routing inputs between them yields more robust overall performance than training a single mixed model, as it better preserves their complementary strengths.
format Preprint
id arxiv_https___arxiv_org_abs_2506_15480
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Instruction Tuning with and without Context: Behavioral Shifts and Downstream Impact
Lee, Hyunji
Yoon, Seunghyun
Won, Yunjae
Oh, Hanseok
Kim, Geewook
Bui, Trung
Dernoncourt, Franck
Stengel-Eskin, Elias
Bansal, Mohit
Seo, Minjoon
Computation and Language
Artificial Intelligence
Instruction tuning is a widely used approach to improve the instruction-following ability of large language models (LLMs). Instruction-tuning datasets typically include a mixture of context-augmented and context-free examples, yet prior work has largely combined these data types without examining their distinct effects. In this paper, we investigate how training LLMs with or without context affects model behavior and downstream performance. First, in the text domain, we show that LLMs trained with context attend more strongly to the provided knowledge, achieving better grounding. We also observe that context-augmented training shifts how LLMs use knowledge: models store and leverage less on parametric knowledge and instead depend more on the provided context. Second, we observe that using LLM trained with context-augmented data as the backbone for vision-language models reduces hallucination and improves grounding in the visual domain. Finally, we explore practical strategies for real-world deployments where context availability varies. We show that maintaining separate context-augmented and context-free models and routing inputs between them yields more robust overall performance than training a single mixed model, as it better preserves their complementary strengths.
title Instruction Tuning with and without Context: Behavioral Shifts and Downstream Impact
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2506.15480