Skills-in-Context Prompting: Unlocking Compositionality in Large Language Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Chen, Jiaao, Pan, Xiaoman, Yu, Dian, Song, Kaiqiang, Wang, Xiaoyang, Yu, Dong, Chen, Jianshu
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911957526773760
author Chen, Jiaao
Pan, Xiaoman
Yu, Dian
Song, Kaiqiang
Wang, Xiaoyang
Yu, Dong
Chen, Jianshu
author_facet Chen, Jiaao
Pan, Xiaoman
Yu, Dian
Song, Kaiqiang
Wang, Xiaoyang
Yu, Dong
Chen, Jianshu
contents We investigate how to elicit compositional generalization capabilities in large language models (LLMs). Compositional generalization empowers LLMs to solve complex problems by combining foundational skills, a critical reasoning ability akin to human intelligence. However, even the most advanced LLMs currently struggle with this form of reasoning. We examine this problem within the framework of in-context learning and find that demonstrating both foundational skills and compositional examples grounded in these skills within the same prompt context is crucial. We refer to this prompt structure as skills-in-context (SKiC). With as few as two exemplars, this in-context learning structure enables LLMs to tackle more challenging problems requiring innovative skill combinations, achieving near-perfect systematic generalization across a broad range of tasks. Intriguingly, SKiC also unlocks the latent potential of LLMs, allowing them to more actively utilize pre-existing internal skills acquired during earlier pretraining stages to solve complex reasoning problems. The SKiC structure is robust across different skill constructions and exemplar choices and demonstrates strong transferability to new tasks. Finally, inspired by our in-context learning study, we show that fine-tuning LLMs with SKiC-style data can elicit zero-shot weak-to-strong generalization, enabling the models to solve much harder problems directly with standard prompting.
format Preprint
id arxiv_https___arxiv_org_abs_2308_00304
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Skills-in-Context Prompting: Unlocking Compositionality in Large Language Models
Chen, Jiaao
Pan, Xiaoman
Yu, Dian
Song, Kaiqiang
Wang, Xiaoyang
Yu, Dong
Chen, Jianshu
Computation and Language
We investigate how to elicit compositional generalization capabilities in large language models (LLMs). Compositional generalization empowers LLMs to solve complex problems by combining foundational skills, a critical reasoning ability akin to human intelligence. However, even the most advanced LLMs currently struggle with this form of reasoning. We examine this problem within the framework of in-context learning and find that demonstrating both foundational skills and compositional examples grounded in these skills within the same prompt context is crucial. We refer to this prompt structure as skills-in-context (SKiC). With as few as two exemplars, this in-context learning structure enables LLMs to tackle more challenging problems requiring innovative skill combinations, achieving near-perfect systematic generalization across a broad range of tasks. Intriguingly, SKiC also unlocks the latent potential of LLMs, allowing them to more actively utilize pre-existing internal skills acquired during earlier pretraining stages to solve complex reasoning problems. The SKiC structure is robust across different skill constructions and exemplar choices and demonstrates strong transferability to new tasks. Finally, inspired by our in-context learning study, we show that fine-tuning LLMs with SKiC-style data can elicit zero-shot weak-to-strong generalization, enabling the models to solve much harder problems directly with standard prompting.
title Skills-in-Context Prompting: Unlocking Compositionality in Large Language Models
topic Computation and Language
url https://arxiv.org/abs/2308.00304