From Compression to Expression: A Layerwise Analysis of In-Context Learning

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Jiang, Jiachen, Dong, Yuxin, Zhou, Jinxin, Zhu, Zhihui
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866915533242236928
author Jiang, Jiachen
Dong, Yuxin
Zhou, Jinxin
Zhu, Zhihui
author_facet Jiang, Jiachen
Dong, Yuxin
Zhou, Jinxin
Zhu, Zhihui
contents In-context learning (ICL) enables large language models (LLMs) to adapt to new tasks without weight updates by learning from demonstration sequences. While ICL shows strong empirical performance, its internal representational mechanisms are not yet well understood. In this work, we conduct a statistical geometric analysis of ICL representations to investigate how task-specific information is captured across layers. Our analysis reveals an intriguing phenomenon, which we term *Layerwise Compression-Expression*: early layers progressively produce compact and discriminative representations that encode task information from the input demonstrations, while later layers express these representations to incorporate the query and generate the prediction. This phenomenon is observed consistently across diverse tasks and a range of contemporary LLM architectures. We demonstrate that it has important implications for ICL performance -- improving with model size and the number of demonstrations -- and for robustness in the presence of noisy examples. To further understand the effect of the compact task representation, we propose a bias-variance decomposition and provide a theoretical analysis showing how attention mechanisms contribute to reducing both variance and bias, thereby enhancing performance as the number of demonstrations increases. Our findings reveal an intriguing layerwise dynamic in ICL, highlight how structured representations emerge within LLMs, and showcase that analyzing internal representations can facilitate a deeper understanding of model behavior.
format Preprint
id arxiv_https___arxiv_org_abs_2505_17322
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle From Compression to Expression: A Layerwise Analysis of In-Context Learning
Jiang, Jiachen
Dong, Yuxin
Zhou, Jinxin
Zhu, Zhihui
Computation and Language
Artificial Intelligence
Machine Learning
In-context learning (ICL) enables large language models (LLMs) to adapt to new tasks without weight updates by learning from demonstration sequences. While ICL shows strong empirical performance, its internal representational mechanisms are not yet well understood. In this work, we conduct a statistical geometric analysis of ICL representations to investigate how task-specific information is captured across layers. Our analysis reveals an intriguing phenomenon, which we term *Layerwise Compression-Expression*: early layers progressively produce compact and discriminative representations that encode task information from the input demonstrations, while later layers express these representations to incorporate the query and generate the prediction. This phenomenon is observed consistently across diverse tasks and a range of contemporary LLM architectures. We demonstrate that it has important implications for ICL performance -- improving with model size and the number of demonstrations -- and for robustness in the presence of noisy examples. To further understand the effect of the compact task representation, we propose a bias-variance decomposition and provide a theoretical analysis showing how attention mechanisms contribute to reducing both variance and bias, thereby enhancing performance as the number of demonstrations increases. Our findings reveal an intriguing layerwise dynamic in ICL, highlight how structured representations emerge within LLMs, and showcase that analyzing internal representations can facilitate a deeper understanding of model behavior.
title From Compression to Expression: A Layerwise Analysis of In-Context Learning
topic Computation and Language
Artificial Intelligence
Machine Learning
url https://arxiv.org/abs/2505.17322