Sequences of Logits Reveal the Low Rank Structure of Language Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Golowich, Noah, Liu, Allen, Shetty, Abhishek
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866908618585014272
author Golowich, Noah
Liu, Allen
Shetty, Abhishek
author_facet Golowich, Noah
Liu, Allen
Shetty, Abhishek
contents A major problem in the study of large language models is to understand their inherent low-dimensional structure. We introduce an approach to study the low-dimensional structure of language models at a model-agnostic level: as sequential probabilistic models. We first empirically demonstrate that a wide range of modern language models exhibit low-rank structure: in particular, matrices built from the model's logits for varying sets of prompts and responses have low approximate rank. We then show that this low-rank structure can be leveraged for generation -- in particular, we can generate a response to a target prompt using a linear combination of the model's outputs on unrelated, or even nonsensical prompts. On the theoretical front, we observe that studying the approximate rank of language models in the sense discussed above yields a simple universal abstraction whose theoretical predictions parallel our experiments. We then analyze the representation power of the abstraction and give provable learning guarantees.
format Preprint
id arxiv_https___arxiv_org_abs_2510_24966
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Sequences of Logits Reveal the Low Rank Structure of Language Models
Golowich, Noah
Liu, Allen
Shetty, Abhishek
Machine Learning
Artificial Intelligence
Computation and Language
A major problem in the study of large language models is to understand their inherent low-dimensional structure. We introduce an approach to study the low-dimensional structure of language models at a model-agnostic level: as sequential probabilistic models. We first empirically demonstrate that a wide range of modern language models exhibit low-rank structure: in particular, matrices built from the model's logits for varying sets of prompts and responses have low approximate rank. We then show that this low-rank structure can be leveraged for generation -- in particular, we can generate a response to a target prompt using a linear combination of the model's outputs on unrelated, or even nonsensical prompts. On the theoretical front, we observe that studying the approximate rank of language models in the sense discussed above yields a simple universal abstraction whose theoretical predictions parallel our experiments. We then analyze the representation power of the abstraction and give provable learning guarantees.
title Sequences of Logits Reveal the Low Rank Structure of Language Models
topic Machine Learning
Artificial Intelligence
Computation and Language
url https://arxiv.org/abs/2510.24966