Rethinking the Role of LLMs in Time Series Forecasting

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Qiu, Xin, Tong, Junlong, Sun, Yirong, Ma, Yunpu, Zhang, Wei, Shen, Xiaoyu
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910038689316864
author Qiu, Xin
Tong, Junlong
Sun, Yirong
Ma, Yunpu
Zhang, Wei
Shen, Xiaoyu
author_facet Qiu, Xin
Tong, Junlong
Sun, Yirong
Ma, Yunpu
Zhang, Wei
Shen, Xiaoyu
contents Large language models (LLMs) have been introduced to time series forecasting (TSF) to incorporate contextual knowledge beyond numerical signals. However, existing studies question whether LLMs provide genuine benefits, often reporting comparable performance without LLMs. We show that such conclusions stem from limited evaluation settings and do not hold at scale. We conduct a large-scale study of LLM-based TSF (LLM4TSF) across 8 billion observations, 17 forecasting scenarios, 4 horizons, multiple alignment strategies, and both in-domain and out-of-domain settings. Our results demonstrate that \emph{LLM4TS indeed improves forecasting performance}, with especially large gains in cross-domain generalization. Pre-alignment outperforming post-alignment in over 90\% of tasks. Both pretrained knowledge and model architecture of LLMs contribute and play complementary roles: pretraining is critical under distribution shifts, while architecture excels at modeling complex temporal dynamics. Moreover, under large-scale mixed distributions, a fully intact LLM becomes indispensable, as confirmed by token-level routing analysis and prompt-based improvements. Overall, Our findings overturn prior negative assessments, establish clear conditions under which LLMs are not only useful, and provide practical guidance for effective model design. We release our code at https://github.com/EIT-NLP/LLM4TSF.
format Preprint
id arxiv_https___arxiv_org_abs_2602_14744
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Rethinking the Role of LLMs in Time Series Forecasting
Qiu, Xin
Tong, Junlong
Sun, Yirong
Ma, Yunpu
Zhang, Wei
Shen, Xiaoyu
Computation and Language
Large language models (LLMs) have been introduced to time series forecasting (TSF) to incorporate contextual knowledge beyond numerical signals. However, existing studies question whether LLMs provide genuine benefits, often reporting comparable performance without LLMs. We show that such conclusions stem from limited evaluation settings and do not hold at scale. We conduct a large-scale study of LLM-based TSF (LLM4TSF) across 8 billion observations, 17 forecasting scenarios, 4 horizons, multiple alignment strategies, and both in-domain and out-of-domain settings. Our results demonstrate that \emph{LLM4TS indeed improves forecasting performance}, with especially large gains in cross-domain generalization. Pre-alignment outperforming post-alignment in over 90\% of tasks. Both pretrained knowledge and model architecture of LLMs contribute and play complementary roles: pretraining is critical under distribution shifts, while architecture excels at modeling complex temporal dynamics. Moreover, under large-scale mixed distributions, a fully intact LLM becomes indispensable, as confirmed by token-level routing analysis and prompt-based improvements. Overall, Our findings overturn prior negative assessments, establish clear conditions under which LLMs are not only useful, and provide practical guidance for effective model design. We release our code at https://github.com/EIT-NLP/LLM4TSF.
title Rethinking the Role of LLMs in Time Series Forecasting
topic Computation and Language
url https://arxiv.org/abs/2602.14744