Random Initialization Can't Catch Up: The Advantage of Language Model Transfer for Time Series Forecasting
Fuente:
arXiv
Saved in:
| Main Authors: | Riachi, Roland, Rasul, Kashif, Ashok, Arjun, Humane, Prateek, Roger, Alexis, Williams, Andrew R., Nevmyvaka, Yuriy, Rish, Irina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Small Vocabularies, Big Gains: Pretraining and Tokenization in Time Series Models
by: Roger, Alexis, et al.
Published: (2025)
by: Roger, Alexis, et al.
Published: (2025)
LLM Pretraining Shapes a Generalizable Manifold: Insights into Cross-Modal Transfer to Time Series
by: Roger, Alexis, et al.
Published: (2026)
by: Roger, Alexis, et al.
Published: (2026)
Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting
by: Rasul, Kashif, et al.
Published: (2023)
by: Rasul, Kashif, et al.
Published: (2023)
Structural Knowledge Informed Continual Multivariate Time Series Forecasting
by: Pan, Zijie, et al.
Published: (2024)
by: Pan, Zijie, et al.
Published: (2024)
Deep Generative Sampling in the Dual Divergence Space: A Data-efficient & Interpretative Approach for Generative AI
by: Garg, Sahil, et al.
Published: (2024)
by: Garg, Sahil, et al.
Published: (2024)
Chart-RVR: Reinforcement Learning with Verifiable Rewards for Explainable Chart Reasoning
by: Sinha, Sanchit, et al.
Published: (2025)
by: Sinha, Sanchit, et al.
Published: (2025)
Context is Key: A Benchmark for Forecasting with Essential Textual Information
by: Williams, Andrew Robert, et al.
Published: (2024)
by: Williams, Andrew Robert, et al.
Published: (2024)
Influence Functions for Efficient Data Selection in Reasoning
by: Humane, Prateek, et al.
Published: (2025)
by: Humane, Prateek, et al.
Published: (2025)
TS-RAG: Retrieval-Augmented Generation based Time Series Foundation Models are Stronger Zero-Shot Forecaster
by: Ning, Kanghui, et al.
Published: (2025)
by: Ning, Kanghui, et al.
Published: (2025)
Forecasting with Hyper-Trees
by: März, Alexander, et al.
Published: (2024)
by: März, Alexander, et al.
Published: (2024)
Improving Reasoning for Diffusion Language Models via Group Diffusion Policy Optimization
by: Rojas, Kevin, et al.
Published: (2025)
by: Rojas, Kevin, et al.
Published: (2025)
Towards ethical multimodal systems
by: Roger, Alexis, et al.
Published: (2023)
by: Roger, Alexis, et al.
Published: (2023)
CHIRP: A Fine-Grained Benchmark for Open-Ended Response Evaluation in Vision-Language Models
by: Roger, Alexis, et al.
Published: (2025)
by: Roger, Alexis, et al.
Published: (2025)
Beyond Naïve Prompting: Strategies for Improved Context-aided Forecasting with LLMs
by: Ashok, Arjun, et al.
Published: (2025)
by: Ashok, Arjun, et al.
Published: (2025)
Privacy Amplification by Structured Subsampling for Deep Differentially Private Time Series Forecasting
by: Schuchardt, Jan, et al.
Published: (2025)
by: Schuchardt, Jan, et al.
Published: (2025)
AlphaLab: Autonomous Multi-Agent Research Across Optimization Domains with Frontier LLMs
by: Hogan, Brendan R., et al.
Published: (2026)
by: Hogan, Brendan R., et al.
Published: (2026)
Image Tiling for High-Resolution Reasoning: Balancing Local Detail with Global Context
by: de Margerie, Anatole Jacquin, et al.
Published: (2025)
by: de Margerie, Anatole Jacquin, et al.
Published: (2025)
Spectra 1.1: Scaling Laws and Efficient Inference for Ternary Language Models
by: Vaidhya, Tejas, et al.
Published: (2025)
by: Vaidhya, Tejas, et al.
Published: (2025)
$\textbf{S}^2$IP-LLM: Semantic Space Informed Prompt Learning with LLM for Time Series Forecasting
by: Pan, Zijie, et al.
Published: (2024)
by: Pan, Zijie, et al.
Published: (2024)
Warming Up for Zeroth-Order Federated Pre-Training with Low Resource Clients
by: Legate, Gwen, et al.
Published: (2025)
by: Legate, Gwen, et al.
Published: (2025)
Stochastic Kimura Equations
by: Riachi, Roland, et al.
Published: (2024)
by: Riachi, Roland, et al.
Published: (2024)
Speculative Sampling for Parametric Temporal Point Processes
by: Biloš, Marin, et al.
Published: (2025)
by: Biloš, Marin, et al.
Published: (2025)
Safety Net: Weaving a Web of Resources to Catch What One-Shots Can't
by: Gross, Brooke
Published: (2023)
by: Gross, Brooke
Published: (2023)
Towards Interpretable and Trustworthy Time Series Reasoning: A BlueSky Vision
by: Ning, Kanghui, et al.
Published: (2025)
by: Ning, Kanghui, et al.
Published: (2025)
The Dilemma of Random Parameter Initialization and Barren Plateaus in Variational Quantum Algorithms
by: Kashif, Muhammad, et al.
Published: (2024)
by: Kashif, Muhammad, et al.
Published: (2024)
Graph Partitioning With Limited Moves
by: Behbahani, Majid, et al.
Published: (2024)
by: Behbahani, Majid, et al.
Published: (2024)
A Mechanistic Study of Tabular Foundation Models
by: Biloš, Marin, et al.
Published: (2026)
by: Biloš, Marin, et al.
Published: (2026)
Can’t Touch This
Published: (2024)
Published: (2024)
Catch-Up Mix: Catch-Up Class for Struggling Filters in CNN
by: Kang, Minsoo, et al.
Published: (2024)
by: Kang, Minsoo, et al.
Published: (2024)
Proactive Statistical Process Control Using AI: A Time Series Forecasting Approach for Semiconductor Manufacturing
by: Seeam, Mohammad Iqbal Rasul
Published: (2025)
by: Seeam, Mohammad Iqbal Rasul
Published: (2025)
Empowering Time Series Analysis with Large Language Models: A Survey
by: Jiang, Yushan, et al.
Published: (2024)
by: Jiang, Yushan, et al.
Published: (2024)
Impermanent: A Live Benchmark for Temporal Generalization in Time Series Forecasting
by: Garza, Azul, et al.
Published: (2026)
by: Garza, Azul, et al.
Published: (2026)
Reflexiones sobre la enseñanza de la Química
by: Susana Martinez Riachi
Published: (2007)
by: Susana Martinez Riachi
Published: (2007)
They Can't Hear Us Does Not Mean We Can't Serve Them.
by: McDaniel, Julie Ann
Published: (1992)
by: McDaniel, Julie Ann
Published: (1992)
Databases: Catching Up and Keeping Up.
by: Tenopir, Carol
Published: (1983)
by: Tenopir, Carol
Published: (1983)
Recurrent Interpolants for Probabilistic Time Series Prediction
by: Chen, Yu, et al.
Published: (2024)
by: Chen, Yu, et al.
Published: (2024)
Towards Adversarially Robust Vision-Language Models: Insights from Design Choices and Prompt Formatting Techniques
by: Bhagwatkar, Rishika, et al.
Published: (2024)
by: Bhagwatkar, Rishika, et al.
Published: (2024)
Variational Schrödinger Momentum Diffusion
by: Rojas, Kevin, et al.
Published: (2025)
by: Rojas, Kevin, et al.
Published: (2025)
Technical Report: Full-Stack Fine-Tuning for the Q Programming Language
by: Hogan, Brendan R., et al.
Published: (2025)
by: Hogan, Brendan R., et al.
Published: (2025)
Hiding in Plain Text: Detecting Concealed Jailbreaks via Activation Disentanglement
by: Farzam, Amirhossein, et al.
Published: (2026)
by: Farzam, Amirhossein, et al.
Published: (2026)
Similar Items
-
Small Vocabularies, Big Gains: Pretraining and Tokenization in Time Series Models
by: Roger, Alexis, et al.
Published: (2025) -
LLM Pretraining Shapes a Generalizable Manifold: Insights into Cross-Modal Transfer to Time Series
by: Roger, Alexis, et al.
Published: (2026) -
Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting
by: Rasul, Kashif, et al.
Published: (2023) -
Structural Knowledge Informed Continual Multivariate Time Series Forecasting
by: Pan, Zijie, et al.
Published: (2024) -
Deep Generative Sampling in the Dual Divergence Space: A Data-efficient & Interpretative Approach for Generative AI
by: Garg, Sahil, et al.
Published: (2024)