Latte: Transfering LLMs` Latent-level Knowledge for Few-shot Tabular Learning

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Shi, Ruxue, Gu, Hengrui, Ye, Hangting, Dai, Yiwei, Shen, Xu, Wang, Xin
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866910932922269696
author Shi, Ruxue
Gu, Hengrui
Ye, Hangting
Dai, Yiwei
Shen, Xu
Wang, Xin
author_facet Shi, Ruxue
Gu, Hengrui
Ye, Hangting
Dai, Yiwei
Shen, Xu
Wang, Xin
contents Few-shot tabular learning, in which machine learning models are trained with a limited amount of labeled data, provides a cost-effective approach to addressing real-world challenges. The advent of Large Language Models (LLMs) has sparked interest in leveraging their pre-trained knowledge for few-shot tabular learning. Despite promising results, existing approaches either rely on test-time knowledge extraction, which introduces undesirable latency, or text-level knowledge, which leads to unreliable feature engineering. To overcome these limitations, we propose Latte, a training-time knowledge extraction framework that transfers the latent prior knowledge within LLMs to optimize a more generalized downstream model. Latte enables general knowledge-guided downstream tabular learning, facilitating the weighted fusion of information across different feature values while reducing the risk of overfitting to limited labeled data. Furthermore, Latte is compatible with existing unsupervised pre-training paradigms and effectively utilizes available unlabeled samples to overcome the performance limitations imposed by an extremely small labeled dataset. Extensive experiments on various few-shot tabular learning benchmarks demonstrate the superior performance of Latte, establishing it as a state-of-the-art approach in this domain
format Preprint
id arxiv_https___arxiv_org_abs_2505_05237
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Latte: Transfering LLMs` Latent-level Knowledge for Few-shot Tabular Learning
Shi, Ruxue
Gu, Hengrui
Ye, Hangting
Dai, Yiwei
Shen, Xu
Wang, Xin
Machine Learning
Few-shot tabular learning, in which machine learning models are trained with a limited amount of labeled data, provides a cost-effective approach to addressing real-world challenges. The advent of Large Language Models (LLMs) has sparked interest in leveraging their pre-trained knowledge for few-shot tabular learning. Despite promising results, existing approaches either rely on test-time knowledge extraction, which introduces undesirable latency, or text-level knowledge, which leads to unreliable feature engineering. To overcome these limitations, we propose Latte, a training-time knowledge extraction framework that transfers the latent prior knowledge within LLMs to optimize a more generalized downstream model. Latte enables general knowledge-guided downstream tabular learning, facilitating the weighted fusion of information across different feature values while reducing the risk of overfitting to limited labeled data. Furthermore, Latte is compatible with existing unsupervised pre-training paradigms and effectively utilizes available unlabeled samples to overcome the performance limitations imposed by an extremely small labeled dataset. Extensive experiments on various few-shot tabular learning benchmarks demonstrate the superior performance of Latte, establishing it as a state-of-the-art approach in this domain
title Latte: Transfering LLMs` Latent-level Knowledge for Few-shot Tabular Learning
topic Machine Learning
url https://arxiv.org/abs/2505.05237