Internformat: :: Library Catalog

Gespeichert in:

Bibliographische Detailangaben
Hauptverfasser:	Emer, Lorenzo, Lippi, Marco, Mina, Andrea, Vandin, Andrea
Format:	Preprint
Veröffentlicht:	2026
Schlagworte:	Computational Engineering, Finance, and Science 68T07 (Primary) I.2.7; I.5.1; H.3.1
Online-Zugang:	https://arxiv.org/abs/2601.23200
Tags:	Tag hinzufügen Keine Tags, Fügen Sie den ersten Tag hinzu!

_version_	1866918315995168768
author	Emer, Lorenzo Lippi, Marco Mina, Andrea Vandin, Andrea
author_facet	Emer, Lorenzo Lippi, Marco Mina, Andrea Vandin, Andrea
contents	Patent classification into CPC codes underpins large scale analyses of technological change but remains challenging due to its hierarchical, multi label, and highly imbalanced structure. While pre Generative AI supervised encoder based models became the de facto standard for large scale patent classification, recent advances in large language models (LLMs) raise questions about whether they can provide complementary capabilities, particularly for rare or weakly represented technological categories. In this work, we perform a systematic comparison of encoder based classifiers (BERT, SciBERT, and PatentSBERTa) and open weight LLMs on a highly imbalanced benchmark dataset (USPTO 70k). We evaluate LLMs under zero shot, few shot, and retrieval augmented prompting, and further assess parameter efficient fine tuning of the best performing model. Our results show that encoder based models achieve higher aggregate performance, driven by strong results on frequent CPC subclasses, but struggle on rare ones. In contrast, LLMs achieve relatively higher performance on infrequent subclasses, often associated with early stage, cross domain, or weakly institutionalised technologies, particularly at higher hierarchical levels. These findings indicate that encoder based and LLM based approaches play complementary roles in patent classification. We additionally quantify inference time and energy consumption, showing that encoder based models are up to three orders of magnitude more efficient than LLMs. Overall, our results inform responsible patentometrics and technology mapping, and motivate hybrid classification approaches that combine encoder efficiency with the long tail coverage of LLMs under computational and environmental constraints.
format	Preprint
id	arxiv_https___arxiv_org_abs_2601_23200
institution	arXiv
publishDate	2026
record_format	arxiv
spellingShingle	Large Language Models for Patent Classification: Strengths, Trade-offs, and the Long Tail Effect Emer, Lorenzo Lippi, Marco Mina, Andrea Vandin, Andrea Computational Engineering, Finance, and Science 68T07 (Primary) I.2.7; I.5.1; H.3.1 Patent classification into CPC codes underpins large scale analyses of technological change but remains challenging due to its hierarchical, multi label, and highly imbalanced structure. While pre Generative AI supervised encoder based models became the de facto standard for large scale patent classification, recent advances in large language models (LLMs) raise questions about whether they can provide complementary capabilities, particularly for rare or weakly represented technological categories. In this work, we perform a systematic comparison of encoder based classifiers (BERT, SciBERT, and PatentSBERTa) and open weight LLMs on a highly imbalanced benchmark dataset (USPTO 70k). We evaluate LLMs under zero shot, few shot, and retrieval augmented prompting, and further assess parameter efficient fine tuning of the best performing model. Our results show that encoder based models achieve higher aggregate performance, driven by strong results on frequent CPC subclasses, but struggle on rare ones. In contrast, LLMs achieve relatively higher performance on infrequent subclasses, often associated with early stage, cross domain, or weakly institutionalised technologies, particularly at higher hierarchical levels. These findings indicate that encoder based and LLM based approaches play complementary roles in patent classification. We additionally quantify inference time and energy consumption, showing that encoder based models are up to three orders of magnitude more efficient than LLMs. Overall, our results inform responsible patentometrics and technology mapping, and motivate hybrid classification approaches that combine encoder efficiency with the long tail coverage of LLMs under computational and environmental constraints.
title	Large Language Models for Patent Classification: Strengths, Trade-offs, and the Long Tail Effect
topic	Computational Engineering, Finance, and Science 68T07 (Primary) I.2.7; I.5.1; H.3.1
url	https://arxiv.org/abs/2601.23200

Ähnliche Einträge