Small Language Models are Good Too: An Empirical Study of Zero-Shot Classification
Fuente:
arXiv
Salvato in:
| Autori principali: | Lepagnol, Pierre, Gerald, Thomas, Ghannay, Sahar, Servan, Christophe, Rosset, Sophie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Leveraging Information Retrieval to Enhance Spoken Language Understanding Prompts in Few-Shot Learning
di: Lepagnol, Pierre, et al.
Pubblicazione: (2025)
di: Lepagnol, Pierre, et al.
Pubblicazione: (2025)
mALBERT: Is a Compact Multilingual BERT Model Still Worth It?
di: Servan, Christophe, et al.
Pubblicazione: (2024)
di: Servan, Christophe, et al.
Pubblicazione: (2024)
New Semantic Task for the French Spoken Language Understanding MEDIA Benchmark
di: Alavoine, Nadège, et al.
Pubblicazione: (2024)
di: Alavoine, Nadège, et al.
Pubblicazione: (2024)
LLM-based Atomic Propositions help weak extractors: Evaluation of a Propositioner for triplet extraction
di: Pommeret, Luc, et al.
Pubblicazione: (2026)
di: Pommeret, Luc, et al.
Pubblicazione: (2026)
Enhancing Small Language Models for Cross-Lingual Generalized Zero-Shot Classification with Soft Prompt Tuning
di: Philippy, Fred, et al.
Pubblicazione: (2025)
di: Philippy, Fred, et al.
Pubblicazione: (2025)
Performance of Small Language Model Pretraining on FABRIC: An Empirical Study
di: Rao, Praveen
Pubblicazione: (2026)
di: Rao, Praveen
Pubblicazione: (2026)
Zero-Shot Classification of Crisis Tweets Using Instruction-Finetuned Large Language Models
di: McDaniel, Emma, et al.
Pubblicazione: (2024)
di: McDaniel, Emma, et al.
Pubblicazione: (2024)
Small Models, Big Tasks: An Exploratory Empirical Study on Small Language Models for Function Calling
di: Kavathekar, Ishan, et al.
Pubblicazione: (2025)
di: Kavathekar, Ishan, et al.
Pubblicazione: (2025)
Enabling Small Models for Zero-Shot Selection and Reuse through Model Label Learning
di: Zhang, Jia, et al.
Pubblicazione: (2024)
di: Zhang, Jia, et al.
Pubblicazione: (2024)
AnyMatch -- Efficient Zero-Shot Entity Matching with a Small Language Model
di: Zhang, Zeyu, et al.
Pubblicazione: (2024)
di: Zhang, Zeyu, et al.
Pubblicazione: (2024)
Zero-Shot Goal Recognition with Large Language Models
di: Gusmão, Kin Max Piamolini, et al.
Pubblicazione: (2026)
di: Gusmão, Kin Max Piamolini, et al.
Pubblicazione: (2026)
LLM meets Vision-Language Models for Zero-Shot One-Class Classification
di: Bendou, Yassir, et al.
Pubblicazione: (2024)
di: Bendou, Yassir, et al.
Pubblicazione: (2024)
An Empirical Study of SFT-DPO Interaction and Parameterization in Small Language Models
di: Feng, Yuming, et al.
Pubblicazione: (2026)
di: Feng, Yuming, et al.
Pubblicazione: (2026)
A Benchmark Evaluation of Clinical Named Entity Recognition in French
di: Bannour, Nesrine, et al.
Pubblicazione: (2024)
di: Bannour, Nesrine, et al.
Pubblicazione: (2024)
Zero-Shot Robustification of Zero-Shot Models
di: Adila, Dyah, et al.
Pubblicazione: (2023)
di: Adila, Dyah, et al.
Pubblicazione: (2023)
Too Good to be Bad: On the Failure of LLMs to Role-Play Villains
di: Yi, Zihao, et al.
Pubblicazione: (2025)
di: Yi, Zihao, et al.
Pubblicazione: (2025)
Transductive Zero-Shot and Few-Shot CLIP
di: Martin, Ségolène, et al.
Pubblicazione: (2024)
di: Martin, Ségolène, et al.
Pubblicazione: (2024)
Breaking the Myth: Can Small Models Infer Postconditions Too?
di: Zhang, Gehao, et al.
Pubblicazione: (2025)
di: Zhang, Gehao, et al.
Pubblicazione: (2025)
Large Language Models as Universal Predictors? An Empirical Study on Small Tabular Datasets
di: Pavlidis, Nikolaos, et al.
Pubblicazione: (2025)
di: Pavlidis, Nikolaos, et al.
Pubblicazione: (2025)
Evaluating Zero-Shot and One-Shot Adaptation of Small Language Models in Leader-Follower Interaction
di: Baptista, Rafael R., et al.
Pubblicazione: (2026)
di: Baptista, Rafael R., et al.
Pubblicazione: (2026)
Small or Large? Zero-Shot or Finetuned? Guiding Language Model Choice for Specialized Applications in Healthcare
di: Gondara, Lovedeep, et al.
Pubblicazione: (2025)
di: Gondara, Lovedeep, et al.
Pubblicazione: (2025)
Benchmarking Small Language Models and Small Reasoning Language Models on System Log Severity Classification
di: Masri, Yahya, et al.
Pubblicazione: (2026)
di: Masri, Yahya, et al.
Pubblicazione: (2026)
Fine-Tuned 'Small' LLMs (Still) Significantly Outperform Zero-Shot Generative AI Models in Text Classification
di: Bucher, Martin Juan José, et al.
Pubblicazione: (2024)
di: Bucher, Martin Juan José, et al.
Pubblicazione: (2024)
Feasibility with Language Models for Open-World Compositional Zero-Shot Learning
di: Kim, Jae Myung, et al.
Pubblicazione: (2025)
di: Kim, Jae Myung, et al.
Pubblicazione: (2025)
Are Video Models Ready as Zero-Shot Reasoners? An Empirical Study with the MME-CoF Benchmark
di: Guo, Ziyu, et al.
Pubblicazione: (2025)
di: Guo, Ziyu, et al.
Pubblicazione: (2025)
Zero-Shot Spam Email Classification Using Pre-trained Large Language Models
di: Rojas-Galeano, Sergio
Pubblicazione: (2024)
di: Rojas-Galeano, Sergio
Pubblicazione: (2024)
Too Good to be True? Turn Any Model Differentially Private With DP-Weights
di: Zagardo, David
Pubblicazione: (2024)
di: Zagardo, David
Pubblicazione: (2024)
Evaluating Vision-Language Models for Zero-Shot Detection, Classification, and Association of Motorcycles, Passengers, and Helmets
di: Choi, Lucas, et al.
Pubblicazione: (2024)
di: Choi, Lucas, et al.
Pubblicazione: (2024)
On the Applicability of Zero-Shot Cross-Lingual Transfer Learning for Sentiment Classification in Distant Language Pairs
di: Rusli, Andre, et al.
Pubblicazione: (2024)
di: Rusli, Andre, et al.
Pubblicazione: (2024)
Data Generation Using Large Language Models for Text Classification: An Empirical Case Study
di: Li, Yinheng, et al.
Pubblicazione: (2024)
di: Li, Yinheng, et al.
Pubblicazione: (2024)
Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning
di: Rocamonde, Juan, et al.
Pubblicazione: (2023)
di: Rocamonde, Juan, et al.
Pubblicazione: (2023)
Benchmarking Open-Source Large Language Models for Persian in Zero-Shot and Few-Shot Learning
di: Cherakhloo, Mahdi, et al.
Pubblicazione: (2025)
di: Cherakhloo, Mahdi, et al.
Pubblicazione: (2025)
Bayesian Modeling of Zero-Shot Classifications for Urban Flood Detection
di: Franchi, Matt, et al.
Pubblicazione: (2025)
di: Franchi, Matt, et al.
Pubblicazione: (2025)
Too Good To Be True: performance overestimation in (re)current practices for Human Activity Recognition
di: Tello, Andrés, et al.
Pubblicazione: (2023)
di: Tello, Andrés, et al.
Pubblicazione: (2023)
MALMM: Multi-Agent Large Language Models for Zero-Shot Robotics Manipulation
di: Singh, Harsh, et al.
Pubblicazione: (2024)
di: Singh, Harsh, et al.
Pubblicazione: (2024)
Prevalent Frequency of Emotional and Physical Symptoms in Social Anxiety using Zero Shot Classification: An Observational Study
di: Rizwan, Muhammad, et al.
Pubblicazione: (2024)
di: Rizwan, Muhammad, et al.
Pubblicazione: (2024)
Language Model Representations for Efficient Few-Shot Tabular Classification
di: Kang, Inwon, et al.
Pubblicazione: (2026)
di: Kang, Inwon, et al.
Pubblicazione: (2026)
Utilizing Large Language Models for Zero-Shot Medical Ontology Extension from Clinical Notes
di: Wu, Guanchen, et al.
Pubblicazione: (2025)
di: Wu, Guanchen, et al.
Pubblicazione: (2025)
Memory-Free Continual Learning with Null Space Adaptation for Zero-Shot Vision-Language Models
di: Jo, Yujin, et al.
Pubblicazione: (2025)
di: Jo, Yujin, et al.
Pubblicazione: (2025)
ZeroShotOpt: Towards Zero-Shot Pretrained Models for Efficient Black-Box Optimization
di: Meindl, Jamison, et al.
Pubblicazione: (2025)
di: Meindl, Jamison, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Leveraging Information Retrieval to Enhance Spoken Language Understanding Prompts in Few-Shot Learning
di: Lepagnol, Pierre, et al.
Pubblicazione: (2025) -
mALBERT: Is a Compact Multilingual BERT Model Still Worth It?
di: Servan, Christophe, et al.
Pubblicazione: (2024) -
New Semantic Task for the French Spoken Language Understanding MEDIA Benchmark
di: Alavoine, Nadège, et al.
Pubblicazione: (2024) -
LLM-based Atomic Propositions help weak extractors: Evaluation of a Propositioner for triplet extraction
di: Pommeret, Luc, et al.
Pubblicazione: (2026) -
Enhancing Small Language Models for Cross-Lingual Generalized Zero-Shot Classification with Soft Prompt Tuning
di: Philippy, Fred, et al.
Pubblicazione: (2025)