Task Contamination: Language Models May Not Be Few-Shot Anymore
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Changmao, Flanigan, Jeffrey |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Future Language Modeling from Temporal Document History
von: Li, Changmao, et al.
Veröffentlicht: (2024)
von: Li, Changmao, et al.
Veröffentlicht: (2024)
RAC: Efficient LLM Factuality Correction with Retrieval Augmentation
von: Li, Changmao, et al.
Veröffentlicht: (2024)
von: Li, Changmao, et al.
Veröffentlicht: (2024)
Active Few-Shot Learning for Text Classification
von: Ahmadnia, Saeed, et al.
Veröffentlicht: (2025)
von: Ahmadnia, Saeed, et al.
Veröffentlicht: (2025)
Fane at SemEval-2025 Task 10: Zero-Shot Entity Framing with Large Language Models
von: Fane, Enfa, et al.
Veröffentlicht: (2025)
von: Fane, Enfa, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models for Zero-Shot Disease Labeling in CT Radiology Reports Across Organ Systems
von: Garcia-Alcoser, Michael E., et al.
Veröffentlicht: (2025)
von: Garcia-Alcoser, Michael E., et al.
Veröffentlicht: (2025)
Language Models Can Resolve Reference Compositionally, But It's Not Their Native Strength: The Case of the Personal Relation Task
von: Evelo, Bart, et al.
Veröffentlicht: (2026)
von: Evelo, Bart, et al.
Veröffentlicht: (2026)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
A Case Study of Cross-Lingual Zero-Shot Generalization for Classical Languages in LLMs
von: Akavarapu, V. S. D. S. Mahesh, et al.
Veröffentlicht: (2025)
von: Akavarapu, V. S. D. S. Mahesh, et al.
Veröffentlicht: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
A Multi-Task Benchmark for Abusive Language Detection in Low-Resource Settings
von: Gaim, Fitsum, et al.
Veröffentlicht: (2025)
von: Gaim, Fitsum, et al.
Veröffentlicht: (2025)
Adapting Multilingual Models to Code-Mixed Tasks via Model Merging
von: Kodali, Prashant, et al.
Veröffentlicht: (2025)
von: Kodali, Prashant, et al.
Veröffentlicht: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
Efficient Few-shot Learning for Multi-label Classification of Scientific Documents with Many Classes
von: Schopf, Tim, et al.
Veröffentlicht: (2024)
von: Schopf, Tim, et al.
Veröffentlicht: (2024)
Heidelberg-Boston @ SIGTYP 2024 Shared Task: Enhancing Low-Resource Language Analysis With Character-Aware Hierarchical Transformers
von: Riemenschneider, Frederick, et al.
Veröffentlicht: (2024)
von: Riemenschneider, Frederick, et al.
Veröffentlicht: (2024)
RUQuant: Towards Refining Uniform Quantization for Large Language Models
von: Liu, Han, et al.
Veröffentlicht: (2026)
von: Liu, Han, et al.
Veröffentlicht: (2026)
OPOR-Bench: Evaluating Large Language Models on Online Public Opinion Report Generation
von: Yu, Jinzheng, et al.
Veröffentlicht: (2025)
von: Yu, Jinzheng, et al.
Veröffentlicht: (2025)
Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NER
von: Ewais, Ahmed, et al.
Veröffentlicht: (2026)
von: Ewais, Ahmed, et al.
Veröffentlicht: (2026)
ConPET: Continual Parameter-Efficient Tuning for Large Language Models
von: Song, Chenyang, et al.
Veröffentlicht: (2023)
von: Song, Chenyang, et al.
Veröffentlicht: (2023)
LoRS: Efficient Low-Rank Adaptation for Sparse Large Language Model
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Hu, Yuxuan, et al.
Veröffentlicht: (2025)
Precise Length Control in Large Language Models
von: Butcher, Bradley, et al.
Veröffentlicht: (2024)
von: Butcher, Bradley, et al.
Veröffentlicht: (2024)
What Drives Performance in Multilingual Language Models?
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
Large Language Models for Biomedical Article Classification
von: Proboszcz, Jakub, et al.
Veröffentlicht: (2026)
von: Proboszcz, Jakub, et al.
Veröffentlicht: (2026)
Induce, Align, Predict: Zero-Shot Stance Detection via Cognitive Inductive Reasoning
von: Zhang, Bowen, et al.
Veröffentlicht: (2025)
von: Zhang, Bowen, et al.
Veröffentlicht: (2025)
SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models
von: Liu, Han, et al.
Veröffentlicht: (2026)
von: Liu, Han, et al.
Veröffentlicht: (2026)
Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents
von: Wang, Renxi, et al.
Veröffentlicht: (2024)
von: Wang, Renxi, et al.
Veröffentlicht: (2024)
Lisbon Computational Linguists at SemEval-2024 Task 2: Using A Mistral 7B Model and Data Augmentation
von: Guimarães, Artur, et al.
Veröffentlicht: (2024)
von: Guimarães, Artur, et al.
Veröffentlicht: (2024)
Automatic Task Detection and Heterogeneous LLM Speculative Decoding
von: Ge, Danying, et al.
Veröffentlicht: (2025)
von: Ge, Danying, et al.
Veröffentlicht: (2025)
PLM: Efficient Peripheral Language Models Hardware-Co-Designed for Ubiquitous Computing
von: Deng, Cheng, et al.
Veröffentlicht: (2025)
von: Deng, Cheng, et al.
Veröffentlicht: (2025)
After Retrieval, Before Generation: Enhancing the Trustworthiness of Large Language Models in Retrieval-Augmented Generation
von: Dai, Xinbang, et al.
Veröffentlicht: (2025)
von: Dai, Xinbang, et al.
Veröffentlicht: (2025)
Strategy Adaptation in Large Language Model Werewolf Agents
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
PL-Guard: Benchmarking Language Model Safety for Polish
von: Krasnodębska, Aleksandra, et al.
Veröffentlicht: (2025)
von: Krasnodębska, Aleksandra, et al.
Veröffentlicht: (2025)
Socially Responsible Data for Large Multilingual Language Models
von: Smart, Andrew, et al.
Veröffentlicht: (2024)
von: Smart, Andrew, et al.
Veröffentlicht: (2024)
Large Language Models for Persian $ \leftrightarrow $ English Idiom Translation
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2024)
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2024)
Qomhra: A Bilingual Irish and English Large Language Model
von: McInerney, Joseph, et al.
Veröffentlicht: (2025)
von: McInerney, Joseph, et al.
Veröffentlicht: (2025)
Dialect Normalization using Large Language Models and Morphological Rules
von: Dimakis, Antonios, et al.
Veröffentlicht: (2025)
von: Dimakis, Antonios, et al.
Veröffentlicht: (2025)
Towards Human Understanding of Paraphrase Types in Large Language Models
von: Meier, Dominik, et al.
Veröffentlicht: (2024)
von: Meier, Dominik, et al.
Veröffentlicht: (2024)
Few-Shot Optimization for Sensor Data Using Large Language Models: A Case Study on Fatigue Detection
von: Ronando, Elsen, et al.
Veröffentlicht: (2025)
von: Ronando, Elsen, et al.
Veröffentlicht: (2025)
LinkNER: Linking Local Named Entity Recognition Models to Large Language Models using Uncertainty
von: Zhang, Zhen, et al.
Veröffentlicht: (2024)
von: Zhang, Zhen, et al.
Veröffentlicht: (2024)
A Domain-Based Taxonomy of Jailbreak Vulnerabilities in Large Language Models
von: Peláez-González, Carlos, et al.
Veröffentlicht: (2025)
von: Peláez-González, Carlos, et al.
Veröffentlicht: (2025)
Linguistic Interpretability of Transformer-based Language Models: a systematic review
von: López-Otal, Miguel, et al.
Veröffentlicht: (2025)
von: López-Otal, Miguel, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Future Language Modeling from Temporal Document History
von: Li, Changmao, et al.
Veröffentlicht: (2024) -
RAC: Efficient LLM Factuality Correction with Retrieval Augmentation
von: Li, Changmao, et al.
Veröffentlicht: (2024) -
Active Few-Shot Learning for Text Classification
von: Ahmadnia, Saeed, et al.
Veröffentlicht: (2025) -
Fane at SemEval-2025 Task 10: Zero-Shot Entity Framing with Large Language Models
von: Fane, Enfa, et al.
Veröffentlicht: (2025) -
Evaluating Large Language Models for Zero-Shot Disease Labeling in CT Radiology Reports Across Organ Systems
von: Garcia-Alcoser, Michael E., et al.
Veröffentlicht: (2025)