Using tournaments to calculate AUROC for zero-shot classification with LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Yoon, WonJin, Bulovic, Ian, Miller, Timothy A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation
por: Yoon, WonJin, et al.
Publicado: (2026)
por: Yoon, WonJin, et al.
Publicado: (2026)
Aspect-Oriented Summarization for Psychiatric Short-Term Readmission Prediction
por: Yoon, WonJin, et al.
Publicado: (2025)
por: Yoon, WonJin, et al.
Publicado: (2025)
LAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classification
por: Du, Zhicheng, et al.
Publicado: (2024)
por: Du, Zhicheng, et al.
Publicado: (2024)
A comparative study of zero-shot inference with large language models and supervised modeling in breast cancer pathology classification
por: Sushil, Madhumita, et al.
Publicado: (2024)
por: Sushil, Madhumita, et al.
Publicado: (2024)
CCQA: Generating Question from Solution Can Improve Inference-Time Reasoning in SLMs
por: Kim, Jin Young, et al.
Publicado: (2025)
por: Kim, Jin Young, et al.
Publicado: (2025)
Zero-shot data citation function classification using transformer-based large language models (LLMs)
por: Byers, Neil, et al.
Publicado: (2025)
por: Byers, Neil, et al.
Publicado: (2025)
Zero-shot Sentiment Analysis in Low-Resource Languages Using a Multilingual Sentiment Lexicon
por: Koto, Fajri, et al.
Publicado: (2024)
por: Koto, Fajri, et al.
Publicado: (2024)
DiZiNER: Disagreement-guided Instruction Refinement via Pilot Annotation Simulation for Zero-shot Named Entity Recognition
por: Kim, Siun, et al.
Publicado: (2026)
por: Kim, Siun, et al.
Publicado: (2026)
Empirical study of pretrained multilingual language models for zero-shot cross-lingual knowledge transfer in generation
por: Chirkova, Nadezhda, et al.
Publicado: (2023)
por: Chirkova, Nadezhda, et al.
Publicado: (2023)
Effectiveness of Zero-shot-CoT in Japanese Prompts
por: Takayama, Shusuke, et al.
Publicado: (2025)
por: Takayama, Shusuke, et al.
Publicado: (2025)
Zero-shot Slot Filling in the Age of LLMs for Dialogue Systems
por: Rana, Mansi, et al.
Publicado: (2024)
por: Rana, Mansi, et al.
Publicado: (2024)
Evaluating and explaining training strategies for zero-shot cross-lingual news sentiment analysis
por: Andrenšek, Luka, et al.
Publicado: (2024)
por: Andrenšek, Luka, et al.
Publicado: (2024)
Key ingredients for effective zero-shot cross-lingual knowledge transfer in generative tasks
por: Chirkova, Nadezhda, et al.
Publicado: (2024)
por: Chirkova, Nadezhda, et al.
Publicado: (2024)
Analysing zero-shot temporal relation extraction on clinical notes using temporal consistency
por: Kougia, Vasiliki, et al.
Publicado: (2024)
por: Kougia, Vasiliki, et al.
Publicado: (2024)
Prefill-Guided Thinking for zero-shot detection of AI-generated images
por: Kachwala, Zoher, et al.
Publicado: (2025)
por: Kachwala, Zoher, et al.
Publicado: (2025)
Scalable and consistent few-shot classification of survey responses using text embeddings
por: Mjaaland, Jonas Timmann, et al.
Publicado: (2025)
por: Mjaaland, Jonas Timmann, et al.
Publicado: (2025)
Scaling few-shot spoken word classification with generative meta-continual learning
por: Beyers, Louise, et al.
Publicado: (2026)
por: Beyers, Louise, et al.
Publicado: (2026)
PETapter: Leveraging PET-style classification heads for modular few-shot parameter-efficient fine-tuning
por: Rieger, Jonas, et al.
Publicado: (2024)
por: Rieger, Jonas, et al.
Publicado: (2024)
Can Prompting LLMs Unlock Hate Speech Detection across Languages? A Zero-shot and Few-shot Study
por: Ghorbanpour, Faeze, et al.
Publicado: (2025)
por: Ghorbanpour, Faeze, et al.
Publicado: (2025)
Zero-shot and Few-shot Learning with Instruction-following LLMs for Claim Matching in Automated Fact-checking
por: Pisarevskaya, Dina, et al.
Publicado: (2025)
por: Pisarevskaya, Dina, et al.
Publicado: (2025)
Assessing LLMs for Zero-shot Abstractive Summarization Through the Lens of Relevance Paraphrasing
por: Askari, Hadi, et al.
Publicado: (2024)
por: Askari, Hadi, et al.
Publicado: (2024)
SEPS: A Separability Measure for Robust Unlearning in LLMs
por: Jeung, Wonje, et al.
Publicado: (2025)
por: Jeung, Wonje, et al.
Publicado: (2025)
Zero-shot prompt-based classification: topic labeling in times of foundation models in German Tweets
por: Münker, Simon, et al.
Publicado: (2024)
por: Münker, Simon, et al.
Publicado: (2024)
Team LA at SCIDOCA shared task 2025: Citation Discovery via relation-based zero-shot retrieval
por: An, Trieu, et al.
Publicado: (2025)
por: An, Trieu, et al.
Publicado: (2025)
KG-HTC: Integrating Knowledge Graphs into LLMs for Effective Zero-shot Hierarchical Text Classification
por: Zang, Qianbo, et al.
Publicado: (2025)
por: Zang, Qianbo, et al.
Publicado: (2025)
Evaluating the fairness of task-adaptive pretraining on unlabeled test data before few-shot text classification
por: Dubey, Kush
Publicado: (2024)
por: Dubey, Kush
Publicado: (2024)
Can Video LLMs Refuse to Answer? Alignment for Answerability in Video Large Language Models
por: Yoon, Eunseop, et al.
Publicado: (2025)
por: Yoon, Eunseop, et al.
Publicado: (2025)
SpeechLLMs for Large-scale Contextualized Zero-shot Slot Filling
por: Hacioglu, Kadri, et al.
Publicado: (2025)
por: Hacioglu, Kadri, et al.
Publicado: (2025)
Zero-shot Graph Reasoning via Retrieval Augmented Framework with LLMs
por: Li, Hanqing, et al.
Publicado: (2025)
por: Li, Hanqing, et al.
Publicado: (2025)
Surrogate modeling for interpreting black-box LLMs in medical predictions
por: Han, Changho, et al.
Publicado: (2026)
por: Han, Changho, et al.
Publicado: (2026)
Few-shot Personalization of LLMs with Mis-aligned Responses
por: Kim, Jaehyung, et al.
Publicado: (2024)
por: Kim, Jaehyung, et al.
Publicado: (2024)
LLM economicus? Mapping the Behavioral Biases of LLMs via Utility Theory
por: Ross, Jillian, et al.
Publicado: (2024)
por: Ross, Jillian, et al.
Publicado: (2024)
AdpQ: A Zero-shot Calibration Free Adaptive Post Training Quantization Method for LLMs
por: Ghaffari, Alireza, et al.
Publicado: (2024)
por: Ghaffari, Alireza, et al.
Publicado: (2024)
Few shot chain-of-thought driven reasoning to prompt LLMs for open ended medical question answering
por: Nachane, Saeel Sandeep, et al.
Publicado: (2024)
por: Nachane, Saeel Sandeep, et al.
Publicado: (2024)
Consistency Guided Knowledge Retrieval and Denoising in LLMs for Zero-shot Document-level Relation Triplet Extraction
por: Sun, Qi, et al.
Publicado: (2024)
por: Sun, Qi, et al.
Publicado: (2024)
Identifying Task Groupings for Multi-Task Learning Using Pointwise V-Usable Information
por: Li, Yingya, et al.
Publicado: (2024)
por: Li, Yingya, et al.
Publicado: (2024)
Social Construction of Urban Space: Using LLMs to Identify Neighborhood Boundaries From Craigslist Ads
por: Visokay, Adam, et al.
Publicado: (2025)
por: Visokay, Adam, et al.
Publicado: (2025)
EHRNoteQA: An LLM Benchmark for Real-World Clinical Practice Using Discharge Summaries
por: Kweon, Sunjun, et al.
Publicado: (2024)
por: Kweon, Sunjun, et al.
Publicado: (2024)
Exploring the Potential of LLMs as Personalized Assistants: Dataset, Evaluation, and Analysis
por: Mok, Jisoo, et al.
Publicado: (2025)
por: Mok, Jisoo, et al.
Publicado: (2025)
Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAG
por: Jin, Bowen, et al.
Publicado: (2024)
por: Jin, Bowen, et al.
Publicado: (2024)
Ejemplares similares
-
CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation
por: Yoon, WonJin, et al.
Publicado: (2026) -
Aspect-Oriented Summarization for Psychiatric Short-Term Readmission Prediction
por: Yoon, WonJin, et al.
Publicado: (2025) -
LAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classification
por: Du, Zhicheng, et al.
Publicado: (2024) -
A comparative study of zero-shot inference with large language models and supervised modeling in breast cancer pathology classification
por: Sushil, Madhumita, et al.
Publicado: (2024) -
CCQA: Generating Question from Solution Can Improve Inference-Time Reasoning in SLMs
por: Kim, Jin Young, et al.
Publicado: (2025)