TALEC: Teach Your LLM to Evaluate in Specific Domain with In-house Criteria by Criteria Division and Zero-shot Plus Few-shot
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Kaiqi, Yuan, Shuai, Zhao, Honghan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Revisiting Chain-of-Thought Prompting: Zero-shot Can Be Stronger than Few-shot
por: Cheng, Xiang, et al.
Publicado: (2025)
por: Cheng, Xiang, et al.
Publicado: (2025)
Qworld: Question-Specific Evaluation Criteria for LLMs
por: Gao, Shanghua, et al.
Publicado: (2026)
por: Gao, Shanghua, et al.
Publicado: (2026)
Zero-shot and Few-shot Learning with Instruction-following LLMs for Claim Matching in Automated Fact-checking
por: Pisarevskaya, Dina, et al.
Publicado: (2025)
por: Pisarevskaya, Dina, et al.
Publicado: (2025)
Few-shot LLM Synthetic Data with Distribution Matching
por: Ren, Jiyuan, et al.
Publicado: (2025)
por: Ren, Jiyuan, et al.
Publicado: (2025)
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
por: Ramprasad, Sanjana, et al.
Publicado: (2024)
por: Ramprasad, Sanjana, et al.
Publicado: (2024)
LELA: An End-to-end LLM-based Entity Linking Framework with Zero-shot Domain Adaptation
por: Haffoudhi, Samy, et al.
Publicado: (2026)
por: Haffoudhi, Samy, et al.
Publicado: (2026)
Explainable Few-shot Knowledge Tracing
por: Li, Haoxuan, et al.
Publicado: (2024)
por: Li, Haoxuan, et al.
Publicado: (2024)
A Zero-shot and Few-shot Study of Instruction-Finetuned Large Language Models Applied to Clinical and Biomedical Tasks
por: Labrak, Yanis, et al.
Publicado: (2023)
por: Labrak, Yanis, et al.
Publicado: (2023)
Beyond Pointwise Scores: Decomposed Criteria-Based Evaluation of LLM Responses
por: Yu, Fangyi, et al.
Publicado: (2025)
por: Yu, Fangyi, et al.
Publicado: (2025)
Few-shot Name Entity Recognition on StackOverflow
por: Chen, Xinwei, et al.
Publicado: (2024)
por: Chen, Xinwei, et al.
Publicado: (2024)
Generating Synthetic Datasets for Few-shot Prompt Tuning
por: Guo, Xu, et al.
Publicado: (2024)
por: Guo, Xu, et al.
Publicado: (2024)
Zero- and Few-shot Named Entity Recognition and Text Expansion in Medication Prescriptions using ChatGPT
por: Isaradech, Natthanaphop, et al.
Publicado: (2024)
por: Isaradech, Natthanaphop, et al.
Publicado: (2024)
AHP-Powered LLM Reasoning for Multi-Criteria Evaluation of Open-Ended Responses
por: Lu, Xiaotian, et al.
Publicado: (2024)
por: Lu, Xiaotian, et al.
Publicado: (2024)
Generation-driven Contrastive Self-training for Zero-shot Text Classification with Instruction-following LLM
por: Zhang, Ruohong, et al.
Publicado: (2023)
por: Zhang, Ruohong, et al.
Publicado: (2023)
LLMs as Zero-shot Graph Learners: Alignment of GNN Representations with LLM Token Embeddings
por: Wang, Duo, et al.
Publicado: (2024)
por: Wang, Duo, et al.
Publicado: (2024)
Effectiveness of Zero-shot-CoT in Japanese Prompts
por: Takayama, Shusuke, et al.
Publicado: (2025)
por: Takayama, Shusuke, et al.
Publicado: (2025)
From Zero to Hero: Harnessing Transformers for Biomedical Named Entity Recognition in Zero- and Few-shot Contexts
por: Košprdić, Miloš, et al.
Publicado: (2023)
por: Košprdić, Miloš, et al.
Publicado: (2023)
MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning
por: Tang, Xiangru, et al.
Publicado: (2023)
por: Tang, Xiangru, et al.
Publicado: (2023)
UniGen: Universal Domain Generalization for Sentiment Classification via Zero-shot Dataset Generation
por: Choi, Juhwan, et al.
Publicado: (2024)
por: Choi, Juhwan, et al.
Publicado: (2024)
Few-shot Policy (de)composition in Conversational Question Answering
por: Erwin, Kyle, et al.
Publicado: (2025)
por: Erwin, Kyle, et al.
Publicado: (2025)
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
por: Bhattacharjee, Amrita, et al.
Publicado: (2024)
por: Bhattacharjee, Amrita, et al.
Publicado: (2024)
Confidence-guided Refinement Reasoning for Zero-shot Question Answering
por: Jang, Youwon, et al.
Publicado: (2025)
por: Jang, Youwon, et al.
Publicado: (2025)
Zero-shot Commonsense Reasoning over Machine Imagination
por: Park, Hyuntae, et al.
Publicado: (2024)
por: Park, Hyuntae, et al.
Publicado: (2024)
Irony Detection, Reasoning and Understanding in Zero-shot Learning
por: Yi, Peiling, et al.
Publicado: (2025)
por: Yi, Peiling, et al.
Publicado: (2025)
Autonomous Data Selection with Zero-shot Generative Classifiers for Mathematical Texts
por: Zhang, Yifan, et al.
Publicado: (2024)
por: Zhang, Yifan, et al.
Publicado: (2024)
Zero-shot Benchmarking: A Framework for Flexible and Scalable Automatic Evaluation of Language Models
por: Pombal, José, et al.
Publicado: (2025)
por: Pombal, José, et al.
Publicado: (2025)
Iterative Repair with Weak Verifiers for Few-shot Transfer in KBQA with Unanswerability
por: Sawhney, Riya, et al.
Publicado: (2024)
por: Sawhney, Riya, et al.
Publicado: (2024)
Intent-driven In-context Learning for Few-shot Dialogue State Tracking
por: Yi, Zihao, et al.
Publicado: (2024)
por: Yi, Zihao, et al.
Publicado: (2024)
Preserving Generalization of Language models in Few-shot Continual Relation Extraction
por: Tran, Quyen, et al.
Publicado: (2024)
por: Tran, Quyen, et al.
Publicado: (2024)
MFORT-QA: Multi-hop Few-shot Open Rich Table Question Answering
por: Guan, Che, et al.
Publicado: (2024)
por: Guan, Che, et al.
Publicado: (2024)
The Behavior Gap: Evaluating Zero-shot LLM Agents in Complex Task-Oriented Dialogs
por: Baidya, Avinash, et al.
Publicado: (2025)
por: Baidya, Avinash, et al.
Publicado: (2025)
TEGEE: Task dEfinition Guided Expert Ensembling for Generalizable and Few-shot Learning
por: Qu, Xingwei, et al.
Publicado: (2024)
por: Qu, Xingwei, et al.
Publicado: (2024)
Anti-LM Decoding for Zero-shot In-context Machine Translation
por: Sia, Suzanna, et al.
Publicado: (2023)
por: Sia, Suzanna, et al.
Publicado: (2023)
Evaluating Large Language Models on the Frame and Symbol Grounding Problems: A Zero-shot Benchmark
por: Oka, Shoko
Publicado: (2025)
por: Oka, Shoko
Publicado: (2025)
Few-shot Personalization of LLMs with Mis-aligned Responses
por: Kim, Jaehyung, et al.
Publicado: (2024)
por: Kim, Jaehyung, et al.
Publicado: (2024)
BayesPrompt: Prompting Large-Scale Pre-Trained Language Models on Few-shot Inference via Debiased Domain Abstraction
por: Li, Jiangmeng, et al.
Publicado: (2024)
por: Li, Jiangmeng, et al.
Publicado: (2024)
Label-Guided Prompt for Multi-label Few-shot Aspect Category Detection
por: Guan, ChaoFeng, et al.
Publicado: (2024)
por: Guan, ChaoFeng, et al.
Publicado: (2024)
Cross-lingual Few-shot Learning for Persian Sentiment Analysis with Incremental Adaptation
por: Majidi, Farideh, et al.
Publicado: (2025)
por: Majidi, Farideh, et al.
Publicado: (2025)
Adaptive Few-shot Prompting for Machine Translation with Pre-trained Language Models
por: Tang, Lei, et al.
Publicado: (2025)
por: Tang, Lei, et al.
Publicado: (2025)
RAGs to Riches: RAG-like Few-shot Learning for Large Language Model Role-playing
por: Rupprecht, Timothy, et al.
Publicado: (2025)
por: Rupprecht, Timothy, et al.
Publicado: (2025)
Ejemplares similares
-
Revisiting Chain-of-Thought Prompting: Zero-shot Can Be Stronger than Few-shot
por: Cheng, Xiang, et al.
Publicado: (2025) -
Qworld: Question-Specific Evaluation Criteria for LLMs
por: Gao, Shanghua, et al.
Publicado: (2026) -
Zero-shot and Few-shot Learning with Instruction-following LLMs for Claim Matching in Automated Fact-checking
por: Pisarevskaya, Dina, et al.
Publicado: (2025) -
Few-shot LLM Synthetic Data with Distribution Matching
por: Ren, Jiyuan, et al.
Publicado: (2025) -
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
por: Ramprasad, Sanjana, et al.
Publicado: (2024)