Selective Annotation via Data Allocation: These Data Should Be Triaged to Experts for Annotation Rather Than the Model
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Chen, Deng, Yang, Lei, Wenqiang, Lv, Jiancheng, Dagan, Ido |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ARAIDA: Analogical Reasoning-Augmented Interactive Data Annotation
by: Huang, Chen, et al.
Published: (2024)
by: Huang, Chen, et al.
Published: (2024)
EventFull: Complete and Consistent Event Relation Annotation
by: Eirew, Alon, et al.
Published: (2024)
by: Eirew, Alon, et al.
Published: (2024)
CoAnnotating: Uncertainty-Guided Work Allocation between Human and Large Language Models for Data Annotation
by: Li, Minzhi, et al.
Published: (2023)
by: Li, Minzhi, et al.
Published: (2023)
DREditor: An Time-efficient Approach for Building a Domain-specific Dense Retrieval Model
by: Huang, Chen, et al.
Published: (2024)
by: Huang, Chen, et al.
Published: (2024)
Effective QA-driven Annotation of Predicate-Argument Relations Across Languages
by: Davidov, Jonathan, et al.
Published: (2026)
by: Davidov, Jonathan, et al.
Published: (2026)
Can Large Language Models Understand Internet Buzzwords Through User-Generated Content
by: Huang, Chen, et al.
Published: (2025)
by: Huang, Chen, et al.
Published: (2025)
Cross-model Transferability among Large Language Models on the Platonic Representations of Concepts
by: Huang, Youcheng, et al.
Published: (2025)
by: Huang, Youcheng, et al.
Published: (2025)
Are Expert-Level Language Models Expert-Level Annotators?
by: Tseng, Yu-Min, et al.
Published: (2024)
by: Tseng, Yu-Min, et al.
Published: (2024)
How to Enable Effective Cooperation Between Humans and NLP Models: A Survey of Principles, Formalizations, and Beyond
by: Huang, Chen, et al.
Published: (2025)
by: Huang, Chen, et al.
Published: (2025)
Routing with Generated Data: Annotation-Free LLM Skill Estimation and Expert Selection
by: Niu, Tianyi, et al.
Published: (2026)
by: Niu, Tianyi, et al.
Published: (2026)
Evaluating Large Language Models as Expert Annotators
by: Tseng, Yu-Min, et al.
Published: (2025)
by: Tseng, Yu-Min, et al.
Published: (2025)
The ALCHEmist: Automated Labeling 500x CHEaper Than LLM Data Annotators
by: Huang, Tzu-Heng, et al.
Published: (2024)
by: Huang, Tzu-Heng, et al.
Published: (2024)
Concept -- An Evaluation Protocol on Conversational Recommender Systems with System-centric and User-centric Factors
by: Huang, Chen, et al.
Published: (2024)
by: Huang, Chen, et al.
Published: (2024)
Legend: Leveraging Representation Engineering to Annotate Safety Margin for Preference Datasets
by: Feng, Duanyu, et al.
Published: (2024)
by: Feng, Duanyu, et al.
Published: (2024)
A Unifying Scheme for Extractive Content Selection Tasks
by: Amar, Shmuel, et al.
Published: (2025)
by: Amar, Shmuel, et al.
Published: (2025)
See the Unseen: Better Context-Consistent Knowledge-Editing by Noises
by: Huang, Youcheng, et al.
Published: (2024)
by: Huang, Youcheng, et al.
Published: (2024)
LakeHopper: Cross Data Lakes Column Type Annotation through Model Adaptation
by: Sun, Yushi, et al.
Published: (2026)
by: Sun, Yushi, et al.
Published: (2026)
Dont Add, dont Miss: Effective Content Preserving Generation from Pre-Selected Text Spans
by: Slobodkin, Aviv, et al.
Published: (2023)
by: Slobodkin, Aviv, et al.
Published: (2023)
Text Annotation via Inductive Coding: Comparing Human Experts to LLMs in Qualitative Data Analysis
by: Parfenova, Angelina, et al.
Published: (2025)
by: Parfenova, Angelina, et al.
Published: (2025)
Beyond Pairwise: Global Zero-shot Temporal Graph Generation
by: Eirew, Alon, et al.
Published: (2025)
by: Eirew, Alon, et al.
Published: (2025)
Superlatives in Context: Modeling the Implicit Semantics of Superlatives
by: Pyatkin, Valentina, et al.
Published: (2024)
by: Pyatkin, Valentina, et al.
Published: (2024)
Modeling Annotator Disagreement with Demographic-Aware Experts and Synthetic Perspectives
by: Xu, Yinuo, et al.
Published: (2025)
by: Xu, Yinuo, et al.
Published: (2025)
Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
by: Cheang, Chi Seng, et al.
Published: (2025)
by: Cheang, Chi Seng, et al.
Published: (2025)
Large Language Models for Data Annotation and Synthesis: A Survey
by: Tan, Zhen, et al.
Published: (2024)
by: Tan, Zhen, et al.
Published: (2024)
Effective and Efficient Adversarial Detection for Vision-Language Models via A Single Vector
by: Huang, Youcheng, et al.
Published: (2024)
by: Huang, Youcheng, et al.
Published: (2024)
TeacherLM: Teaching to Fish Rather Than Giving the Fish, Language Modeling Likewise
by: He, Nan, et al.
Published: (2023)
by: He, Nan, et al.
Published: (2023)
Annotations on a Budget: Leveraging Geo-Data Similarity to Balance Model Performance and Annotation Cost
by: Ignat, Oana, et al.
Published: (2024)
by: Ignat, Oana, et al.
Published: (2024)
ACORD: An Expert-Annotated Retrieval Dataset for Legal Contract Drafting
by: Wang, Steven H., et al.
Published: (2025)
by: Wang, Steven H., et al.
Published: (2025)
If in a Crowdsourced Data Annotation Pipeline, a GPT-4
by: He, Zeyu, et al.
Published: (2024)
by: He, Zeyu, et al.
Published: (2024)
Large Language Model Selection with Limited Annotations
by: Durmazkeser, Yavuz, et al.
Published: (2026)
by: Durmazkeser, Yavuz, et al.
Published: (2026)
QA-Noun: Representing Nominal Semantics via Natural Language Question-Answer Pairs
by: Tseytlin, Maria, et al.
Published: (2025)
by: Tseytlin, Maria, et al.
Published: (2025)
Large Language Models as Financial Data Annotators: A Study on Effectiveness and Efficiency
by: Aguda, Toyin, et al.
Published: (2024)
by: Aguda, Toyin, et al.
Published: (2024)
BiaSWE: An Expert Annotated Dataset for Misogyny Detection in Swedish
by: Kukk, Kätriin, et al.
Published: (2025)
by: Kukk, Kätriin, et al.
Published: (2025)
Personas with Attitudes: Controlling LLMs for Diverse Data Annotation
by: Fröhling, Leon, et al.
Published: (2024)
by: Fröhling, Leon, et al.
Published: (2024)
Multi-Review Fusion-in-Context
by: Slobodkin, Aviv, et al.
Published: (2024)
by: Slobodkin, Aviv, et al.
Published: (2024)
Accurate and Data-Efficient Toxicity Prediction when Annotators Disagree
by: Jaggi, Harbani, et al.
Published: (2024)
by: Jaggi, Harbani, et al.
Published: (2024)
Tasks and Roles in Legal AI: Data Curation, Annotation, and Verification
by: Koenecke, Allison, et al.
Published: (2025)
by: Koenecke, Allison, et al.
Published: (2025)
LLMs as Data Annotators: How Close Are We to Human Performance
by: Haq, Muhammad Uzair Ul, et al.
Published: (2025)
by: Haq, Muhammad Uzair Ul, et al.
Published: (2025)
Dishonesty in Helpful and Harmless Alignment
by: Huang, Youcheng, et al.
Published: (2024)
by: Huang, Youcheng, et al.
Published: (2024)
TIGQA:An Expert Annotated Question Answering Dataset in Tigrinya
by: Teklehaymanot, Hailay, et al.
Published: (2024)
by: Teklehaymanot, Hailay, et al.
Published: (2024)
Similar Items
-
ARAIDA: Analogical Reasoning-Augmented Interactive Data Annotation
by: Huang, Chen, et al.
Published: (2024) -
EventFull: Complete and Consistent Event Relation Annotation
by: Eirew, Alon, et al.
Published: (2024) -
CoAnnotating: Uncertainty-Guided Work Allocation between Human and Large Language Models for Data Annotation
by: Li, Minzhi, et al.
Published: (2023) -
DREditor: An Time-efficient Approach for Building a Domain-specific Dense Retrieval Model
by: Huang, Chen, et al.
Published: (2024) -
Effective QA-driven Annotation of Predicate-Argument Relations Across Languages
by: Davidov, Jonathan, et al.
Published: (2026)