VariErr NLI: Separating Annotation Error from Human Label Variation
Fuente:
arXiv
Salvato in:
| Autori principali: | Weber-Genzel, Leon, Peng, Siyao, de Marneffe, Marie-Catherine, Plank, Barbara |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
LiTEx: A Linguistic Taxonomy of Explanations for Understanding Within-Label Variation in Natural Language Inference
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
EVADE: LLM-Based Explanation Generation and Validation for Error Detection in NLI
di: Zuo, Longfei, et al.
Pubblicazione: (2025)
di: Zuo, Longfei, et al.
Pubblicazione: (2025)
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
di: Peng, Siyao, et al.
Pubblicazione: (2024)
di: Peng, Siyao, et al.
Pubblicazione: (2024)
Donkii: Can Annotation Error Detection Methods Find Errors in Instruction-Tuning Datasets?
di: Weber-Genzel, Leon, et al.
Pubblicazione: (2023)
di: Weber-Genzel, Leon, et al.
Pubblicazione: (2023)
A Rose by Any Other Name: LLM-Generated Explanations Are Good Proxies for Human Explanations to Collect Label Distributions on NLI
di: Chen, Beiduo, et al.
Pubblicazione: (2024)
di: Chen, Beiduo, et al.
Pubblicazione: (2024)
"Seeing the Big through the Small": Can LLMs Approximate Human Judgment Distributions on NLI from a Few Explanations?
di: Chen, Beiduo, et al.
Pubblicazione: (2024)
di: Chen, Beiduo, et al.
Pubblicazione: (2024)
MaiBaam Annotation Guidelines
di: Blaschke, Verena, et al.
Pubblicazione: (2024)
di: Blaschke, Verena, et al.
Pubblicazione: (2024)
Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
EEVEE: An Easy Annotation Tool for Natural Language Processing
di: Sorensen, Axel, et al.
Pubblicazione: (2024)
di: Sorensen, Axel, et al.
Pubblicazione: (2024)
To Err Is Human; To Annotate, SILICON? Toward Robust Reproducibility in LLM Annotation
di: Cheng, Xiang, et al.
Pubblicazione: (2024)
di: Cheng, Xiang, et al.
Pubblicazione: (2024)
BoN Appetit Team at LeWiDi-2025: Best-of-N Test-time Scaling Can Not Stomach Annotation Disagreements (Yet)
di: Ruiz, Tomas, et al.
Pubblicazione: (2025)
di: Ruiz, Tomas, et al.
Pubblicazione: (2025)
CLIMATELI: Evaluating Entity Linking on Climate Change Data
di: Zhou, Shijia, et al.
Pubblicazione: (2024)
di: Zhou, Shijia, et al.
Pubblicazione: (2024)
Interpreting Predictive Probabilities: Model Confidence or Human Label Variation?
di: Baan, Joris, et al.
Pubblicazione: (2024)
di: Baan, Joris, et al.
Pubblicazione: (2024)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
di: Chen, Beiduo, et al.
Pubblicazione: (2025)
di: Chen, Beiduo, et al.
Pubblicazione: (2025)
References Matter: Investigating the Impact of Reference Set Variation on Summarization Evaluation
di: Casola, Silvia, et al.
Pubblicazione: (2025)
di: Casola, Silvia, et al.
Pubblicazione: (2025)
Humans and LLMs Diverge on Probabilistic Inferences
di: Kamath, Gaurav, et al.
Pubblicazione: (2026)
di: Kamath, Gaurav, et al.
Pubblicazione: (2026)
Revisiting Active Learning under (Human) Label Variation
di: Gruber, Cornelia, et al.
Pubblicazione: (2025)
di: Gruber, Cornelia, et al.
Pubblicazione: (2025)
Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
MaiBaam: A Multi-Dialectal Bavarian Universal Dependency Treebank
di: Blaschke, Verena, et al.
Pubblicazione: (2024)
di: Blaschke, Verena, et al.
Pubblicazione: (2024)
MultiClimate: Multimodal Stance Detection on Climate Change Videos
di: Wang, Jiawen, et al.
Pubblicazione: (2024)
di: Wang, Jiawen, et al.
Pubblicazione: (2024)
MedErrBench: A Fine-Grained Multilingual Benchmark for Medical Error Detection and Correction with Clinical Expert Annotations
di: Ma, Congbo, et al.
Pubblicazione: (2026)
di: Ma, Congbo, et al.
Pubblicazione: (2026)
To Err Is Human, but Llamas Can Learn It Too
di: Luhtaru, Agnes, et al.
Pubblicazione: (2024)
di: Luhtaru, Agnes, et al.
Pubblicazione: (2024)
To Err Is Human: Systematic Quantification of Errors in Published AI Papers via LLM Analysis
di: Bianchi, Federico, et al.
Pubblicazione: (2025)
di: Bianchi, Federico, et al.
Pubblicazione: (2025)
Information Asymmetry across Language Varieties: A Case Study on Cantonese-Mandarin and Bavarian-German QA
di: Pei, Renhao, et al.
Pubblicazione: (2026)
di: Pei, Renhao, et al.
Pubblicazione: (2026)
"My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
di: Wang, Xinpeng, et al.
Pubblicazione: (2024)
di: Wang, Xinpeng, et al.
Pubblicazione: (2024)
Sebastian, Basti, Wastl?! Recognizing Named Entities in Bavarian Dialectal Data
di: Peng, Siyao, et al.
Pubblicazione: (2024)
di: Peng, Siyao, et al.
Pubblicazione: (2024)
A survey of diversity quantification in natural language processing: The why, what, where and how
di: Estève, Louis, et al.
Pubblicazione: (2025)
di: Estève, Louis, et al.
Pubblicazione: (2025)
Neural Text Normalization for Luxembourgish using Real-Life Variation Data
di: Lutgen, Anne-Marie, et al.
Pubblicazione: (2024)
di: Lutgen, Anne-Marie, et al.
Pubblicazione: (2024)
Variation is the Norm: Embracing Sociolinguistics in NLP
di: Lutgen, Anne-Marie, et al.
Pubblicazione: (2026)
di: Lutgen, Anne-Marie, et al.
Pubblicazione: (2026)
The Ecological Fallacy in Annotation: Modelling Human Label Variation goes beyond Sociodemographics
di: Orlikowski, Matthias, et al.
Pubblicazione: (2023)
di: Orlikowski, Matthias, et al.
Pubblicazione: (2023)
Explanation sensitivity to the randomness of large language models: the case of journalistic text classification
di: Bogaert, Jeremie, et al.
Pubblicazione: (2024)
di: Bogaert, Jeremie, et al.
Pubblicazione: (2024)
Comparing Inferential Strategies of Humans and Large Language Models in Deductive Reasoning
di: Mondorf, Philipp, et al.
Pubblicazione: (2024)
di: Mondorf, Philipp, et al.
Pubblicazione: (2024)
Lost in Variation? Evaluating NLI Performance in Basque and Spanish Geographical Variants
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2025)
di: Bengoetxea, Jaione, et al.
Pubblicazione: (2025)
Humans Hallucinate Too: Language Models Identify and Correct Subjective Annotation Errors With Label-in-a-Haystack Prompts
di: Chochlakis, Georgios, et al.
Pubblicazione: (2025)
di: Chochlakis, Georgios, et al.
Pubblicazione: (2025)
Are Non-English Papers Reviewed Fairly? Language-of-Study Bias in NLP Peer Reviews
di: Barkhordar, Ehsan, et al.
Pubblicazione: (2026)
di: Barkhordar, Ehsan, et al.
Pubblicazione: (2026)
Make Every Letter Count: Building Dialect Variation Dictionaries from Monolingual Corpora
di: Litschko, Robert, et al.
Pubblicazione: (2025)
di: Litschko, Robert, et al.
Pubblicazione: (2025)
Probing LLMs for Multilingual Discourse Generalization Through a Unified Label Set
di: Eichin, Florian, et al.
Pubblicazione: (2025)
di: Eichin, Florian, et al.
Pubblicazione: (2025)
What Media Frames Reveal About Stance: A Dataset and Study about Memes in Climate Change Discourse
di: Zhou, Shijia, et al.
Pubblicazione: (2025)
di: Zhou, Shijia, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations
di: Hong, Pingjun, et al.
Pubblicazione: (2025) -
LiTEx: A Linguistic Taxonomy of Explanations for Understanding Within-Label Variation in Natural Language Inference
di: Hong, Pingjun, et al.
Pubblicazione: (2025) -
EVADE: LLM-Based Explanation Generation and Validation for Error Detection in NLI
di: Zuo, Longfei, et al.
Pubblicazione: (2025) -
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
di: Peng, Siyao, et al.
Pubblicazione: (2024) -
Donkii: Can Annotation Error Detection Methods Find Errors in Instruction-Tuning Datasets?
di: Weber-Genzel, Leon, et al.
Pubblicazione: (2023)