NUTMEG: Separating Signal From Noise in Annotator Disagreement
Fuente:
arXiv
Guardado en:
| Autores principales: | Ivey, Jonathan, Gauch, Susan, Jurgens, David |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Beyond Consensus: Perspectivist Modeling and Evaluation of Annotator Disagreement in NLP
por: Xu, Yinuo, et al.
Publicado: (2026)
por: Xu, Yinuo, et al.
Publicado: (2026)
Modeling Annotator Disagreement with Demographic-Aware Experts and Synthetic Perspectives
por: Xu, Yinuo, et al.
Publicado: (2025)
por: Xu, Yinuo, et al.
Publicado: (2025)
Leveraging Annotator Disagreement for Text Classification
por: Xu, Jin, et al.
Publicado: (2024)
por: Xu, Jin, et al.
Publicado: (2024)
Intermediate-Task Transfer Learning: Leveraging Sarcasm Detection for Stance Detection
por: Nkhata, Gibson, et al.
Publicado: (2025)
por: Nkhata, Gibson, et al.
Publicado: (2025)
Boolean-aware Attention for Dense Retrieval
por: Mai, Quan, et al.
Publicado: (2025)
por: Mai, Quan, et al.
Publicado: (2025)
SetBERT: Enhancing Retrieval Performance for Boolean Logic and Set Operation Queries
por: Mai, Quan, et al.
Publicado: (2024)
por: Mai, Quan, et al.
Publicado: (2024)
Incorporating Classifier-Free Guidance in Diffusion Model-Based Recommendation
por: Buchanan, Noah, et al.
Publicado: (2024)
por: Buchanan, Noah, et al.
Publicado: (2024)
Crowd-Calibrator: Can Annotator Disagreement Inform Calibration in Subjective Tasks?
por: Khurana, Urja, et al.
Publicado: (2024)
por: Khurana, Urja, et al.
Publicado: (2024)
A Decomposition-Based Approach for Evaluating and Analyzing Inter-Annotator Disagreement
por: Levi, Effi, et al.
Publicado: (2022)
por: Levi, Effi, et al.
Publicado: (2022)
Sequence Graph Network for Online Debate Analysis
por: Mai, Quan, et al.
Publicado: (2024)
por: Mai, Quan, et al.
Publicado: (2024)
When the Majority is Wrong: Modeling Annotator Disagreement for Subjective Tasks
por: Fleisig, Eve, et al.
Publicado: (2023)
por: Fleisig, Eve, et al.
Publicado: (2023)
Mixed Signals: Understanding Model Disagreement in Multimodal Empathy Detection
por: Srikanth, Maya, et al.
Publicado: (2025)
por: Srikanth, Maya, et al.
Publicado: (2025)
Sarcasm Detection as a Catalyst: Improving Stance Detection with Cross-Target Capabilities
por: Hong, Gibson Nkhata Shi Yin, et al.
Publicado: (2025)
por: Hong, Gibson Nkhata Shi Yin, et al.
Publicado: (2025)
Fine-tuning BERT with Bidirectional LSTM for Fine-grained Movie Reviews Sentiment Analysis
por: Nkhata, Gibson, et al.
Publicado: (2025)
por: Nkhata, Gibson, et al.
Publicado: (2025)
Are Rules Meant to be Broken? Understanding Multilingual Moral Reasoning as a Computational Pipeline with UniMoral
por: Kumar, Shivani, et al.
Publicado: (2025)
por: Kumar, Shivani, et al.
Publicado: (2025)
Modeling Empathetic Alignment in Conversation
por: Yang, Jiamin, et al.
Publicado: (2024)
por: Yang, Jiamin, et al.
Publicado: (2024)
What Makes a Good Response? An Empirical Analysis of Quality in Qualitative Interviews
por: Ivey, Jonathan, et al.
Publicado: (2026)
por: Ivey, Jonathan, et al.
Publicado: (2026)
Structured Disagreement in Health-Literacy Annotation: Epistemic Stability, Conceptual Difficulty, and Agreement-Stratified Inference
por: Kellert, Olga, et al.
Publicado: (2026)
por: Kellert, Olga, et al.
Publicado: (2026)
JuniperLiu at CoMeDi Shared Task: Models as Annotators in Lexical Semantics Disagreements
por: Liu, Zhu, et al.
Publicado: (2024)
por: Liu, Zhu, et al.
Publicado: (2024)
Can Reasoning Help Large Language Models Capture Human Annotator Disagreement?
por: Ni, Jingwei, et al.
Publicado: (2025)
por: Ni, Jingwei, et al.
Publicado: (2025)
From Disagreement to Understanding: The Case for Ambiguity Detection in NLI
por: Jayaweera, Chathuri, et al.
Publicado: (2025)
por: Jayaweera, Chathuri, et al.
Publicado: (2025)
Dealing with Annotator Disagreement in Hate Speech Classification
por: Dehghan, Somaiyeh, et al.
Publicado: (2025)
por: Dehghan, Somaiyeh, et al.
Publicado: (2025)
Verifying Rumors via Stance-Aware Structural Modeling
por: Nkhata, Gibson, et al.
Publicado: (2025)
por: Nkhata, Gibson, et al.
Publicado: (2025)
Funzac at CoMeDi Shared Task: Modeling Annotator Disagreement from Word-In-Context Perspectives
por: Sarumi, Olufunke O., et al.
Publicado: (2025)
por: Sarumi, Olufunke O., et al.
Publicado: (2025)
Is LLM an Overconfident Judge? Unveiling the Capabilities of LLMs in Detecting Offensive Language with Annotation Disagreement
por: Lu, Junyu, et al.
Publicado: (2025)
por: Lu, Junyu, et al.
Publicado: (2025)
Think Multilingual, Not Harder: A Data-Efficient Framework for Teaching Reasoning Models to Code-Switch
por: Lin, Eleanor M., et al.
Publicado: (2026)
por: Lin, Eleanor M., et al.
Publicado: (2026)
SANTA: Separate Strategies for Inaccurate and Incomplete Annotation Noise in Distantly-Supervised Named Entity Recognition
por: Si, Shuzheng, et al.
Publicado: (2023)
por: Si, Shuzheng, et al.
Publicado: (2023)
The Role of Syntactic Span Preferences in Post-Hoc Explanation Disagreement
por: Kamp, Jonathan, et al.
Publicado: (2024)
por: Kamp, Jonathan, et al.
Publicado: (2024)
Tokenization is Sensitive to Language Variation
por: Wegmann, Anna, et al.
Publicado: (2025)
por: Wegmann, Anna, et al.
Publicado: (2025)
Leveraging Multilingual Training for Authorship Representation: Enhancing Generalization across Languages and Domains
por: Kim, Junghwan, et al.
Publicado: (2025)
por: Kim, Junghwan, et al.
Publicado: (2025)
The Noisy Path from Source to Citation: Measuring How Scholars Engage with Past Research
por: Chen, Hong, et al.
Publicado: (2025)
por: Chen, Hong, et al.
Publicado: (2025)
Are Economists Always More Introverted? Analyzing Consistency in Persona-Assigned LLMs
por: Reusens, Manon, et al.
Publicado: (2025)
por: Reusens, Manon, et al.
Publicado: (2025)
ClusterRAG: Cluster-Based Collaborative Filtering for Personalized Retrieval-Augmented Generation
por: Nkhata, Gibson, et al.
Publicado: (2026)
por: Nkhata, Gibson, et al.
Publicado: (2026)
Mapping the Podcast Ecosystem with the Structured Podcast Research Corpus
por: Litterer, Benjamin, et al.
Publicado: (2024)
por: Litterer, Benjamin, et al.
Publicado: (2024)
From Dissonance to Insights: Dissecting Disagreements in Rationale Construction for Case Outcome Classification
por: Xu, Shanshan, et al.
Publicado: (2023)
por: Xu, Shanshan, et al.
Publicado: (2023)
When Disagreements Elicit Robustness: Investigating Self-Repair Capabilities under LLM Multi-Agent Disagreements
por: Ju, Tianjie, et al.
Publicado: (2025)
por: Ju, Tianjie, et al.
Publicado: (2025)
Vicarious Offense and Noise Audit of Offensive Speech Classifiers: Unifying Human and Machine Disagreement on What is Offensive
por: Weerasooriya, Tharindu Cyril, et al.
Publicado: (2023)
por: Weerasooriya, Tharindu Cyril, et al.
Publicado: (2023)
Mind2: Mind-to-Mind Emotional Support System with Bidirectional Cognitive Discourse Analysis
por: Hong, Shi Yin, et al.
Publicado: (2025)
por: Hong, Shi Yin, et al.
Publicado: (2025)
Predicting Disagreement with Human Raters in LLM-as-a-Judge Difficulty Assessment without Using Generation-Time Probability Signals
por: Ehara, Yo
Publicado: (2026)
por: Ehara, Yo
Publicado: (2026)
From Noise to Signal: When Outliers Seed New Topics
por: Zve, Evangelia, et al.
Publicado: (2026)
por: Zve, Evangelia, et al.
Publicado: (2026)
Ejemplares similares
-
Beyond Consensus: Perspectivist Modeling and Evaluation of Annotator Disagreement in NLP
por: Xu, Yinuo, et al.
Publicado: (2026) -
Modeling Annotator Disagreement with Demographic-Aware Experts and Synthetic Perspectives
por: Xu, Yinuo, et al.
Publicado: (2025) -
Leveraging Annotator Disagreement for Text Classification
por: Xu, Jin, et al.
Publicado: (2024) -
Intermediate-Task Transfer Learning: Leveraging Sarcasm Detection for Stance Detection
por: Nkhata, Gibson, et al.
Publicado: (2025) -
Boolean-aware Attention for Dense Retrieval
por: Mai, Quan, et al.
Publicado: (2025)