To Aggregate or Not to Aggregate. That is the Question: A Case Study on Annotation Subjectivity in Span Prediction
Fuente:
arXiv
Salvato in:
| Autori principali: | Kurniawan, Kemal, Mistica, Meladel, Baldwin, Timothy, Lau, Jey Han |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Training and Evaluating with Human Label Variation: An Empirical Study
di: Kurniawan, Kemal, et al.
Pubblicazione: (2025)
di: Kurniawan, Kemal, et al.
Pubblicazione: (2025)
On the Interplay between Human Label Variation and Model Fairness
di: Kurniawan, Kemal, et al.
Pubblicazione: (2025)
di: Kurniawan, Kemal, et al.
Pubblicazione: (2025)
MoDEM: Mixture of Domain Expert Models
di: Simonds, Toby, et al.
Pubblicazione: (2024)
di: Simonds, Toby, et al.
Pubblicazione: (2024)
Evaluating Evidence Attribution in Generated Fact Checking Explanations
di: Xing, Rui, et al.
Pubblicazione: (2024)
di: Xing, Rui, et al.
Pubblicazione: (2024)
A Joint Multitask Model for Morpho-Syntactic Parsing
di: Inostroza, Demian, et al.
Pubblicazione: (2025)
di: Inostroza, Demian, et al.
Pubblicazione: (2025)
COMMUNITYNOTES: A Dataset for Exploring the Helpfulness of Fact-Checking Explanations
di: Xing, Rui, et al.
Pubblicazione: (2025)
di: Xing, Rui, et al.
Pubblicazione: (2025)
An Analytical Emotion Framework of Rumour Threads on Social Media
di: Xing, Rui, et al.
Pubblicazione: (2025)
di: Xing, Rui, et al.
Pubblicazione: (2025)
A Sentiment Consolidation Framework for Meta-Review Generation
di: Li, Miao, et al.
Pubblicazione: (2024)
di: Li, Miao, et al.
Pubblicazione: (2024)
Span-Aggregatable, Contextualized Word Embeddings for Effective Phrase Mining
di: Orbach, Eyal, et al.
Pubblicazione: (2024)
di: Orbach, Eyal, et al.
Pubblicazione: (2024)
Robustness of Neurosymbolic Reasoners on First-Order Logic Problems
di: Bansal, Hannah, et al.
Pubblicazione: (2025)
di: Bansal, Hannah, et al.
Pubblicazione: (2025)
CMA-R:Causal Mediation Analysis for Explaining Rumour Detection
di: Tian, Lin, et al.
Pubblicazione: (2024)
di: Tian, Lin, et al.
Pubblicazione: (2024)
Beyond Seen Data: Improving KBQA Generalization Through Schema-Guided Logical Form Generation
di: Gao, Shengxiang, et al.
Pubblicazione: (2025)
di: Gao, Shengxiang, et al.
Pubblicazione: (2025)
Predicting Sentence Acceptability Judgments in Multimodal Contexts
di: Jang, Hyewon, et al.
Pubblicazione: (2026)
di: Jang, Hyewon, et al.
Pubblicazione: (2026)
Interaction Matters: An Evaluation Framework for Interactive Dialogue Assessment on English Second Language Conversations
di: Gao, Rena, et al.
Pubblicazione: (2024)
di: Gao, Rena, et al.
Pubblicazione: (2024)
WHoW: A Cross-domain Approach for Analysing Conversation Moderation
di: Chen, Ming-Bin, et al.
Pubblicazione: (2024)
di: Chen, Ming-Bin, et al.
Pubblicazione: (2024)
A Comparative Study on Annotation Quality of Crowdsourcing and LLM via Label Aggregation
di: Li, Jiyi
Pubblicazione: (2024)
di: Li, Jiyi
Pubblicazione: (2024)
Who Wrote the Book? Detecting and Attributing LLM Ghostwriters
di: Shetty, Anudeex, et al.
Pubblicazione: (2026)
di: Shetty, Anudeex, et al.
Pubblicazione: (2026)
WET: Overcoming Paraphrasing Vulnerabilities in Embeddings-as-a-Service with Linear Transformation Watermarks
di: Shetty, Anudeex, et al.
Pubblicazione: (2024)
di: Shetty, Anudeex, et al.
Pubblicazione: (2024)
Generating bilingual example sentences with large language models as lexicography assistants
di: Merx, Raphael, et al.
Pubblicazione: (2024)
di: Merx, Raphael, et al.
Pubblicazione: (2024)
Context Volume Drives Performance: Tackling Domain Shift in Extremely Low-Resource Translation via RAG
di: Setiawan, David Samuel, et al.
Pubblicazione: (2026)
di: Setiawan, David Samuel, et al.
Pubblicazione: (2026)
Structured RAG for Answering Aggregative Questions
di: Koshorek, Omri, et al.
Pubblicazione: (2025)
di: Koshorek, Omri, et al.
Pubblicazione: (2025)
Location Aware Modular Biencoder for Tourism Question Answering
di: Li, Haonan, et al.
Pubblicazione: (2024)
di: Li, Haonan, et al.
Pubblicazione: (2024)
LLMs as Span Annotators: A Comparative Study of LLMs and Humans
di: Kasner, Zdeněk, et al.
Pubblicazione: (2025)
di: Kasner, Zdeněk, et al.
Pubblicazione: (2025)
Moderation Matters:Measuring Conversational Moderation Impact in English as a Second Language Group Discussion
di: Gao, Rena, et al.
Pubblicazione: (2025)
di: Gao, Rena, et al.
Pubblicazione: (2025)
Factual Dialogue Summarization via Learning from Large Language Models
di: Zhu, Rongxin, et al.
Pubblicazione: (2024)
di: Zhu, Rongxin, et al.
Pubblicazione: (2024)
Decomposed Opinion Summarization with Verified Aspect-Aware Modules
di: Li, Miao, et al.
Pubblicazione: (2025)
di: Li, Miao, et al.
Pubblicazione: (2025)
Human-LLM Hybrid Text Answer Aggregation for Crowd Annotations
di: Li, Jiyi
Pubblicazione: (2024)
di: Li, Jiyi
Pubblicazione: (2024)
Aggregation Artifacts in Subjective Tasks Collapse Large Language Models' Posteriors
di: Chochlakis, Georgios, et al.
Pubblicazione: (2024)
di: Chochlakis, Georgios, et al.
Pubblicazione: (2024)
Inclusion-of-Thoughts: Mitigating Preference Instability via Purifying the Decision Space
di: Madani, Mohammad Reza Ghasemi, et al.
Pubblicazione: (2026)
di: Madani, Mohammad Reza Ghasemi, et al.
Pubblicazione: (2026)
When Reasoning Meets Information Aggregation: A Case Study with Sports Narratives
di: Hu, Yebowen, et al.
Pubblicazione: (2024)
di: Hu, Yebowen, et al.
Pubblicazione: (2024)
CIG: Measuring Conversational Information Gain in Deliberative Dialogues with Semantic Memory Dynamics
di: Chen, Ming-Bin, et al.
Pubblicazione: (2026)
di: Chen, Ming-Bin, et al.
Pubblicazione: (2026)
Aggregating Soft Labels from Crowd Annotations Improves Uncertainty Estimation Under Distribution Shift
di: Wright, Dustin, et al.
Pubblicazione: (2022)
di: Wright, Dustin, et al.
Pubblicazione: (2022)
From Chat Logs to Collective Insights: Aggregative Question Answering
di: Zhang, Wentao, et al.
Pubblicazione: (2025)
di: Zhang, Wentao, et al.
Pubblicazione: (2025)
SkillAggregation: Reference-free LLM-Dependent Aggregation
di: Sun, Guangzhi, et al.
Pubblicazione: (2024)
di: Sun, Guangzhi, et al.
Pubblicazione: (2024)
A Little Leak Will Sink a Great Ship: Survey of Transparency for Large Language Models from Start to Finish
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
di: Kaneko, Masahiro, et al.
Pubblicazione: (2024)
Psychometric Predictive Power of Large Language Models
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2023)
di: Kuribayashi, Tatsuki, et al.
Pubblicazione: (2023)
CommonMorph: Participatory Morphological Documentation Platform
di: Mahmudi, Aso, et al.
Pubblicazione: (2026)
di: Mahmudi, Aso, et al.
Pubblicazione: (2026)
An Interpretable and Crosslingual Method for Evaluating Second-Language Dialogues
di: Gao, Rena, et al.
Pubblicazione: (2024)
di: Gao, Rena, et al.
Pubblicazione: (2024)
Controlling Distributional Bias in Multi-Round LLM Generation via KL-Optimized Fine-Tuning
di: Jiang, Yanbei, et al.
Pubblicazione: (2026)
di: Jiang, Yanbei, et al.
Pubblicazione: (2026)
Error Span Annotation: A Balanced Approach for Human Evaluation of Machine Translation
di: Kocmi, Tom, et al.
Pubblicazione: (2024)
di: Kocmi, Tom, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Training and Evaluating with Human Label Variation: An Empirical Study
di: Kurniawan, Kemal, et al.
Pubblicazione: (2025) -
On the Interplay between Human Label Variation and Model Fairness
di: Kurniawan, Kemal, et al.
Pubblicazione: (2025) -
MoDEM: Mixture of Domain Expert Models
di: Simonds, Toby, et al.
Pubblicazione: (2024) -
Evaluating Evidence Attribution in Generated Fact Checking Explanations
di: Xing, Rui, et al.
Pubblicazione: (2024) -
A Joint Multitask Model for Morpho-Syntactic Parsing
di: Inostroza, Demian, et al.
Pubblicazione: (2025)