On the Interplay between Human Label Variation and Model Fairness
Fuente:
arXiv
Saved in:
| Main Authors: | Kurniawan, Kemal, Mistica, Meladel, Baldwin, Timothy, Lau, Jey Han |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Training and Evaluating with Human Label Variation: An Empirical Study
by: Kurniawan, Kemal, et al.
Published: (2025)
by: Kurniawan, Kemal, et al.
Published: (2025)
To Aggregate or Not to Aggregate. That is the Question: A Case Study on Annotation Subjectivity in Span Prediction
by: Kurniawan, Kemal, et al.
Published: (2024)
by: Kurniawan, Kemal, et al.
Published: (2024)
MoDEM: Mixture of Domain Expert Models
by: Simonds, Toby, et al.
Published: (2024)
by: Simonds, Toby, et al.
Published: (2024)
A Joint Multitask Model for Morpho-Syntactic Parsing
by: Inostroza, Demian, et al.
Published: (2025)
by: Inostroza, Demian, et al.
Published: (2025)
Evaluating Evidence Attribution in Generated Fact Checking Explanations
by: Xing, Rui, et al.
Published: (2024)
by: Xing, Rui, et al.
Published: (2024)
COMMUNITYNOTES: A Dataset for Exploring the Helpfulness of Fact-Checking Explanations
by: Xing, Rui, et al.
Published: (2025)
by: Xing, Rui, et al.
Published: (2025)
An Analytical Emotion Framework of Rumour Threads on Social Media
by: Xing, Rui, et al.
Published: (2025)
by: Xing, Rui, et al.
Published: (2025)
Robustness of Neurosymbolic Reasoners on First-Order Logic Problems
by: Bansal, Hannah, et al.
Published: (2025)
by: Bansal, Hannah, et al.
Published: (2025)
Beyond Seen Data: Improving KBQA Generalization Through Schema-Guided Logical Form Generation
by: Gao, Shengxiang, et al.
Published: (2025)
by: Gao, Shengxiang, et al.
Published: (2025)
CMA-R:Causal Mediation Analysis for Explaining Rumour Detection
by: Tian, Lin, et al.
Published: (2024)
by: Tian, Lin, et al.
Published: (2024)
A Sentiment Consolidation Framework for Meta-Review Generation
by: Li, Miao, et al.
Published: (2024)
by: Li, Miao, et al.
Published: (2024)
Factual Dialogue Summarization via Learning from Large Language Models
by: Zhu, Rongxin, et al.
Published: (2024)
by: Zhu, Rongxin, et al.
Published: (2024)
Interaction Matters: An Evaluation Framework for Interactive Dialogue Assessment on English Second Language Conversations
by: Gao, Rena, et al.
Published: (2024)
by: Gao, Rena, et al.
Published: (2024)
Who Wrote the Book? Detecting and Attributing LLM Ghostwriters
by: Shetty, Anudeex, et al.
Published: (2026)
by: Shetty, Anudeex, et al.
Published: (2026)
WET: Overcoming Paraphrasing Vulnerabilities in Embeddings-as-a-Service with Linear Transformation Watermarks
by: Shetty, Anudeex, et al.
Published: (2024)
by: Shetty, Anudeex, et al.
Published: (2024)
Generating bilingual example sentences with large language models as lexicography assistants
by: Merx, Raphael, et al.
Published: (2024)
by: Merx, Raphael, et al.
Published: (2024)
Context Volume Drives Performance: Tackling Domain Shift in Extremely Low-Resource Translation via RAG
by: Setiawan, David Samuel, et al.
Published: (2026)
by: Setiawan, David Samuel, et al.
Published: (2026)
WHoW: A Cross-domain Approach for Analysing Conversation Moderation
by: Chen, Ming-Bin, et al.
Published: (2024)
by: Chen, Ming-Bin, et al.
Published: (2024)
Moderation Matters:Measuring Conversational Moderation Impact in English as a Second Language Group Discussion
by: Gao, Rena, et al.
Published: (2025)
by: Gao, Rena, et al.
Published: (2025)
Decomposed Opinion Summarization with Verified Aspect-Aware Modules
by: Li, Miao, et al.
Published: (2025)
by: Li, Miao, et al.
Published: (2025)
Inclusion-of-Thoughts: Mitigating Preference Instability via Purifying the Decision Space
by: Madani, Mohammad Reza Ghasemi, et al.
Published: (2026)
by: Madani, Mohammad Reza Ghasemi, et al.
Published: (2026)
Inference-Time Selective Debiasing to Enhance Fairness in Text Classification Models
by: Kuzmin, Gleb, et al.
Published: (2024)
by: Kuzmin, Gleb, et al.
Published: (2024)
Predicting Sentence Acceptability Judgments in Multimodal Contexts
by: Jang, Hyewon, et al.
Published: (2026)
by: Jang, Hyewon, et al.
Published: (2026)
The Gaps between Pre-train and Downstream Settings in Bias Evaluation and Debiasing
by: Kaneko, Masahiro, et al.
Published: (2024)
by: Kaneko, Masahiro, et al.
Published: (2024)
CIG: Measuring Conversational Information Gain in Deliberative Dialogues with Semantic Memory Dynamics
by: Chen, Ming-Bin, et al.
Published: (2026)
by: Chen, Ming-Bin, et al.
Published: (2026)
Interpreting Predictive Probabilities: Model Confidence or Human Label Variation?
by: Baan, Joris, et al.
Published: (2024)
by: Baan, Joris, et al.
Published: (2024)
A Little Leak Will Sink a Great Ship: Survey of Transparency for Large Language Models from Start to Finish
by: Kaneko, Masahiro, et al.
Published: (2024)
by: Kaneko, Masahiro, et al.
Published: (2024)
Large Language Models Are Human-Like Internally
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
by: Kuribayashi, Tatsuki, et al.
Published: (2025)
Benchmarking Gender and Political Bias in Large Language Models
by: Yang, Jinrui, et al.
Published: (2025)
by: Yang, Jinrui, et al.
Published: (2025)
CommonMorph: Participatory Morphological Documentation Platform
by: Mahmudi, Aso, et al.
Published: (2026)
by: Mahmudi, Aso, et al.
Published: (2026)
Fine-grained Fallacy Detection with Human Label Variation
by: Ramponi, Alan, et al.
Published: (2025)
by: Ramponi, Alan, et al.
Published: (2025)
Human Label Variation in Implicit Discourse Relation Recognition
by: Yung, Frances, et al.
Published: (2026)
by: Yung, Frances, et al.
Published: (2026)
Controlling Distributional Bias in Multi-Round LLM Generation via KL-Optimized Fine-Tuning
by: Jiang, Yanbei, et al.
Published: (2026)
by: Jiang, Yanbei, et al.
Published: (2026)
An Interpretable and Crosslingual Method for Evaluating Second-Language Dialogues
by: Gao, Rena, et al.
Published: (2024)
by: Gao, Rena, et al.
Published: (2024)
The Ecological Fallacy in Annotation: Modelling Human Label Variation goes beyond Sociodemographics
by: Orlikowski, Matthias, et al.
Published: (2023)
by: Orlikowski, Matthias, et al.
Published: (2023)
Does Vision Accelerate Hierarchical Generalization in Neural Language Learners?
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
On the Interplay of Human-AI Alignment,Fairness, and Performance Trade-offs in Medical Imaging
by: Luo, Haozhe, et al.
Published: (2025)
by: Luo, Haozhe, et al.
Published: (2025)
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models
by: Kaneko, Masahiro, et al.
Published: (2025)
by: Kaneko, Masahiro, et al.
Published: (2025)
Psychometric Predictive Power of Large Language Models
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
by: Kuribayashi, Tatsuki, et al.
Published: (2023)
Revisiting Active Learning under (Human) Label Variation
by: Gruber, Cornelia, et al.
Published: (2025)
by: Gruber, Cornelia, et al.
Published: (2025)
Similar Items
-
Training and Evaluating with Human Label Variation: An Empirical Study
by: Kurniawan, Kemal, et al.
Published: (2025) -
To Aggregate or Not to Aggregate. That is the Question: A Case Study on Annotation Subjectivity in Span Prediction
by: Kurniawan, Kemal, et al.
Published: (2024) -
MoDEM: Mixture of Domain Expert Models
by: Simonds, Toby, et al.
Published: (2024) -
A Joint Multitask Model for Morpho-Syntactic Parsing
by: Inostroza, Demian, et al.
Published: (2025) -
Evaluating Evidence Attribution in Generated Fact Checking Explanations
by: Xing, Rui, et al.
Published: (2024)