Exploring the Influence of Label Aggregation on Minority Voices: Implications for Dataset Bias and Model Training
Fuente:
arXiv
Salvato in:
| Autori principali: | Pandya, Mugdha, Moosavi, Nafise Sadat, Maynard, Diana |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hostility Detection in UK Politics: A Dataset on Online Abuse Targeting MPs
di: Pandya, Mugdha, et al.
Pubblicazione: (2024)
di: Pandya, Mugdha, et al.
Pubblicazione: (2024)
More or Less Wrong: A Benchmark for Directional Bias in LLM Comparative Reasoning
di: Shafiei, Mohammadamin, et al.
Pubblicazione: (2025)
di: Shafiei, Mohammadamin, et al.
Pubblicazione: (2025)
MultiHoax: A Dataset of Multi-hop False-Premise Questions
di: Shafiei, Mohammadamin, et al.
Pubblicazione: (2025)
di: Shafiei, Mohammadamin, et al.
Pubblicazione: (2025)
From Input Perception to Predictive Insight: Modeling Model Blind Spots Before They Become Errors
di: Mi, Maggie, et al.
Pubblicazione: (2025)
di: Mi, Maggie, et al.
Pubblicazione: (2025)
How to Leverage Digit Embeddings to Represent Numbers?
di: Sivakumar, Jasivan Alex, et al.
Pubblicazione: (2024)
di: Sivakumar, Jasivan Alex, et al.
Pubblicazione: (2024)
Faithful Summarisation under Disagreement via Belief-Level Aggregation
di: Aghaebe, Favour Yahdii, et al.
Pubblicazione: (2026)
di: Aghaebe, Favour Yahdii, et al.
Pubblicazione: (2026)
Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantization
di: Kennedy, Ian W., et al.
Pubblicazione: (2026)
di: Kennedy, Ian W., et al.
Pubblicazione: (2026)
Deconstructing Attention: Investigating Design Principles for Effective Language Modeling
di: Xue, Huiyin, et al.
Pubblicazione: (2025)
di: Xue, Huiyin, et al.
Pubblicazione: (2025)
LLMs Do Not See Age: Assessing Demographic Bias in Automated Systematic Review Synthesis
di: Aghaebe, Favour Yahdii, et al.
Pubblicazione: (2025)
di: Aghaebe, Favour Yahdii, et al.
Pubblicazione: (2025)
Rolling the DICE on Idiomaticity: How LLMs Fail to Grasp Context
di: Mi, Maggie, et al.
Pubblicazione: (2024)
di: Mi, Maggie, et al.
Pubblicazione: (2024)
LLMs as Narcissistic Evaluators: When Ego Inflates Evaluation Scores
di: Liu, Yiqi, et al.
Pubblicazione: (2023)
di: Liu, Yiqi, et al.
Pubblicazione: (2023)
Decoding News Narratives: A Critical Analysis of Large Language Models in Framing Detection
di: Pastorino, Valeria, et al.
Pubblicazione: (2024)
di: Pastorino, Valeria, et al.
Pubblicazione: (2024)
Exploring Gender Disparities in Automatic Speech Recognition Technology
di: ElGhazaly, Hend, et al.
Pubblicazione: (2025)
di: ElGhazaly, Hend, et al.
Pubblicazione: (2025)
No Shortcuts to Culture: Indonesian Multi-hop Question Answering for Complex Cultural Understanding
di: Permadi, Vynska Amalia, et al.
Pubblicazione: (2026)
di: Permadi, Vynska Amalia, et al.
Pubblicazione: (2026)
Beyond Hate Speech: NLP's Challenges and Opportunities in Uncovering Dehumanizing Language
di: Saffari, Hamidreza, et al.
Pubblicazione: (2024)
di: Saffari, Hamidreza, et al.
Pubblicazione: (2024)
RIGOURATE: Quantifying Scientific Exaggeration with Evidence-Aligned Claim Evaluation
di: James, Joseph, et al.
Pubblicazione: (2026)
di: James, Joseph, et al.
Pubblicazione: (2026)
Hidden Failures in Robustness: Why Supervised Uncertainty Quantification Needs Better Evaluation
di: Stacey, Joe, et al.
Pubblicazione: (2026)
di: Stacey, Joe, et al.
Pubblicazione: (2026)
ContrastScore: Towards Higher Quality, Less Biased, More Efficient Evaluation Metrics with Contrastive Evaluation
di: Wang, Xiao, et al.
Pubblicazione: (2025)
di: Wang, Xiao, et al.
Pubblicazione: (2025)
Detecting Bias and Enhancing Diagnostic Accuracy in Large Language Models for Healthcare
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2024)
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2024)
Analyzing Bias in Swiss Federal Supreme Court Judgments Using Facebook's Holistic Bias Dataset: Implications for Language Model Training
di: Wehnert, Sabine, et al.
Pubblicazione: (2025)
di: Wehnert, Sabine, et al.
Pubblicazione: (2025)
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2025)
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2025)
FAIR Enough: How Can We Develop and Assess a FAIR-Compliant Dataset for Large Language Models' Training?
di: Raza, Shaina, et al.
Pubblicazione: (2024)
di: Raza, Shaina, et al.
Pubblicazione: (2024)
Sonos Voice Control Bias Assessment Dataset: A Methodology for Demographic Bias Assessment in Voice Assistants
di: Sekkat, Chloé, et al.
Pubblicazione: (2024)
di: Sekkat, Chloé, et al.
Pubblicazione: (2024)
I Am Aligned, But With Whom? MENA Values Benchmark for Evaluating Cultural Alignment and Multilingual Bias in LLMs
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2025)
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2025)
Mitigating Label Length Bias in Large Language Models
di: Sanz-Guerrero, Mario, et al.
Pubblicazione: (2025)
di: Sanz-Guerrero, Mario, et al.
Pubblicazione: (2025)
The Promises and Pitfalls of LLM Annotations in Dataset Labeling: a Case Study on Media Bias Detection
di: Horych, Tomas, et al.
Pubblicazione: (2024)
di: Horych, Tomas, et al.
Pubblicazione: (2024)
VoiceBBQ: Investigating Effect of Content and Acoustics in Social Bias of Spoken Language Model
di: Choi, Junhyuk, et al.
Pubblicazione: (2025)
di: Choi, Junhyuk, et al.
Pubblicazione: (2025)
Transforming Science with Large Language Models: A Survey on AI-assisted Scientific Discovery, Experimentation, Content Generation, and Evaluation
di: Eger, Steffen, et al.
Pubblicazione: (2025)
di: Eger, Steffen, et al.
Pubblicazione: (2025)
Thinking into the Future: Latent Lookahead Training for Transformers
di: Noci, Lorenzo, et al.
Pubblicazione: (2026)
di: Noci, Lorenzo, et al.
Pubblicazione: (2026)
RuBia: A Russian Language Bias Detection Dataset
di: Grigoreva, Veronika, et al.
Pubblicazione: (2024)
di: Grigoreva, Veronika, et al.
Pubblicazione: (2024)
Race, Ethnicity and Their Implication on Bias in Large Language Models
di: Hu, Shiyue, et al.
Pubblicazione: (2026)
di: Hu, Shiyue, et al.
Pubblicazione: (2026)
Balanced Aggregation: Understanding and Fixing Aggregation Bias in GRPO
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2026)
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2026)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
di: Ma, Mingyu Derek, et al.
Pubblicazione: (2023)
di: Ma, Mingyu Derek, et al.
Pubblicazione: (2023)
Finding A Voice: Exploring the Potential of African American Dialect and Voice Generation for Chatbots
di: Finch, Sarah E., et al.
Pubblicazione: (2025)
di: Finch, Sarah E., et al.
Pubblicazione: (2025)
Beyond Performance: Quantifying and Mitigating Label Bias in LLMs
di: Reif, Yuval, et al.
Pubblicazione: (2024)
di: Reif, Yuval, et al.
Pubblicazione: (2024)
Gender Bias in Large Language Models for Healthcare: Assignment Consistency and Clinical Implications
di: Liu, Mingxuan, et al.
Pubblicazione: (2025)
di: Liu, Mingxuan, et al.
Pubblicazione: (2025)
FRACTAL: Fine-Grained Scoring from Aggregate Text Labels
di: Makhija, Yukti, et al.
Pubblicazione: (2024)
di: Makhija, Yukti, et al.
Pubblicazione: (2024)
Exploring the Effects of Alignment on Numerical Bias in Large Language Models
di: Sato, Ayako, et al.
Pubblicazione: (2026)
di: Sato, Ayako, et al.
Pubblicazione: (2026)
Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation
di: Conti, Lina, et al.
Pubblicazione: (2025)
di: Conti, Lina, et al.
Pubblicazione: (2025)
Cross-Care: Assessing the Healthcare Implications of Pre-training Data on Language Model Bias
di: Chen, Shan, et al.
Pubblicazione: (2024)
di: Chen, Shan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Hostility Detection in UK Politics: A Dataset on Online Abuse Targeting MPs
di: Pandya, Mugdha, et al.
Pubblicazione: (2024) -
More or Less Wrong: A Benchmark for Directional Bias in LLM Comparative Reasoning
di: Shafiei, Mohammadamin, et al.
Pubblicazione: (2025) -
MultiHoax: A Dataset of Multi-hop False-Premise Questions
di: Shafiei, Mohammadamin, et al.
Pubblicazione: (2025) -
From Input Perception to Predictive Insight: Modeling Model Blind Spots Before They Become Errors
di: Mi, Maggie, et al.
Pubblicazione: (2025) -
How to Leverage Digit Embeddings to Represent Numbers?
di: Sivakumar, Jasivan Alex, et al.
Pubblicazione: (2024)