Side-by-side Comparison Amplifies Dialect Bias in Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kondapally, Kritee, Smerdon, Claire J., Patel, Pooja C., Akoni, Ogheneyoma, Torres, Jevon, Ranjit, Jaspreet, Finlayson, Matthew, Swayamdipta, Swabha |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Logits of API-Protected LLMs Leak Proprietary Information
by: Finlayson, Matthew, et al.
Published: (2024)
by: Finlayson, Matthew, et al.
Published: (2024)
Better Language Model Inversion by Compactly Representing Next-Token Distributions
by: Nazir, Murtaza, et al.
Published: (2025)
by: Nazir, Murtaza, et al.
Published: (2025)
Every Language Model Has a Forgery-Resistant Signature
by: Finlayson, Matthew, et al.
Published: (2025)
by: Finlayson, Matthew, et al.
Published: (2025)
Teaching Models to Understand (but not Generate) High-risk Data
by: Wang, Ryan, et al.
Published: (2025)
by: Wang, Ryan, et al.
Published: (2025)
Are We Automating the Joy Out of Work? Designing AI to Augment Work, Not Meaning
by: Ranjit, Jaspreet, et al.
Published: (2026)
by: Ranjit, Jaspreet, et al.
Published: (2026)
Annotating FrameNet via Structure-Conditioned Language Generation
by: Cui, Xinyue, et al.
Published: (2024)
by: Cui, Xinyue, et al.
Published: (2024)
Uncovering Intervention Opportunities for Suicide Prevention with Language Model Assistants
by: Ranjit, Jaspreet, et al.
Published: (2025)
by: Ranjit, Jaspreet, et al.
Published: (2025)
How Reliable is Language Model Micro-Benchmarking?
by: Yauney, Gregory, et al.
Published: (2025)
by: Yauney, Gregory, et al.
Published: (2025)
OATH-Frames: Characterizing Online Attitudes Towards Homelessness with LLM Assistants
by: Ranjit, Jaspreet, et al.
Published: (2024)
by: Ranjit, Jaspreet, et al.
Published: (2024)
Disentangling Geometry, Performance, and Training in Language Models
by: Kulkarni, Atharva, et al.
Published: (2026)
by: Kulkarni, Atharva, et al.
Published: (2026)
Compare without Despair: Reliable Preference Evaluation with Generation Separability
by: Ghosh, Sayan, et al.
Published: (2024)
by: Ghosh, Sayan, et al.
Published: (2024)
Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
by: Ethayarajh, Kawin, et al.
Published: (2021)
by: Ethayarajh, Kawin, et al.
Published: (2021)
Improving Language Model Personas via Rationalization with Psychological Scaffolds
by: Joshi, Brihi, et al.
Published: (2025)
by: Joshi, Brihi, et al.
Published: (2025)
Robust Data Watermarking in Language Models by Injecting Fictitious Knowledge
by: Cui, Xinyue, et al.
Published: (2025)
by: Cui, Xinyue, et al.
Published: (2025)
Crowd-Calibrator: Can Annotator Disagreement Inform Calibration in Subjective Tasks?
by: Khurana, Urja, et al.
Published: (2024)
by: Khurana, Urja, et al.
Published: (2024)
Evaluation Under Imperfect Benchmarks and Ratings: A Case Study in Text Simplification
by: Liu, Joseph, et al.
Published: (2025)
by: Liu, Joseph, et al.
Published: (2025)
Generative Explanations for Program Synthesizers
by: Nazari, Amirmohammad, et al.
Published: (2024)
by: Nazari, Amirmohammad, et al.
Published: (2024)
BenchBrowser: Retrieving Evidence for Evaluating Benchmark Validity
by: Diddee, Harshita, et al.
Published: (2026)
by: Diddee, Harshita, et al.
Published: (2026)
Believing without Seeing: Quality Scores for Contextualizing Vision-Language Model Explanations
by: He, Keyu, et al.
Published: (2025)
by: He, Keyu, et al.
Published: (2025)
Sample, Align, Synthesize: Graph-Based Response Synthesis with ConGrs
by: Ghosh, Sayan, et al.
Published: (2025)
by: Ghosh, Sayan, et al.
Published: (2025)
ChEmREF: Evaluating Language Model Readiness for Chemical Emergency Response
by: Surana, Risha, et al.
Published: (2025)
by: Surana, Risha, et al.
Published: (2025)
ELI-Why: Evaluating the Pedagogical Utility of Language Model Explanations
by: Joshi, Brihi, et al.
Published: (2025)
by: Joshi, Brihi, et al.
Published: (2025)
NeuroComparatives: Neuro-Symbolic Distillation of Comparative Knowledge
by: Howard, Phillip, et al.
Published: (2023)
by: Howard, Phillip, et al.
Published: (2023)
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination
by: Fleisig, Eve, et al.
Published: (2024)
by: Fleisig, Eve, et al.
Published: (2024)
Harmful Speech Detection by Language Models Exhibits Gender-Queer Dialect Bias
by: Dorn, Rebecca, et al.
Published: (2024)
by: Dorn, Rebecca, et al.
Published: (2024)
DialUp! Modeling the Language Continuum by Adapting Models to Dialects and Dialects to Models
by: Bafna, Niyati, et al.
Published: (2025)
by: Bafna, Niyati, et al.
Published: (2025)
Steering Large Language Models to Evaluate and Amplify Creativity
by: Olson, Matthew Lyle, et al.
Published: (2024)
by: Olson, Matthew Lyle, et al.
Published: (2024)
Why Fine-Tuning Encourages Hallucinations and How to Fix It
by: Kaplan, Guy, et al.
Published: (2026)
by: Kaplan, Guy, et al.
Published: (2026)
Mitigating Harmful Erraticism in LLMs Through Dialectical Behavior Therapy Based De-Escalation Strategies
by: Rangarajan, Pooja, et al.
Published: (2025)
by: Rangarajan, Pooja, et al.
Published: (2025)
Dialect and Gender Bias in YouTube's Spanish Captioning System
by: Jimenez, Iris Dania, et al.
Published: (2026)
by: Jimenez, Iris Dania, et al.
Published: (2026)
DialectalArabicMMLU: Benchmarking Dialectal Capabilities in Arabic and Multilingual Language Models
by: Altakrori, Malik H., et al.
Published: (2025)
by: Altakrori, Malik H., et al.
Published: (2025)
Evaluating Evaluation Metrics -- The Mirage of Hallucination Detection
by: Kulkarni, Atharva, et al.
Published: (2025)
by: Kulkarni, Atharva, et al.
Published: (2025)
AlcLaM: Arabic Dialectal Language Model
by: Ahmed, Murtadha, et al.
Published: (2024)
by: Ahmed, Murtadha, et al.
Published: (2024)
Low-Resource Dialect Adaptation of Large Language Models: A French Dialect Case-Study
by: Khan, Eeham, et al.
Published: (2025)
by: Khan, Eeham, et al.
Published: (2025)
Disentangling Dialect from Social Bias via Multitask Learning to Improve Fairness
by: Spliethöver, Maximilian, et al.
Published: (2024)
by: Spliethöver, Maximilian, et al.
Published: (2024)
Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement
by: Xu, Wenda, et al.
Published: (2024)
by: Xu, Wenda, et al.
Published: (2024)
Images Amplify Misinformation Sharing in Vision-Language Models
by: Plebe, Alice, et al.
Published: (2025)
by: Plebe, Alice, et al.
Published: (2025)
Evaluating Dialect Robustness of Language Models via Conversation Understanding
by: Srirag, Dipankar, et al.
Published: (2024)
by: Srirag, Dipankar, et al.
Published: (2024)
Large Language Models Discriminate Against Speakers of German Dialects
by: Bui, Minh Duc, et al.
Published: (2025)
by: Bui, Minh Duc, et al.
Published: (2025)
Invisible Filters: Cultural Bias in Hiring Evaluations Using Large Language Models
by: Rao, Pooja S. B., et al.
Published: (2025)
by: Rao, Pooja S. B., et al.
Published: (2025)
Similar Items
-
Logits of API-Protected LLMs Leak Proprietary Information
by: Finlayson, Matthew, et al.
Published: (2024) -
Better Language Model Inversion by Compactly Representing Next-Token Distributions
by: Nazir, Murtaza, et al.
Published: (2025) -
Every Language Model Has a Forgery-Resistant Signature
by: Finlayson, Matthew, et al.
Published: (2025) -
Teaching Models to Understand (but not Generate) High-risk Data
by: Wang, Ryan, et al.
Published: (2025) -
Are We Automating the Joy Out of Work? Designing AI to Augment Work, Not Meaning
by: Ranjit, Jaspreet, et al.
Published: (2026)