Investigating Intersectional Bias in Large Language Models using Confidence Disparities in Coreference Resolution
Fuente:
arXiv
Saved in:
| Main Authors: | Khan, Falaah Arif, Sivakumar, Nivedha, Wang, Yinong Oliver, Metcalf, Katherine, Camacho, Cezanne, Theobald, Barry-John, Zappella, Luca, Apostoloff, Nicholas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
by: Wang, Yinong Oliver, et al.
Published: (2025)
by: Wang, Yinong Oliver, et al.
Published: (2025)
DSO: Direct Steering Optimization for Bias Mitigation
by: Paes, Lucas Monteiro, et al.
Published: (2025)
by: Paes, Lucas Monteiro, et al.
Published: (2025)
Fairness Dynamics During Training
by: Patel, Krishna, et al.
Published: (2025)
by: Patel, Krishna, et al.
Published: (2025)
Bias after Prompting: Persistent Discrimination in Large Language Models
by: Sivakumar, Nivedha, et al.
Published: (2025)
by: Sivakumar, Nivedha, et al.
Published: (2025)
Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models
by: Mackraz, Natalie, et al.
Published: (2024)
by: Mackraz, Natalie, et al.
Published: (2024)
Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models
by: Suau, Xavier, et al.
Published: (2024)
by: Suau, Xavier, et al.
Published: (2024)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
by: Metcalf, Katherine, et al.
Published: (2024)
by: Metcalf, Katherine, et al.
Published: (2024)
Aligning LLMs by Predicting Preferences from User Writing Samples
by: Aroca-Ouellette, Stéphane, et al.
Published: (2025)
by: Aroca-Ouellette, Stéphane, et al.
Published: (2025)
PREDICT: Preference Reasoning by Evaluating Decomposed preferences Inferred from Candidate Trajectories
by: Aroca-Ouellette, Stephane, et al.
Published: (2024)
by: Aroca-Ouellette, Stephane, et al.
Published: (2024)
Controlling Language and Diffusion Models by Transporting Activations
by: Rodriguez, Pau, et al.
Published: (2024)
by: Rodriguez, Pau, et al.
Published: (2024)
An Epistemic and Aleatoric Decomposition of Arbitrariness to Constrain the Set of Good Models
by: Khan, Falaah Arif, et al.
Published: (2023)
by: Khan, Falaah Arif, et al.
Published: (2023)
Still More Shades of Null: An Evaluation Suite for Responsible Missing Value Imputation
by: Khan, Falaah Arif, et al.
Published: (2024)
by: Khan, Falaah Arif, et al.
Published: (2024)
ExpertLens: Activation steering features are highly interpretable
by: Fedzechkina, Masha, et al.
Published: (2025)
by: Fedzechkina, Masha, et al.
Published: (2025)
Steering into New Embedding Spaces: Analyzing Cross-Lingual Alignment Induced by Model Interventions in Multilingual Language Models
by: Sundar, Anirudh, et al.
Published: (2025)
by: Sundar, Anirudh, et al.
Published: (2025)
STRUCTURAL FRAGMENTATION IN INDIAN ENVIRONMENTAL LAW: ENFORCEMENT DEFICITS AND THE IMPERATIVE FOR A UNIFIED ENVIRONMENTAL CODE
by: Joan Nivedha S
Published: (2026)
by: Joan Nivedha S
Published: (2026)
Coreference Resolution for Vietnamese Narrative Texts
by: Tran, Hieu-Dai, et al.
Published: (2025)
by: Tran, Hieu-Dai, et al.
Published: (2025)
BOOKCOREF: Coreference Resolution at Book Scale
by: Martinelli, Giuliano, et al.
Published: (2025)
by: Martinelli, Giuliano, et al.
Published: (2025)
Improving LLMs' Learning for Coreference Resolution
by: Gan, Yujian, et al.
Published: (2025)
by: Gan, Yujian, et al.
Published: (2025)
Detecting Anaphoricity and Antecedenthood for Coreference Resolution
by: Olga Uryupina
Published: (2009)
by: Olga Uryupina
Published: (2009)
Argument-Centric Causal Intervention Method for Mitigating Bias in Cross-Document Event Coreference Resolution
by: Yao, Long, et al.
Published: (2025)
by: Yao, Long, et al.
Published: (2025)
Feature Importance Disparities for Data Bias Investigations
by: Chang, Peter W., et al.
Published: (2023)
by: Chang, Peter W., et al.
Published: (2023)
We Are AI: Taking Control of Technology
by: Stoyanovich, Julia, et al.
Published: (2025)
by: Stoyanovich, Julia, et al.
Published: (2025)
Comparing Humans and Models on a Similar Scale: Towards Cognitive Gender Bias Evaluation in Coreference Resolution
by: Lior, Gili, et al.
Published: (2023)
by: Lior, Gili, et al.
Published: (2023)
ThaiCoref: Thai Coreference Resolution Dataset
by: Trakuekul, Pontakorn, et al.
Published: (2024)
by: Trakuekul, Pontakorn, et al.
Published: (2024)
A Controlled Reevaluation of Coreference Resolution Models
by: Porada, Ian, et al.
Published: (2024)
by: Porada, Ian, et al.
Published: (2024)
Alcohol Use Disparities Among Transgender and Nonbinary Adults: An Intersectional Investigation
by: Ryan C. Shorey, et al.
Published: (2025)
by: Ryan C. Shorey, et al.
Published: (2025)
Light Coreference Resolution for Russian with Hierarchical Discourse Features
by: Chistova, Elena, et al.
Published: (2023)
by: Chistova, Elena, et al.
Published: (2023)
Cross-Document Contextual Coreference Resolution in Knowledge Graphs
by: Dong, Zhang, et al.
Published: (2025)
by: Dong, Zhang, et al.
Published: (2025)
CorefInst: Leveraging LLMs for Multilingual Coreference Resolution
by: Arslan, Tuğba Pamay, et al.
Published: (2025)
by: Arslan, Tuğba Pamay, et al.
Published: (2025)
BioCoref: Benchmarking Biomedical Coreference Resolution with LLMs
by: Salem, Nourah M, et al.
Published: (2025)
by: Salem, Nourah M, et al.
Published: (2025)
Findings of the Third Shared Task on Multilingual Coreference Resolution
by: Novák, Michal, et al.
Published: (2024)
by: Novák, Michal, et al.
Published: (2024)
Interpretable Coreference Resolution Evaluation Using Explicit Semantics
by: Gatti, Bruno, et al.
Published: (2026)
by: Gatti, Bruno, et al.
Published: (2026)
Autophagy—from yeast to humans: Thirty years of molecular autophagy
by: Tassula Proikas‐Cezanne, et al.
Published: (2024)
by: Tassula Proikas‐Cezanne, et al.
Published: (2024)
Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation
by: Conti, Lina, et al.
Published: (2025)
by: Conti, Lina, et al.
Published: (2025)
Investigating Bias in LLM-Based Bias Detection: Disparities between LLMs and Human Perception
by: Lin, Luyang, et al.
Published: (2024)
by: Lin, Luyang, et al.
Published: (2024)
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
by: Danieli, Federico, et al.
Published: (2025)
by: Danieli, Federico, et al.
Published: (2025)
Multilingual Coreference Resolution in Low-resource South Asian Languages
by: Mishra, Ritwik, et al.
Published: (2024)
by: Mishra, Ritwik, et al.
Published: (2024)
Maverick: Efficient and Accurate Coreference Resolution Defying Recent Trends
by: Martinelli, Giuliano, et al.
Published: (2024)
by: Martinelli, Giuliano, et al.
Published: (2024)
Linear Cross-document Event Coreference Resolution with X-AMR
by: Ahmed, Shafiuddin Rehan, et al.
Published: (2024)
by: Ahmed, Shafiuddin Rehan, et al.
Published: (2024)
Major Entity Identification: A Generalizable Alternative to Coreference Resolution
by: Manikantan, Kawshik, et al.
Published: (2024)
by: Manikantan, Kawshik, et al.
Published: (2024)
Similar Items
-
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
by: Wang, Yinong Oliver, et al.
Published: (2025) -
DSO: Direct Steering Optimization for Bias Mitigation
by: Paes, Lucas Monteiro, et al.
Published: (2025) -
Fairness Dynamics During Training
by: Patel, Krishna, et al.
Published: (2025) -
Bias after Prompting: Persistent Discrimination in Large Language Models
by: Sivakumar, Nivedha, et al.
Published: (2025) -
Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models
by: Mackraz, Natalie, et al.
Published: (2024)