Comparing Humans and Models on a Similar Scale: Towards Cognitive Gender Bias Evaluation in Coreference Resolution
Fuente:
arXiv
Saved in:
| Main Authors: | Lior, Gili, Stanovsky, Gabriel |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
by: Lior, Gili, et al.
Published: (2025)
by: Lior, Gili, et al.
Published: (2025)
Leveraging Collection-Wide Similarities for Unsupervised Document Structure Extraction
by: Lior, Gili, et al.
Published: (2024)
by: Lior, Gili, et al.
Published: (2024)
PromptSuite: A Task-Agnostic Framework for Multi-Prompt Generation
by: Habba, Eliya, et al.
Published: (2025)
by: Habba, Eliya, et al.
Published: (2025)
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments
by: Lior, Gili, et al.
Published: (2025)
by: Lior, Gili, et al.
Published: (2025)
SEAM: A Stochastic Benchmark for Multi-Document Tasks
by: Lior, Gili, et al.
Published: (2024)
by: Lior, Gili, et al.
Published: (2024)
Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition
by: Goldstein, Ariel, et al.
Published: (2024)
by: Goldstein, Ariel, et al.
Published: (2024)
BOOKCOREF: Coreference Resolution at Book Scale
by: Martinelli, Giuliano, et al.
Published: (2025)
by: Martinelli, Giuliano, et al.
Published: (2025)
Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation
by: Conti, Lina, et al.
Published: (2025)
by: Conti, Lina, et al.
Published: (2025)
Inferring Scientific Cross-Document Coreference and Hierarchy with Definition-Augmented Relational Reasoning
by: Forer, Lior, et al.
Published: (2024)
by: Forer, Lior, et al.
Published: (2024)
Challenges to Evaluating the Generalization of Coreference Resolution Models: A Measurement Modeling Perspective
by: Porada, Ian, et al.
Published: (2023)
by: Porada, Ian, et al.
Published: (2023)
A Controlled Reevaluation of Coreference Resolution Models
by: Porada, Ian, et al.
Published: (2024)
by: Porada, Ian, et al.
Published: (2024)
Investigating Intersectional Bias in Large Language Models using Confidence Disparities in Coreference Resolution
by: Khan, Falaah Arif, et al.
Published: (2025)
by: Khan, Falaah Arif, et al.
Published: (2025)
Interpretable Coreference Resolution Evaluation Using Explicit Semantics
by: Gatti, Bruno, et al.
Published: (2026)
by: Gatti, Bruno, et al.
Published: (2026)
Coreference Resolution for Vietnamese Narrative Texts
by: Tran, Hieu-Dai, et al.
Published: (2025)
by: Tran, Hieu-Dai, et al.
Published: (2025)
Improving LLMs' Learning for Coreference Resolution
by: Gan, Yujian, et al.
Published: (2025)
by: Gan, Yujian, et al.
Published: (2025)
The State and Fate of Summarization Datasets: A Survey
by: Dahan, Noam, et al.
Published: (2024)
by: Dahan, Noam, et al.
Published: (2024)
Planted in Pretraining, Swayed by Finetuning: A Case Study on the Origins of Cognitive Biases in LLMs
by: Itzhak, Itay, et al.
Published: (2025)
by: Itzhak, Itay, et al.
Published: (2025)
ThaiCoref: Thai Coreference Resolution Dataset
by: Trakuekul, Pontakorn, et al.
Published: (2024)
by: Trakuekul, Pontakorn, et al.
Published: (2024)
CorPipe at CRAC 2025: Evaluating Multilingual Encoders for Multilingual Coreference Resolution
by: Straka, Milan
Published: (2025)
by: Straka, Milan
Published: (2025)
DOVE: A Large-Scale Multi-Dimensional Predictions Dataset Towards Meaningful LLM Evaluation
by: Habba, Eliya, et al.
Published: (2025)
by: Habba, Eliya, et al.
Published: (2025)
Argument-Centric Causal Intervention Method for Mitigating Bias in Cross-Document Event Coreference Resolution
by: Yao, Long, et al.
Published: (2025)
by: Yao, Long, et al.
Published: (2025)
EasyECR: A Library for Easy Implementation and Evaluation of Event Coreference Resolution Models
by: Li, Yuncong, et al.
Published: (2024)
by: Li, Yuncong, et al.
Published: (2024)
Anticipatory Evaluation of Language Models
by: Park, Jungsoo, et al.
Published: (2025)
by: Park, Jungsoo, et al.
Published: (2025)
Light Coreference Resolution for Russian with Hierarchical Discourse Features
by: Chistova, Elena, et al.
Published: (2023)
by: Chistova, Elena, et al.
Published: (2023)
Findings of the Third Shared Task on Multilingual Coreference Resolution
by: Novák, Michal, et al.
Published: (2024)
by: Novák, Michal, et al.
Published: (2024)
Reverse Probing: Evaluating Knowledge Transfer via Finetuned Task Embeddings for Coreference Resolution
by: Anikina, Tatiana, et al.
Published: (2025)
by: Anikina, Tatiana, et al.
Published: (2025)
End-to-End Dialog Neural Coreference Resolution: Balancing Efficiency and Accuracy in Large-Scale Systems
by: Dong, Zhang, et al.
Published: (2025)
by: Dong, Zhang, et al.
Published: (2025)
In-Context Learning on a Budget: A Case Study in Token Classification
by: Berger, Uri, et al.
Published: (2024)
by: Berger, Uri, et al.
Published: (2024)
ÚFAL CorPipe at CRAC 2022: Effectivity of Multilingual Models for Coreference Resolution
by: Straka, Milan, et al.
Published: (2022)
by: Straka, Milan, et al.
Published: (2022)
Improving Image Captioning by Mimicking Human Reformulation Feedback at Inference-time
by: Berger, Uri, et al.
Published: (2025)
by: Berger, Uri, et al.
Published: (2025)
Evaluating Gender Bias in Large Language Models
by: Döll, Michael, et al.
Published: (2024)
by: Döll, Michael, et al.
Published: (2024)
Cross-Document Contextual Coreference Resolution in Knowledge Graphs
by: Dong, Zhang, et al.
Published: (2025)
by: Dong, Zhang, et al.
Published: (2025)
CorefInst: Leveraging LLMs for Multilingual Coreference Resolution
by: Arslan, Tuğba Pamay, et al.
Published: (2025)
by: Arslan, Tuğba Pamay, et al.
Published: (2025)
BioCoref: Benchmarking Biomedical Coreference Resolution with LLMs
by: Salem, Nourah M, et al.
Published: (2025)
by: Salem, Nourah M, et al.
Published: (2025)
Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias
by: Itzhak, Itay, et al.
Published: (2023)
by: Itzhak, Itay, et al.
Published: (2023)
Leveraging Digitized Newspapers to Collect Summarization Data in Low-Resource Languages
by: Dahan, Noam, et al.
Published: (2025)
by: Dahan, Noam, et al.
Published: (2025)
Beyond Memorization: Distinguishing between Reductive and Epistemic Reasoning in LLMs using Classic Logic Puzzles
by: Gabay, Adi, et al.
Published: (2026)
by: Gabay, Adi, et al.
Published: (2026)
Efficient Seq2seq Coreference Resolution Using Entity Representations
by: Grenander, Matt, et al.
Published: (2025)
by: Grenander, Matt, et al.
Published: (2025)
Surveying the Landscape of Image Captioning Evaluation: A Comprehensive Taxonomy, Trends and Metrics Analysis
by: Berger, Uri, et al.
Published: (2024)
by: Berger, Uri, et al.
Published: (2024)
Multilingual Coreference Resolution in Low-resource South Asian Languages
by: Mishra, Ritwik, et al.
Published: (2024)
by: Mishra, Ritwik, et al.
Published: (2024)
Similar Items
-
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
by: Lior, Gili, et al.
Published: (2025) -
Leveraging Collection-Wide Similarities for Unsupervised Document Structure Extraction
by: Lior, Gili, et al.
Published: (2024) -
PromptSuite: A Task-Agnostic Framework for Multi-Prompt Generation
by: Habba, Eliya, et al.
Published: (2025) -
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments
by: Lior, Gili, et al.
Published: (2025) -
SEAM: A Stochastic Benchmark for Multi-Document Tasks
by: Lior, Gili, et al.
Published: (2024)