QueerBench: Quantifying Discrimination in Language Models Toward Queer Identities
Fuente:
arXiv
Salvato in:
| Autori principali: | Sosto, Mae, Barrón-Cedeño, Alberto |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
QueerGen: How LLMs Reflect Societal Norms on Gender and Sexuality in Sentence Completion Tasks
di: Sosto, Mae, et al.
Pubblicazione: (2026)
di: Sosto, Mae, et al.
Pubblicazione: (2026)
Unequal Voices: How LLMs Construct Constrained Queer Narratives
di: Ghosal, Atreya, et al.
Pubblicazione: (2025)
di: Ghosal, Atreya, et al.
Pubblicazione: (2025)
DarkBench: Benchmarking Dark Patterns in Large Language Models
di: Kran, Esben, et al.
Pubblicazione: (2025)
di: Kran, Esben, et al.
Pubblicazione: (2025)
CBT-Bench: Evaluating Large Language Models on Assisting Cognitive Behavior Therapy
di: Zhang, Mian, et al.
Pubblicazione: (2024)
di: Zhang, Mian, et al.
Pubblicazione: (2024)
Can Large Language Models Replace Human Coders? Introducing ContentBench
di: Haman, Michael
Pubblicazione: (2026)
di: Haman, Michael
Pubblicazione: (2026)
Information Suppression in Large Language Models: Auditing, Quantifying, and Characterizing Censorship in DeepSeek
di: Qiu, Peiran, et al.
Pubblicazione: (2025)
di: Qiu, Peiran, et al.
Pubblicazione: (2025)
Harmful Speech Detection by Language Models Exhibits Gender-Queer Dialect Bias
di: Dorn, Rebecca, et al.
Pubblicazione: (2024)
di: Dorn, Rebecca, et al.
Pubblicazione: (2024)
Nunchi-Bench: Benchmarking Language Models on Cultural Reasoning with a Focus on Korean Superstition
di: Kim, Kyuhee, et al.
Pubblicazione: (2025)
di: Kim, Kyuhee, et al.
Pubblicazione: (2025)
Quantifying Risk Propensities of Large Language Models: Ethical Focus and Bias Detection through Role-Play
di: Zeng, Yifan, et al.
Pubblicazione: (2024)
di: Zeng, Yifan, et al.
Pubblicazione: (2024)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
di: Sakhawat, Adib, et al.
Pubblicazione: (2026)
di: Sakhawat, Adib, et al.
Pubblicazione: (2026)
Whose Journey Matters? Investigating Identity Biases in Large Language Models (LLMs) for Travel Planning Assistance
di: Ren, Ruiping, et al.
Pubblicazione: (2024)
di: Ren, Ruiping, et al.
Pubblicazione: (2024)
WinoQueer: A Community-in-the-Loop Benchmark for Anti-LGBTQ+ Bias in Large Language Models
di: Felkner, Virginia K., et al.
Pubblicazione: (2023)
di: Felkner, Virginia K., et al.
Pubblicazione: (2023)
LocalValueBench: A Collaboratively Built and Extensible Benchmark for Evaluating Localized Value Alignment and Ethical Safety in Large Language Models
di: Meadows, Gwenyth Isobel, et al.
Pubblicazione: (2024)
di: Meadows, Gwenyth Isobel, et al.
Pubblicazione: (2024)
SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors
di: Hu, Tiancheng, et al.
Pubblicazione: (2025)
di: Hu, Tiancheng, et al.
Pubblicazione: (2025)
The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
di: Schmucker, Robin, et al.
Pubblicazione: (2025)
di: Schmucker, Robin, et al.
Pubblicazione: (2025)
HypoBench: Towards Systematic and Principled Benchmarking for Hypothesis Generation
di: Liu, Haokun, et al.
Pubblicazione: (2025)
di: Liu, Haokun, et al.
Pubblicazione: (2025)
The Colorful Future of LLMs: Evaluating and Improving LLMs as Emotional Supporters for Queer Youth
di: Lissak, Shir, et al.
Pubblicazione: (2024)
di: Lissak, Shir, et al.
Pubblicazione: (2024)
Simulating Students with Large Language Models: A Review of Architecture, Mechanisms, and Role Modelling in Education with Generative AI
di: Marquez-Carpintero, Luis, et al.
Pubblicazione: (2025)
di: Marquez-Carpintero, Luis, et al.
Pubblicazione: (2025)
Denevil: Towards Deciphering and Navigating the Ethical Values of Large Language Models via Instruction Learning
di: Duan, Shitong, et al.
Pubblicazione: (2023)
di: Duan, Shitong, et al.
Pubblicazione: (2023)
A Unified Framework to Quantify Cultural Intelligence of AI
di: Dev, Sunipa, et al.
Pubblicazione: (2026)
di: Dev, Sunipa, et al.
Pubblicazione: (2026)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
di: Xu, Xin, et al.
Pubblicazione: (2025)
di: Xu, Xin, et al.
Pubblicazione: (2025)
WorldView-Bench: A Benchmark for Evaluating Global Cultural Perspectives in Large Language Models
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
ASCenD-BDS: Adaptable, Stochastic and Context-aware framework for Detection of Bias, Discrimination and Stereotyping
di: Bahl, Rajiv, et al.
Pubblicazione: (2025)
di: Bahl, Rajiv, et al.
Pubblicazione: (2025)
LocalBench: Benchmarking LLMs on County-Level Local Knowledge and Reasoning
di: Gao, Zihan, et al.
Pubblicazione: (2025)
di: Gao, Zihan, et al.
Pubblicazione: (2025)
Queer NLP: A Critical Survey on Literature Gaps, Biases and Trends
di: Weber, Sabine, et al.
Pubblicazione: (2026)
di: Weber, Sabine, et al.
Pubblicazione: (2026)
XCR-Bench: A Multi-Task Benchmark for Evaluating Cultural Reasoning in LLMs
di: Kabir, Mohsinul, et al.
Pubblicazione: (2026)
di: Kabir, Mohsinul, et al.
Pubblicazione: (2026)
Computational Phenomenology of Temporal Experience in Autism: Quantifying the Emotional and Narrative Characteristics of Lived Unpredictability
di: Dudzic, Kacper, et al.
Pubblicazione: (2026)
di: Dudzic, Kacper, et al.
Pubblicazione: (2026)
LLM-Driven Robots Risk Enacting Discrimination, Violence, and Unlawful Actions
di: Hundt, Andrew, et al.
Pubblicazione: (2024)
di: Hundt, Andrew, et al.
Pubblicazione: (2024)
A Tale of Two Identities: An Ethical Audit of Human and AI-Crafted Personas
di: Venkit, Pranav Narayanan, et al.
Pubblicazione: (2025)
di: Venkit, Pranav Narayanan, et al.
Pubblicazione: (2025)
Towards Measuring and Modeling "Culture" in LLMs: A Survey
di: Adilazuarda, Muhammad Farid, et al.
Pubblicazione: (2024)
di: Adilazuarda, Muhammad Farid, et al.
Pubblicazione: (2024)
Cross-Language Bias Examination in Large Language Models
di: Liang, Yuxuan, et al.
Pubblicazione: (2025)
di: Liang, Yuxuan, et al.
Pubblicazione: (2025)
ClinBench-HPB: A Clinical Benchmark for Evaluating LLMs in Hepato-Pancreato-Biliary Diseases
di: Li, Yuchong, et al.
Pubblicazione: (2025)
di: Li, Yuchong, et al.
Pubblicazione: (2025)
RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity
di: Shin, Jisu, et al.
Pubblicazione: (2025)
di: Shin, Jisu, et al.
Pubblicazione: (2025)
PLawBench: A Rubric-Based Benchmark for Evaluating LLMs in Real-World Legal Practice
di: Shi, Yuzhen, et al.
Pubblicazione: (2026)
di: Shi, Yuzhen, et al.
Pubblicazione: (2026)
NLP Meets the World: Toward Improving Conversations With the Public About Natural Language Processing Research
di: Wilson, Shomir
Pubblicazione: (2025)
di: Wilson, Shomir
Pubblicazione: (2025)
On the Creativity of Large Language Models
di: Franceschelli, Giorgio, et al.
Pubblicazione: (2023)
di: Franceschelli, Giorgio, et al.
Pubblicazione: (2023)
Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench
di: Wang, Tianyu, et al.
Pubblicazione: (2026)
di: Wang, Tianyu, et al.
Pubblicazione: (2026)
SomaliBench Eval: Measuring English-to-Somali Refusal Gaps in Open-Weight Language Models
di: Dahir, Khalid Yusuf
Pubblicazione: (2026)
di: Dahir, Khalid Yusuf
Pubblicazione: (2026)
Towards Large Language Models that Benefit for All: Benchmarking Group Fairness in Reward Models
di: Song, Kefan, et al.
Pubblicazione: (2025)
di: Song, Kefan, et al.
Pubblicazione: (2025)
Open-Ended Wargames with Large Language Models
di: Hogan, Daniel P., et al.
Pubblicazione: (2024)
di: Hogan, Daniel P., et al.
Pubblicazione: (2024)
Documenti analoghi
-
QueerGen: How LLMs Reflect Societal Norms on Gender and Sexuality in Sentence Completion Tasks
di: Sosto, Mae, et al.
Pubblicazione: (2026) -
Unequal Voices: How LLMs Construct Constrained Queer Narratives
di: Ghosal, Atreya, et al.
Pubblicazione: (2025) -
DarkBench: Benchmarking Dark Patterns in Large Language Models
di: Kran, Esben, et al.
Pubblicazione: (2025) -
CBT-Bench: Evaluating Large Language Models on Assisting Cognitive Behavior Therapy
di: Zhang, Mian, et al.
Pubblicazione: (2024) -
Can Large Language Models Replace Human Coders? Introducing ContentBench
di: Haman, Michael
Pubblicazione: (2026)