Ask LLMs Directly, "What shapes your bias?": Measuring Social Bias in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shin, Jisu, Song, Hoyun, Lee, Huije, Jeong, Soyeong, Park, Jong C. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Different Bias Under Different Criteria: Assessing Bias in LLMs with a Fact-Based Approach
von: Ko, Changgeon, et al.
Veröffentlicht: (2024)
von: Ko, Changgeon, et al.
Veröffentlicht: (2024)
Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation
von: Lee, Huije, et al.
Veröffentlicht: (2026)
von: Lee, Huije, et al.
Veröffentlicht: (2026)
Does Rationale Quality Matter? Enhancing Mental Disorder Detection via Selective Reasoning Distillation
von: Song, Hoyun, et al.
Veröffentlicht: (2025)
von: Song, Hoyun, et al.
Veröffentlicht: (2025)
Social Dynamics as Critical Vulnerabilities that Undermine Objective Decision-Making in LLM Collectives
von: Ko, Changgeon, et al.
Veröffentlicht: (2026)
von: Ko, Changgeon, et al.
Veröffentlicht: (2026)
Towards Effective Counter-Responses: Aligning Human Preferences with Strategies to Combat Online Trolling
von: Lee, Huije, et al.
Veröffentlicht: (2024)
von: Lee, Huije, et al.
Veröffentlicht: (2024)
Lossless Acceleration of Large Language Models with Hierarchical Drafting based on Temporal Locality in Speculative Decoding
von: Cho, Sukmin, et al.
Veröffentlicht: (2025)
von: Cho, Sukmin, et al.
Veröffentlicht: (2025)
RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
A Multi-Task Benchmark for Abusive Language Detection in Low-Resource Settings
von: Gaim, Fitsum, et al.
Veröffentlicht: (2025)
von: Gaim, Fitsum, et al.
Veröffentlicht: (2025)
EXIT: Context-Aware Extractive Compression for Enhancing Retrieval-Augmented Generation
von: Hwang, Taeho, et al.
Veröffentlicht: (2024)
von: Hwang, Taeho, et al.
Veröffentlicht: (2024)
Measuring Implicit Bias in Explicitly Unbiased Large Language Models
von: Bai, Xuechunzi, et al.
Veröffentlicht: (2024)
von: Bai, Xuechunzi, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models to Measure Gender Representation Bias in Gendered Language Corpora
von: Derner, Erik, et al.
Veröffentlicht: (2024)
von: Derner, Erik, et al.
Veröffentlicht: (2024)
LLMs are Biased Teachers: Evaluating LLM Bias in Personalized Education
von: Weissburg, Iain, et al.
Veröffentlicht: (2024)
von: Weissburg, Iain, et al.
Veröffentlicht: (2024)
Inference-Time Reasoning Selectively Reduces Implicit Social Bias in Large Language Models
von: Apsel, Molly, et al.
Veröffentlicht: (2026)
von: Apsel, Molly, et al.
Veröffentlicht: (2026)
Spotting Out-of-Character Behavior: Atomic-Level Evaluation of Persona Fidelity in Open-Ended Generation
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
Only a Little to the Left: A Theory-grounded Measure of Political Bias in Large Language Models
von: Faulborn, Mats, et al.
Veröffentlicht: (2025)
von: Faulborn, Mats, et al.
Veröffentlicht: (2025)
Source framing triggers systematic evaluation bias in Large Language Models
von: Germani, Federico, et al.
Veröffentlicht: (2025)
von: Germani, Federico, et al.
Veröffentlicht: (2025)
LIBRA: Measuring Bias of Large Language Model from a Local Context
von: Pang, Bo, et al.
Veröffentlicht: (2025)
von: Pang, Bo, et al.
Veröffentlicht: (2025)
The Language You Ask In: Language-Conditioned Ideological Divergence in LLM Analysis of Contested Political Documents
von: Smirnov, Oleg
Veröffentlicht: (2026)
von: Smirnov, Oleg
Veröffentlicht: (2026)
MentalBench: A DSM-Grounded Benchmark for Evaluating Psychiatric Diagnostic Capability of Large Language Models
von: Song, Hoyun, et al.
Veröffentlicht: (2026)
von: Song, Hoyun, et al.
Veröffentlicht: (2026)
Characterizing Selective Refusal Bias in Large Language Models
von: Khorramrouz, Adel, et al.
Veröffentlicht: (2025)
von: Khorramrouz, Adel, et al.
Veröffentlicht: (2025)
Gender Bias in Emotion Recognition by Large Language Models
von: Herbert, Maureen, et al.
Veröffentlicht: (2025)
von: Herbert, Maureen, et al.
Veröffentlicht: (2025)
Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs)
von: Neumann, Anna, et al.
Veröffentlicht: (2025)
von: Neumann, Anna, et al.
Veröffentlicht: (2025)
Adaptive-RAG: Learning to Adapt Retrieval-Augmented Large Language Models through Question Complexity
von: Jeong, Soyeong, et al.
Veröffentlicht: (2024)
von: Jeong, Soyeong, et al.
Veröffentlicht: (2024)
Protected group bias and stereotypes in Large Language Models
von: Kotek, Hadas, et al.
Veröffentlicht: (2024)
von: Kotek, Hadas, et al.
Veröffentlicht: (2024)
Self-Knowledge Distillation for Learning Ambiguity
von: Park, Hancheol, et al.
Veröffentlicht: (2024)
von: Park, Hancheol, et al.
Veröffentlicht: (2024)
What's in a Name? Auditing Large Language Models for Race and Gender Bias
von: Salinas, Alejandro, et al.
Veröffentlicht: (2024)
von: Salinas, Alejandro, et al.
Veröffentlicht: (2024)
Cross-Language Bias Examination in Large Language Models
von: Liang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Liang, Yuxuan, et al.
Veröffentlicht: (2025)
Assessing Judging Bias in Large Reasoning Models: An Empirical Study
von: Wang, Qian, et al.
Veröffentlicht: (2025)
von: Wang, Qian, et al.
Veröffentlicht: (2025)
Towards Equitable AI: Detecting Bias in Using Large Language Models for Marketing
von: Yilmaz, Berk, et al.
Veröffentlicht: (2025)
von: Yilmaz, Berk, et al.
Veröffentlicht: (2025)
Characterizing Bias: Benchmarking Large Language Models in Simplified versus Traditional Chinese
von: Lyu, Hanjia, et al.
Veröffentlicht: (2025)
von: Lyu, Hanjia, et al.
Veröffentlicht: (2025)
Unsupervised Concept Vector Extraction for Bias Control in LLMs
von: Cyberey, Hannah, et al.
Veröffentlicht: (2025)
von: Cyberey, Hannah, et al.
Veröffentlicht: (2025)
A Comprehensive Survey of Bias in LLMs: Current Landscape and Future Directions
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2024)
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2024)
Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
Harmful Speech Detection by Language Models Exhibits Gender-Queer Dialect Bias
von: Dorn, Rebecca, et al.
Veröffentlicht: (2024)
von: Dorn, Rebecca, et al.
Veröffentlicht: (2024)
Covert Bias: The Severity of Social Views' Unalignment in Language Models Towards Implicit and Explicit Opinion
von: Aldayel, Abeer, et al.
Veröffentlicht: (2024)
von: Aldayel, Abeer, et al.
Veröffentlicht: (2024)
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations
von: Zahraei, Pardis Sadat, et al.
Veröffentlicht: (2025)
von: Zahraei, Pardis Sadat, et al.
Veröffentlicht: (2025)
Evaluating LLMs for Demographic-Targeted Social Bias Detection: A Comprehensive Benchmark Study
von: Majumdar, Ayan, et al.
Veröffentlicht: (2025)
von: Majumdar, Ayan, et al.
Veröffentlicht: (2025)
A Spatio-Temporal Representation Learning as an Alternative to Traditional Glosses in Sign Language Translation and Production
von: Hwang, Eui Jun, et al.
Veröffentlicht: (2024)
von: Hwang, Eui Jun, et al.
Veröffentlicht: (2024)
Dual-Metric Evaluation of Social Bias in Large Language Models: Evidence from an Underrepresented Nepali Cultural Context
von: Pandey, Ashish, et al.
Veröffentlicht: (2026)
von: Pandey, Ashish, et al.
Veröffentlicht: (2026)
From Individuals to Interactions: Benchmarking Gender Bias in Multimodal Large Language Models from the Lens of Social Relationship
von: Xu, Yue, et al.
Veröffentlicht: (2025)
von: Xu, Yue, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Different Bias Under Different Criteria: Assessing Bias in LLMs with a Fact-Based Approach
von: Ko, Changgeon, et al.
Veröffentlicht: (2024) -
Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation
von: Lee, Huije, et al.
Veröffentlicht: (2026) -
Does Rationale Quality Matter? Enhancing Mental Disorder Detection via Selective Reasoning Distillation
von: Song, Hoyun, et al.
Veröffentlicht: (2025) -
Social Dynamics as Critical Vulnerabilities that Undermine Objective Decision-Making in LLM Collectives
von: Ko, Changgeon, et al.
Veröffentlicht: (2026) -
Towards Effective Counter-Responses: Aligning Human Preferences with Strategies to Combat Online Trolling
von: Lee, Huije, et al.
Veröffentlicht: (2024)