Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yiran, Yang, Ke, Qi, Zehan, Liu, Xiao, Yu, Yang, Zhai, ChengXiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TinyHelen's First Curriculum: Training and Evaluating Tiny Language Models in a Simpler Language Environment
von: Yang, Ke, et al.
Veröffentlicht: (2024)
von: Yang, Ke, et al.
Veröffentlicht: (2024)
Evaluating Chinese Large Language Models: The Influence of Persona Assignment on Stereotypes and Safeguards
von: Liu, Geng, et al.
Veröffentlicht: (2025)
von: Liu, Geng, et al.
Veröffentlicht: (2025)
$\texttt{ModSCAN}$: Measuring Stereotypical Bias in Large Vision-Language Models from Vision and Language Modalities
von: Jiang, Yukun, et al.
Veröffentlicht: (2024)
von: Jiang, Yukun, et al.
Veröffentlicht: (2024)
Evaluation of Bias Towards Medical Professionals in Large Language Models
von: Chen, Xi, et al.
Veröffentlicht: (2024)
von: Chen, Xi, et al.
Veröffentlicht: (2024)
Walking in Others' Shoes: How Perspective-Taking Guides Large Language Models in Reducing Toxicity and Bias
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024)
von: Plaza-del-Arco, Flor Miriam, et al.
Veröffentlicht: (2024)
DECASTE: Unveiling Caste Stereotypes in Large Language Models through Multi-Dimensional Bias Analysis
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2025)
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2025)
Analyzing the Safety of Japanese Large Language Models in Stereotype-Triggering Prompts
von: Nakanishi, Akito, et al.
Veröffentlicht: (2025)
von: Nakanishi, Akito, et al.
Veröffentlicht: (2025)
User Simulation in the Era of Generative AI: User Modeling, Synthetic Data Generation, and System Evaluation
von: Balog, Krisztian, et al.
Veröffentlicht: (2025)
von: Balog, Krisztian, et al.
Veröffentlicht: (2025)
SESGO: Spanish Evaluation of Stereotypical Generative Outputs
von: Robles, Melissa, et al.
Veröffentlicht: (2025)
von: Robles, Melissa, et al.
Veröffentlicht: (2025)
AesBiasBench: Evaluating Bias and Alignment in Multimodal Language Models for Personalized Image Aesthetic Assessment
von: Li, Kun, et al.
Veröffentlicht: (2025)
von: Li, Kun, et al.
Veröffentlicht: (2025)
Accuracy and Political Bias of News Source Credibility Ratings by Large Language Models
von: Yang, Kai-Cheng, et al.
Veröffentlicht: (2023)
von: Yang, Kai-Cheng, et al.
Veröffentlicht: (2023)
User Simulation for Evaluating Information Access Systems
von: Balog, Krisztian, et al.
Veröffentlicht: (2023)
von: Balog, Krisztian, et al.
Veröffentlicht: (2023)
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations
von: Davani, Aida, et al.
Veröffentlicht: (2025)
von: Davani, Aida, et al.
Veröffentlicht: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Who is better at math, Jenny or Jingzhen? Uncovering Stereotypes in Large Language Models
von: Siddique, Zara, et al.
Veröffentlicht: (2024)
von: Siddique, Zara, et al.
Veröffentlicht: (2024)
Exploring Social Desirability Response Bias in Large Language Models: Evidence from GPT-4 Simulations
von: Lee, Sanguk, et al.
Veröffentlicht: (2024)
von: Lee, Sanguk, et al.
Veröffentlicht: (2024)
Generative Large Language Models for Knowledge Representation: A Systematic Review of Concept Map Generation
von: Zhai, Xiaoming
Veröffentlicht: (2025)
von: Zhai, Xiaoming
Veröffentlicht: (2025)
Competence-Based Analysis of Language Models
von: Davies, Adam, et al.
Veröffentlicht: (2023)
von: Davies, Adam, et al.
Veröffentlicht: (2023)
Different Demographic Cues Yield Inconsistent Conclusions About LLM Personalization and Bias
von: Tonneau, Manuel, et al.
Veröffentlicht: (2026)
von: Tonneau, Manuel, et al.
Veröffentlicht: (2026)
A Taxonomy of Stereotype Content in Large Language Models
von: Nicolas, Gandalf, et al.
Veröffentlicht: (2024)
von: Nicolas, Gandalf, et al.
Veröffentlicht: (2024)
Addressing Stereotypes in Large Language Models: A Critical Examination and Mitigation
von: Kazi, Fatima
Veröffentlicht: (2025)
von: Kazi, Fatima
Veröffentlicht: (2025)
AI Gender Bias, Disparities, and Fairness: Does Training Data Matter?
von: Latif, Ehsan, et al.
Veröffentlicht: (2023)
von: Latif, Ehsan, et al.
Veröffentlicht: (2023)
Artificial Intelligence Bias on English Language Learners in Automatic Scoring
von: Guo, Shuchen, et al.
Veröffentlicht: (2025)
von: Guo, Shuchen, et al.
Veröffentlicht: (2025)
A Survey on Stereotype Detection in Natural Language Processing
von: Cignarella, Alessandra Teresa, et al.
Veröffentlicht: (2025)
von: Cignarella, Alessandra Teresa, et al.
Veröffentlicht: (2025)
Listen and Speak Fairly: A Study on Semantic Gender Bias in Speech Integrated Large Language Models
von: Lin, Yi-Cheng, et al.
Veröffentlicht: (2024)
von: Lin, Yi-Cheng, et al.
Veröffentlicht: (2024)
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2024)
von: Gallegos, Isabel O., et al.
Veröffentlicht: (2024)
Hidden Bias in the Machine: Stereotypes in Text-to-Image Models
von: Porikli, Sedat, et al.
Veröffentlicht: (2025)
von: Porikli, Sedat, et al.
Veröffentlicht: (2025)
The LLM Wears Prada: Analysing Gender Bias and Stereotypes through Online Shopping Data
von: Luca, Massimiliano, et al.
Veröffentlicht: (2025)
von: Luca, Massimiliano, et al.
Veröffentlicht: (2025)
StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs
von: Jeune, Pierre Le, et al.
Veröffentlicht: (2026)
von: Jeune, Pierre Le, et al.
Veröffentlicht: (2026)
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes
von: Nejadgholi, Isar, et al.
Veröffentlicht: (2024)
von: Nejadgholi, Isar, et al.
Veröffentlicht: (2024)
An Empirical Investigation of Gender Stereotype Representation in Large Language Models: The Italian Case
von: Giachino, Gioele, et al.
Veröffentlicht: (2025)
von: Giachino, Gioele, et al.
Veröffentlicht: (2025)
Globally Optimal Training of Spiking Neural Networks via Parameter Reconstruction
von: Udupi, Himanshu, et al.
Veröffentlicht: (2026)
von: Udupi, Himanshu, et al.
Veröffentlicht: (2026)
Validated Hypotheses as a Lens for Human-Likeness Evaluation in AI Agents
von: Liu, Xuan, et al.
Veröffentlicht: (2026)
von: Liu, Xuan, et al.
Veröffentlicht: (2026)
Down the Toxicity Rabbit Hole: A Novel Framework to Bias Audit Large Language Models
von: Dutta, Arka, et al.
Veröffentlicht: (2023)
von: Dutta, Arka, et al.
Veröffentlicht: (2023)
ASCenD-BDS: Adaptable, Stochastic and Context-aware framework for Detection of Bias, Discrimination and Stereotyping
von: Bahl, Rajiv, et al.
Veröffentlicht: (2025)
von: Bahl, Rajiv, et al.
Veröffentlicht: (2025)
Surfacing Subtle Stereotypes: A Multilingual, Debate-Oriented Evaluation of Modern LLMs
von: Saeed, Muhammed, et al.
Veröffentlicht: (2025)
von: Saeed, Muhammed, et al.
Veröffentlicht: (2025)
Beyond Partisan Leaning: A Comparative Analysis of Political Bias in Large Language Models
von: Peng, Tai-Quan, et al.
Veröffentlicht: (2024)
von: Peng, Tai-Quan, et al.
Veröffentlicht: (2024)
Using Large Language Models to Assess Teachers' Pedagogical Content Knowledge
von: Yang, Yaxuan, et al.
Veröffentlicht: (2025)
von: Yang, Yaxuan, et al.
Veröffentlicht: (2025)
Interpretations, Representations, and Stereotypes of Caste within Text-to-Image Generators
von: Ghosh, Sourojit
Veröffentlicht: (2024)
von: Ghosh, Sourojit
Veröffentlicht: (2024)
Ähnliche Einträge
-
TinyHelen's First Curriculum: Training and Evaluating Tiny Language Models in a Simpler Language Environment
von: Yang, Ke, et al.
Veröffentlicht: (2024) -
Evaluating Chinese Large Language Models: The Influence of Persona Assignment on Stereotypes and Safeguards
von: Liu, Geng, et al.
Veröffentlicht: (2025) -
$\texttt{ModSCAN}$: Measuring Stereotypical Bias in Large Vision-Language Models from Vision and Language Modalities
von: Jiang, Yukun, et al.
Veröffentlicht: (2024) -
Evaluation of Bias Towards Medical Professionals in Large Language Models
von: Chen, Xi, et al.
Veröffentlicht: (2024) -
Walking in Others' Shoes: How Perspective-Taking Guides Large Language Models in Reducing Toxicity and Bias
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)