Salvato in:
| Autori principali: | Pelosio, Giulio, Batra, Devesh, Bovey, Noémie, Hankache, Robert, Iglesias, Cristovao, Cowan, Greig, Khraishi, Raad |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2507.16989 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Helping Customers in Distress: An LLM-powered Agent that Converses, Probes, and Routes
di: Atreya, Alankar, et al.
Pubblicazione: (2026)
di: Atreya, Alankar, et al.
Pubblicazione: (2026)
Evaluating the Sensitivity of LLMs to Prior Context
di: Hankache, Robert, et al.
Pubblicazione: (2025)
di: Hankache, Robert, et al.
Pubblicazione: (2025)
Evaluating Performance Drift from Model Switching in Multi-Turn LLM Systems
di: Khraishi, Raad, et al.
Pubblicazione: (2026)
di: Khraishi, Raad, et al.
Pubblicazione: (2026)
How Personality Traits Shape LLM Risk-Taking Behaviour
di: Hartley, John, et al.
Pubblicazione: (2025)
di: Hartley, John, et al.
Pubblicazione: (2025)
Measuring Mechanistic Independence: Can Bias Be Removed Without Erasing Demographics?
di: Shan, Zhengyang, et al.
Pubblicazione: (2025)
di: Shan, Zhengyang, et al.
Pubblicazione: (2025)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
di: Hanif, Ikhlasul Akmal, et al.
Pubblicazione: (2026)
di: Hanif, Ikhlasul Akmal, et al.
Pubblicazione: (2026)
When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation
di: Xu, Wenda, et al.
Pubblicazione: (2025)
di: Xu, Wenda, et al.
Pubblicazione: (2025)
Do Bias Benchmarks Generalise? Evidence from Voice-based Evaluation of Gender Bias in SpeechLLMs
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2025)
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2025)
The Bias is in the Details: An Assessment of Cognitive Bias in LLMs
di: Knipper, R. Alexander, et al.
Pubblicazione: (2025)
di: Knipper, R. Alexander, et al.
Pubblicazione: (2025)
A Study of Nationality Bias in Names and Perplexity using Off-the-Shelf Affect-related Tweet Classifiers
di: Barriere, Valentin, et al.
Pubblicazione: (2024)
di: Barriere, Valentin, et al.
Pubblicazione: (2024)
DIF: A Framework for Benchmarking and Verifying Implicit Bias in LLMs
di: Yin, Lake, et al.
Pubblicazione: (2025)
di: Yin, Lake, et al.
Pubblicazione: (2025)
Evaluating Gender Bias of LLMs in Making Morality Judgements
di: Bajaj, Divij, et al.
Pubblicazione: (2024)
di: Bajaj, Divij, et al.
Pubblicazione: (2024)
Language Bias under Conflicting Information in Multilingual LLMs
di: Östling, Robert, et al.
Pubblicazione: (2026)
di: Östling, Robert, et al.
Pubblicazione: (2026)
A Japanese Benchmark for Evaluating Social Bias in Reasoning Based on Attribution Theory
di: Shiotani, Taihei, et al.
Pubblicazione: (2026)
di: Shiotani, Taihei, et al.
Pubblicazione: (2026)
Robust Bias Evaluation with FilBBQ: A Filipino Bias Benchmark for Question-Answering Language Models
di: Gamboa, Lance Calvin Lim, et al.
Pubblicazione: (2026)
di: Gamboa, Lance Calvin Lim, et al.
Pubblicazione: (2026)
Evaluating LLMs for Demographic-Targeted Social Bias Detection: A Comprehensive Benchmark Study
di: Majumdar, Ayan, et al.
Pubblicazione: (2025)
di: Majumdar, Ayan, et al.
Pubblicazione: (2025)
What is in a name? Mitigating Name Bias in Text Embeddings via Anonymization
di: Manchanda, Sahil, et al.
Pubblicazione: (2025)
di: Manchanda, Sahil, et al.
Pubblicazione: (2025)
Enabling Scalable Evaluation of Bias Patterns in Medical LLMs
di: Fayyaz, Hamed, et al.
Pubblicazione: (2024)
di: Fayyaz, Hamed, et al.
Pubblicazione: (2024)
I Am Aligned, But With Whom? MENA Values Benchmark for Evaluating Cultural Alignment and Multilingual Bias in LLMs
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2025)
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2025)
Evaluating the Bias in LLMs for Surveying Opinion and Decision Making in Healthcare
di: Khaokaew, Yonchanok, et al.
Pubblicazione: (2025)
di: Khaokaew, Yonchanok, et al.
Pubblicazione: (2025)
A Dual-Layered Evaluation of Geopolitical and Cultural Bias in LLMs
di: Kim, Sean, et al.
Pubblicazione: (2025)
di: Kim, Sean, et al.
Pubblicazione: (2025)
The power of Prompts: Evaluating and Mitigating Gender Bias in MT with LLMs
di: Sant, Aleix, et al.
Pubblicazione: (2024)
di: Sant, Aleix, et al.
Pubblicazione: (2024)
Social Bias Benchmark for Generation: A Comparison of Generation and QA-Based Evaluations
di: Jin, Jiho, et al.
Pubblicazione: (2025)
di: Jin, Jiho, et al.
Pubblicazione: (2025)
Evaluating Metrics for Bias in Word Embeddings
di: Schröder, Sarah, et al.
Pubblicazione: (2021)
di: Schröder, Sarah, et al.
Pubblicazione: (2021)
Template-Based Probes Are Imperfect Lenses for Counterfactual Bias Evaluation in LLMs
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
Investigating Subtler Biases in LLMs: Ageism, Beauty, Institutional, and Nationality Bias in Generative Models
di: Kamruzzaman, Mahammed, et al.
Pubblicazione: (2023)
di: Kamruzzaman, Mahammed, et al.
Pubblicazione: (2023)
BiasAlert: A Plug-and-play Tool for Social Bias Detection in LLMs
di: Fan, Zhiting, et al.
Pubblicazione: (2024)
di: Fan, Zhiting, et al.
Pubblicazione: (2024)
Relative Bias: A Comparative Framework for Quantifying Bias in LLMs
di: Arbabi, Alireza, et al.
Pubblicazione: (2025)
di: Arbabi, Alireza, et al.
Pubblicazione: (2025)
Capturing Bias Diversity in LLMs
di: Gosavi, Purva Prasad, et al.
Pubblicazione: (2024)
di: Gosavi, Purva Prasad, et al.
Pubblicazione: (2024)
FairCoder: Evaluating Social Bias of LLMs in Code Generation
di: Du, Yongkang, et al.
Pubblicazione: (2025)
di: Du, Yongkang, et al.
Pubblicazione: (2025)
LLMs are Biased Teachers: Evaluating LLM Bias in Personalized Education
di: Weissburg, Iain, et al.
Pubblicazione: (2024)
di: Weissburg, Iain, et al.
Pubblicazione: (2024)
Different Bias Under Different Criteria: Assessing Bias in LLMs with a Fact-Based Approach
di: Ko, Changgeon, et al.
Pubblicazione: (2024)
di: Ko, Changgeon, et al.
Pubblicazione: (2024)
A Comprehensive Study of Gender Bias in Chemical Named Entity Recognition Models
di: Zhao, Xingmeng, et al.
Pubblicazione: (2022)
di: Zhao, Xingmeng, et al.
Pubblicazione: (2022)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
di: Yamamoto, Taisei, et al.
Pubblicazione: (2025)
di: Yamamoto, Taisei, et al.
Pubblicazione: (2025)
Evaluating Bias in LLMs for Job-Resume Matching: Gender, Race, and Education
di: Iso, Hayate, et al.
Pubblicazione: (2025)
di: Iso, Hayate, et al.
Pubblicazione: (2025)
Disclosure and Mitigation of Gender Bias in LLMs
di: Dong, Xiangjue, et al.
Pubblicazione: (2024)
di: Dong, Xiangjue, et al.
Pubblicazione: (2024)
Breaking Bias, Building Bridges: Evaluation and Mitigation of Social Biases in LLMs via Contact Hypothesis
di: Raj, Chahat, et al.
Pubblicazione: (2024)
di: Raj, Chahat, et al.
Pubblicazione: (2024)
In the Name of Fairness: Assessing the Bias in Clinical Record De-identification
di: Xiao, Yuxin, et al.
Pubblicazione: (2023)
di: Xiao, Yuxin, et al.
Pubblicazione: (2023)
Causally Testing Gender Bias in LLMs: A Case Study on Occupational Bias
di: Chen, Yuen, et al.
Pubblicazione: (2022)
di: Chen, Yuen, et al.
Pubblicazione: (2022)
Bias-Augmented Consistency Training Reduces Biased Reasoning in Chain-of-Thought
di: Chua, James, et al.
Pubblicazione: (2024)
di: Chua, James, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Helping Customers in Distress: An LLM-powered Agent that Converses, Probes, and Routes
di: Atreya, Alankar, et al.
Pubblicazione: (2026) -
Evaluating the Sensitivity of LLMs to Prior Context
di: Hankache, Robert, et al.
Pubblicazione: (2025) -
Evaluating Performance Drift from Model Switching in Multi-Turn LLM Systems
di: Khraishi, Raad, et al.
Pubblicazione: (2026) -
How Personality Traits Shape LLM Risk-Taking Behaviour
di: Hartley, John, et al.
Pubblicazione: (2025) -
Measuring Mechanistic Independence: Can Bias Be Removed Without Erasing Demographics?
di: Shan, Zhengyang, et al.
Pubblicazione: (2025)