A database to support the evaluation of gender biases in GPT-4o output
Fuente:
arXiv
Salvato in:
| Autori principali: | Mehner, Luise, Fiedler, Lena Alicija Philine, Ammon, Sabine, Kolossa, Dorothea |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MisinfoTeleGraph: Network-driven Misinformation Detection for German Telegram Messages
di: Kalkbrenner, Lu, et al.
Pubblicazione: (2025)
di: Kalkbrenner, Lu, et al.
Pubblicazione: (2025)
How desirable is alignment between LLMs and linguistically diverse human users?
di: Knoeferle, Pia, et al.
Pubblicazione: (2025)
di: Knoeferle, Pia, et al.
Pubblicazione: (2025)
Identifying the sources of ideological bias in GPT models through linguistic variation in output
di: Walker, Christina, et al.
Pubblicazione: (2024)
di: Walker, Christina, et al.
Pubblicazione: (2024)
DistriBlock: Identifying adversarial audio samples by leveraging characteristics of the output distribution
di: Pizarro, Matías, et al.
Pubblicazione: (2023)
di: Pizarro, Matías, et al.
Pubblicazione: (2023)
Emergence of a phonological bias in ChatGPT
di: Toro, Juan Manuel
Pubblicazione: (2023)
di: Toro, Juan Manuel
Pubblicazione: (2023)
Extending Information Bottleneck Attribution to Video Sequences
di: Solopova, Veronika, et al.
Pubblicazione: (2025)
di: Solopova, Veronika, et al.
Pubblicazione: (2025)
Anthropocentric bias in language model evaluation
di: Millière, Raphaël, et al.
Pubblicazione: (2024)
di: Millière, Raphaël, et al.
Pubblicazione: (2024)
The high dimensional psychological profile and cultural bias of ChatGPT
di: Yuan, Hang, et al.
Pubblicazione: (2024)
di: Yuan, Hang, et al.
Pubblicazione: (2024)
Surprising gender biases in GPT
di: Fulgu, Raluca Alexandra, et al.
Pubblicazione: (2024)
di: Fulgu, Raluca Alexandra, et al.
Pubblicazione: (2024)
An investigation of structures responsible for gender bias in BERT and DistilBERT
di: Leteno, Thibaud, et al.
Pubblicazione: (2024)
di: Leteno, Thibaud, et al.
Pubblicazione: (2024)
Addressing speaker gender bias in large scale speech translation systems
di: Bansal, Shubham, et al.
Pubblicazione: (2025)
di: Bansal, Shubham, et al.
Pubblicazione: (2025)
Optimal Calibration of the endpoint-corrected Hilbert Transform
di: Osmers, Eike, et al.
Pubblicazione: (2026)
di: Osmers, Eike, et al.
Pubblicazione: (2026)
Multi-Intent Recognition in Dialogue Understanding: A Comparison Between Smaller Open-Source LLMs
di: Ahmad, Adnan, et al.
Pubblicazione: (2025)
di: Ahmad, Adnan, et al.
Pubblicazione: (2025)
BIPOLAR: Polarization-based granular framework for LLM bias evaluation
di: Pavlíček, Martin, et al.
Pubblicazione: (2025)
di: Pavlíček, Martin, et al.
Pubblicazione: (2025)
Source framing triggers systematic evaluation bias in Large Language Models
di: Germani, Federico, et al.
Pubblicazione: (2025)
di: Germani, Federico, et al.
Pubblicazione: (2025)
CBEval: A framework for evaluating and interpreting cognitive biases in LLMs
di: Shaikh, Ammar, et al.
Pubblicazione: (2024)
di: Shaikh, Ammar, et al.
Pubblicazione: (2024)
An evaluation of LLMs for generating movie reviews: GPT-4o, Gemini-2.0 and DeepSeek-V3
di: Sands, Brendan, et al.
Pubblicazione: (2025)
di: Sands, Brendan, et al.
Pubblicazione: (2025)
Aligning Large Language Models with Diverse Political Viewpoints
di: Stammbach, Dominik, et al.
Pubblicazione: (2024)
di: Stammbach, Dominik, et al.
Pubblicazione: (2024)
The Silicon Ceiling: Auditing GPT's Race and Gender Biases in Hiring
di: Armstrong, Lena, et al.
Pubblicazione: (2024)
di: Armstrong, Lena, et al.
Pubblicazione: (2024)
Exploring the features used for summary evaluation by Human and GPT
di: Sadeghi, Zahra, et al.
Pubblicazione: (2025)
di: Sadeghi, Zahra, et al.
Pubblicazione: (2025)
Is GPT-4 Less Politically Biased than GPT-3.5? A Renewed Investigation of ChatGPT's Political Biases
di: Weber, Erik, et al.
Pubblicazione: (2024)
di: Weber, Erik, et al.
Pubblicazione: (2024)
GPTEval: A Survey on Assessments of ChatGPT and GPT-4
di: Mao, Rui, et al.
Pubblicazione: (2023)
di: Mao, Rui, et al.
Pubblicazione: (2023)
Conversational Assistants to support Heart Failure Patients: comparing a Neurosymbolic Architecture with ChatGPT
di: Tayal, Anuja, et al.
Pubblicazione: (2025)
di: Tayal, Anuja, et al.
Pubblicazione: (2025)
A word association network methodology for evaluating implicit biases in LLMs compared to humans
di: Abramski, Katherine, et al.
Pubblicazione: (2025)
di: Abramski, Katherine, et al.
Pubblicazione: (2025)
Evaluation of GPT-4o and GPT-4o-mini's Vision Capabilities for Compositional Analysis from Dried Solution Drops
di: Dangi, Deven B., et al.
Pubblicazione: (2024)
di: Dangi, Deven B., et al.
Pubblicazione: (2024)
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo
di: Wood, Michael C., et al.
Pubblicazione: (2024)
di: Wood, Michael C., et al.
Pubblicazione: (2024)
LLMs left, right, and center: Assessing GPT's capabilities to label political bias from web domains
di: Hernandes, Raphael, et al.
Pubblicazione: (2024)
di: Hernandes, Raphael, et al.
Pubblicazione: (2024)
Media Slant is Contagious
di: Widmer, Philine, et al.
Pubblicazione: (2022)
di: Widmer, Philine, et al.
Pubblicazione: (2022)
Can we trust the evaluation on ChatGPT?
di: Aiyappa, Rachith, et al.
Pubblicazione: (2023)
di: Aiyappa, Rachith, et al.
Pubblicazione: (2023)
Math anxiety and associative knowledge structure are entwined in psychology students but not in Large Language Models like GPT-3.5 and GPT-4o
di: Ciringione, Luciana, et al.
Pubblicazione: (2025)
di: Ciringione, Luciana, et al.
Pubblicazione: (2025)
From Theory to Comprehension: A Comparative Study of Differential Privacy and $k$-Anonymity
di: von Voigt, Saskia Nuñez, et al.
Pubblicazione: (2024)
di: von Voigt, Saskia Nuñez, et al.
Pubblicazione: (2024)
DeID-GPT: Zero-shot Medical Text De-Identification by GPT-4
di: Liu, Zhengliang, et al.
Pubblicazione: (2023)
di: Liu, Zhengliang, et al.
Pubblicazione: (2023)
A comparison of Human, GPT-3.5, and GPT-4 Performance in a University-Level Coding Course
di: Yeadon, Will, et al.
Pubblicazione: (2024)
di: Yeadon, Will, et al.
Pubblicazione: (2024)
A Preliminary Exploration with GPT-4o Voice Mode
di: Lin, Yu-Xiang, et al.
Pubblicazione: (2025)
di: Lin, Yu-Xiang, et al.
Pubblicazione: (2025)
An evaluation of LLMs for political bias in Western media: Israel-Hamas and Ukraine-Russia wars
di: Chandra, Rohitash, et al.
Pubblicazione: (2026)
di: Chandra, Rohitash, et al.
Pubblicazione: (2026)
Building A Proof-Oriented Programmer That Is 64% Better Than GPT-4o Under Data Scarcity
di: Zhang, Dylan, et al.
Pubblicazione: (2025)
di: Zhang, Dylan, et al.
Pubblicazione: (2025)
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
di: Stewart, Ian, et al.
Pubblicazione: (2024)
di: Stewart, Ian, et al.
Pubblicazione: (2024)
Notes on Applicability of GPT-4 to Document Understanding
di: Borchmann, Łukasz
Pubblicazione: (2024)
di: Borchmann, Łukasz
Pubblicazione: (2024)
Text Understanding in GPT-4 vs Humans
di: Shultz, Thomas R., et al.
Pubblicazione: (2024)
di: Shultz, Thomas R., et al.
Pubblicazione: (2024)
Is GPT-4 a reliable rater? Evaluating Consistency in GPT-4 Text Ratings
di: Hackl, Veronika, et al.
Pubblicazione: (2023)
di: Hackl, Veronika, et al.
Pubblicazione: (2023)
Documenti analoghi
-
MisinfoTeleGraph: Network-driven Misinformation Detection for German Telegram Messages
di: Kalkbrenner, Lu, et al.
Pubblicazione: (2025) -
How desirable is alignment between LLMs and linguistically diverse human users?
di: Knoeferle, Pia, et al.
Pubblicazione: (2025) -
Identifying the sources of ideological bias in GPT models through linguistic variation in output
di: Walker, Christina, et al.
Pubblicazione: (2024) -
DistriBlock: Identifying adversarial audio samples by leveraging characteristics of the output distribution
di: Pizarro, Matías, et al.
Pubblicazione: (2023) -
Emergence of a phonological bias in ChatGPT
di: Toro, Juan Manuel
Pubblicazione: (2023)