HESEIA: A community-based dataset for evaluating social biases in large language models, co-designed in real school settings in Latin America
Fuente:
arXiv
Saved in:
| Main Authors: | Ivetta, Guido, Gomez, Marcos J., Martinelli, Sofía, Palombini, Pietro, Echeveste, M. Emilia, Mazzeo, Nair Carolina, Busaniche, Beatriz, Benotti, Luciana |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Data Collection for Latin-American Community-sourced Evaluation of Stereotypes (LACES)
by: Ivetta, Guido, et al.
Published: (2025)
by: Ivetta, Guido, et al.
Published: (2025)
Selectively Answering Visual Questions
by: Eisenschlos, Julian Martin, et al.
Published: (2024)
by: Eisenschlos, Julian Martin, et al.
Published: (2024)
Towards culturally-appropriate conversational AI for health in the majority world: An exploratory study with citizens and professionals in Latin America
by: Peters, Dorian, et al.
Published: (2025)
by: Peters, Dorian, et al.
Published: (2025)
ROSA: Addressing text understanding challenges in photographs via ROtated SAmpling
by: Maina, Hernán, et al.
Published: (2025)
by: Maina, Hernán, et al.
Published: (2025)
Low-resource domain adaptation while minimizing energy and hardware resource consumption
by: Maina, Hernán, et al.
Published: (2025)
by: Maina, Hernán, et al.
Published: (2025)
Synthetic social data: trials and tribulations
by: Ivetta, Guido, et al.
Published: (2025)
by: Ivetta, Guido, et al.
Published: (2025)
CCI3.0-HQ: a large-scale Chinese dataset of high quality designed for pre-training large language models
by: Wang, Liangdong, et al.
Published: (2024)
by: Wang, Liangdong, et al.
Published: (2024)
DPA: A one-stop metric to measure bias amplification in classification datasets
by: Tokas, Bhanu, et al.
Published: (2024)
by: Tokas, Bhanu, et al.
Published: (2024)
Special issue on Argentine symposium on articial intelligence (ASAI 2013)
by: Luciana Benotti
Published: (2014)
by: Luciana Benotti
Published: (2014)
How do datasets, developers, and models affect biases in a low-resourced language?: The Case of the Bengali Language
by: Das, Dipto, et al.
Published: (2025)
by: Das, Dipto, et al.
Published: (2025)
A closer look at how large language models trust humans: patterns and biases
by: Lerman, Valeria, et al.
Published: (2025)
by: Lerman, Valeria, et al.
Published: (2025)
Inducing anxiety in large language models can induce bias
by: Coda-Forno, Julian, et al.
Published: (2023)
by: Coda-Forno, Julian, et al.
Published: (2023)
Evaluating the capability of large language models to personalize science texts for diverse middle-school-age learners
by: Vaccaro Jr, Michael, et al.
Published: (2024)
by: Vaccaro Jr, Michael, et al.
Published: (2024)
Sentiment trading with large language models
by: Kirtac, Kemal, et al.
Published: (2024)
by: Kirtac, Kemal, et al.
Published: (2024)
Open Challenges on Fairness of Artificial Intelligence in Medical Imaging Applications
by: Ferrante, Enzo, et al.
Published: (2024)
by: Ferrante, Enzo, et al.
Published: (2024)
B-score: Detecting biases in large language models using response history
by: Vo, An, et al.
Published: (2025)
by: Vo, An, et al.
Published: (2025)
A benchmark dataset for evaluating Syndrome Differentiation and Treatment in large language models
by: Li, Kunning, et al.
Published: (2025)
by: Li, Kunning, et al.
Published: (2025)
Addressing cognitive bias in medical language models
by: Schmidgall, Samuel, et al.
Published: (2024)
by: Schmidgall, Samuel, et al.
Published: (2024)
'Neural howlround' in large language models: a self-reinforcing bias phenomenon, and a dynamic attenuation solution
by: Drake, Seth
Published: (2025)
by: Drake, Seth
Published: (2025)
The role of large language models in UI/UX design: A systematic literature review
by: Ahmed, Ammar, et al.
Published: (2025)
by: Ahmed, Ammar, et al.
Published: (2025)
Anthropocentric bias in language model evaluation
by: Millière, Raphaël, et al.
Published: (2024)
by: Millière, Raphaël, et al.
Published: (2024)
MindScope: Exploring cognitive biases in large language models through Multi-Agent Systems
by: Xie, Zhentao, et al.
Published: (2024)
by: Xie, Zhentao, et al.
Published: (2024)
TAGLAS: An atlas of text-attributed graph datasets in the era of large graph and language models
by: Feng, Jiarui, et al.
Published: (2024)
by: Feng, Jiarui, et al.
Published: (2024)
Evaluation of large language models for discovery of gene set function
by: Hu, Mengzhou, et al.
Published: (2023)
by: Hu, Mengzhou, et al.
Published: (2023)
A dataset and benchmark for hospital course summarization with adapted large language models
by: Aali, Asad, et al.
Published: (2024)
by: Aali, Asad, et al.
Published: (2024)
A benchmark multimodal oro-dental dataset for large vision-language models
by: Lv, Haoxin, et al.
Published: (2025)
by: Lv, Haoxin, et al.
Published: (2025)
Question answering system of bridge design specification based on large language model
by: Zhang, Leye, et al.
Published: (2024)
by: Zhang, Leye, et al.
Published: (2024)
PhDGPT: Introducing a psychometric and linguistic dataset about how large language models perceive graduate students and professors in psychology
by: De Duro, Edoardo Sebastiano, et al.
Published: (2024)
by: De Duro, Edoardo Sebastiano, et al.
Published: (2024)
Auditing demographic bias in AI-based emergency police dispatch: a cross-lingual evaluation of eleven large language models
by: Guey, William, et al.
Published: (2026)
by: Guey, William, et al.
Published: (2026)
Modeling reputation-based behavioral biases in school choice
by: Kleinberg, Jon, et al.
Published: (2024)
by: Kleinberg, Jon, et al.
Published: (2024)
AI-AI Bias: large language models favor communications generated by large language models
by: Laurito, Walter, et al.
Published: (2024)
by: Laurito, Walter, et al.
Published: (2024)
Finnish primary school students' conceptions of machine learning
by: Mertala, Pekka, et al.
Published: (2024)
by: Mertala, Pekka, et al.
Published: (2024)
Lean Workbook: A large-scale Lean problem set formalized from natural language math problems
by: Ying, Huaiyuan, et al.
Published: (2024)
by: Ying, Huaiyuan, et al.
Published: (2024)
The in-context inductive biases of vision-language models differ across modalities
by: Allen, Kelsey, et al.
Published: (2025)
by: Allen, Kelsey, et al.
Published: (2025)
Dissociating language and thought in large language models
by: Mahowald, Kyle, et al.
Published: (2023)
by: Mahowald, Kyle, et al.
Published: (2023)
Designing large language model prompts to extract scores from messy text: A shared dataset and challenge
by: Thelwall, Mike
Published: (2026)
by: Thelwall, Mike
Published: (2026)
Retrieval augmentation of large language models for lay language generation
by: Guo, Yue, et al.
Published: (2022)
by: Guo, Yue, et al.
Published: (2022)
Do large language models resemble humans in language use?
by: Cai, Zhenguang G., et al.
Published: (2023)
by: Cai, Zhenguang G., et al.
Published: (2023)
Failure of contextual invariance in large language models
by: Kumar, Sagar, et al.
Published: (2026)
by: Kumar, Sagar, et al.
Published: (2026)
Large-scale moral machine experiment on large language models
by: Ahmad, Muhammad Shahrul Zaim bin, et al.
Published: (2024)
by: Ahmad, Muhammad Shahrul Zaim bin, et al.
Published: (2024)
Similar Items
-
Adaptive Data Collection for Latin-American Community-sourced Evaluation of Stereotypes (LACES)
by: Ivetta, Guido, et al.
Published: (2025) -
Selectively Answering Visual Questions
by: Eisenschlos, Julian Martin, et al.
Published: (2024) -
Towards culturally-appropriate conversational AI for health in the majority world: An exploratory study with citizens and professionals in Latin America
by: Peters, Dorian, et al.
Published: (2025) -
ROSA: Addressing text understanding challenges in photographs via ROtated SAmpling
by: Maina, Hernán, et al.
Published: (2025) -
Low-resource domain adaptation while minimizing energy and hardware resource consumption
by: Maina, Hernán, et al.
Published: (2025)