Uncovering Bias in Foundation Models: Impact, Testing, Harm, and Mitigation
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Shuzhou, Liu, Li, Liu, Yongxiang, Liu, Zhen, Zhang, Shuanghui, Heikkilä, Janne, Li, Xiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mitigating Gender Bias in Depression Detection via Counterfactual Inference
by: Hu, Mingxuan, et al.
Published: (2025)
by: Hu, Mingxuan, et al.
Published: (2025)
A Causal Adjustment Module for Debiasing Scene Graph Generation
by: Liu, Li, et al.
Published: (2025)
by: Liu, Li, et al.
Published: (2025)
Harm Amplification in Text-to-Image Models
by: Hao, Susan, et al.
Published: (2024)
by: Hao, Susan, et al.
Published: (2024)
Backdoor for Debias: Mitigating Model Bias with Backdoor Attack-based Artificial Bias
by: Wu, Shangxi, et al.
Published: (2023)
by: Wu, Shangxi, et al.
Published: (2023)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
by: Ma, Mingyu Derek, et al.
Published: (2023)
by: Ma, Mingyu Derek, et al.
Published: (2023)
On the Societal Impact of Open Foundation Models
by: Kapoor, Sayash, et al.
Published: (2024)
by: Kapoor, Sayash, et al.
Published: (2024)
Beyond Behaviorist Representational Harms: A Plan for Measurement and Mitigation
by: Chien, Jennifer, et al.
Published: (2024)
by: Chien, Jennifer, et al.
Published: (2024)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Demographic Bias of Expert-Level Vision-Language Foundation Models in Medical Imaging
by: Yang, Yuzhe, et al.
Published: (2024)
by: Yang, Yuzhe, et al.
Published: (2024)
Unmasking Bias in AI: A Systematic Review of Bias Detection and Mitigation Strategies in Electronic Health Record-based Models
by: Chen, Feng, et al.
Published: (2023)
by: Chen, Feng, et al.
Published: (2023)
Effective Controllable Bias Mitigation for Classification and Retrieval using Gate Adapters
by: Masoudian, Shahed, et al.
Published: (2024)
by: Masoudian, Shahed, et al.
Published: (2024)
Different Horses for Different Courses: Comparing Bias Mitigation Algorithms in ML
by: Ganesh, Prakhar, et al.
Published: (2024)
by: Ganesh, Prakhar, et al.
Published: (2024)
Towards Urban General Intelligence: A Review and Outlook of Urban Foundation Models
by: Zhang, Weijia, et al.
Published: (2024)
by: Zhang, Weijia, et al.
Published: (2024)
Personalized Knowledge Tracing through Student Representation Reconstruction and Class Imbalance Mitigation
by: Chen, Zhiyu, et al.
Published: (2024)
by: Chen, Zhiyu, et al.
Published: (2024)
Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias
by: Itzhak, Itay, et al.
Published: (2023)
by: Itzhak, Itay, et al.
Published: (2023)
Contextual StereoSet: Stress-Testing Bias Alignment Robustness in Large Language Models
by: Basu, Abhinaba, et al.
Published: (2026)
by: Basu, Abhinaba, et al.
Published: (2026)
TestAgent: An Adaptive and Intelligent Expert for Human Assessment
by: Yu, Junhao, et al.
Published: (2025)
by: Yu, Junhao, et al.
Published: (2025)
Beyond AlphaEarth: Toward Human-Centered Geospatial Foundation Models via POI-Guided Contrastive Learning
by: Liu, Junyuan, et al.
Published: (2025)
by: Liu, Junyuan, et al.
Published: (2025)
Laissez-Faire Harms: Algorithmic Biases in Generative Language Models
by: Shieh, Evan, et al.
Published: (2024)
by: Shieh, Evan, et al.
Published: (2024)
OpenCity: Open Spatio-Temporal Foundation Models for Traffic Prediction
by: Li, Zhonghang, et al.
Published: (2024)
by: Li, Zhonghang, et al.
Published: (2024)
A Question-centric Multi-experts Contrastive Learning Framework for Improving the Accuracy and Interpretability of Deep Sequential Knowledge Tracing Models
by: Zhang, Hengyuan, et al.
Published: (2024)
by: Zhang, Hengyuan, et al.
Published: (2024)
Addressing Selection Bias in Computerized Adaptive Testing: A User-Wise Aggregate Influence Function Approach
by: Kwon, Soonwoo, et al.
Published: (2023)
by: Kwon, Soonwoo, et al.
Published: (2023)
Foundation Model Transparency Reports
by: Bommasani, Rishi, et al.
Published: (2024)
by: Bommasani, Rishi, et al.
Published: (2024)
Impacts of Racial Bias in Historical Training Data for News AI
by: Bhargava, Rahul, et al.
Published: (2025)
by: Bhargava, Rahul, et al.
Published: (2025)
Synthetic Data Augmentation for Enhancing Harmful Algal Bloom Detection with Machine Learning
by: Huang, Tianyi
Published: (2025)
by: Huang, Tianyi
Published: (2025)
On Catastrophic Inheritance of Large Foundation Models
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
The 2024 Foundation Model Transparency Index
by: Bommasani, Rishi, et al.
Published: (2024)
by: Bommasani, Rishi, et al.
Published: (2024)
The 2025 Foundation Model Transparency Index
by: Wan, Alexander, et al.
Published: (2025)
by: Wan, Alexander, et al.
Published: (2025)
Say My Name: a Model's Bias Discovery Framework
by: Ciranni, Massimiliano, et al.
Published: (2024)
by: Ciranni, Massimiliano, et al.
Published: (2024)
Epistemic Uncertainty-Weighted Loss for Visual Bias Mitigation
by: Stone, Rebecca S, et al.
Published: (2022)
by: Stone, Rebecca S, et al.
Published: (2022)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Mitigating Unfairness via Evolutionary Multi-objective Ensemble Learning
by: Zhang, Qingquan, et al.
Published: (2022)
by: Zhang, Qingquan, et al.
Published: (2022)
Ecosystem Graphs: The Social Footprint of Foundation Models
by: Bommasani, Rishi, et al.
Published: (2023)
by: Bommasani, Rishi, et al.
Published: (2023)
Pre-trained Transformer Uncovers Meaningful Patterns in Human Mobility Data
by: Najjar, Alameen
Published: (2024)
by: Najjar, Alameen
Published: (2024)
Empowering Epidemic Response: The Role of Reinforcement Learning in Infectious Disease Control
by: Liu, Mutong, et al.
Published: (2026)
by: Liu, Mutong, et al.
Published: (2026)
Bias and Fairness in Large Language Models: A Survey
by: Gallegos, Isabel O., et al.
Published: (2023)
by: Gallegos, Isabel O., et al.
Published: (2023)
Integrating Social Determinants of Health into Knowledge Graphs: Evaluating Prediction Bias and Fairness in Healthcare
by: Shang, Tianqi, et al.
Published: (2024)
by: Shang, Tianqi, et al.
Published: (2024)
Training Foundation Models as Data Compression: On Information, Model Weights and Copyright Law
by: Franceschelli, Giorgio, et al.
Published: (2024)
by: Franceschelli, Giorgio, et al.
Published: (2024)
Data-driven Energy Consumption Modelling for Electric Micromobility using an Open Dataset
by: Ding, Yue, et al.
Published: (2024)
by: Ding, Yue, et al.
Published: (2024)
Explainable AI for Predicting and Understanding Mathematics Achievement: A Cross-National Analysis of PISA 2018
by: Liu, Liu, et al.
Published: (2025)
by: Liu, Liu, et al.
Published: (2025)
Similar Items
-
Mitigating Gender Bias in Depression Detection via Counterfactual Inference
by: Hu, Mingxuan, et al.
Published: (2025) -
A Causal Adjustment Module for Debiasing Scene Graph Generation
by: Liu, Li, et al.
Published: (2025) -
Harm Amplification in Text-to-Image Models
by: Hao, Susan, et al.
Published: (2024) -
Backdoor for Debias: Mitigating Model Bias with Backdoor Attack-based Artificial Bias
by: Wu, Shangxi, et al.
Published: (2023) -
Mitigating Bias for Question Answering Models by Tracking Bias Influence
by: Ma, Mingyu Derek, et al.
Published: (2023)