Effective Controllable Bias Mitigation for Classification and Retrieval using Gate Adapters
Fuente:
arXiv
Salvato in:
| Autori principali: | Masoudian, Shahed, Volaucnik, Cornelia, Schedl, Markus, Rekabsaz, Navid |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Unlabeled Debiasing in Downstream Tasks via Class-wise Low Variance Regularization
di: Masoudian, Shahed, et al.
Pubblicazione: (2024)
di: Masoudian, Shahed, et al.
Pubblicazione: (2024)
ScaLearn: Simple and Highly Parameter-Efficient Task Transfer by Learning to Scale
di: Frohmann, Markus, et al.
Pubblicazione: (2023)
di: Frohmann, Markus, et al.
Pubblicazione: (2023)
Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes
di: Masoudian, Shahed, et al.
Pubblicazione: (2025)
di: Masoudian, Shahed, et al.
Pubblicazione: (2025)
Explanatory Interactive Machine Learning for Bias Mitigation in Visual Gender Classification
di: Satriani, Nathanya, et al.
Pubblicazione: (2026)
di: Satriani, Nathanya, et al.
Pubblicazione: (2026)
Backdoor for Debias: Mitigating Model Bias with Backdoor Attack-based Artificial Bias
di: Wu, Shangxi, et al.
Pubblicazione: (2023)
di: Wu, Shangxi, et al.
Pubblicazione: (2023)
Mitigating Bias for Question Answering Models by Tracking Bias Influence
di: Ma, Mingyu Derek, et al.
Pubblicazione: (2023)
di: Ma, Mingyu Derek, et al.
Pubblicazione: (2023)
Mitigating Gender Bias in Depression Detection via Counterfactual Inference
di: Hu, Mingxuan, et al.
Pubblicazione: (2025)
di: Hu, Mingxuan, et al.
Pubblicazione: (2025)
Uncovering Bias in Foundation Models: Impact, Testing, Harm, and Mitigation
di: Sun, Shuzhou, et al.
Pubblicazione: (2025)
di: Sun, Shuzhou, et al.
Pubblicazione: (2025)
Different Horses for Different Courses: Comparing Bias Mitigation Algorithms in ML
di: Ganesh, Prakhar, et al.
Pubblicazione: (2024)
di: Ganesh, Prakhar, et al.
Pubblicazione: (2024)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
di: Xu, Xin, et al.
Pubblicazione: (2025)
di: Xu, Xin, et al.
Pubblicazione: (2025)
Scaling Legal AI: Benchmarking Mamba and Transformers for Statutory Classification and Case Law Retrieval
di: Maurya, Anuraj
Pubblicazione: (2025)
di: Maurya, Anuraj
Pubblicazione: (2025)
Unmasking Bias in AI: A Systematic Review of Bias Detection and Mitigation Strategies in Electronic Health Record-based Models
di: Chen, Feng, et al.
Pubblicazione: (2023)
di: Chen, Feng, et al.
Pubblicazione: (2023)
Epistemic Uncertainty-Weighted Loss for Visual Bias Mitigation
di: Stone, Rebecca S, et al.
Pubblicazione: (2022)
di: Stone, Rebecca S, et al.
Pubblicazione: (2022)
Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias
di: Itzhak, Itay, et al.
Pubblicazione: (2023)
di: Itzhak, Itay, et al.
Pubblicazione: (2023)
Say My Name: a Model's Bias Discovery Framework
di: Ciranni, Massimiliano, et al.
Pubblicazione: (2024)
di: Ciranni, Massimiliano, et al.
Pubblicazione: (2024)
BiasGuard: Guardrailing Fairness in Machine Learning Production Systems
di: Cohen-Inger, Nurit, et al.
Pubblicazione: (2025)
di: Cohen-Inger, Nurit, et al.
Pubblicazione: (2025)
Wisdom from Diversity: Bias Mitigation Through Hybrid Human-LLM Crowds
di: Abels, Axel, et al.
Pubblicazione: (2025)
di: Abels, Axel, et al.
Pubblicazione: (2025)
Automated Toll Management System Using RFID and Image Processing
di: Ahmed, Raihan, et al.
Pubblicazione: (2024)
di: Ahmed, Raihan, et al.
Pubblicazione: (2024)
Random Initialization of Gated Sparse Adapters
di: Retault, Vi, et al.
Pubblicazione: (2025)
di: Retault, Vi, et al.
Pubblicazione: (2025)
Forecasting and Mitigating Disruptions in Public Bus Transit Services
di: Han, Chaeeun, et al.
Pubblicazione: (2024)
di: Han, Chaeeun, et al.
Pubblicazione: (2024)
Integrating Social Determinants of Health into Knowledge Graphs: Evaluating Prediction Bias and Fairness in Healthcare
di: Shang, Tianqi, et al.
Pubblicazione: (2024)
di: Shang, Tianqi, et al.
Pubblicazione: (2024)
Mapping the Media Landscape: Predicting Factual Reporting and Political Bias Through Web Interactions
di: Sánchez-Cortés, Dairazalia, et al.
Pubblicazione: (2024)
di: Sánchez-Cortés, Dairazalia, et al.
Pubblicazione: (2024)
AI Companies Should Report Pre- and Post-Mitigation Safety Evaluations
di: Bowen, Dillon, et al.
Pubblicazione: (2025)
di: Bowen, Dillon, et al.
Pubblicazione: (2025)
Personalized Knowledge Tracing through Student Representation Reconstruction and Class Imbalance Mitigation
di: Chen, Zhiyu, et al.
Pubblicazione: (2024)
di: Chen, Zhiyu, et al.
Pubblicazione: (2024)
Multi-Layer Personalized Federated Learning for Mitigating Biases in Student Predictive Analytics
di: Chu, Yun-Wei, et al.
Pubblicazione: (2022)
di: Chu, Yun-Wei, et al.
Pubblicazione: (2022)
Towards Effective Discrimination Testing for Generative AI
di: Zollo, Thomas P., et al.
Pubblicazione: (2024)
di: Zollo, Thomas P., et al.
Pubblicazione: (2024)
Addressing Selection Bias in Computerized Adaptive Testing: A User-Wise Aggregate Influence Function Approach
di: Kwon, Soonwoo, et al.
Pubblicazione: (2023)
di: Kwon, Soonwoo, et al.
Pubblicazione: (2023)
One-vs.-One Mitigation of Intersectional Bias: A General Method to Extend Fairness-Aware Binary Classification
di: Kobayashi, Kenji, et al.
Pubblicazione: (2020)
di: Kobayashi, Kenji, et al.
Pubblicazione: (2020)
Mitigating Extrinsic Gender Bias for Bangla Classification Tasks
di: Joy, Sajib Kumar Saha, et al.
Pubblicazione: (2024)
di: Joy, Sajib Kumar Saha, et al.
Pubblicazione: (2024)
Segment Any Text: A Universal Approach for Robust, Efficient and Adaptable Sentence Segmentation
di: Frohmann, Markus, et al.
Pubblicazione: (2024)
di: Frohmann, Markus, et al.
Pubblicazione: (2024)
Balancing Fairness and Accuracy in Data-Restricted Binary Classification
di: Lazri, Zachary McBride, et al.
Pubblicazione: (2024)
di: Lazri, Zachary McBride, et al.
Pubblicazione: (2024)
Ordinal Behavior Classification of Student Online Course Interactions
di: Trask, Thomas
Pubblicazione: (2024)
di: Trask, Thomas
Pubblicazione: (2024)
Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval
di: Wilson, Kyra, et al.
Pubblicazione: (2024)
di: Wilson, Kyra, et al.
Pubblicazione: (2024)
Breaking Down Bias: On The Limits of Generalizable Pruning Strategies
di: Ma, Sibo, et al.
Pubblicazione: (2025)
di: Ma, Sibo, et al.
Pubblicazione: (2025)
Bias and Fairness in Large Language Models: A Survey
di: Gallegos, Isabel O., et al.
Pubblicazione: (2023)
di: Gallegos, Isabel O., et al.
Pubblicazione: (2023)
What the Weight?! A Unified Framework for Zero-Shot Knowledge Composition
di: Holtermann, Carolin, et al.
Pubblicazione: (2024)
di: Holtermann, Carolin, et al.
Pubblicazione: (2024)
Arbitrariness and Social Prediction: The Confounding Role of Variance in Fair Classification
di: Cooper, A. Feder, et al.
Pubblicazione: (2023)
di: Cooper, A. Feder, et al.
Pubblicazione: (2023)
PakBBQ: A Culturally Adapted Bias Benchmark for QA
di: Hashmat, Abdullah, et al.
Pubblicazione: (2025)
di: Hashmat, Abdullah, et al.
Pubblicazione: (2025)
Impacts of Racial Bias in Historical Training Data for News AI
di: Bhargava, Rahul, et al.
Pubblicazione: (2025)
di: Bhargava, Rahul, et al.
Pubblicazione: (2025)
Fair Classification with Partial Feedback: An Exploration-Based Data Collection Approach
di: Keswani, Vijay, et al.
Pubblicazione: (2024)
di: Keswani, Vijay, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Unlabeled Debiasing in Downstream Tasks via Class-wise Low Variance Regularization
di: Masoudian, Shahed, et al.
Pubblicazione: (2024) -
ScaLearn: Simple and Highly Parameter-Efficient Task Transfer by Learning to Scale
di: Frohmann, Markus, et al.
Pubblicazione: (2023) -
Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes
di: Masoudian, Shahed, et al.
Pubblicazione: (2025) -
Explanatory Interactive Machine Learning for Bias Mitigation in Visual Gender Classification
di: Satriani, Nathanya, et al.
Pubblicazione: (2026) -
Backdoor for Debias: Mitigating Model Bias with Backdoor Attack-based Artificial Bias
di: Wu, Shangxi, et al.
Pubblicazione: (2023)