Fair Mixed Effects Support Vector Machine

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Burgard, Jan Pablo, Pamplona, João Vitor
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916501994340352
author Burgard, Jan Pablo
Pamplona, João Vitor
author_facet Burgard, Jan Pablo
Pamplona, João Vitor
contents To ensure unbiased and ethical automated predictions, fairness must be a core principle in machine learning applications. Fairness in machine learning aims to mitigate biases present in the training data and model imperfections that could lead to discriminatory outcomes. This is achieved by preventing the model from making decisions based on sensitive characteristics like ethnicity or sexual orientation. A fundamental assumption in machine learning is the independence of observations. However, this assumption often does not hold true for data describing social phenomena, where data points are often clustered based. Hence, if the machine learning models do not account for the cluster correlations, the results may be biased. Especially high is the bias in cases where the cluster assignment is correlated to the variable of interest. We present a fair mixed effects support vector machine algorithm that can handle both problems simultaneously. With a reproducible simulation study we demonstrate the impact of clustered data on the quality of fair machine learning predictions.
format Preprint
id arxiv_https___arxiv_org_abs_2405_06433
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Fair Mixed Effects Support Vector Machine
Burgard, Jan Pablo
Pamplona, João Vitor
Machine Learning
Computers and Society
Optimization and Control
To ensure unbiased and ethical automated predictions, fairness must be a core principle in machine learning applications. Fairness in machine learning aims to mitigate biases present in the training data and model imperfections that could lead to discriminatory outcomes. This is achieved by preventing the model from making decisions based on sensitive characteristics like ethnicity or sexual orientation. A fundamental assumption in machine learning is the independence of observations. However, this assumption often does not hold true for data describing social phenomena, where data points are often clustered based. Hence, if the machine learning models do not account for the cluster correlations, the results may be biased. Especially high is the bias in cases where the cluster assignment is correlated to the variable of interest. We present a fair mixed effects support vector machine algorithm that can handle both problems simultaneously. With a reproducible simulation study we demonstrate the impact of clustered data on the quality of fair machine learning predictions.
title Fair Mixed Effects Support Vector Machine
topic Machine Learning
Computers and Society
Optimization and Control
url https://arxiv.org/abs/2405.06433