Fast algorithms to improve fair information access in networks

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Windham, Dennis Robert, Wendt, Caroline J., Crane, Alex, Warr, Madelyn J, Shi, Freda, Friedler, Sorelle A., Sullivan, Blair D., Clauset, Aaron
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866913697708900352
author Windham, Dennis Robert
Wendt, Caroline J.
Crane, Alex
Warr, Madelyn J
Shi, Freda
Friedler, Sorelle A.
Sullivan, Blair D.
Clauset, Aaron
author_facet Windham, Dennis Robert
Wendt, Caroline J.
Crane, Alex
Warr, Madelyn J
Shi, Freda
Friedler, Sorelle A.
Sullivan, Blair D.
Clauset, Aaron
contents We consider the problem of selecting $k$ seed nodes in a network to maximize the minimum probability of activation under an independent cascade beginning at these seeds. The motivation is to promote fairness by ensuring that even the least advantaged members of the network have good access to information. Our problem can be viewed as a variant of the classic influence maximization objective, but it appears somewhat more difficult to solve: only heuristics are known. Moreover, the scalability of these methods is sharply constrained by the need to repeatedly estimate access probabilities. We design and evaluate a suite of $10$ new scalable algorithms which crucially do not require probability estimation. To facilitate comparison with the state-of-the-art, we make three more contributions which may be of broader interest. We introduce a principled method of selecting a pairwise information transmission parameter used in experimental evaluations, as well as a new performance metric which allows for comparison of algorithms across a range of values for the parameter $k$. Finally, we provide a new benchmark corpus of $174$ networks drawn from $6$ domains. Our algorithms retain most of the performance of the state-of-the-art while reducing running time by orders of magnitude. Specifically, a meta-learner approach is on average only $20\%$ less effective than the state-of-the-art on held-out data, but about $75-130$ times faster. Further, the meta-learner's performance exceeds the state-of the-art on about $20\%$ of networks, and the magnitude of its running time advantage is maintained on much larger networks.
format Preprint
id arxiv_https___arxiv_org_abs_2409_03127
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Fast algorithms to improve fair information access in networks
Windham, Dennis Robert
Wendt, Caroline J.
Crane, Alex
Warr, Madelyn J
Shi, Freda
Friedler, Sorelle A.
Sullivan, Blair D.
Clauset, Aaron
Social and Information Networks
Computers and Society
Physics and Society
We consider the problem of selecting $k$ seed nodes in a network to maximize the minimum probability of activation under an independent cascade beginning at these seeds. The motivation is to promote fairness by ensuring that even the least advantaged members of the network have good access to information. Our problem can be viewed as a variant of the classic influence maximization objective, but it appears somewhat more difficult to solve: only heuristics are known. Moreover, the scalability of these methods is sharply constrained by the need to repeatedly estimate access probabilities. We design and evaluate a suite of $10$ new scalable algorithms which crucially do not require probability estimation. To facilitate comparison with the state-of-the-art, we make three more contributions which may be of broader interest. We introduce a principled method of selecting a pairwise information transmission parameter used in experimental evaluations, as well as a new performance metric which allows for comparison of algorithms across a range of values for the parameter $k$. Finally, we provide a new benchmark corpus of $174$ networks drawn from $6$ domains. Our algorithms retain most of the performance of the state-of-the-art while reducing running time by orders of magnitude. Specifically, a meta-learner approach is on average only $20\%$ less effective than the state-of-the-art on held-out data, but about $75-130$ times faster. Further, the meta-learner's performance exceeds the state-of the-art on about $20\%$ of networks, and the magnitude of its running time advantage is maintained on much larger networks.
title Fast algorithms to improve fair information access in networks
topic Social and Information Networks
Computers and Society
Physics and Society
url https://arxiv.org/abs/2409.03127