On the Vulnerability of Fairness Constrained Learning to Malicious Noise
Fuente:
arXiv
Salvato in:
| Autori principali: | Blum, Avrim, Okoroafor, Princewill, Saha, Aadirupa, Stangl, Kevin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Training AI to be Loyal
di: Oh, Sewoong, et al.
Pubblicazione: (2025)
di: Oh, Sewoong, et al.
Pubblicazione: (2025)
Agnostic Multi-Robust Learning Using ERM
di: Ahmadi, Saba, et al.
Pubblicazione: (2023)
di: Ahmadi, Saba, et al.
Pubblicazione: (2023)
PUFFLE: Balancing Privacy, Utility, and Fairness in Federated Learning
di: Corbucci, Luca, et al.
Pubblicazione: (2024)
di: Corbucci, Luca, et al.
Pubblicazione: (2024)
Recovering from Biased Data: Can Fairness Constraints Improve Accuracy?
di: Blum, Avrim, et al.
Pubblicazione: (2019)
di: Blum, Avrim, et al.
Pubblicazione: (2019)
FAIRPLAI: A Human-in-the-Loop Approach to Fair and Private Machine Learning
di: Sanchez Jr., David, et al.
Pubblicazione: (2025)
di: Sanchez Jr., David, et al.
Pubblicazione: (2025)
Revisiting the Performance of Deep Learning-Based Vulnerability Detection on Realistic Datasets
di: Chakraborty, Partha, et al.
Pubblicazione: (2024)
di: Chakraborty, Partha, et al.
Pubblicazione: (2024)
VISION: Robust and Interpretable Code Vulnerability Detection Leveraging Counterfactual Augmentation
di: Egea, David, et al.
Pubblicazione: (2025)
di: Egea, David, et al.
Pubblicazione: (2025)
Inferring Discussion Topics about Exploitation of Vulnerabilities from Underground Hacking Forums
di: Moreno-Vera, Felipe
Pubblicazione: (2024)
di: Moreno-Vera, Felipe
Pubblicazione: (2024)
Privacy at a Price: Exploring its Dual Impact on AI Fairness
di: Yang, Mengmeng, et al.
Pubblicazione: (2024)
di: Yang, Mengmeng, et al.
Pubblicazione: (2024)
Combating Phone Scams with LLM-based Detection: Where Do We Stand?
di: Shen, Zitong, et al.
Pubblicazione: (2024)
di: Shen, Zitong, et al.
Pubblicazione: (2024)
Safeguarding Efficacy in Large Language Models: Evaluating Resistance to Human-Written and Algorithmic Adversarial Prompts
di: Downey-Webb, Tiarnaigh, et al.
Pubblicazione: (2025)
di: Downey-Webb, Tiarnaigh, et al.
Pubblicazione: (2025)
Dynamic Relation-Attentive Graph Neural Networks for Fraud Detection
di: Kim, Heehyeon, et al.
Pubblicazione: (2023)
di: Kim, Heehyeon, et al.
Pubblicazione: (2023)
Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain
di: Boisvert, Léo, et al.
Pubblicazione: (2025)
di: Boisvert, Léo, et al.
Pubblicazione: (2025)
MalDataGen: A Modular Framework for Synthetic Tabular Data Generation in Malware Detection
di: Paim, Kayua Oleques, et al.
Pubblicazione: (2025)
di: Paim, Kayua Oleques, et al.
Pubblicazione: (2025)
DP-Dueling: Learning from Preference Feedback without Compromising User Privacy
di: Saha, Aadirupa, et al.
Pubblicazione: (2024)
di: Saha, Aadirupa, et al.
Pubblicazione: (2024)
Predicting Likely-Vulnerable Code Changes: Machine Learning-based Vulnerability Protections for Android Open Source Project
di: Yim, Keun Soo
Pubblicazione: (2024)
di: Yim, Keun Soo
Pubblicazione: (2024)
MathLedger: A Verifiable Learning Substrate with Ledger-Attested Feedback
di: Abdullah, Ismail Ahmad
Pubblicazione: (2025)
di: Abdullah, Ismail Ahmad
Pubblicazione: (2025)
Mitigating Label Flipping Attacks in Malicious URL Detectors Using Ensemble Trees
di: Nowroozi, Ehsan, et al.
Pubblicazione: (2024)
di: Nowroozi, Ehsan, et al.
Pubblicazione: (2024)
Exploiting Novel GPT-4 APIs
di: Pelrine, Kellin, et al.
Pubblicazione: (2023)
di: Pelrine, Kellin, et al.
Pubblicazione: (2023)
Scaling Trends in Language Model Robustness
di: Howe, Nikolaus, et al.
Pubblicazione: (2024)
di: Howe, Nikolaus, et al.
Pubblicazione: (2024)
Adversarial Attacks on Large Language Models Using Regularized Relaxation
di: Chacko, Samuel Jacob, et al.
Pubblicazione: (2024)
di: Chacko, Samuel Jacob, et al.
Pubblicazione: (2024)
Unlocking the Potential of Large Language Models for Clinical Text Anonymization: A Comparative Study
di: Pissarra, David, et al.
Pubblicazione: (2024)
di: Pissarra, David, et al.
Pubblicazione: (2024)
David and Goliath: An Empirical Evaluation of Attacks and Defenses for QNNs at the Deep Edge
di: Costa, Miguel, et al.
Pubblicazione: (2024)
di: Costa, Miguel, et al.
Pubblicazione: (2024)
FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF
di: Fan, Flint Xiaofeng, et al.
Pubblicazione: (2024)
di: Fan, Flint Xiaofeng, et al.
Pubblicazione: (2024)
TrojanTime: Backdoor Attacks on Time Series Classification
di: Dong, Chang, et al.
Pubblicazione: (2025)
di: Dong, Chang, et al.
Pubblicazione: (2025)
Towards Independence Criterion in Machine Unlearning of Features and Labels
di: Han, Ling, et al.
Pubblicazione: (2024)
di: Han, Ling, et al.
Pubblicazione: (2024)
Privacy-preserving Universal Adversarial Defense for Black-box Models
di: Li, Qiao, et al.
Pubblicazione: (2024)
di: Li, Qiao, et al.
Pubblicazione: (2024)
Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems
di: Hackett, William, et al.
Pubblicazione: (2025)
di: Hackett, William, et al.
Pubblicazione: (2025)
JavelinGuard: Low-Cost Transformer Architectures for LLM Security
di: Datta, Yash, et al.
Pubblicazione: (2025)
di: Datta, Yash, et al.
Pubblicazione: (2025)
A Novel Self-Attention-Enabled Weighted Ensemble-Based Convolutional Neural Network Framework for Distributed Denial of Service Attack Classification
di: S, Kanthimathi, et al.
Pubblicazione: (2024)
di: S, Kanthimathi, et al.
Pubblicazione: (2024)
Unlearning at Scale: Implementing the Right to be Forgotten in Large Language Models
di: X, Abdullah
Pubblicazione: (2025)
di: X, Abdullah
Pubblicazione: (2025)
Privacy Constrained Fairness Estimation for Decision Trees
di: van der Steen, Florian, et al.
Pubblicazione: (2023)
di: van der Steen, Florian, et al.
Pubblicazione: (2023)
Defense Against the Dark Prompts: Mitigating Best-of-N Jailbreaking with Prompt Evaluation
di: Armstrong, Stuart, et al.
Pubblicazione: (2025)
di: Armstrong, Stuart, et al.
Pubblicazione: (2025)
An Information-Flow Perspective on Algorithmic Fairness
di: Teuber, Samuel, et al.
Pubblicazione: (2023)
di: Teuber, Samuel, et al.
Pubblicazione: (2023)
Variables are a Curse in Software Vulnerability Prediction
di: Groppe, Jinghua, et al.
Pubblicazione: (2024)
di: Groppe, Jinghua, et al.
Pubblicazione: (2024)
Keep on Swimming: Real Attackers Only Need Partial Knowledge of a Multi-Model System
di: Collado, Julian, et al.
Pubblicazione: (2024)
di: Collado, Julian, et al.
Pubblicazione: (2024)
Playing Devil's Advocate: Unmasking Toxicity and Vulnerabilities in Large Vision-Language Models
di: Erol, Abdulkadir, et al.
Pubblicazione: (2025)
di: Erol, Abdulkadir, et al.
Pubblicazione: (2025)
Evaluating Query Efficiency and Accuracy of Transfer Learning-based Model Extraction Attack in Federated Learning
di: Ahamed, Sayyed Farid, et al.
Pubblicazione: (2025)
di: Ahamed, Sayyed Farid, et al.
Pubblicazione: (2025)
The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation
di: Brundage, Miles, et al.
Pubblicazione: (2018)
di: Brundage, Miles, et al.
Pubblicazione: (2018)
Federated Learning Under Attack: Exposing Vulnerabilities through Data Poisoning Attacks in Computer Networks
di: Nowroozi, Ehsan, et al.
Pubblicazione: (2024)
di: Nowroozi, Ehsan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Training AI to be Loyal
di: Oh, Sewoong, et al.
Pubblicazione: (2025) -
Agnostic Multi-Robust Learning Using ERM
di: Ahmadi, Saba, et al.
Pubblicazione: (2023) -
PUFFLE: Balancing Privacy, Utility, and Fairness in Federated Learning
di: Corbucci, Luca, et al.
Pubblicazione: (2024) -
Recovering from Biased Data: Can Fairness Constraints Improve Accuracy?
di: Blum, Avrim, et al.
Pubblicazione: (2019) -
FAIRPLAI: A Human-in-the-Loop Approach to Fair and Private Machine Learning
di: Sanchez Jr., David, et al.
Pubblicazione: (2025)