Leveraging Imperfect Sources to Detect Fairwashing in Black-Box Auditing
Fuente:
arXiv
Salvato in:
| Autori principali: | Bourrée, Jade Garcia, Merrer, Erwan Le, Tredan, Gilles, Rottembourg, Benoît |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
P2NIA: Privacy-Preserving Non-Iterative Auditing
di: Bourrée, Jade Garcia, et al.
Pubblicazione: (2025)
di: Bourrée, Jade Garcia, et al.
Pubblicazione: (2025)
Fairness Auditing with Multi-Agent Collaboration
di: de Vos, Martijn, et al.
Pubblicazione: (2024)
di: de Vos, Martijn, et al.
Pubblicazione: (2024)
Robust ML Auditing using Prior Knowledge
di: Bourrée, Jade Garcia, et al.
Pubblicazione: (2025)
di: Bourrée, Jade Garcia, et al.
Pubblicazione: (2025)
CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs
di: Guo, Hanxi, et al.
Pubblicazione: (2025)
di: Guo, Hanxi, et al.
Pubblicazione: (2025)
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
di: Jewitt, James, et al.
Pubblicazione: (2026)
di: Jewitt, James, et al.
Pubblicazione: (2026)
Carbon Footprint Evaluation of Code Generation through LLM as a Service
di: Vartziotis, Tina, et al.
Pubblicazione: (2025)
di: Vartziotis, Tina, et al.
Pubblicazione: (2025)
Leveraging Generative AI for Enhancing Automated Assessment in Programming Education Contests
di: Dascalescu, Stefan, et al.
Pubblicazione: (2025)
di: Dascalescu, Stefan, et al.
Pubblicazione: (2025)
Choosing the Right Path for AI Integration in Engineering Companies: A Strategic Guide
di: Dzhusupova, Rimma, et al.
Pubblicazione: (2023)
di: Dzhusupova, Rimma, et al.
Pubblicazione: (2023)
Risk Management for Mitigating Benchmark Failure Modes: BenchRisk
di: McGregor, Sean, et al.
Pubblicazione: (2025)
di: McGregor, Sean, et al.
Pubblicazione: (2025)
Contexts Matter: An Empirical Study on Contextual Influence in Fairness Testing for Deep Learning Systems
di: Du, Chengwen, et al.
Pubblicazione: (2024)
di: Du, Chengwen, et al.
Pubblicazione: (2024)
To Err is AI : A Case Study Informing LLM Flaw Reporting Practices
di: McGregor, Sean, et al.
Pubblicazione: (2024)
di: McGregor, Sean, et al.
Pubblicazione: (2024)
Fairpriori: Improving Biased Subgroup Discovery for Deep Neural Network Fairness
di: Zhou, Kacy, et al.
Pubblicazione: (2024)
di: Zhou, Kacy, et al.
Pubblicazione: (2024)
FairLay-ML: Intuitive Debugging of Fairness in Data-Driven Social-Critical Software
di: Yu, Normen, et al.
Pubblicazione: (2024)
di: Yu, Normen, et al.
Pubblicazione: (2024)
Mitigating Attrition: Data-Driven Approach Using Machine Learning and Data Engineering
di: Vijayan, Naveen Edapurath
Pubblicazione: (2025)
di: Vijayan, Naveen Edapurath
Pubblicazione: (2025)
FairSense: Long-Term Fairness Analysis of ML-Enabled Systems
di: She, Yining, et al.
Pubblicazione: (2025)
di: She, Yining, et al.
Pubblicazione: (2025)
BacPrep: Lessons from Deploying an LLM-Based Bacalaureat Assessment Platform
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
Whence Is A Model Fair? Fixing Fairness Bugs via Propensity Score Matching
di: Peng, Kewen, et al.
Pubblicazione: (2025)
di: Peng, Kewen, et al.
Pubblicazione: (2025)
Automated Reproducibility Has a Problem Statement Problem
di: Snelleman, Thijs, et al.
Pubblicazione: (2025)
di: Snelleman, Thijs, et al.
Pubblicazione: (2025)
Mining patterns in syntax trees to automate code reviews of student solutions for programming exercises
di: Van Petegem, Charlotte, et al.
Pubblicazione: (2024)
di: Van Petegem, Charlotte, et al.
Pubblicazione: (2024)
Maturity Framework for Enhancing Machine Learning Quality
di: Castelli, Angelantonio, et al.
Pubblicazione: (2025)
di: Castelli, Angelantonio, et al.
Pubblicazione: (2025)
Data vs. Model Machine Learning Fairness Testing: An Empirical Study
di: Shome, Arumoy, et al.
Pubblicazione: (2024)
di: Shome, Arumoy, et al.
Pubblicazione: (2024)
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
di: Taraghi, Mina, et al.
Pubblicazione: (2024)
di: Taraghi, Mina, et al.
Pubblicazione: (2024)
Machine Learning Models for the Early Detection of Burnout in Software Engineering: a Systematic Literature Review
di: Tulili, Tien Rahayu, et al.
Pubblicazione: (2026)
di: Tulili, Tien Rahayu, et al.
Pubblicazione: (2026)
Log Probability Tracking of LLM APIs
di: Chauvin, Timothée, et al.
Pubblicazione: (2025)
di: Chauvin, Timothée, et al.
Pubblicazione: (2025)
The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence
di: White, Matt, et al.
Pubblicazione: (2024)
di: White, Matt, et al.
Pubblicazione: (2024)
I hope we don't do to trust what advertising has done to love
di: Alglave, Jade
Pubblicazione: (2026)
di: Alglave, Jade
Pubblicazione: (2026)
Queries, Representation & Detection: The Next 100 Model Fingerprinting Schemes
di: Godinot, Augustin, et al.
Pubblicazione: (2024)
di: Godinot, Augustin, et al.
Pubblicazione: (2024)
Predicting Likely-Vulnerable Code Changes: Machine Learning-based Vulnerability Protections for Android Open Source Project
di: Yim, Keun Soo
Pubblicazione: (2024)
di: Yim, Keun Soo
Pubblicazione: (2024)
Latent Imitator: Generating Natural Individual Discriminatory Instances for Black-Box Fairness Testing
di: Xiao, Yisong, et al.
Pubblicazione: (2023)
di: Xiao, Yisong, et al.
Pubblicazione: (2023)
Privacy and Copyright Protection in Generative AI: A Lifecycle Perspective
di: Zhang, Dawen, et al.
Pubblicazione: (2023)
di: Zhang, Dawen, et al.
Pubblicazione: (2023)
Fairness Improvement with Multiple Protected Attributes: How Far Are We?
di: Chen, Zhenpeng, et al.
Pubblicazione: (2023)
di: Chen, Zhenpeng, et al.
Pubblicazione: (2023)
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
di: Vishwarupe, Varad, et al.
Pubblicazione: (2026)
di: Vishwarupe, Varad, et al.
Pubblicazione: (2026)
An AI System Evaluation Framework for Advancing AI Safety: Terminology, Taxonomy, Lifecycle Mapping
di: Xia, Boming, et al.
Pubblicazione: (2024)
di: Xia, Boming, et al.
Pubblicazione: (2024)
Fairness Testing through Extreme Value Theory
di: Monjezi, Verya, et al.
Pubblicazione: (2025)
di: Monjezi, Verya, et al.
Pubblicazione: (2025)
Federated Data Analytics for Cancer Immunotherapy: A Privacy-Preserving Collaborative Platform for Patient Management
di: Raheem, Mira, et al.
Pubblicazione: (2025)
di: Raheem, Mira, et al.
Pubblicazione: (2025)
Rethinking Technological Readiness in the Era of AI Uncertainty
di: Browne, S. Tucker, et al.
Pubblicazione: (2025)
di: Browne, S. Tucker, et al.
Pubblicazione: (2025)
Predicting Fairness of ML Software Configurations
di: Herrera, Salvador Robles, et al.
Pubblicazione: (2024)
di: Herrera, Salvador Robles, et al.
Pubblicazione: (2024)
MAFT: Efficient Model-Agnostic Fairness Testing for Deep Neural Networks via Zero-Order Gradient Search
di: Wang, Zhaohui, et al.
Pubblicazione: (2024)
di: Wang, Zhaohui, et al.
Pubblicazione: (2024)
A Conceptual Framework for Ethical Evaluation of Machine Learning Systems
di: Gupta, Neha R., et al.
Pubblicazione: (2024)
di: Gupta, Neha R., et al.
Pubblicazione: (2024)
Language model developers should report train-test overlap
di: Zhang, Andy K, et al.
Pubblicazione: (2024)
di: Zhang, Andy K, et al.
Pubblicazione: (2024)
Documenti analoghi
-
P2NIA: Privacy-Preserving Non-Iterative Auditing
di: Bourrée, Jade Garcia, et al.
Pubblicazione: (2025) -
Fairness Auditing with Multi-Agent Collaboration
di: de Vos, Martijn, et al.
Pubblicazione: (2024) -
Robust ML Auditing using Prior Knowledge
di: Bourrée, Jade Garcia, et al.
Pubblicazione: (2025) -
CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs
di: Guo, Hanxi, et al.
Pubblicazione: (2025) -
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
di: Jewitt, James, et al.
Pubblicazione: (2026)