Maturity Framework for Enhancing Machine Learning Quality
Fuente:
arXiv
Saved in:
| Main Authors: | Castelli, Angelantonio, Chouliaras, Georgios Christos, Goldenberg, Dmitri |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Data vs. Model Machine Learning Fairness Testing: An Empirical Study
by: Shome, Arumoy, et al.
Published: (2024)
by: Shome, Arumoy, et al.
Published: (2024)
Mitigating Attrition: Data-Driven Approach Using Machine Learning and Data Engineering
by: Vijayan, Naveen Edapurath
Published: (2025)
by: Vijayan, Naveen Edapurath
Published: (2025)
A Conceptual Framework for Ethical Evaluation of Machine Learning Systems
by: Gupta, Neha R., et al.
Published: (2024)
by: Gupta, Neha R., et al.
Published: (2024)
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
by: Taraghi, Mina, et al.
Published: (2024)
by: Taraghi, Mina, et al.
Published: (2024)
Contexts Matter: An Empirical Study on Contextual Influence in Fairness Testing for Deep Learning Systems
by: Du, Chengwen, et al.
Published: (2024)
by: Du, Chengwen, et al.
Published: (2024)
Machine Learning Models for the Early Detection of Burnout in Software Engineering: a Systematic Literature Review
by: Tulili, Tien Rahayu, et al.
Published: (2026)
by: Tulili, Tien Rahayu, et al.
Published: (2026)
Risk Management for Mitigating Benchmark Failure Modes: BenchRisk
by: McGregor, Sean, et al.
Published: (2025)
by: McGregor, Sean, et al.
Published: (2025)
Carbon Footprint Evaluation of Code Generation through LLM as a Service
by: Vartziotis, Tina, et al.
Published: (2025)
by: Vartziotis, Tina, et al.
Published: (2025)
FairSense: Long-Term Fairness Analysis of ML-Enabled Systems
by: She, Yining, et al.
Published: (2025)
by: She, Yining, et al.
Published: (2025)
BacPrep: Lessons from Deploying an LLM-Based Bacalaureat Assessment Platform
by: Dumitran, Adrian-Marius, et al.
Published: (2025)
by: Dumitran, Adrian-Marius, et al.
Published: (2025)
Whence Is A Model Fair? Fixing Fairness Bugs via Propensity Score Matching
by: Peng, Kewen, et al.
Published: (2025)
by: Peng, Kewen, et al.
Published: (2025)
Automated Reproducibility Has a Problem Statement Problem
by: Snelleman, Thijs, et al.
Published: (2025)
by: Snelleman, Thijs, et al.
Published: (2025)
Choosing the Right Path for AI Integration in Engineering Companies: A Strategic Guide
by: Dzhusupova, Rimma, et al.
Published: (2023)
by: Dzhusupova, Rimma, et al.
Published: (2023)
To Err is AI : A Case Study Informing LLM Flaw Reporting Practices
by: McGregor, Sean, et al.
Published: (2024)
by: McGregor, Sean, et al.
Published: (2024)
Fairpriori: Improving Biased Subgroup Discovery for Deep Neural Network Fairness
by: Zhou, Kacy, et al.
Published: (2024)
by: Zhou, Kacy, et al.
Published: (2024)
FairLay-ML: Intuitive Debugging of Fairness in Data-Driven Social-Critical Software
by: Yu, Normen, et al.
Published: (2024)
by: Yu, Normen, et al.
Published: (2024)
Leveraging Imperfect Sources to Detect Fairwashing in Black-Box Auditing
by: Bourrée, Jade Garcia, et al.
Published: (2023)
by: Bourrée, Jade Garcia, et al.
Published: (2023)
Mining patterns in syntax trees to automate code reviews of student solutions for programming exercises
by: Van Petegem, Charlotte, et al.
Published: (2024)
by: Van Petegem, Charlotte, et al.
Published: (2024)
Leveraging Generative AI for Enhancing Automated Assessment in Programming Education Contests
by: Dascalescu, Stefan, et al.
Published: (2025)
by: Dascalescu, Stefan, et al.
Published: (2025)
An AI System Evaluation Framework for Advancing AI Safety: Terminology, Taxonomy, Lifecycle Mapping
by: Xia, Boming, et al.
Published: (2024)
by: Xia, Boming, et al.
Published: (2024)
The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence
by: White, Matt, et al.
Published: (2024)
by: White, Matt, et al.
Published: (2024)
Predicting Likely-Vulnerable Code Changes: Machine Learning-based Vulnerability Protections for Android Open Source Project
by: Yim, Keun Soo
Published: (2024)
by: Yim, Keun Soo
Published: (2024)
CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs
by: Guo, Hanxi, et al.
Published: (2025)
by: Guo, Hanxi, et al.
Published: (2025)
Fairness Testing through Extreme Value Theory
by: Monjezi, Verya, et al.
Published: (2025)
by: Monjezi, Verya, et al.
Published: (2025)
Federated Data Analytics for Cancer Immunotherapy: A Privacy-Preserving Collaborative Platform for Patient Management
by: Raheem, Mira, et al.
Published: (2025)
by: Raheem, Mira, et al.
Published: (2025)
Rethinking Technological Readiness in the Era of AI Uncertainty
by: Browne, S. Tucker, et al.
Published: (2025)
by: Browne, S. Tucker, et al.
Published: (2025)
Measuring Agents in Production
by: Pan, Melissa Z., et al.
Published: (2025)
by: Pan, Melissa Z., et al.
Published: (2025)
Breaking the ICE: Exploring promises and challenges of benchmarks for Inference Carbon & Energy estimation for LLMs
by: Sikand, Samarth, et al.
Published: (2025)
by: Sikand, Samarth, et al.
Published: (2025)
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
by: Vishwarupe, Varad, et al.
Published: (2026)
by: Vishwarupe, Varad, et al.
Published: (2026)
Predicting Fairness of ML Software Configurations
by: Herrera, Salvador Robles, et al.
Published: (2024)
by: Herrera, Salvador Robles, et al.
Published: (2024)
MAFT: Efficient Model-Agnostic Fairness Testing for Deep Neural Networks via Zero-Order Gradient Search
by: Wang, Zhaohui, et al.
Published: (2024)
by: Wang, Zhaohui, et al.
Published: (2024)
Language model developers should report train-test overlap
by: Zhang, Andy K, et al.
Published: (2024)
by: Zhang, Andy K, et al.
Published: (2024)
Green AI in Action: Strategic Model Selection for Ensembles in Production
by: Nijkamp, Nienke, et al.
Published: (2024)
by: Nijkamp, Nienke, et al.
Published: (2024)
DC-Check: A Data-Centric AI checklist to guide the development of reliable machine learning systems
by: Seedat, Nabeel, et al.
Published: (2022)
by: Seedat, Nabeel, et al.
Published: (2022)
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
by: Jewitt, James, et al.
Published: (2026)
by: Jewitt, James, et al.
Published: (2026)
Privacy and Copyright Protection in Generative AI: A Lifecycle Perspective
by: Zhang, Dawen, et al.
Published: (2023)
by: Zhang, Dawen, et al.
Published: (2023)
Fairness Improvement with Multiple Protected Attributes: How Far Are We?
by: Chen, Zhenpeng, et al.
Published: (2023)
by: Chen, Zhenpeng, et al.
Published: (2023)
Legal Aspects for Software Developers Interested in Generative AI Applications
by: Herbold, Steffen, et al.
Published: (2024)
by: Herbold, Steffen, et al.
Published: (2024)
Optimizing Travel Itineraries with AI Algorithms in a Microservices Architecture: Balancing Cost, Time, Preferences, and Sustainability
by: Barua, Biman, et al.
Published: (2024)
by: Barua, Biman, et al.
Published: (2024)
Unlocking Mental Health: Exploring College Students' Well-being through Smartphone Behaviors
by: Xuan, Wei, et al.
Published: (2025)
by: Xuan, Wei, et al.
Published: (2025)
Similar Items
-
Data vs. Model Machine Learning Fairness Testing: An Empirical Study
by: Shome, Arumoy, et al.
Published: (2024) -
Mitigating Attrition: Data-Driven Approach Using Machine Learning and Data Engineering
by: Vijayan, Naveen Edapurath
Published: (2025) -
A Conceptual Framework for Ethical Evaluation of Machine Learning Systems
by: Gupta, Neha R., et al.
Published: (2024) -
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
by: Taraghi, Mina, et al.
Published: (2024) -
Contexts Matter: An Empirical Study on Contextual Influence in Fairness Testing for Deep Learning Systems
by: Du, Chengwen, et al.
Published: (2024)