Automated Trustworthiness Testing for Machine Learning Classifiers
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Cho, Steven, Cousins-Baxter, Seaton, Ruberto, Stefano, Terragni, Valerio |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
LLMORPH: Automated Metamorphic Testing of Large Language Models
par: Cho, Steven, et autres
Publié: (2026)
par: Cho, Steven, et autres
Publié: (2026)
Automated Trustworthiness Oracle Generation for Machine Learning Text Classifiers
par: Tung, Lam Nguyen, et autres
Publié: (2024)
par: Tung, Lam Nguyen, et autres
Publié: (2024)
Metamorphic Testing of Large Language Models for Natural Language Processing
par: Cho, Steven, et autres
Publié: (2025)
par: Cho, Steven, et autres
Publié: (2025)
Understanding LLM-Driven Test Oracle Generation
par: Bodicoat, Adam, et autres
Publié: (2026)
par: Bodicoat, Adam, et autres
Publié: (2026)
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
par: Ravi, Ravin, et autres
Publié: (2026)
par: Ravi, Ravin, et autres
Publié: (2026)
From Untestable to Testable: Metamorphic Testing in the Age of LLMs
par: Terragni, Valerio
Publié: (2026)
par: Terragni, Valerio
Publié: (2026)
MR-Scout: Automated Synthesis of Metamorphic Relations from Existing Test Cases
par: Xu, Congying, et autres
Publié: (2023)
par: Xu, Congying, et autres
Publié: (2023)
An Evolutionary Approach to Adapt Tests Across Mobile Apps
par: Mariani, Leonardo, et autres
Publié: (2021)
par: Mariani, Leonardo, et autres
Publié: (2021)
Automated Modernization of Machine Learning Engineering Notebooks for Reproducibility
par: Jin, Bihui, et autres
Publié: (2026)
par: Jin, Bihui, et autres
Publié: (2026)
Constrained Adversarial Learning for Automated Software Testing: a literature review
par: Vitorino, João, et autres
Publié: (2023)
par: Vitorino, João, et autres
Publié: (2023)
Evaluating Reinforcement Learning Safety and Trustworthiness in Cyber-Physical Systems
par: Dearstyne, Katherine, et autres
Publié: (2025)
par: Dearstyne, Katherine, et autres
Publié: (2025)
Automating the Training and Deployment of Models in MLOps by Integrating Systems with Machine Learning
par: Liang, Penghao, et autres
Publié: (2024)
par: Liang, Penghao, et autres
Publié: (2024)
Analysing Python Machine Learning Notebooks with Moose
par: Mignard, Marius, et autres
Publié: (2025)
par: Mignard, Marius, et autres
Publié: (2025)
Identifying Flaky Tests in Quantum Code: A Machine Learning Approach
par: Kaur, Khushdeep, et autres
Publié: (2025)
par: Kaur, Khushdeep, et autres
Publié: (2025)
Improving LLM-Driven Test Generation by Learning from Mocking Information
par: Lee, Jamie, et autres
Publié: (2026)
par: Lee, Jamie, et autres
Publié: (2026)
Inferring Data Preconditions from Deep Learning Models for Trustworthy Prediction in Deployment
par: Ahmed, Shibbir, et autres
Publié: (2024)
par: Ahmed, Shibbir, et autres
Publié: (2024)
Adoption and Evolution of Code Style and Best Programming Practices in Open-Source Projects
par: Kupari, Alvari, et autres
Publié: (2026)
par: Kupari, Alvari, et autres
Publié: (2026)
Towards Trustworthy AI Software Development Assistance
par: Maninger, Daniel, et autres
Publié: (2023)
par: Maninger, Daniel, et autres
Publié: (2023)
Automating REST API Postman Test Cases Using LLM
par: Sri, S Deepika, et autres
Publié: (2024)
par: Sri, S Deepika, et autres
Publié: (2024)
CA2: Code-Aware Agent for Automated Game Testing
par: Adaikkappan, Valliappan Chidambaram, et autres
Publié: (2026)
par: Adaikkappan, Valliappan Chidambaram, et autres
Publié: (2026)
MR-Coupler: Automated Metamorphic Test Generation via Functional Coupling Analysis
par: Xu, Congying, et autres
Publié: (2026)
par: Xu, Congying, et autres
Publié: (2026)
PrismaDV: Automated Task-Aware Data Unit Test Generation
par: Chen, Hao, et autres
Publié: (2026)
par: Chen, Hao, et autres
Publié: (2026)
Workflow-Level Design Principles for Trustworthy GenAI in Automotive System Engineering
par: Cheng, Chih-Hong, et autres
Publié: (2026)
par: Cheng, Chih-Hong, et autres
Publié: (2026)
Understanding and Characterizing Mock Assertions in Unit Tests
par: Zhu, Hengcheng, et autres
Publié: (2025)
par: Zhu, Hengcheng, et autres
Publié: (2025)
From Theory to Practice: Real-World Use Cases on Trustworthy LLM-Driven Process Modeling, Prediction and Automation
par: Pfeiffer, Peter, et autres
Publié: (2025)
par: Pfeiffer, Peter, et autres
Publié: (2025)
Geospatial Machine Learning Libraries
par: Stewart, Adam J., et autres
Publié: (2025)
par: Stewart, Adam J., et autres
Publié: (2025)
Data Virtualization for Machine Learning
par: Khan, Saiful, et autres
Publié: (2025)
par: Khan, Saiful, et autres
Publié: (2025)
Automating Formal Verification with Reinforcement Learning and Recursive Inference
par: Tan, Max
Publié: (2026)
par: Tan, Max
Publié: (2026)
On Extending the Automatic Test Markup Language (ATML) for Machine Learning
par: Cody, Tyler, et autres
Publié: (2024)
par: Cody, Tyler, et autres
Publié: (2024)
The Impact of Software Testing with Quantum Optimization Meets Machine Learning
par: Bandarupalli, Gopichand
Publié: (2025)
par: Bandarupalli, Gopichand
Publié: (2025)
Machine Learning Systems are Bloated and Vulnerable
par: Zhang, Huaifeng, et autres
Publié: (2022)
par: Zhang, Huaifeng, et autres
Publié: (2022)
Targeted Deep Learning System Boundary Testing
par: Weißl, Oliver, et autres
Publié: (2024)
par: Weißl, Oliver, et autres
Publié: (2024)
Data vs. Model Machine Learning Fairness Testing: An Empirical Study
par: Shome, Arumoy, et autres
Publié: (2024)
par: Shome, Arumoy, et autres
Publié: (2024)
How do Machine Learning Models Change?
par: Castaño, Joel, et autres
Publié: (2024)
par: Castaño, Joel, et autres
Publié: (2024)
Machine Learning Operations: A Mapping Study
par: Chakraborty, Abhijit, et autres
Publié: (2024)
par: Chakraborty, Abhijit, et autres
Publié: (2024)
Optimising for Energy Efficiency and Performance in Machine Learning
par: Ferreira, Emile Dos Santos, et autres
Publié: (2026)
par: Ferreira, Emile Dos Santos, et autres
Publié: (2026)
DEM: A Method for Certifying Deep Neural Network Classifier Outputs in Aerospace
par: Katz, Guy, et autres
Publié: (2024)
par: Katz, Guy, et autres
Publié: (2024)
Automated Machine Learning: A Case Study on Non-Intrusive Appliance Load Monitoring
par: Moin, Armin, et autres
Publié: (2022)
par: Moin, Armin, et autres
Publié: (2022)
JetTrain: IDE-Native Machine Learning Experiments
par: Trofimov, Artem, et autres
Publié: (2024)
par: Trofimov, Artem, et autres
Publié: (2024)
Understanding Practitioners Perspectives on Monitoring Machine Learning Systems
par: Naveed, Hira, et autres
Publié: (2025)
par: Naveed, Hira, et autres
Publié: (2025)
Documents similaires
-
LLMORPH: Automated Metamorphic Testing of Large Language Models
par: Cho, Steven, et autres
Publié: (2026) -
Automated Trustworthiness Oracle Generation for Machine Learning Text Classifiers
par: Tung, Lam Nguyen, et autres
Publié: (2024) -
Metamorphic Testing of Large Language Models for Natural Language Processing
par: Cho, Steven, et autres
Publié: (2025) -
Understanding LLM-Driven Test Oracle Generation
par: Bodicoat, Adam, et autres
Publié: (2026) -
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
par: Ravi, Ravin, et autres
Publié: (2026)