An empirical study of testing machine learning in the wild
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Openja, Moses, Khomh, Foutse, Foundjem, Armstrong, Ming, Zhen, Jiang, Abidi, Mouna, Hassan, Ahmed E. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FairFLRep: Fairness aware fault localization and repair of Deep Neural Networks
von: Openja, Moses, et al.
Veröffentlicht: (2025)
von: Openja, Moses, et al.
Veröffentlicht: (2025)
Tracing Stereotypes in Pre-trained Transformers: From Biased Neurons to Fairer Models
von: Voria, Gianmario, et al.
Veröffentlicht: (2026)
von: Voria, Gianmario, et al.
Veröffentlicht: (2026)
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
von: Taraghi, Mina, et al.
Veröffentlicht: (2024)
von: Taraghi, Mina, et al.
Veröffentlicht: (2024)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
Quality Issues in Machine Learning Software Systems
von: Côté, Pierre-Olivier, et al.
Veröffentlicht: (2023)
von: Côté, Pierre-Olivier, et al.
Veröffentlicht: (2023)
Improving the Robustness of Large Language Models for Code Tasks via Fine-tuning with Perturbed Data
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
Structural Anchors and Reasoning Fragility:Understanding CoT Robustness in LLM4Code
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
On the Effectiveness of Log Representation for Log-based Anomaly Detection
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
GIST: Generated Inputs Sets Transferability in Deep Learning
von: Tambon, Florian, et al.
Veröffentlicht: (2023)
von: Tambon, Florian, et al.
Veröffentlicht: (2023)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
Fault Localization in Deep Learning-based Software: A System-level Approach
von: Morovati, Mohammad Mehdi, et al.
Veröffentlicht: (2024)
von: Morovati, Mohammad Mehdi, et al.
Veröffentlicht: (2024)
Machine Learning Robustness: A Primer
von: Braiek, Houssem Ben, et al.
Veröffentlicht: (2024)
von: Braiek, Houssem Ben, et al.
Veröffentlicht: (2024)
Mining Action Rules for Defect Reduction Planning
von: Oueslati, Khouloud, et al.
Veröffentlicht: (2024)
von: Oueslati, Khouloud, et al.
Veröffentlicht: (2024)
An Efficient Model Maintenance Approach for MLOps
von: Majidi, Forough, et al.
Veröffentlicht: (2024)
von: Majidi, Forough, et al.
Veröffentlicht: (2024)
Towards Enhancing the Reproducibility of Deep Learning Bugs: An Empirical Study
von: Shah, Mehil B., et al.
Veröffentlicht: (2024)
von: Shah, Mehil B., et al.
Veröffentlicht: (2024)
Prism: Dynamic and Flexible Benchmarking of LLMs Code Generation with Monte Carlo Tree Search
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2025)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2025)
What Information Contributes to Log-based Anomaly Detection? Insights from a Configurable Transformer-Based Approach
von: Wu, Xingfang, et al.
Veröffentlicht: (2024)
von: Wu, Xingfang, et al.
Veröffentlicht: (2024)
Imitation Game: Reproducing Deep Learning Bugs Leveraging an Intelligent Agent
von: Shah, Mehil B, et al.
Veröffentlicht: (2025)
von: Shah, Mehil B, et al.
Veröffentlicht: (2025)
An Empirical Study of Self-Admitted Technical Debt in Machine Learning Software
von: Bhatia, Aaditya, et al.
Veröffentlicht: (2023)
von: Bhatia, Aaditya, et al.
Veröffentlicht: (2023)
Refining GPT-3 Embeddings with a Siamese Structure for Technical Post Duplicate Detection
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer
von: Bouchoucha, Rached, et al.
Veröffentlicht: (2024)
von: Bouchoucha, Rached, et al.
Veröffentlicht: (2024)
Towards Assessing Deep Learning Test Input Generators
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
Characterizing and Classifying Developer Forum Posts with their Intentions
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
Predicting post-release defects with knowledge units (KUs) of programming languages: an empirical study
von: Ahasanuzzaman, Md, et al.
Veröffentlicht: (2024)
von: Ahasanuzzaman, Md, et al.
Veröffentlicht: (2024)
DeepCodeProbe: Towards Understanding What Models Trained on Code Learn
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
Trimming the Risk: Towards Reliable Continuous Training for Deep Learning Inspection Systems
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2024)
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2024)
Protecting Privacy in Software Logs: What Should Be Anonymized?
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
Real Faults in Model Context Protocol (MCP) Software: a Comprehensive Taxonomy
von: Taraghi, Mina, et al.
Veröffentlicht: (2026)
von: Taraghi, Mina, et al.
Veröffentlicht: (2026)
Aggregating empirical evidence from data strategy studies: a case on model quantization
von: del Rey, Santiago, et al.
Veröffentlicht: (2025)
von: del Rey, Santiago, et al.
Veröffentlicht: (2025)
DeepTSF: Codeless machine learning operations for time series forecasting
von: Pelekis, Sotiris, et al.
Veröffentlicht: (2023)
von: Pelekis, Sotiris, et al.
Veröffentlicht: (2023)
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4
von: Abukhalaf, Seif, et al.
Veröffentlicht: (2024)
von: Abukhalaf, Seif, et al.
Veröffentlicht: (2024)
RefAgent: A Multi-agent LLM-based Framework for Automatic Software Refactoring
von: Oueslati, Khouloud, et al.
Veröffentlicht: (2025)
von: Oueslati, Khouloud, et al.
Veröffentlicht: (2025)
Leveraging Data Characteristics for Bug Localization in Deep Learning Programs
von: Manke, Ruchira, et al.
Veröffentlicht: (2024)
von: Manke, Ruchira, et al.
Veröffentlicht: (2024)
Towards Understanding the Impact of Data Bugs on Deep Learning Models in Software Engineering
von: Shah, Mehil B, et al.
Veröffentlicht: (2024)
von: Shah, Mehil B, et al.
Veröffentlicht: (2024)
Understanding Web Application Workloads and Their Applications: Systematic Literature Review and Characterization
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
Reputation Gaming in Stack Overflow
von: Mazloomzadeh, Iren, et al.
Veröffentlicht: (2021)
von: Mazloomzadeh, Iren, et al.
Veröffentlicht: (2021)
Mock Deep Testing: Toward Separate Development of Data and Models for Deep Learning
von: Manke, Ruchira, et al.
Veröffentlicht: (2025)
von: Manke, Ruchira, et al.
Veröffentlicht: (2025)
Risk Management for Mitigating Benchmark Failure Modes: BenchRisk
von: McGregor, Sean, et al.
Veröffentlicht: (2025)
von: McGregor, Sean, et al.
Veröffentlicht: (2025)
BloomAPR: A Bloom's Taxonomy-based Framework for Assessing the Capabilities of LLM-Powered APR Solutions
von: Ma, Yinghang, et al.
Veröffentlicht: (2025)
von: Ma, Yinghang, et al.
Veröffentlicht: (2025)
Towards Refining Developer Questions using LLM-Based Named Entity Recognition for Developer Chatroom Conversations
von: Fathollahzadeh, Pouya, et al.
Veröffentlicht: (2025)
von: Fathollahzadeh, Pouya, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FairFLRep: Fairness aware fault localization and repair of Deep Neural Networks
von: Openja, Moses, et al.
Veröffentlicht: (2025) -
Tracing Stereotypes in Pre-trained Transformers: From Biased Neurons to Fairer Models
von: Voria, Gianmario, et al.
Veröffentlicht: (2026) -
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
von: Taraghi, Mina, et al.
Veröffentlicht: (2024) -
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
von: Liu, Yang, et al.
Veröffentlicht: (2025) -
Quality Issues in Machine Learning Software Systems
von: Côté, Pierre-Olivier, et al.
Veröffentlicht: (2023)