230,439 Test Failures Later: An Empirical Evaluation of Flaky Failure Classifiers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Alshammari, Abdulrahman, Ammann, Paul, Hilton, Michael, Bell, Jonathan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Systemic Flakiness: An Empirical Analysis of Co-Occurring Flaky Test Failures
von: Parry, Owain, et al.
Veröffentlicht: (2025)
von: Parry, Owain, et al.
Veröffentlicht: (2025)
A Dataset of Reproducible Flaky-Test Failures
von: Rafi, Suzzana, et al.
Veröffentlicht: (2026)
von: Rafi, Suzzana, et al.
Veröffentlicht: (2026)
On the Diagnosis of Flaky Job Failures: Understanding and Prioritizing Failure Categories
von: Aïdasso, Henri, et al.
Veröffentlicht: (2025)
von: Aïdasso, Henri, et al.
Veröffentlicht: (2025)
Can We Classify Flaky Tests Using Only Test Code? An LLM-Based Empirical Study
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
The Effects of Computational Resources on Flaky Tests
von: Silva, Denini, et al.
Veröffentlicht: (2023)
von: Silva, Denini, et al.
Veröffentlicht: (2023)
FlakeRanker: Automated Identification and Prioritization of Flaky Job Failure Categories
von: Aïdasso, Henri
Veröffentlicht: (2025)
von: Aïdasso, Henri
Veröffentlicht: (2025)
A Practical Framework for Flaky Failure Triage in Distributed Database Continuous Integration
von: Zhu, Jun-Peng, et al.
Veröffentlicht: (2026)
von: Zhu, Jun-Peng, et al.
Veröffentlicht: (2026)
Do Test and Environmental Complexity Increase Flakiness? An Empirical Study of SAP HANA
von: Berndt, Alexander, et al.
Veröffentlicht: (2024)
von: Berndt, Alexander, et al.
Veröffentlicht: (2024)
A Systematic Evaluation of Environmental Flakiness in JavaScript Tests
von: Hashemi, Negar, et al.
Veröffentlicht: (2026)
von: Hashemi, Negar, et al.
Veröffentlicht: (2026)
Detecting and Evaluating Order-Dependent Flaky Tests in JavaScript
von: Hashemi, Negar, et al.
Veröffentlicht: (2025)
von: Hashemi, Negar, et al.
Veröffentlicht: (2025)
Taming Timeout Flakiness: An Empirical Study of SAP HANA
von: Berndt, Alexander, et al.
Veröffentlicht: (2024)
von: Berndt, Alexander, et al.
Veröffentlicht: (2024)
The Vocabulary of Flaky Tests in the Context of SAP HANA
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
Automated Test Validators for Flaky Cyber-Physical System Simulators: Approach and Evaluation
von: Jodat, Baharin A., et al.
Veröffentlicht: (2025)
von: Jodat, Baharin A., et al.
Veröffentlicht: (2025)
Is this Build Failure Related to my Patch? An Empirical Study of Unrelated Build Failures in Continuous Integration
von: Huang, Andie, et al.
Veröffentlicht: (2026)
von: Huang, Andie, et al.
Veröffentlicht: (2026)
Flaky Tests in a Large Industrial Database Management System: An Empirical Study of Fixed Issue Reports for SAP HANA
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
JS-TOD: Detecting Order-Dependent Flaky Tests in Jest
von: Hashemi, Negar, et al.
Veröffentlicht: (2025)
von: Hashemi, Negar, et al.
Veröffentlicht: (2025)
Detecting Flaky Tests in Quantum Software: A Dynamic Approach
von: Kim, Dongchan, et al.
Veröffentlicht: (2025)
von: Kim, Dongchan, et al.
Veröffentlicht: (2025)
Reduction of Test Re-runs by Prioritizing Potential Order Dependent Flaky Tests
von: Iqbal, Hasnain, et al.
Veröffentlicht: (2025)
von: Iqbal, Hasnain, et al.
Veröffentlicht: (2025)
A Generic Approach to Fix Test Flakiness in Real-World Projects
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
Dockerfile Flakiness: Characterization and Repair
von: Shabani, Taha, et al.
Veröffentlicht: (2024)
von: Shabani, Taha, et al.
Veröffentlicht: (2024)
On the Flakiness of LLM-Generated Tests for Industrial and Open-Source Database Management Systems
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
NeuroFlake: A Neuro-Symbolic LLM Framework for Flaky Test Classification
von: Hoque, Khondaker Tasnia, et al.
Veröffentlicht: (2026)
von: Hoque, Khondaker Tasnia, et al.
Veröffentlicht: (2026)
A Preliminary Study of Fixed Flaky Tests in Rust Projects on GitHub
von: Schroeder, Tom, et al.
Veröffentlicht: (2025)
von: Schroeder, Tom, et al.
Veröffentlicht: (2025)
FlakyGuard: Automatically Fixing Flaky Tests at Industry Scale
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
Dissecting Bug Triggers and Failure Modes in Modern Agentic Frameworks: An Empirical Study
von: Zhang, Xiaowen, et al.
Veröffentlicht: (2026)
von: Zhang, Xiaowen, et al.
Veröffentlicht: (2026)
On the Illusion of Success: An Empirical Study of Build Reruns and Silent Failures in Industrial CI
von: Aïdasso, Henri, et al.
Veröffentlicht: (2025)
von: Aïdasso, Henri, et al.
Veröffentlicht: (2025)
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair
von: Fatima, Sakina, et al.
Veröffentlicht: (2023)
von: Fatima, Sakina, et al.
Veröffentlicht: (2023)
WEFix: Intelligent Automatic Generation of Explicit Waits for Efficient Web End-to-End Flaky Tests
von: Liu, Xinyue, et al.
Veröffentlicht: (2024)
von: Liu, Xinyue, et al.
Veröffentlicht: (2024)
Detecting and Mitigating Flakiness in REST API Fuzzing
von: Zhang, Man, et al.
Veröffentlicht: (2026)
von: Zhang, Man, et al.
Veröffentlicht: (2026)
Identifying Flaky Tests in Quantum Code: A Machine Learning Approach
von: Kaur, Khushdeep, et al.
Veröffentlicht: (2025)
von: Kaur, Khushdeep, et al.
Veröffentlicht: (2025)
Easy over Hard: A Simple Baseline for Test Failures Causes Prediction
von: Gao, Zhipeng, et al.
Veröffentlicht: (2024)
von: Gao, Zhipeng, et al.
Veröffentlicht: (2024)
An Empirical Study on Failures in Automated Issue Solving
von: Liu, Simiao, et al.
Veröffentlicht: (2025)
von: Liu, Simiao, et al.
Veröffentlicht: (2025)
Failure-Aware Enhancements for Large Language Model (LLM) Code Generation: An Empirical Study on Decision Framework
von: Shen, Jianru, et al.
Veröffentlicht: (2026)
von: Shen, Jianru, et al.
Veröffentlicht: (2026)
Understanding Code Agent Behaviour: An Empirical Study of Success and Failure Trajectories
von: Majgaonkar, Oorja, et al.
Veröffentlicht: (2025)
von: Majgaonkar, Oorja, et al.
Veröffentlicht: (2025)
Learning From Lessons Learned: Preliminary Findings From a Study of Learning From Failure
von: Sillito, Jonathan, et al.
Veröffentlicht: (2024)
von: Sillito, Jonathan, et al.
Veröffentlicht: (2024)
Systematic Evaluation of Deep Learning Models for Log-based Failure Prediction
von: Hadadi, Fatemeh, et al.
Veröffentlicht: (2023)
von: Hadadi, Fatemeh, et al.
Veröffentlicht: (2023)
LLMorpheus: Mutation Testing using Large Language Models
von: Tip, Frank, et al.
Veröffentlicht: (2024)
von: Tip, Frank, et al.
Veröffentlicht: (2024)
FlaKat: A Machine Learning-Based Categorization Framework for Flaky Tests
von: Lin, Shizhe, et al.
Veröffentlicht: (2024)
von: Lin, Shizhe, et al.
Veröffentlicht: (2024)
LLM-Based Automated Diagnosis Of Integration Test Failures At Google
von: Ziftci, Celal, et al.
Veröffentlicht: (2026)
von: Ziftci, Celal, et al.
Veröffentlicht: (2026)
A Large Language Model Approach to Identify Flakiness in C++ Projects
von: Sun, Xin, et al.
Veröffentlicht: (2024)
von: Sun, Xin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Systemic Flakiness: An Empirical Analysis of Co-Occurring Flaky Test Failures
von: Parry, Owain, et al.
Veröffentlicht: (2025) -
A Dataset of Reproducible Flaky-Test Failures
von: Rafi, Suzzana, et al.
Veröffentlicht: (2026) -
On the Diagnosis of Flaky Job Failures: Understanding and Prioritizing Failure Categories
von: Aïdasso, Henri, et al.
Veröffentlicht: (2025) -
Can We Classify Flaky Tests Using Only Test Code? An LLM-Based Empirical Study
von: Berndt, Alexander, et al.
Veröffentlicht: (2026) -
The Effects of Computational Resources on Flaky Tests
von: Silva, Denini, et al.
Veröffentlicht: (2023)