On the Illusion of Success: An Empirical Study of Build Reruns and Silent Failures in Industrial CI
Fuente:
arXiv
Saved in:
| Main Authors: | Aïdasso, Henri, Bordeleau, Francis, Tizghadam, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Diagnosis of Flaky Job Failures: Understanding and Prioritizing Failure Categories
by: Aïdasso, Henri, et al.
Published: (2025)
by: Aïdasso, Henri, et al.
Published: (2025)
Towards Build Optimization Using Digital Twins
by: Aïdasso, Henri, et al.
Published: (2025)
by: Aïdasso, Henri, et al.
Published: (2025)
Efficient Detection of Intermittent Job Failures Using Few-Shot Learning
by: Aïdasso, Henri, et al.
Published: (2025)
by: Aïdasso, Henri, et al.
Published: (2025)
Predicting Intermittent Job Failure Categories for Diagnosis Using Few-Shot Fine-Tuned Language Models
by: Aïdasso, Henri, et al.
Published: (2026)
by: Aïdasso, Henri, et al.
Published: (2026)
Build Optimization: A Systematic Literature Review
by: Aïdasso, Henri, et al.
Published: (2025)
by: Aïdasso, Henri, et al.
Published: (2025)
An Empirical Study on the Amount of Changes Required for Merge Request Acceptance
by: Kansab, Samah, et al.
Published: (2025)
by: Kansab, Samah, et al.
Published: (2025)
RevMine: An LLM-Assisted Tool for Code Review Mining and Analysis Across Git Platforms
by: Kansab, Samah, et al.
Published: (2025)
by: Kansab, Samah, et al.
Published: (2025)
FlakeRanker: Automated Identification and Prioritization of Flaky Job Failure Categories
by: Aïdasso, Henri
Published: (2025)
by: Aïdasso, Henri
Published: (2025)
On The Impact of Merge Request Deviations on Code Review Practices
by: Kansab, Samah, et al.
Published: (2025)
by: Kansab, Samah, et al.
Published: (2025)
Analyzing DevOps Practices Through Merge Request Data: A Case Study in Networking Software Company
by: Kansab, Samah, et al.
Published: (2025)
by: Kansab, Samah, et al.
Published: (2025)
The Promise and Reality of Continuous Integration Caching: An Empirical Study of Travis CI Builds
by: Ghaleb, Taher A., et al.
Published: (2026)
by: Ghaleb, Taher A., et al.
Published: (2026)
Role of CI Adoption in Mobile App Success: An Empirical Study of Open-Source Android Projects
by: Zhou, Xiaoxin, et al.
Published: (2026)
by: Zhou, Xiaoxin, et al.
Published: (2026)
A Model-Driven Digital Twin for the Systematic Improvement of DevOps Pipelines
by: Samoud, Achref, et al.
Published: (2026)
by: Samoud, Achref, et al.
Published: (2026)
Is this Build Failure Related to my Patch? An Empirical Study of Unrelated Build Failures in Continuous Integration
by: Huang, Andie, et al.
Published: (2026)
by: Huang, Andie, et al.
Published: (2026)
"Good" and "Bad" Failures in Industrial CI/CD -- Balancing Cost and Quality Assurance
by: Sun, Simin, et al.
Published: (2025)
by: Sun, Simin, et al.
Published: (2025)
Practitioners' Challenges and Perceptions of CI Build Failure Predictions at Atlassian
by: Hong, Yang, et al.
Published: (2024)
by: Hong, Yang, et al.
Published: (2024)
LogSage: An LLM-Based Framework for CI/CD Failure Detection and Remediation with Industrial Validation
by: Xu, Weiyuan, et al.
Published: (2025)
by: Xu, Weiyuan, et al.
Published: (2025)
Understanding Code Agent Behaviour: An Empirical Study of Success and Failure Trajectories
by: Majgaonkar, Oorja, et al.
Published: (2025)
by: Majgaonkar, Oorja, et al.
Published: (2025)
CI/CD Configuration Practices in Open-Source Android Apps: An Empirical Study
by: Ghaleb, Taher A., et al.
Published: (2024)
by: Ghaleb, Taher A., et al.
Published: (2024)
When AI Agents Touch CI/CD Configurations: Frequency and Success
by: Ghaleb, Taher A.
Published: (2026)
by: Ghaleb, Taher A.
Published: (2026)
Portable and Secure CI/CD for COBOL: Lessons from an Industrial Migration
by: Askholm, Andreas, et al.
Published: (2026)
by: Askholm, Andreas, et al.
Published: (2026)
Empirical Analysis on CI/CD Pipeline Evolution in Machine Learning Projects
by: Rzig, Dhia Elhaq, et al.
Published: (2024)
by: Rzig, Dhia Elhaq, et al.
Published: (2024)
How Developers Adopt, Use, and Evolve CI/CD Caching: An Empirical Study on GitHub Actions
by: Hasan, Kazi Amit, et al.
Published: (2026)
by: Hasan, Kazi Amit, et al.
Published: (2026)
Continuous Evolution of Digital Twins using the DarTwin Notation
by: Mertens, Joost, et al.
Published: (2024)
by: Mertens, Joost, et al.
Published: (2024)
Using Large Language Models to Support Automation of Failure Management in CI/CD Pipelines: A Case Study in SAP HANA
by: Bui, Duong, et al.
Published: (2026)
by: Bui, Duong, et al.
Published: (2026)
Diagnosing and Resolving Android Applications Building Issues: An Empirical Study
by: Bodepudi, Lakshmi Priya, et al.
Published: (2025)
by: Bodepudi, Lakshmi Priya, et al.
Published: (2025)
Dissecting Bug Triggers and Failure Modes in Modern Agentic Frameworks: An Empirical Study
by: Zhang, Xiaowen, et al.
Published: (2026)
by: Zhang, Xiaowen, et al.
Published: (2026)
"Silent Is Not Actually Silent": An Investigation of Toxicity on Bug Report Discussion
by: Imran, Mia Mohammad, et al.
Published: (2025)
by: Imran, Mia Mohammad, et al.
Published: (2025)
Cache-Related Smells in GitLab CI/CD: Comprehensive Catalog, Automated Detection, and Empirical Evidence
by: Urdih, Francesco, et al.
Published: (2026)
by: Urdih, Francesco, et al.
Published: (2026)
Fixseeker: An Empirical Driven Graph-based Approach for Detecting Silent Vulnerability Fixes in Open Source Software
by: Cheng, Yiran, et al.
Published: (2025)
by: Cheng, Yiran, et al.
Published: (2025)
230,439 Test Failures Later: An Empirical Evaluation of Flaky Failure Classifiers
by: Alshammari, Abdulrahman, et al.
Published: (2024)
by: Alshammari, Abdulrahman, et al.
Published: (2024)
Scalable CI/CD for Legacy Modernization: An Industrial Experience Addressing Internal Challenges Related to the 2025 Japan Cliff
by: Kudo, Kuniaki, et al.
Published: (2025)
by: Kudo, Kuniaki, et al.
Published: (2025)
AI builds, We Analyze: An Empirical Study of AI-Generated Build Code Quality
by: Ghammam, Anwar, et al.
Published: (2026)
by: Ghammam, Anwar, et al.
Published: (2026)
An Empirical Analysis of Compatibility Issues for Industrial Mobile Games
by: Song, Zihe, et al.
Published: (2025)
by: Song, Zihe, et al.
Published: (2025)
Failure-Aware Enhancements for Large Language Model (LLM) Code Generation: An Empirical Study on Decision Framework
by: Shen, Jianru, et al.
Published: (2026)
by: Shen, Jianru, et al.
Published: (2026)
CI-Repair-Bench: A Repository-Aware Benchmark for Automated Patch Validation via CI Workflows
by: Muna, Rabeya Khatun, et al.
Published: (2026)
by: Muna, Rabeya Khatun, et al.
Published: (2026)
Beyond Resolution Rates: Behavioral Drivers of Coding Agent Success and Failure
by: Mehtiyev, Tural, et al.
Published: (2026)
by: Mehtiyev, Tural, et al.
Published: (2026)
Aligning Academia with Industry: An Empirical Study of Industrial Needs and Academic Capabilities in AI-Driven Software Engineering
by: Yu, Hang, et al.
Published: (2025)
by: Yu, Hang, et al.
Published: (2025)
CI at Scale: Lean, Green, and Fast
by: Juloori, Dhruva, et al.
Published: (2025)
by: Juloori, Dhruva, et al.
Published: (2025)
Systemic Flakiness: An Empirical Analysis of Co-Occurring Flaky Test Failures
by: Parry, Owain, et al.
Published: (2025)
by: Parry, Owain, et al.
Published: (2025)
Similar Items
-
On the Diagnosis of Flaky Job Failures: Understanding and Prioritizing Failure Categories
by: Aïdasso, Henri, et al.
Published: (2025) -
Towards Build Optimization Using Digital Twins
by: Aïdasso, Henri, et al.
Published: (2025) -
Efficient Detection of Intermittent Job Failures Using Few-Shot Learning
by: Aïdasso, Henri, et al.
Published: (2025) -
Predicting Intermittent Job Failure Categories for Diagnosis Using Few-Shot Fine-Tuned Language Models
by: Aïdasso, Henri, et al.
Published: (2026) -
Build Optimization: A Systematic Literature Review
by: Aïdasso, Henri, et al.
Published: (2025)