Unveiling Practical Shortcomings of Patch Overfitting Detection Techniques
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Williams, David, Avraam, Ioakim, Aleti, Aldeida, Martinez, Matias, Petke, Justyna, Sarro, Federica |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Test-based Patch Clustering for Automatically-Generated Patches Assessment
von: Martinez, Matias, et al.
Veröffentlicht: (2022)
von: Martinez, Matias, et al.
Veröffentlicht: (2022)
Empirical and Sustainability Aspects of Software Engineering Research in the Era of Large Language Models: A Reflection
von: Williams, David, et al.
Veröffentlicht: (2025)
von: Williams, David, et al.
Veröffentlicht: (2025)
Hot Fixing Software: A Comprehensive Review of Terminology, Techniques, and Applications
von: Hanna, Carol, et al.
Veröffentlicht: (2024)
von: Hanna, Carol, et al.
Veröffentlicht: (2024)
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
von: Hanna, Carol, et al.
Veröffentlicht: (2025)
von: Hanna, Carol, et al.
Veröffentlicht: (2025)
A Comprehensive Survey of Benchmarks for Automated Improvement of Software's Non-Functional Properties
von: Blot, Aymeric, et al.
Veröffentlicht: (2022)
von: Blot, Aymeric, et al.
Veröffentlicht: (2022)
PAFOT: A Position-Based Approach for Finding Optimal Tests of Autonomous Vehicles
von: Crespo-Rodriguez, Victor, et al.
Veröffentlicht: (2024)
von: Crespo-Rodriguez, Victor, et al.
Veröffentlicht: (2024)
Experimental evaluation of architectural software performance design patterns in microservices
von: Meijer, Willem, et al.
Veröffentlicht: (2024)
von: Meijer, Willem, et al.
Veröffentlicht: (2024)
Hot Fixing in the Wild
von: Hanna, Carol, et al.
Veröffentlicht: (2026)
von: Hanna, Carol, et al.
Veröffentlicht: (2026)
LLM-Guided Genetic Improvement: Envisioning Semantic Aware Automated Software Evolution
von: Even-Mendoza, Karine, et al.
Veröffentlicht: (2025)
von: Even-Mendoza, Karine, et al.
Veröffentlicht: (2025)
Neuron Patching: Semantic-based Neuron-level Language Model Repair for Code Generation
von: Gu, Jian, et al.
Veröffentlicht: (2023)
von: Gu, Jian, et al.
Veröffentlicht: (2023)
UntrustVul: An Automated Approach for Identifying Untrustworthy Alerts in Vulnerability Detection Models
von: Tung, Lam Nguyen, et al.
Veröffentlicht: (2025)
von: Tung, Lam Nguyen, et al.
Veröffentlicht: (2025)
Trustworthy AI Software Engineers
von: Aleti, Aldeida, et al.
Veröffentlicht: (2026)
von: Aleti, Aldeida, et al.
Veröffentlicht: (2026)
Reinforcement Learning for Mutation Operator Selection in Automated Program Repair
von: Hanna, Carol, et al.
Veröffentlicht: (2023)
von: Hanna, Carol, et al.
Veröffentlicht: (2023)
BayesInsights: Modelling Software Delivery and Developer Experience with Bayesian Networks at Bloomberg
von: Kirbas, Serkan, et al.
Veröffentlicht: (2026)
von: Kirbas, Serkan, et al.
Veröffentlicht: (2026)
The Role of Road Features and Vehicle Dynamics in Cost-Effective Autonomous Vehicles Safety Testing: Insights from Instance Space Analysis
von: Crespo-Rodriguez, Victor, et al.
Veröffentlicht: (2026)
von: Crespo-Rodriguez, Victor, et al.
Veröffentlicht: (2026)
Requirements-Driven Automated Software Testing: A Systematic Review
von: Wang, Fanyu, et al.
Veröffentlicht: (2025)
von: Wang, Fanyu, et al.
Veröffentlicht: (2025)
From Domain Documents to Requirements: Retrieval-Augmented Generation in the Space Industry
von: Arora, Chetan, et al.
Veröffentlicht: (2025)
von: Arora, Chetan, et al.
Veröffentlicht: (2025)
Enhancing Large Language Models for Text-to-Testcase Generation
von: Alagarsamy, Saranya, et al.
Veröffentlicht: (2024)
von: Alagarsamy, Saranya, et al.
Veröffentlicht: (2024)
Unveiling Overlooked Performance Variance in Serverless Computing
von: Wen, Jinfeng, et al.
Veröffentlicht: (2023)
von: Wen, Jinfeng, et al.
Veröffentlicht: (2023)
Comparing AI Coding Agents: A Task-Stratified Analysis of Pull Request Acceptance
von: Pinna, Giovanni, et al.
Veröffentlicht: (2026)
von: Pinna, Giovanni, et al.
Veröffentlicht: (2026)
A Semantic-based Optimization Approach for Repairing LLMs: Case Study on Code Generation
von: Gu, Jian, et al.
Veröffentlicht: (2025)
von: Gu, Jian, et al.
Veröffentlicht: (2025)
ViBR: Automated Bug Replay from Video-based Reports using Vision-Language Models
von: Feng, Sidong, et al.
Veröffentlicht: (2026)
von: Feng, Sidong, et al.
Veröffentlicht: (2026)
JMigBench: A Benchmark for Evaluating LLMs on Source Code Migration (Java 8 to Java 11)
von: Amin, Nishil, et al.
Veröffentlicht: (2026)
von: Amin, Nishil, et al.
Veröffentlicht: (2026)
From Research to Practice: An Interactive Rapid Review of Autonomous Driving System Testing in Industry
von: Song, Qunying, et al.
Veröffentlicht: (2026)
von: Song, Qunying, et al.
Veröffentlicht: (2026)
Multi-Modal Requirements Data-based Acceptance Criteria Generation using LLMs
von: Wang, Fanyu, et al.
Veröffentlicht: (2025)
von: Wang, Fanyu, et al.
Veröffentlicht: (2025)
Understanding Fairness in Software Engineering: Insights from Stack Exchange
von: Sesari, Emeralda, et al.
Veröffentlicht: (2024)
von: Sesari, Emeralda, et al.
Veröffentlicht: (2024)
Psychological Safety Framework in Pull-based Open Source Projects
von: Sesari, Emeralda, et al.
Veröffentlicht: (2025)
von: Sesari, Emeralda, et al.
Veröffentlicht: (2025)
It is Giving Major Satisfaction: Why Fairness Matters for Software Practitioners
von: Sesari, Emeralda, et al.
Veröffentlicht: (2024)
von: Sesari, Emeralda, et al.
Veröffentlicht: (2024)
MORTAR: Multi-turn Metamorphic Testing for LLM-based Dialogue Systems
von: Guo, Guoxiang, et al.
Veröffentlicht: (2024)
von: Guo, Guoxiang, et al.
Veröffentlicht: (2024)
Enabling Cost-Effective UI Automation Testing with Retrieval-Based LLMs: A Case Study in WeChat
von: Feng, Sidong, et al.
Veröffentlicht: (2024)
von: Feng, Sidong, et al.
Veröffentlicht: (2024)
SafeTune: Search-based Harmfulness Minimisation for Large Language Models
von: d'Aloisio, Giordano, et al.
Veröffentlicht: (2026)
von: d'Aloisio, Giordano, et al.
Veröffentlicht: (2026)
FairRF: Multi-Objective Search for Single and Intersectional Software Fairness
von: d'Alosio, Giordano, et al.
Veröffentlicht: (2026)
von: d'Alosio, Giordano, et al.
Veröffentlicht: (2026)
LLM-Based Misconfiguration Detection for AWS Serverless Computing
von: Wen, Jinfeng, et al.
Veröffentlicht: (2024)
von: Wen, Jinfeng, et al.
Veröffentlicht: (2024)
On the Influence of Data Resampling for Deep Learning-Based Log Anomaly Detection: Insights and Recommendations
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2024)
Automated Repair of Ambiguous Problem Descriptions for LLM-Based Code Generation
von: Jia, Haoxiang, et al.
Veröffentlicht: (2025)
von: Jia, Haoxiang, et al.
Veröffentlicht: (2025)
Practitioners' Expectations on Log Anomaly Detection
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2024)
Automated Trustworthiness Oracle Generation for Machine Learning Text Classifiers
von: Tung, Lam Nguyen, et al.
Veröffentlicht: (2024)
von: Tung, Lam Nguyen, et al.
Veröffentlicht: (2024)
HerAgent: Rethinking the Automated Environment Deployment via Hierarchical Test Pyramid
von: Li, Xiang, et al.
Veröffentlicht: (2026)
von: Li, Xiang, et al.
Veröffentlicht: (2026)
Fairness Testing: A Comprehensive Survey and Analysis of Trends
von: Chen, Zhenpeng, et al.
Veröffentlicht: (2022)
von: Chen, Zhenpeng, et al.
Veröffentlicht: (2022)
Generative AI for Testing of Autonomous Driving Systems: A Survey
von: Song, Qunying, et al.
Veröffentlicht: (2025)
von: Song, Qunying, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Test-based Patch Clustering for Automatically-Generated Patches Assessment
von: Martinez, Matias, et al.
Veröffentlicht: (2022) -
Empirical and Sustainability Aspects of Software Engineering Research in the Era of Large Language Models: A Reflection
von: Williams, David, et al.
Veröffentlicht: (2025) -
Hot Fixing Software: A Comprehensive Review of Terminology, Techniques, and Applications
von: Hanna, Carol, et al.
Veröffentlicht: (2024) -
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
von: Hanna, Carol, et al.
Veröffentlicht: (2025) -
A Comprehensive Survey of Benchmarks for Automated Improvement of Software's Non-Functional Properties
von: Blot, Aymeric, et al.
Veröffentlicht: (2022)