Mining Subscenario Refactoring Opportunities in Behaviour-Driven Software Test Suites: ML Classifiers and LLM-Judge Baselines
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mughal, Ali Hassaan, Fatima, Noor, Bilal, Muhammad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reducing Maintenance Burden in Behaviour-Driven Development: A Paraphrase-Robust Duplicate-Step Detector with a 1.1M-Step Open Benchmark
von: Mughal, Ali Hassaan, et al.
Veröffentlicht: (2026)
von: Mughal, Ali Hassaan, et al.
Veröffentlicht: (2026)
Software Testing at the Network Layer: Automated HTTP API Quality Assessment and Security Analysis of Production Web Applications
von: Mughal, Ali Hassaan, et al.
Veröffentlicht: (2026)
von: Mughal, Ali Hassaan, et al.
Veröffentlicht: (2026)
Factors that Contribute to the Success of a Software Organisation's DevOps Environment: A Systematic Review
von: Gwangwadza, Ashley, et al.
Veröffentlicht: (2022)
von: Gwangwadza, Ashley, et al.
Veröffentlicht: (2022)
ConfProBench: A Confidence Evaluation Benchmark for MLLM-Based Process Judges
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
LLMCup: Ranking-Enhanced Comment Updating with LLMs
von: Ge, Hua, et al.
Veröffentlicht: (2025)
von: Ge, Hua, et al.
Veröffentlicht: (2025)
Speculative Automated Refactoring of Imperative Deep Learning Programs to Graph Execution
von: Khatchadourian, Raffi, et al.
Veröffentlicht: (2025)
von: Khatchadourian, Raffi, et al.
Veröffentlicht: (2025)
PyPackIT: Automated Research Software Engineering for Scientific Python Applications on GitHub
von: Ariamajd, Armin, et al.
Veröffentlicht: (2025)
von: Ariamajd, Armin, et al.
Veröffentlicht: (2025)
Automated Bug Triaging using Instruction-Tuned Large Language Models
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
Merge-Bench: Resolve Merge Conflicts with Large Language Models
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
Generative AI and the Transformation of Software Development Practices
von: Acharya, Vivek
Veröffentlicht: (2025)
von: Acharya, Vivek
Veröffentlicht: (2025)
Diagnosing Refactoring Dangers
von: Brinksma, Wouter, et al.
Veröffentlicht: (2024)
von: Brinksma, Wouter, et al.
Veröffentlicht: (2024)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
SPViz: A DSL-Driven Approach for Software Project Visualization Tooling
von: Rentz, Niklas, et al.
Veröffentlicht: (2024)
von: Rentz, Niklas, et al.
Veröffentlicht: (2024)
Code Documentation and Analysis to Secure Software Development
von: Attie, Paul, et al.
Veröffentlicht: (2024)
von: Attie, Paul, et al.
Veröffentlicht: (2024)
DRS-OSS: Practical Diff Risk Scoring with LLMs
von: Sayedsalehi, Ali, et al.
Veröffentlicht: (2025)
von: Sayedsalehi, Ali, et al.
Veröffentlicht: (2025)
Software Defined Vehicle Code Generation: A Few-Shot Prompting Approach
von: Nguyen, Quang-Dung, et al.
Veröffentlicht: (2025)
von: Nguyen, Quang-Dung, et al.
Veröffentlicht: (2025)
Test-driven Software Experimentation with LASSO: an LLM Prompt Benchmarking Example
von: Kessel, Marcus
Veröffentlicht: (2024)
von: Kessel, Marcus
Veröffentlicht: (2024)
Energy-Aware Decision Making in Software Stack Upgrades
von: Stocker, Mirko, et al.
Veröffentlicht: (2026)
von: Stocker, Mirko, et al.
Veröffentlicht: (2026)
Agentic Refactoring: An Empirical Study of AI Coding Agents
von: Horikawa, Kosei, et al.
Veröffentlicht: (2025)
von: Horikawa, Kosei, et al.
Veröffentlicht: (2025)
Recommending Variable Names for Extract Local Variable Refactorings
von: Wang, Taiming, et al.
Veröffentlicht: (2025)
von: Wang, Taiming, et al.
Veröffentlicht: (2025)
Technical Debt Management: The Road Ahead for Successful Software Delivery
von: Avgeriou, Paris, et al.
Veröffentlicht: (2024)
von: Avgeriou, Paris, et al.
Veröffentlicht: (2024)
Causality-Driven Neural Network Repair: Challenges and Opportunities
von: Vares, Fatemeh, et al.
Veröffentlicht: (2025)
von: Vares, Fatemeh, et al.
Veröffentlicht: (2025)
Binary-30K: A Heterogeneous Dataset for Deep Learning in Binary Analysis and Malware Detection
von: Bommarito II, Michael J.
Veröffentlicht: (2025)
von: Bommarito II, Michael J.
Veröffentlicht: (2025)
LLM-FACETS: A Privacy-Preserving Framework for Evaluating LLM Transparency and Accountability
von: Lucas, Tom, et al.
Veröffentlicht: (2026)
von: Lucas, Tom, et al.
Veröffentlicht: (2026)
Optimizing Large Language Models for OpenAPI Code Completion
von: Petryshyn, Bohdan, et al.
Veröffentlicht: (2024)
von: Petryshyn, Bohdan, et al.
Veröffentlicht: (2024)
Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture
von: Iscan, Mehmet
Veröffentlicht: (2026)
von: Iscan, Mehmet
Veröffentlicht: (2026)
Towards a Probabilistic Framework for Analyzing and Improving LLM-Enabled Software
von: Baldonado, Juan Manuel, et al.
Veröffentlicht: (2025)
von: Baldonado, Juan Manuel, et al.
Veröffentlicht: (2025)
Beyond Greenfield: The D3 Framework for AI-Driven Productivity in Brownfield Engineering
von: Sharma, Krishna Kumaar
Veröffentlicht: (2025)
von: Sharma, Krishna Kumaar
Veröffentlicht: (2025)
When Code Smells Meet ML: On the Lifecycle of ML-specific Code Smells in ML-enabled Systems
von: Recupito, Gilberto, et al.
Veröffentlicht: (2024)
von: Recupito, Gilberto, et al.
Veröffentlicht: (2024)
EyeLayer: Integrating Human Attention Patterns into LLM-Based Code Summarization
von: Zhang, Jiahao, et al.
Veröffentlicht: (2026)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2026)
Developer-LLM Conversations: An Empirical Study of Interactions and Generated Code Quality
von: Zhong, Suzhen, et al.
Veröffentlicht: (2025)
von: Zhong, Suzhen, et al.
Veröffentlicht: (2025)
EvoGraph: Hybrid Directed Graph Evolution toward Software 3.0
von: Costa, Igor, et al.
Veröffentlicht: (2025)
von: Costa, Igor, et al.
Veröffentlicht: (2025)
Combining Serverless and High-Performance Computing Paradigms to support ML Data-Intensive Applications
von: Staylor, Mills, et al.
Veröffentlicht: (2025)
von: Staylor, Mills, et al.
Veröffentlicht: (2025)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
TerraFormer: Automated Infrastructure-as-Code with LLMs Fine-Tuned via Policy-Guided Verifier Feedback
von: Jana, Prithwish, et al.
Veröffentlicht: (2026)
von: Jana, Prithwish, et al.
Veröffentlicht: (2026)
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
von: Hu, Yuelin, et al.
Veröffentlicht: (2026)
von: Hu, Yuelin, et al.
Veröffentlicht: (2026)
AgentOps: Enabling Observability of LLM Agents
von: Dong, Liming, et al.
Veröffentlicht: (2024)
von: Dong, Liming, et al.
Veröffentlicht: (2024)
A History Equivalence Algorithm for Dynamic Process Migration
von: Bakshi, Gargi, et al.
Veröffentlicht: (2024)
von: Bakshi, Gargi, et al.
Veröffentlicht: (2024)
Early-Stage Requirements Transformation Approaches: A Systematic Review
von: Letsholo, Keletso J.
Veröffentlicht: (2024)
von: Letsholo, Keletso J.
Veröffentlicht: (2024)
SmellBench: Evaluating LLM Agents on Architectural Code Smell Repair
von: Dinu, Ion George, et al.
Veröffentlicht: (2026)
von: Dinu, Ion George, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Reducing Maintenance Burden in Behaviour-Driven Development: A Paraphrase-Robust Duplicate-Step Detector with a 1.1M-Step Open Benchmark
von: Mughal, Ali Hassaan, et al.
Veröffentlicht: (2026) -
Software Testing at the Network Layer: Automated HTTP API Quality Assessment and Security Analysis of Production Web Applications
von: Mughal, Ali Hassaan, et al.
Veröffentlicht: (2026) -
Factors that Contribute to the Success of a Software Organisation's DevOps Environment: A Systematic Review
von: Gwangwadza, Ashley, et al.
Veröffentlicht: (2022) -
ConfProBench: A Confidence Evaluation Benchmark for MLLM-Based Process Judges
von: Zhou, Yue, et al.
Veröffentlicht: (2025) -
LLMCup: Ranking-Enhanced Comment Updating with LLMs
von: Ge, Hua, et al.
Veröffentlicht: (2025)