A semantic mutation metric for metamorphic relation adequacy in scientific computing programs
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Meng, Yang, Xiaohua, Liu, Jie, Yan, Shiyu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Comparing Human and LLM Generated Code: The Jury is Still Out!
by: Licorish, Sherlock A., et al.
Published: (2025)
by: Licorish, Sherlock A., et al.
Published: (2025)
Framework Matters: Energy Efficiency of UI Automation Testing Frameworks
by: Lagermann, Timmie M. R., et al.
Published: (2025)
by: Lagermann, Timmie M. R., et al.
Published: (2025)
GBM Returns the Best Prediction Performance among Regression Approaches: A Case Study of Stack Overflow Code Quality
by: Licorish, Sherlock A., et al.
Published: (2025)
by: Licorish, Sherlock A., et al.
Published: (2025)
Assessing Reliability of Statistical Maximum Coverage Estimators in Fuzzing
by: Liyanage, Danushka, et al.
Published: (2025)
by: Liyanage, Danushka, et al.
Published: (2025)
Evaluating the Overhead of the Performance Profiler Cloudprofiler With MooBench
by: Yang, Shinhyung, et al.
Published: (2024)
by: Yang, Shinhyung, et al.
Published: (2024)
The ACPATH Metric: Precise Estimation of the Number of Acyclic Paths in C-like Languages
by: Bagnara, Roberto, et al.
Published: (2016)
by: Bagnara, Roberto, et al.
Published: (2016)
Software Testing at the Network Layer: Automated HTTP API Quality Assessment and Security Analysis of Production Web Applications
by: Mughal, Ali Hassaan, et al.
Published: (2026)
by: Mughal, Ali Hassaan, et al.
Published: (2026)
NOETHER: A Constructive Framework for Metamorphic Pattern Discovery from Operator Algebras
by: Li, Meng, et al.
Published: (2026)
by: Li, Meng, et al.
Published: (2026)
InterEvo-TR: Interactive Evolutionary Test Generation With Readability Assessment
by: Delgado-Pérez, Pedro, et al.
Published: (2024)
by: Delgado-Pérez, Pedro, et al.
Published: (2024)
Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture
by: Iscan, Mehmet
Published: (2026)
by: Iscan, Mehmet
Published: (2026)
Making Software Metrics Useful
by: Tempero, Ewan, et al.
Published: (2026)
by: Tempero, Ewan, et al.
Published: (2026)
Stabilization Without Simplification: A Two-Dimensional Model of Software Evolution
by: Furukawa, Masaru
Published: (2026)
by: Furukawa, Masaru
Published: (2026)
From Monolith to Microservices: A Comparative Evaluation of Decomposition Frameworks
by: Weerasinghe, Mineth, et al.
Published: (2026)
by: Weerasinghe, Mineth, et al.
Published: (2026)
A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification
by: Odmark, Joshua, et al.
Published: (2026)
by: Odmark, Joshua, et al.
Published: (2026)
GitHub Copilot and Developer Productivity: An Observational Dose-Response Analysis
by: Heilman, Alex, et al.
Published: (2026)
by: Heilman, Alex, et al.
Published: (2026)
OODEval: Evaluating Large Language Models on Object-Oriented Design
by: Xiao, Bingxu, et al.
Published: (2026)
by: Xiao, Bingxu, et al.
Published: (2026)
Context Engineering for Multi-Agent LLM Code Assistants Using Elicit, NotebookLM, ChatGPT, and Claude Code
by: Haseeb, Muhammad
Published: (2025)
by: Haseeb, Muhammad
Published: (2025)
Toward Efficient Testing of Graph Neural Networks via Test Input Prioritization
by: Yang, Lichen, et al.
Published: (2025)
by: Yang, Lichen, et al.
Published: (2025)
GALA: Multimodal Graph Alignment for Bug Localization in Automated Program Repair
by: Liu, Zhuoyao, et al.
Published: (2026)
by: Liu, Zhuoyao, et al.
Published: (2026)
A Comprehensive Study on Large Language Models for Mutation Testing
by: Wang, Bo, et al.
Published: (2024)
by: Wang, Bo, et al.
Published: (2024)
VLM-Fuzz: Vision Language Model Assisted Recursive Depth-first Search Exploration for Effective UI Testing of Android Apps
by: Demissie, Biniam Fisseha, et al.
Published: (2025)
by: Demissie, Biniam Fisseha, et al.
Published: (2025)
Flow-of-Action: SOP Enhanced LLM-Based Multi-Agent System for Root Cause Analysis
by: Pei, Changhua, et al.
Published: (2025)
by: Pei, Changhua, et al.
Published: (2025)
PinChecker: Identifying Unsound Safe Abstractions of Rust Pinning APIs
by: Dai, Yuxuan, et al.
Published: (2025)
by: Dai, Yuxuan, et al.
Published: (2025)
It's Alive! What a Live Object Environment Changes in Software Engineering Practice
by: Grigera, Julián, et al.
Published: (2026)
by: Grigera, Julián, et al.
Published: (2026)
Comprehensive Evaluation of Large Language Models on Software Engineering Tasks: A Multi-Task Benchmark
by: Gunawan, Go Frendi, et al.
Published: (2026)
by: Gunawan, Go Frendi, et al.
Published: (2026)
How Quickly Do Development Teams Update Their Vulnerable Dependencies?
by: Rahman, Imranur, et al.
Published: (2024)
by: Rahman, Imranur, et al.
Published: (2024)
Synthesizing Performance Constraints for Evaluating and Improving Code Efficiency
by: Yang, Jun, et al.
Published: (2025)
by: Yang, Jun, et al.
Published: (2025)
Analyzing the Adoption of Database Management Systems Throughout the History of Open Source Projects
by: Paiva, Camila A., et al.
Published: (2026)
by: Paiva, Camila A., et al.
Published: (2026)
SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark
by: Yu, Boxi, et al.
Published: (2026)
by: Yu, Boxi, et al.
Published: (2026)
Reliability of AI Bots Footprints in GitHub Actions CI/CD Workflows
by: Shah, Syed Muhammad Ashhar, et al.
Published: (2026)
by: Shah, Syed Muhammad Ashhar, et al.
Published: (2026)
Resilient Microservices: A Systematic Review of Recovery Patterns, Strategies, and Evaluation Frameworks
by: Mohammad, Muzeeb
Published: (2025)
by: Mohammad, Muzeeb
Published: (2025)
ErrorPrism: Reconstructing Error Propagation Paths in Cloud Service Systems
by: Pu, Junsong, et al.
Published: (2025)
by: Pu, Junsong, et al.
Published: (2025)
Contribution Rate Imputation Theory: A Conceptual Model
by: Bishop III, Vincil, et al.
Published: (2024)
by: Bishop III, Vincil, et al.
Published: (2024)
Evaluating Software Contribution Quality: Time-to-Modification Theory
by: Bishop III, Vincil, et al.
Published: (2024)
by: Bishop III, Vincil, et al.
Published: (2024)
Benchmarking Generative AI Models for Deep Learning Test Input Generation
by: Maryam, et al.
Published: (2024)
by: Maryam, et al.
Published: (2024)
Testing SSD Firmware with State Data-Aware Fuzzing: Accelerating Coverage in Nondeterministic I/O Environments
by: Yoon, Gangho, et al.
Published: (2025)
by: Yoon, Gangho, et al.
Published: (2025)
Tests4Py: A Benchmark for System Testing
by: Smytzek, Marius, et al.
Published: (2023)
by: Smytzek, Marius, et al.
Published: (2023)
Combined Program Analysis Techniques: A Systematic Mapping Study
by: Braione, Pietro, et al.
Published: (2026)
by: Braione, Pietro, et al.
Published: (2026)
Automatically Detecting Numerical Instability in Machine Learning Applications via Soft Assertions
by: Sharmin, Shaila, et al.
Published: (2025)
by: Sharmin, Shaila, et al.
Published: (2025)
PICKLES: a Natural Language Framework for Requirement Specification and Model-Based Testing
by: Rodríguez, María Belén, et al.
Published: (2026)
by: Rodríguez, María Belén, et al.
Published: (2026)
Similar Items
-
Comparing Human and LLM Generated Code: The Jury is Still Out!
by: Licorish, Sherlock A., et al.
Published: (2025) -
Framework Matters: Energy Efficiency of UI Automation Testing Frameworks
by: Lagermann, Timmie M. R., et al.
Published: (2025) -
GBM Returns the Best Prediction Performance among Regression Approaches: A Case Study of Stack Overflow Code Quality
by: Licorish, Sherlock A., et al.
Published: (2025) -
Assessing Reliability of Statistical Maximum Coverage Estimators in Fuzzing
by: Liyanage, Danushka, et al.
Published: (2025) -
Evaluating the Overhead of the Performance Profiler Cloudprofiler With MooBench
by: Yang, Shinhyung, et al.
Published: (2024)