Can Search-Based Testing with Pareto Optimization Effectively Cover Failure-Revealing Test Inputs?
Fuente:
arXiv
Saved in:
| Main Authors: | Sorokin, Lev, Safin, Damir, Nejati, Shiva |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bridging the Gap between Real-world and Synthetic Images for Testing Autonomous Driving Systems
by: Amini, Mohammad Hossein, et al.
Published: (2024)
by: Amini, Mohammad Hossein, et al.
Published: (2024)
Guiding the Search Towards Failure-Inducing Test Inputs Using Support Vector Machines
by: Sorokin, Lev, et al.
Published: (2024)
by: Sorokin, Lev, et al.
Published: (2024)
Simulator Ensembles for Trustworthy Autonomous Driving Testing
by: Sorokin, Lev, et al.
Published: (2025)
by: Sorokin, Lev, et al.
Published: (2025)
Enhancing LLM-Based Test Generation by Eliminating Covered Code
by: Xu, WeiZhe, et al.
Published: (2026)
by: Xu, WeiZhe, et al.
Published: (2026)
Generating Minimalist Adversarial Perturbations to Test Object-Detection Models: An Adaptive Multi-Metric Evolutionary Search Approach
by: McIntyre-Garcia, Cristopher, et al.
Published: (2024)
by: McIntyre-Garcia, Cristopher, et al.
Published: (2024)
Automated Factual Benchmarking for In-Car Conversational Systems using Large Language Models
by: Giebisch, Rafael, et al.
Published: (2025)
by: Giebisch, Rafael, et al.
Published: (2025)
The Impact of Software Testing with Quantum Optimization Meets Machine Learning
by: Bandarupalli, Gopichand
Published: (2025)
by: Bandarupalli, Gopichand
Published: (2025)
MASTEST: A LLM-Based Multi-Agent System For RESTful API Tests
by: Han, Xiaoke, et al.
Published: (2025)
by: Han, Xiaoke, et al.
Published: (2025)
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair
by: Fatima, Sakina, et al.
Published: (2023)
by: Fatima, Sakina, et al.
Published: (2023)
Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation
by: Jacopin, Éric
Published: (2026)
by: Jacopin, Éric
Published: (2026)
Generative AI to Generate Test Data Generators
by: Baudry, Benoit, et al.
Published: (2024)
by: Baudry, Benoit, et al.
Published: (2024)
Understanding LLM-Driven Test Oracle Generation
by: Bodicoat, Adam, et al.
Published: (2026)
by: Bodicoat, Adam, et al.
Published: (2026)
Code Generation by Differential Test Time Scaling
by: He, Yifeng, et al.
Published: (2026)
by: He, Yifeng, et al.
Published: (2026)
Effective Large Language Model Debugging with Best-first Tree Search
by: Song, Jialin, et al.
Published: (2024)
by: Song, Jialin, et al.
Published: (2024)
DeepKnowledge: Generalisation-Driven Deep Learning Testing
by: Missaoui, Sondess, et al.
Published: (2024)
by: Missaoui, Sondess, et al.
Published: (2024)
Mutation-Guided LLM-based Test Generation at Meta
by: Foster, Christopher, et al.
Published: (2025)
by: Foster, Christopher, et al.
Published: (2025)
A Theoretical Analysis of Test-Driven Code Generation
by: Menet, Nicolas, et al.
Published: (2026)
by: Menet, Nicolas, et al.
Published: (2026)
AI-driven Java Performance Testing: Balancing Result Quality with Testing Time
by: Traini, Luca, et al.
Published: (2024)
by: Traini, Luca, et al.
Published: (2024)
Methodological Framework for Quantifying Semantic Test Coverage in RAG Systems
by: Broestl, Noah, et al.
Published: (2025)
by: Broestl, Noah, et al.
Published: (2025)
RBT4DNN: Requirements-based Testing of Neural Networks
by: Mozumder, Nusrat Jahan, et al.
Published: (2025)
by: Mozumder, Nusrat Jahan, et al.
Published: (2025)
Exploring the Integration of Large Language Models in Industrial Test Maintenance Processes
by: Liu, Jingxiong, et al.
Published: (2024)
by: Liu, Jingxiong, et al.
Published: (2024)
Using Quality Attribute Scenarios for ML Model Test Case Generation
by: Brower-Sinning, Rachel, et al.
Published: (2024)
by: Brower-Sinning, Rachel, et al.
Published: (2024)
LGMT: Logic-Grounded Metamorphic Testing for Evaluating the Reasoning Reliability of LLMs
by: Zhou, Zenghui, et al.
Published: (2026)
by: Zhou, Zenghui, et al.
Published: (2026)
VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation
by: Bai, Yifan, et al.
Published: (2026)
by: Bai, Yifan, et al.
Published: (2026)
SWE-Replay: Efficient Test-Time Scaling for Software Engineering Agents
by: Ding, Yifeng, et al.
Published: (2026)
by: Ding, Yifeng, et al.
Published: (2026)
Fairness Testing in Retrieval-Augmented Generation: How Small Perturbations Reveal Bias in Small Language Models
by: de Oliveira, Matheus Vinicius da Silva, et al.
Published: (2025)
by: de Oliveira, Matheus Vinicius da Silva, et al.
Published: (2025)
SWT-Bench: Testing and Validating Real-World Bug-Fixes with Code Agents
by: Mündler, Niels, et al.
Published: (2024)
by: Mündler, Niels, et al.
Published: (2024)
CLOVER: A Test Case Generation Benchmark with Coverage, Long-Context, and Verification
by: Xu, Jiacheng, et al.
Published: (2025)
by: Xu, Jiacheng, et al.
Published: (2025)
Zero-Shot Attribution for Large Language Models: A Distribution Testing Approach
by: Canonne, Clément L., et al.
Published: (2025)
by: Canonne, Clément L., et al.
Published: (2025)
MIST-RL: Mutation-based Incremental Suite Testing via Reinforcement Learning
by: Zhu, Sicheng, et al.
Published: (2026)
by: Zhu, Sicheng, et al.
Published: (2026)
High-Dimensional Fault Tolerance Testing of Highly Automated Vehicles Based on Low-Rank Models
by: Mei, Yuewen, et al.
Published: (2024)
by: Mei, Yuewen, et al.
Published: (2024)
Parameter-Efficient Fine-Tuning of Large Language Models for Unit Test Generation: An Empirical Study
by: Storhaug, André, et al.
Published: (2024)
by: Storhaug, André, et al.
Published: (2024)
MAFT: Efficient Model-Agnostic Fairness Testing for Deep Neural Networks via Zero-Order Gradient Search
by: Wang, Zhaohui, et al.
Published: (2024)
by: Wang, Zhaohui, et al.
Published: (2024)
STELLAR: A Search-Based Testing Framework for Large Language Model Applications
by: Sorokin, Lev, et al.
Published: (2026)
by: Sorokin, Lev, et al.
Published: (2026)
LLM-Based Automated Diagnosis Of Integration Test Failures At Google
by: Ziftci, Celal, et al.
Published: (2026)
by: Ziftci, Celal, et al.
Published: (2026)
From paper to benchmark: agentic, framework-based reproduction of under-specified methods in machine health intelligence
by: Theiler, Raffael, et al.
Published: (2026)
by: Theiler, Raffael, et al.
Published: (2026)
LLMs for Test Input Generation for Semantic Caches
by: Rasool, Zafaryab, et al.
Published: (2024)
by: Rasool, Zafaryab, et al.
Published: (2024)
Order Matters! An Empirical Study on Large Language Models' Input Order Bias in Software Fault Localization
by: Rafi, Md Nakhla, et al.
Published: (2024)
by: Rafi, Md Nakhla, et al.
Published: (2024)
Can Coding Agents Be General Agents?
by: Ivanov, Maksim, et al.
Published: (2026)
by: Ivanov, Maksim, et al.
Published: (2026)
Reinforcement Learning for Online Testing of Autonomous Driving Systems: a Replication and Extension Study
by: Giamattei, Luca, et al.
Published: (2024)
by: Giamattei, Luca, et al.
Published: (2024)
Similar Items
-
Bridging the Gap between Real-world and Synthetic Images for Testing Autonomous Driving Systems
by: Amini, Mohammad Hossein, et al.
Published: (2024) -
Guiding the Search Towards Failure-Inducing Test Inputs Using Support Vector Machines
by: Sorokin, Lev, et al.
Published: (2024) -
Simulator Ensembles for Trustworthy Autonomous Driving Testing
by: Sorokin, Lev, et al.
Published: (2025) -
Enhancing LLM-Based Test Generation by Eliminating Covered Code
by: Xu, WeiZhe, et al.
Published: (2026) -
Generating Minimalist Adversarial Perturbations to Test Object-Detection Models: An Adaptive Multi-Metric Evolutionary Search Approach
by: McIntyre-Garcia, Cristopher, et al.
Published: (2024)