Synthesizing Test Cases for Narrowing Specification Candidates
Fuente:
arXiv
Saved in:
| Main Authors: | Cunha, Alcino, Macedo, Nuno |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Validating Formal Specifications with LLM-generated Test Cases
by: Cunha, Alcino, et al.
Published: (2025)
by: Cunha, Alcino, et al.
Published: (2025)
SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark
by: Yu, Boxi, et al.
Published: (2026)
by: Yu, Boxi, et al.
Published: (2026)
Testing SSD Firmware with State Data-Aware Fuzzing: Accelerating Coverage in Nondeterministic I/O Environments
by: Yoon, Gangho, et al.
Published: (2025)
by: Yoon, Gangho, et al.
Published: (2025)
PICKLES: a Natural Language Framework for Requirement Specification and Model-Based Testing
by: Rodríguez, María Belén, et al.
Published: (2026)
by: Rodríguez, María Belén, et al.
Published: (2026)
Model checking of hyperproperties for high-level relational models
by: Macedo, Nuno, et al.
Published: (2025)
by: Macedo, Nuno, et al.
Published: (2025)
Automatically Detecting Numerical Instability in Machine Learning Applications via Soft Assertions
by: Sharmin, Shaila, et al.
Published: (2025)
by: Sharmin, Shaila, et al.
Published: (2025)
Combined Program Analysis Techniques: A Systematic Mapping Study
by: Braione, Pietro, et al.
Published: (2026)
by: Braione, Pietro, et al.
Published: (2026)
The Specification as Quality Gate: Three Hypotheses on AI-Assisted Code Review
by: Zietsman, Christo
Published: (2026)
by: Zietsman, Christo
Published: (2026)
Comparing Human and LLM Generated Code: The Jury is Still Out!
by: Licorish, Sherlock A., et al.
Published: (2025)
by: Licorish, Sherlock A., et al.
Published: (2025)
On the Soundness and Consistency of LLM Agents for Executing Test Cases Written in Natural Language
by: Salva, Sébastien, et al.
Published: (2025)
by: Salva, Sébastien, et al.
Published: (2025)
GBM Returns the Best Prediction Performance among Regression Approaches: A Case Study of Stack Overflow Code Quality
by: Licorish, Sherlock A., et al.
Published: (2025)
by: Licorish, Sherlock A., et al.
Published: (2025)
Evaluating Cryptographic API Misuse Detectors for Go
by: Andersson, Vivi, et al.
Published: (2026)
by: Andersson, Vivi, et al.
Published: (2026)
Tractable Verification of Model Transformations: A Cutoff-Theorem Approach for DSLTrans
by: Lucio, Levi
Published: (2026)
by: Lucio, Levi
Published: (2026)
Understanding and Reusing Test Suites Across Database Systems
by: Zhong, Suyang, et al.
Published: (2024)
by: Zhong, Suyang, et al.
Published: (2024)
Talk is Cheap, Logic is Hard: Benchmarking LLMs on Post-Condition Formalization
by: Prasetya, I. S. W. B., et al.
Published: (2026)
by: Prasetya, I. S. W. B., et al.
Published: (2026)
Trace Validation of Unmodified Concurrent Systems with OmniLink
by: Hackett, Finn, et al.
Published: (2026)
by: Hackett, Finn, et al.
Published: (2026)
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
by: Ravi, Ravin, et al.
Published: (2026)
by: Ravi, Ravin, et al.
Published: (2026)
Path-optimal symbolic execution of heap-manipulating programs
by: Braione, Pietro, et al.
Published: (2024)
by: Braione, Pietro, et al.
Published: (2024)
Recommended Practices for Spreadsheet Testing
by: Panko, Raymond R.
Published: (2007)
by: Panko, Raymond R.
Published: (2007)
Assessing Reliability of Statistical Maximum Coverage Estimators in Fuzzing
by: Liyanage, Danushka, et al.
Published: (2025)
by: Liyanage, Danushka, et al.
Published: (2025)
Validating Solidity Code Defects using Symbolic and Concrete Execution powered by Large Language Models
by: Susan, Ştefan-Claudiu, et al.
Published: (2025)
by: Susan, Ştefan-Claudiu, et al.
Published: (2025)
Good modelling software practices
by: Lemmen, Carsten, et al.
Published: (2024)
by: Lemmen, Carsten, et al.
Published: (2024)
Iterative Audit Convergence in LLM-Managed Multi-Agent Systems: A Case Study in Prompt Engineering Quality Assurance
by: Calboreanu, Elias
Published: (2026)
by: Calboreanu, Elias
Published: (2026)
Adaptive and AI-Augmented Security Testing: A Systematic Survey of Program Analysis, Feedback-Driven Testing, and Hybrid Learning-Based Approaches
by: Wienczkowski, Michael
Published: (2026)
by: Wienczkowski, Michael
Published: (2026)
Automated Vulnerability Detection Using Deep Learning Technique
by: Yang, Guan-Yan, et al.
Published: (2024)
by: Yang, Guan-Yan, et al.
Published: (2024)
Towards a Probabilistic Framework for Analyzing and Improving LLM-Enabled Software
by: Baldonado, Juan Manuel, et al.
Published: (2025)
by: Baldonado, Juan Manuel, et al.
Published: (2025)
Monitoring Agentic Systems Before They're Reliable
by: Boston, Marisa Ferrara, et al.
Published: (2026)
by: Boston, Marisa Ferrara, et al.
Published: (2026)
AIRA: AI-Induced Risk Audit: A Structured Inspection Framework for AI-Generated Code
by: Parris, William M.
Published: (2026)
by: Parris, William M.
Published: (2026)
AgentEval: DAG-Structured Step-Level Evaluation for Agentic Workflows with Error Propagation Tracking
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
Constrained LTL Specification Learning from Examples
by: Zhang, Changjian, et al.
Published: (2024)
by: Zhang, Changjian, et al.
Published: (2024)
Leveraging LLMs for Formal Software Requirements -- Challenges and Prospects
by: Beg, Arshad, et al.
Published: (2025)
by: Beg, Arshad, et al.
Published: (2025)
A Short Survey on Formalising Software Requirements using Large Language Models
by: Beg, Arshad, et al.
Published: (2025)
by: Beg, Arshad, et al.
Published: (2025)
Short Version of VERIFAI2026 Paper -- Learning Infused Formal Reasoning: Contract Synthesis, Artefact Reuse and Semantic Foundations
by: Beg, Arshad, et al.
Published: (2026)
by: Beg, Arshad, et al.
Published: (2026)
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
by: Rehan, Tzafrir
Published: (2026)
by: Rehan, Tzafrir
Published: (2026)
Constitutional Spec-Driven Development: Enforcing Security by Construction in AI-Assisted Code Generation
by: Marri, Srinivas Rao
Published: (2026)
by: Marri, Srinivas Rao
Published: (2026)
SeMA: Extending and Analyzing Storyboards to Develop Secure Android Apps
by: Mitra, Joydeep, et al.
Published: (2020)
by: Mitra, Joydeep, et al.
Published: (2020)
PyPackIT: Automated Research Software Engineering for Scientific Python Applications on GitHub
by: Ariamajd, Armin, et al.
Published: (2025)
by: Ariamajd, Armin, et al.
Published: (2025)
Three Decades of Formal Methods in Business Process Compliance: A Systematic Literature Review
by: López, Hugo A., et al.
Published: (2024)
by: López, Hugo A., et al.
Published: (2024)
Towards Observation Lakehouses: Living, Interactive Archives of Software Behavior
by: Kessel, Marcus
Published: (2025)
by: Kessel, Marcus
Published: (2025)
Structural Quality Gaps in Practitioner AI Governance Prompts: An Empirical Study Using a Five-Principle Evaluation Framework
by: Zietsman, Christo
Published: (2026)
by: Zietsman, Christo
Published: (2026)
Similar Items
-
Validating Formal Specifications with LLM-generated Test Cases
by: Cunha, Alcino, et al.
Published: (2025) -
SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark
by: Yu, Boxi, et al.
Published: (2026) -
Testing SSD Firmware with State Data-Aware Fuzzing: Accelerating Coverage in Nondeterministic I/O Environments
by: Yoon, Gangho, et al.
Published: (2025) -
PICKLES: a Natural Language Framework for Requirement Specification and Model-Based Testing
by: Rodríguez, María Belén, et al.
Published: (2026) -
Model checking of hyperproperties for high-level relational models
by: Macedo, Nuno, et al.
Published: (2025)