Clotho: Measuring Task-Specific Pre-Generation Test Adequacy for LLM Inputs
Fuente:
arXiv
Saved in:
| Main Authors: | Yoon, Juyeon, Kim, Somin, Feldt, Robert, Yoo, Shin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Capturing Semantic Flow of ML-based Systems
by: Yoo, Shin, et al.
Published: (2025)
by: Yoo, Shin, et al.
Published: (2025)
DANDI: Diffusion as Normative Distribution for Deep Neural Network Input
by: Kim, Somin, et al.
Published: (2025)
by: Kim, Somin, et al.
Published: (2025)
Adaptive Testing for LLM-Based Applications: A Diversity-based Approach
by: Yoon, Juyeon, et al.
Published: (2025)
by: Yoon, Juyeon, et al.
Published: (2025)
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
by: Dobslaw, Felix, et al.
Published: (2025)
by: Dobslaw, Felix, et al.
Published: (2025)
Atropos: Improving Cost-Benefit Trade-off of LLM-based Agents under Self-Consistency with Early Termination and Model Hotswap
by: Kim, Naryeong, et al.
Published: (2026)
by: Kim, Naryeong, et al.
Published: (2026)
Latent Regularization in Generative Test Input Generation
by: Merabishvili, Giorgi, et al.
Published: (2026)
by: Merabishvili, Giorgi, et al.
Published: (2026)
LLM Performance for Code Generation on Noisy Tasks
by: Sendyka, Radzim, et al.
Published: (2025)
by: Sendyka, Radzim, et al.
Published: (2025)
Grounding Data Science Code Generation with Input-Output Specifications
by: Wen, Yeming, et al.
Published: (2024)
by: Wen, Yeming, et al.
Published: (2024)
Test Design and Review Argumentation in AI-Assisted Test Generation
by: Enoiu, Eduard Paul, et al.
Published: (2026)
by: Enoiu, Eduard Paul, et al.
Published: (2026)
Generating Realistic, Diverse, and Fault-Revealing Inputs with Latent Space Interpolation for Testing Deep Neural Networks
by: Duan, Bin, et al.
Published: (2025)
by: Duan, Bin, et al.
Published: (2025)
PrismaDV: Automated Task-Aware Data Unit Test Generation
by: Chen, Hao, et al.
Published: (2026)
by: Chen, Hao, et al.
Published: (2026)
COSMosFL: Ensemble of Small Language Models for Fault Localisation
by: Cho, Hyunjoon, et al.
Published: (2025)
by: Cho, Hyunjoon, et al.
Published: (2025)
Predictive Prompt Analysis
by: Lee, Jae Yong, et al.
Published: (2025)
by: Lee, Jae Yong, et al.
Published: (2025)
Understanding on the Edge: LLM-generated Boundary Test Explanations
by: Akbarova, Sabinakhon, et al.
Published: (2026)
by: Akbarova, Sabinakhon, et al.
Published: (2026)
Test Adequacy for Metamorphic Testing: Criteria, Measurement, and Implication
by: Fu, An, et al.
Published: (2024)
by: Fu, An, et al.
Published: (2024)
A Survey on LLM-based Code Generation for Low-Resource and Domain-Specific Programming Languages
by: Joel, Sathvik, et al.
Published: (2024)
by: Joel, Sathvik, et al.
Published: (2024)
PRIMG : Efficient LLM-driven Test Generation Using Mutant Prioritization
by: Bouafif, Mohamed Salah, et al.
Published: (2025)
by: Bouafif, Mohamed Salah, et al.
Published: (2025)
LLM-Powered Test Case Generation for Detecting Bugs in Plausible Programs
by: Liu, Kaibo, et al.
Published: (2024)
by: Liu, Kaibo, et al.
Published: (2024)
Learning-Based Testing for Deep Learning: Enhancing Model Robustness with Adversarial Input Prioritization
by: Rahman, Sheikh Md Mushfiqur, et al.
Published: (2025)
by: Rahman, Sheikh Md Mushfiqur, et al.
Published: (2025)
GIST: Generated Inputs Sets Transferability in Deep Learning
by: Tambon, Florian, et al.
Published: (2023)
by: Tambon, Florian, et al.
Published: (2023)
TopoMap: A Feature-based Semantic Discriminator of the Topographical Regions in the Test Input Space
by: De Vita, Gianmarco, et al.
Published: (2025)
by: De Vita, Gianmarco, et al.
Published: (2025)
Mutation-Guided LLM-based Test Generation at Meta
by: Foster, Christopher, et al.
Published: (2025)
by: Foster, Christopher, et al.
Published: (2025)
Benchmarking Generative AI Models for Deep Learning Test Input Generation
by: Maryam, et al.
Published: (2024)
by: Maryam, et al.
Published: (2024)
On the Usage of Continual Learning for Out-of-Distribution Generalization in Pre-trained Language Models of Code
by: Weyssow, Martin, et al.
Published: (2023)
by: Weyssow, Martin, et al.
Published: (2023)
Adaptive Reinforcement Learning for Dynamic Configuration Allocation in Pre-Production Testing
by: Zhu, Yu
Published: (2025)
by: Zhu, Yu
Published: (2025)
Code-Aware Prompting: A study of Coverage Guided Test Generation in Regression Setting using LLM
by: Ryan, Gabriel, et al.
Published: (2024)
by: Ryan, Gabriel, et al.
Published: (2024)
Cross-Functional AI Task Forces (X-FAITs) for AI Transformation of Software Organizations
by: Gren, Lucas, et al.
Published: (2025)
by: Gren, Lucas, et al.
Published: (2025)
Combining Neuroevolution with the Search for Novelty to Improve the Generation of Test Inputs for Games
by: Feldmeier, Patric, et al.
Published: (2024)
by: Feldmeier, Patric, et al.
Published: (2024)
ExplainFuzz: Explainable and Constraint-Conditioned Test Generation with Probabilistic Circuits
by: Baiget, Annaëlle, et al.
Published: (2026)
by: Baiget, Annaëlle, et al.
Published: (2026)
Real Faults in Deep Learning Fault Benchmarks: How Real Are They?
by: Jahangirova, Gunel, et al.
Published: (2024)
by: Jahangirova, Gunel, et al.
Published: (2024)
An Empirical Study of Fault Localisation Techniques for Deep Learning
by: Humbatova, Nargiz, et al.
Published: (2024)
by: Humbatova, Nargiz, et al.
Published: (2024)
AutoSpec: Automated Generation of Neural Network Specifications
by: Jin, Shuowei, et al.
Published: (2024)
by: Jin, Shuowei, et al.
Published: (2024)
Towards Assessing Deep Learning Test Input Generators
by: Mzoughi, Seif, et al.
Published: (2025)
by: Mzoughi, Seif, et al.
Published: (2025)
Understanding LLM-Driven Test Oracle Generation
by: Bodicoat, Adam, et al.
Published: (2026)
by: Bodicoat, Adam, et al.
Published: (2026)
FeedbackLLM: Metadata driven Multi-Agentic Language Agnostic Test Case Generator with Evolving prompt and Coverage Feedback
by: Jasti, Kushal, et al.
Published: (2026)
by: Jasti, Kushal, et al.
Published: (2026)
Should Code Models Learn Pedagogically? A Preliminary Evaluation of Curriculum Learning for Real-World Software Engineering Tasks
by: Khant, Kyi Shin, et al.
Published: (2025)
by: Khant, Kyi Shin, et al.
Published: (2025)
Automating REST API Postman Test Cases Using LLM
by: Sri, S Deepika, et al.
Published: (2024)
by: Sri, S Deepika, et al.
Published: (2024)
ReqBrain: Task-Specific Instruction Tuning of LLMs for AI-Assisted Requirements Generation
by: Habib, Mohammad Kasra, et al.
Published: (2025)
by: Habib, Mohammad Kasra, et al.
Published: (2025)
Design-Specification Tiling for ICL-based CAD Code Generation
by: Du, Yali, et al.
Published: (2026)
by: Du, Yali, et al.
Published: (2026)
Can Search-Based Testing with Pareto Optimization Effectively Cover Failure-Revealing Test Inputs?
by: Sorokin, Lev, et al.
Published: (2024)
by: Sorokin, Lev, et al.
Published: (2024)
Similar Items
-
Capturing Semantic Flow of ML-based Systems
by: Yoo, Shin, et al.
Published: (2025) -
DANDI: Diffusion as Normative Distribution for Deep Neural Network Input
by: Kim, Somin, et al.
Published: (2025) -
Adaptive Testing for LLM-Based Applications: A Diversity-based Approach
by: Yoon, Juyeon, et al.
Published: (2025) -
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
by: Dobslaw, Felix, et al.
Published: (2025) -
Atropos: Improving Cost-Benefit Trade-off of LLM-based Agents under Self-Consistency with Early Termination and Model Hotswap
by: Kim, Naryeong, et al.
Published: (2026)