Generating executable oracles to check conformance of client code to requirements of JDK Javadocs using LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Shan, Zhu, Chenguang, Khurshid, Sarfraz |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OBsmith: LLM-Powered JavaScript Obfuscator Testing
by: Jiang, Shan, et al.
Published: (2025)
by: Jiang, Shan, et al.
Published: (2025)
APRIL: API Synthesis with Automatic Prompt Optimization and Reinforcement Learning
by: Zhong, Hua, et al.
Published: (2025)
by: Zhong, Hua, et al.
Published: (2025)
An approach for API synthesis using large language models
by: Zhong, Hua, et al.
Published: (2025)
by: Zhong, Hua, et al.
Published: (2025)
Semantic Voting: Execution-Grounded Consensus for LLM Code Generation
by: Jiang, Shan, et al.
Published: (2026)
by: Jiang, Shan, et al.
Published: (2026)
On the Effectiveness of Large Language Models in Writing Alloy Formulas
by: Hong, Yang, et al.
Published: (2025)
by: Hong, Yang, et al.
Published: (2025)
LLMs as verification oracles for Solidity
by: Bartoletti, Massimo, et al.
Published: (2025)
by: Bartoletti, Massimo, et al.
Published: (2025)
Property-Driven Evaluation of GNN Expressiveness at Scale: Datasets, Framework, and Study
by: Che, Sicong, et al.
Published: (2026)
by: Che, Sicong, et al.
Published: (2026)
Sketch-and-Verify: Structured Inference-Time Scaling via Program Sketching
by: Jiang, Shan, et al.
Published: (2026)
by: Jiang, Shan, et al.
Published: (2026)
Generating Energy-efficient code with LLMs
by: Cappendijk, Tom, et al.
Published: (2024)
by: Cappendijk, Tom, et al.
Published: (2024)
REMODEL-LLM: Transforming C code to Java using LLMs
by: Gupta, Aryan, et al.
Published: (2025)
by: Gupta, Aryan, et al.
Published: (2025)
Detecting Multi-Parameter Constraint Inconsistencies in Python Data Science Libraries
by: Xu, Xiufeng, et al.
Published: (2024)
by: Xu, Xiufeng, et al.
Published: (2024)
RAG-Verus: Repository-Level Program Verification with LLMs using Retrieval Augmented Generation
by: Zhong, Sicheng, et al.
Published: (2025)
by: Zhong, Sicheng, et al.
Published: (2025)
Large language models for automated PRISMA 2020 adherence checking
by: Kataoka, Yuki, et al.
Published: (2025)
by: Kataoka, Yuki, et al.
Published: (2025)
Out of style: Misadventures with LLMs and code style transfer
by: Munson, Karl, et al.
Published: (2024)
by: Munson, Karl, et al.
Published: (2024)
When Fuzzing Meets LLMs: Challenges and Opportunities
by: Jiang, Yu, et al.
Published: (2024)
by: Jiang, Yu, et al.
Published: (2024)
Context-Guided Decompilation: A Step Towards Re-executability
by: Wang, Xiaohan, et al.
Published: (2025)
by: Wang, Xiaohan, et al.
Published: (2025)
Insights into resource utilization of code small language models serving with runtime engines and execution providers
by: Durán, Francisco, et al.
Published: (2024)
by: Durán, Francisco, et al.
Published: (2024)
GenAI-based test case generation and execution in SDV platform
by: Zyberaj, Denesa, et al.
Published: (2025)
by: Zyberaj, Denesa, et al.
Published: (2025)
Lifecycle-Aware code generation: Leveraging Software Engineering Phases in LLMs
by: Xing, Xing, et al.
Published: (2025)
by: Xing, Xing, et al.
Published: (2025)
PACIFIC: a framework for generating benchmarks to check Precise Automatically Checked Instruction Following In Code
by: Dreyfuss, Itay, et al.
Published: (2025)
by: Dreyfuss, Itay, et al.
Published: (2025)
Prompt engineering and framework: implementation to increase code reliability based guideline for LLMs
by: Cruz, Rogelio, et al.
Published: (2025)
by: Cruz, Rogelio, et al.
Published: (2025)
LLM assisted web application functional requirements generation: A case study of four popular LLMs over a Mess Management System
by: Gupta, Rashmi, et al.
Published: (2025)
by: Gupta, Rashmi, et al.
Published: (2025)
Software Vulnerability and Functionality Assessment using LLMs
by: Jensen, Rasmus Ingemann Tuffveson, et al.
Published: (2024)
by: Jensen, Rasmus Ingemann Tuffveson, et al.
Published: (2024)
Evaluating perturbation robustness of generative systems that use COBOL code inputs
by: Ackerman, Samuel, et al.
Published: (2025)
by: Ackerman, Samuel, et al.
Published: (2025)
LLMs for Test Input Generation for Semantic Caches
by: Rasool, Zafaryab, et al.
Published: (2024)
by: Rasool, Zafaryab, et al.
Published: (2024)
Generating Structured Plan Representation of Procedures with LLMs
by: Garg, Deepeka, et al.
Published: (2025)
by: Garg, Deepeka, et al.
Published: (2025)
Evaluating the Energy-Efficiency of the Code Generated by LLMs
by: Islam, Md Arman, et al.
Published: (2025)
by: Islam, Md Arman, et al.
Published: (2025)
Enhancing High-Quality Code Generation in Large Language Models with Comparative Prefix-Tuning
by: Jiang, Yuan, et al.
Published: (2025)
by: Jiang, Yuan, et al.
Published: (2025)
Evaluating the Generalizability of LLMs in Automated Program Repair
by: Li, Fengjie, et al.
Published: (2025)
by: Li, Fengjie, et al.
Published: (2025)
Holistic Evaluation of State-of-the-Art LLMs for Code Generation
by: Zhang, Le, et al.
Published: (2025)
by: Zhang, Le, et al.
Published: (2025)
Towards a General Framework for HTN Modeling with LLMs
by: Puerta-Merino, Israel, et al.
Published: (2025)
by: Puerta-Merino, Israel, et al.
Published: (2025)
RuleFlow : Generating Reusable Program Optimizations with LLMs
by: Singh, Avaljot, et al.
Published: (2026)
by: Singh, Avaljot, et al.
Published: (2026)
Can LLMs Generate User Stories and Assess Their Quality?
by: Quattrocchi, Giovanni, et al.
Published: (2025)
by: Quattrocchi, Giovanni, et al.
Published: (2025)
Generating a Low-code Complete Workflow via Task Decomposition and RAG
by: Ayala, Orlando Marquez, et al.
Published: (2024)
by: Ayala, Orlando Marquez, et al.
Published: (2024)
GBQA: A Game Benchmark for Evaluating LLMs as Quality Assurance Engineers
by: Jiang, Shufan, et al.
Published: (2026)
by: Jiang, Shufan, et al.
Published: (2026)
Operational Robustness of LLMs on Code Generation
by: Paul, Debalina Ghosh, et al.
Published: (2026)
by: Paul, Debalina Ghosh, et al.
Published: (2026)
The Potential of LLMs in Automating Software Testing: From Generation to Reporting
by: Sherifi, Betim, et al.
Published: (2024)
by: Sherifi, Betim, et al.
Published: (2024)
Using LLMs in Generating Design Rationale for Software Architecture Decisions
by: Zhou, Xiyu, et al.
Published: (2025)
by: Zhou, Xiyu, et al.
Published: (2025)
ReCatcher: Towards LLMs Regression Testing for Code Generation
by: Abbassi, Altaf Allah, et al.
Published: (2025)
by: Abbassi, Altaf Allah, et al.
Published: (2025)
Benchmark Dataset Generation and Evaluation for Excel Formula Repair with LLMs
by: Singha, Ananya, et al.
Published: (2025)
by: Singha, Ananya, et al.
Published: (2025)
Similar Items
-
OBsmith: LLM-Powered JavaScript Obfuscator Testing
by: Jiang, Shan, et al.
Published: (2025) -
APRIL: API Synthesis with Automatic Prompt Optimization and Reinforcement Learning
by: Zhong, Hua, et al.
Published: (2025) -
An approach for API synthesis using large language models
by: Zhong, Hua, et al.
Published: (2025) -
Semantic Voting: Execution-Grounded Consensus for LLM Code Generation
by: Jiang, Shan, et al.
Published: (2026) -
On the Effectiveness of Large Language Models in Writing Alloy Formulas
by: Hong, Yang, et al.
Published: (2025)