LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ravi, Ravin, Bradshaw, Dylan, Ruberto, Stefano, Jahangirova, Gunel, Terragni, Valerio |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
LLM-Assisted Translation of Legacy FORTRAN Codes to C++: A Cross-Platform Study
par: Ranasinghe, Nishath Rajiv, et autres
Publié: (2025)
par: Ranasinghe, Nishath Rajiv, et autres
Publié: (2025)
LLMORPH: Automated Metamorphic Testing of Large Language Models
par: Cho, Steven, et autres
Publié: (2026)
par: Cho, Steven, et autres
Publié: (2026)
Characterizing JavaScript Security Code Smells
par: Kambhampati, Vikas, et autres
Publié: (2024)
par: Kambhampati, Vikas, et autres
Publié: (2024)
Validating Formal Specifications with LLM-generated Test Cases
par: Cunha, Alcino, et autres
Publié: (2025)
par: Cunha, Alcino, et autres
Publié: (2025)
Comparing Human and LLM Generated Code: The Jury is Still Out!
par: Licorish, Sherlock A., et autres
Publié: (2025)
par: Licorish, Sherlock A., et autres
Publié: (2025)
Code Documentation and Analysis to Secure Software Development
par: Attie, Paul, et autres
Publié: (2024)
par: Attie, Paul, et autres
Publié: (2024)
PyPackIT: Automated Research Software Engineering for Scientific Python Applications on GitHub
par: Ariamajd, Armin, et autres
Publié: (2025)
par: Ariamajd, Armin, et autres
Publié: (2025)
Synthesizing Test Cases for Narrowing Specification Candidates
par: Cunha, Alcino, et autres
Publié: (2025)
par: Cunha, Alcino, et autres
Publié: (2025)
SmellBench: Evaluating LLM Agents on Architectural Code Smell Repair
par: Dinu, Ion George, et autres
Publié: (2026)
par: Dinu, Ion George, et autres
Publié: (2026)
GBM Returns the Best Prediction Performance among Regression Approaches: A Case Study of Stack Overflow Code Quality
par: Licorish, Sherlock A., et autres
Publié: (2025)
par: Licorish, Sherlock A., et autres
Publié: (2025)
A History Equivalence Algorithm for Dynamic Process Migration
par: Bakshi, Gargi, et autres
Publié: (2024)
par: Bakshi, Gargi, et autres
Publié: (2024)
Combating Reentrancy Bugs on Sharded Blockchains
par: Kashitsyn, Roman, et autres
Publié: (2025)
par: Kashitsyn, Roman, et autres
Publié: (2025)
You Don't Need Public Tests to Generate Correct Code
par: Silva, Kaushitha, et autres
Publié: (2026)
par: Silva, Kaushitha, et autres
Publié: (2026)
RustAssure: Differential Symbolic Testing for LLM-Transpiled C-to-Rust Code
par: Bai, Yubo, et autres
Publié: (2025)
par: Bai, Yubo, et autres
Publié: (2025)
The Specification as Quality Gate: Three Hypotheses on AI-Assisted Code Review
par: Zietsman, Christo
Publié: (2026)
par: Zietsman, Christo
Publié: (2026)
From Scientific Texts to Verifiable Code: Automating the Process with Transformers
par: Wang, Changjie, et autres
Publié: (2025)
par: Wang, Changjie, et autres
Publié: (2025)
SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark
par: Yu, Boxi, et autres
Publié: (2026)
par: Yu, Boxi, et autres
Publié: (2026)
Testing SSD Firmware with State Data-Aware Fuzzing: Accelerating Coverage in Nondeterministic I/O Environments
par: Yoon, Gangho, et autres
Publié: (2025)
par: Yoon, Gangho, et autres
Publié: (2025)
Source Code Hotspots: A Diagnostic Method for Quality Issues
par: Muzammil, Saleha, et autres
Publié: (2026)
par: Muzammil, Saleha, et autres
Publié: (2026)
Iterative Audit Convergence in LLM-Managed Multi-Agent Systems: A Case Study in Prompt Engineering Quality Assurance
par: Calboreanu, Elias
Publié: (2026)
par: Calboreanu, Elias
Publié: (2026)
Rango: Adaptive Retrieval-Augmented Proving for Automated Software Verification
par: Thompson, Kyle, et autres
Publié: (2024)
par: Thompson, Kyle, et autres
Publié: (2024)
From Untestable to Testable: Metamorphic Testing in the Age of LLMs
par: Terragni, Valerio
Publié: (2026)
par: Terragni, Valerio
Publié: (2026)
Graph-Based Specification and Automated Construction of ILP Problems
par: Ehmes, Sebastian, et autres
Publié: (2022)
par: Ehmes, Sebastian, et autres
Publié: (2022)
Code Less to Code More: Streamlining Language Server Protocol and Type System Development for Language Families
par: Bruzzone, Federico, et autres
Publié: (2025)
par: Bruzzone, Federico, et autres
Publié: (2025)
Adaptive and AI-Augmented Security Testing: A Systematic Survey of Program Analysis, Feedback-Driven Testing, and Hybrid Learning-Based Approaches
par: Wienczkowski, Michael
Publié: (2026)
par: Wienczkowski, Michael
Publié: (2026)
Combined Program Analysis Techniques: A Systematic Mapping Study
par: Braione, Pietro, et autres
Publié: (2026)
par: Braione, Pietro, et autres
Publié: (2026)
A Study of Undefined Behavior Across Foreign Function Boundaries in Rust Libraries
par: McCormack, Ian, et autres
Publié: (2024)
par: McCormack, Ian, et autres
Publié: (2024)
Automatically Detecting Numerical Instability in Machine Learning Applications via Soft Assertions
par: Sharmin, Shaila, et autres
Publié: (2025)
par: Sharmin, Shaila, et autres
Publié: (2025)
BACE: LLM-based Code Generation through Bayesian Anchored Co-Evolution of Code and Test Populations
par: Silva, Kaushitha, et autres
Publié: (2026)
par: Silva, Kaushitha, et autres
Publié: (2026)
Towards a Probabilistic Framework for Analyzing and Improving LLM-Enabled Software
par: Baldonado, Juan Manuel, et autres
Publié: (2025)
par: Baldonado, Juan Manuel, et autres
Publié: (2025)
SIADAFIX: issue description response for adaptive program repair
par: Cao, Xin, et autres
Publié: (2025)
par: Cao, Xin, et autres
Publié: (2025)
Test-driven Software Experimentation with LASSO: an LLM Prompt Benchmarking Example
par: Kessel, Marcus
Publié: (2024)
par: Kessel, Marcus
Publié: (2024)
A RAG Method for Source Code Inquiry Tailored to Long-Context LLMs
par: Kamiya, Toshihiro
Publié: (2024)
par: Kamiya, Toshihiro
Publié: (2024)
Litrepl: Literate Paper Processor Promoting Transparency More Than Reproducibility
par: Mironov, Sergei
Publié: (2025)
par: Mironov, Sergei
Publié: (2025)
Validating Solidity Code Defects using Symbolic and Concrete Execution powered by Large Language Models
par: Susan, Ştefan-Claudiu, et autres
Publié: (2025)
par: Susan, Ştefan-Claudiu, et autres
Publié: (2025)
Trace Validation of Unmodified Concurrent Systems with OmniLink
par: Hackett, Finn, et autres
Publié: (2026)
par: Hackett, Finn, et autres
Publié: (2026)
Highly Interactive Testing for Uninterrupted Development Flow
par: Tropin, Andrew
Publié: (2025)
par: Tropin, Andrew
Publié: (2025)
Teaching Complex Systems based on Microservices
par: Ferreira, Renato Cordeiro, et autres
Publié: (2025)
par: Ferreira, Renato Cordeiro, et autres
Publié: (2025)
Automated structural testing of LLM-based agents: methods, framework, and case studies
par: Kohl, Jens, et autres
Publié: (2026)
par: Kohl, Jens, et autres
Publié: (2026)
Knowledge-Aware Code Generation with Large Language Models
par: Huang, Tao, et autres
Publié: (2024)
par: Huang, Tao, et autres
Publié: (2024)
Documents similaires
-
LLM-Assisted Translation of Legacy FORTRAN Codes to C++: A Cross-Platform Study
par: Ranasinghe, Nishath Rajiv, et autres
Publié: (2025) -
LLMORPH: Automated Metamorphic Testing of Large Language Models
par: Cho, Steven, et autres
Publié: (2026) -
Characterizing JavaScript Security Code Smells
par: Kambhampati, Vikas, et autres
Publié: (2024) -
Validating Formal Specifications with LLM-generated Test Cases
par: Cunha, Alcino, et autres
Publié: (2025) -
Comparing Human and LLM Generated Code: The Jury is Still Out!
par: Licorish, Sherlock A., et autres
Publié: (2025)