LLM-Based Test-Driven Interactive Code Generation: User Study and Empirical Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Fakhoury, Sarah, Naik, Aaditya, Sakkas, Georgios, Chakraborty, Saikat, Lahiri, Shuvendu K. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can Large Language Models Transform Natural Language Intent into Formal Method Postconditions?
by: Endres, Madeline, et al.
Published: (2023)
by: Endres, Madeline, et al.
Published: (2023)
3DGen: AI-Assisted Generation of Provably Correct Binary Format Parsers
by: Fakhoury, Sarah, et al.
Published: (2024)
by: Fakhoury, Sarah, et al.
Published: (2024)
Towards Neural Synthesis for SMT-Assisted Proof-Oriented Programming
by: Chakraborty, Saikat, et al.
Published: (2024)
by: Chakraborty, Saikat, et al.
Published: (2024)
Ranking LLM-Generated Loop Invariants for Program Verification
by: Chakraborty, Saikat, et al.
Published: (2023)
by: Chakraborty, Saikat, et al.
Published: (2023)
Program Structure Aware Precondition Generation
by: Dinella, Elizabeth, et al.
Published: (2023)
by: Dinella, Elizabeth, et al.
Published: (2023)
Evaluating LLM-driven User-Intent Formalization for Verification-Aware Languages
by: Lahiri, Shuvendu K.
Published: (2024)
by: Lahiri, Shuvendu K.
Published: (2024)
Intent Formalization: A Grand Challenge for Reliable Coding in the Age of AI Agents
by: Lahiri, Shuvendu K.
Published: (2026)
by: Lahiri, Shuvendu K.
Published: (2026)
Scaling Test-Driven Code Generation from Functions to Classes: An Empirical Study
by: Liang, Yunhao, et al.
Published: (2026)
by: Liang, Yunhao, et al.
Published: (2026)
Good Vibrations? A Qualitative Study of Co-Creation, Communication, Flow, and Trust in Vibe Coding
by: Pimenova, Veronica, et al.
Published: (2025)
by: Pimenova, Veronica, et al.
Published: (2025)
ClassInvGen: Class Invariant Synthesis using Large Language Models
by: Sun, Chuyue, et al.
Published: (2025)
by: Sun, Chuyue, et al.
Published: (2025)
Can We Classify Flaky Tests Using Only Test Code? An LLM-Based Empirical Study
by: Berndt, Alexander, et al.
Published: (2026)
by: Berndt, Alexander, et al.
Published: (2026)
DiffSpec: Differential Testing with LLMs using Natural Language Specifications and Code Artifacts
by: Rao, Nikitha, et al.
Published: (2024)
by: Rao, Nikitha, et al.
Published: (2024)
LLM-Vectorizer: LLM-based Verified Loop Vectorizer
by: Taneja, Jubi, et al.
Published: (2024)
by: Taneja, Jubi, et al.
Published: (2024)
An Empirical Study of Interaction Smells in Multi-Turn Human-LLM Collaborative Code Generation
by: Zhang, Binquan, et al.
Published: (2026)
by: Zhang, Binquan, et al.
Published: (2026)
An Empirical Study of LLM-Based Code Clone Detection
by: Zhu, Wenqing, et al.
Published: (2025)
by: Zhu, Wenqing, et al.
Published: (2025)
Are Coding Agents Generating Over-Mocked Tests? An Empirical Study
by: Hora, Andre, et al.
Published: (2026)
by: Hora, Andre, et al.
Published: (2026)
Guiding AI to Fix Its Own Flaws: An Empirical Study on LLM-Driven Secure Code Generation
by: Yan, Hao, et al.
Published: (2025)
by: Yan, Hao, et al.
Published: (2025)
An Empirical Study on Strong-Weak Model Collaboration for Repo-level Code Generation
by: Gandhi, Shubham, et al.
Published: (2025)
by: Gandhi, Shubham, et al.
Published: (2025)
GENCNIPPET: Automated Generation of Code Snippets for Supporting Programming Questions
by: Mondal, Saikat, et al.
Published: (2025)
by: Mondal, Saikat, et al.
Published: (2025)
Understanding Specification-Driven Code Generation with LLMs: An Empirical Study Design
by: Rosa, Giovanni, et al.
Published: (2026)
by: Rosa, Giovanni, et al.
Published: (2026)
Code vs Serialized AST Inputs for LLM-Based Code Summarization: An Empirical Study
by: Dong, Shijia, et al.
Published: (2026)
by: Dong, Shijia, et al.
Published: (2026)
Enhancing User Interaction in ChatGPT: Characterizing and Consolidating Multiple Prompts for Issue Resolution
by: Mondal, Saikat, et al.
Published: (2024)
by: Mondal, Saikat, et al.
Published: (2024)
LLM For Loop Invariant Generation and Fixing: How Far Are We?
by: Akhond, Mostafijur Rahman, et al.
Published: (2025)
by: Akhond, Mostafijur Rahman, et al.
Published: (2025)
Teaching an Old LLM Secure Coding: Localized Preference Optimization on Distilled Preferences
by: Hasan, Mohammad Saqib, et al.
Published: (2025)
by: Hasan, Mohammad Saqib, et al.
Published: (2025)
On the Effectiveness of Training Data Optimization for LLM-based Code Generation: An Empirical Study
by: Kuang, Shiqi, et al.
Published: (2025)
by: Kuang, Shiqi, et al.
Published: (2025)
Studying LLM Performance on Closed- and Open-source Data
by: Ahmed, Toufique, et al.
Published: (2024)
by: Ahmed, Toufique, et al.
Published: (2024)
An Empirical Security Evaluation of LLM-Generated Cryptographic Rust Code
by: Elsayed, Mohamed, et al.
Published: (2026)
by: Elsayed, Mohamed, et al.
Published: (2026)
On the Limitations of Embedding Based Methods for Measuring Functional Correctness for Code Generation
by: Naik, Atharva
Published: (2024)
by: Naik, Atharva
Published: (2024)
AUTOGENICS: Automated Generation of Context-Aware Inline Comments for Code Snippets on Programming Q&A Sites Using LLM
by: Bappon, Suborno Deb, et al.
Published: (2024)
by: Bappon, Suborno Deb, et al.
Published: (2024)
Rethinking Code Review Workflows with LLM Assistance: An Empirical Study
by: Aðalsteinsson, Fannar Steinn, et al.
Published: (2025)
by: Aðalsteinsson, Fannar Steinn, et al.
Published: (2025)
An Empirical Study of Self-Admitted Technical Debt in Machine Learning Software
by: Bhatia, Aaditya, et al.
Published: (2023)
by: Bhatia, Aaditya, et al.
Published: (2023)
Learn to Code Sustainably: An Empirical Study on LLM-based Green Code Generation
by: Vartziotis, Tina, et al.
Published: (2024)
by: Vartziotis, Tina, et al.
Published: (2024)
Quantum-Guided Test Case Minimization for LLM-Based Code Generation
by: Zhang, Huixiang, et al.
Published: (2025)
by: Zhang, Huixiang, et al.
Published: (2025)
Testing the Untestable? An Empirical Study on the Testing Process of LLM-Powered Software Systems
by: Magalhaes, Cleyton, et al.
Published: (2025)
by: Magalhaes, Cleyton, et al.
Published: (2025)
Algorithm-Based Pipeline for Reliable and Intent-Preserving Code Translation with LLMs
by: Dipto, Shahriar Rumi, et al.
Published: (2026)
by: Dipto, Shahriar Rumi, et al.
Published: (2026)
Compact Constraint Encoding for LLM Code Generation: An Empirical Study of Token Economics and Constraint Compliance
by: Tang, Hanzhang
Published: (2026)
by: Tang, Hanzhang
Published: (2026)
Detect Repair Verify for Securing LLM Generated Code: A Multi-Language Empirical Study
by: Cheng, Cheng
Published: (2026)
by: Cheng, Cheng
Published: (2026)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
by: Cui, Yi
Published: (2025)
by: Cui, Yi
Published: (2025)
Understanding Self-Admitted Technical Debt in Test Code: An Empirical Study
by: Nakamura, Ibuki, et al.
Published: (2025)
by: Nakamura, Ibuki, et al.
Published: (2025)
Evaluating LLM-Generated Code: A Benchmark and Developer Study
by: Szych, Joanna, et al.
Published: (2026)
by: Szych, Joanna, et al.
Published: (2026)
Similar Items
-
Can Large Language Models Transform Natural Language Intent into Formal Method Postconditions?
by: Endres, Madeline, et al.
Published: (2023) -
3DGen: AI-Assisted Generation of Provably Correct Binary Format Parsers
by: Fakhoury, Sarah, et al.
Published: (2024) -
Towards Neural Synthesis for SMT-Assisted Proof-Oriented Programming
by: Chakraborty, Saikat, et al.
Published: (2024) -
Ranking LLM-Generated Loop Invariants for Program Verification
by: Chakraborty, Saikat, et al.
Published: (2023) -
Program Structure Aware Precondition Generation
by: Dinella, Elizabeth, et al.
Published: (2023)