Understanding Specification-Driven Code Generation with LLMs: An Empirical Study Design
Fuente:
arXiv
Guardado en:
| Autores principales: | Rosa, Giovanni, Moreno-Lumbreras, David, Robles, Gregorio, González-Barahona, Jesús M. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Role of Code Proficiency in the Era of Generative AI
por: Robles, Gregorio, et al.
Publicado: (2024)
por: Robles, Gregorio, et al.
Publicado: (2024)
Not Only for Developers: Exploring Plugin Maintenance for Knowledge-Centric Communities
por: Rosa, Giovanni, et al.
Publicado: (2026)
por: Rosa, Giovanni, et al.
Publicado: (2026)
Software development in the age of LLMs and XR
por: Gonzalez-Barahona, Jesus M.
Publicado: (2024)
por: Gonzalez-Barahona, Jesus M.
Publicado: (2024)
How Do Code Smells Affect Skill Growth in Scratch Novice Programmers?
por: Aragón, Ricardo Hidalgo, et al.
Publicado: (2025)
por: Aragón, Ricardo Hidalgo, et al.
Publicado: (2025)
Scaling Test-Driven Code Generation from Functions to Classes: An Empirical Study
por: Liang, Yunhao, et al.
Publicado: (2026)
por: Liang, Yunhao, et al.
Publicado: (2026)
LLMs are Bug Replicators: An Empirical Study on LLMs' Capability in Completing Bug-prone Code
por: Guo, Liwei, et al.
Publicado: (2025)
por: Guo, Liwei, et al.
Publicado: (2025)
SmartDoc: A Context-Aware Agentic Method Comment Generation Plugin
por: Etemadi, Vahid, et al.
Publicado: (2025)
por: Etemadi, Vahid, et al.
Publicado: (2025)
LLM-Based Test-Driven Interactive Code Generation: User Study and Empirical Evaluation
por: Fakhoury, Sarah, et al.
Publicado: (2024)
por: Fakhoury, Sarah, et al.
Publicado: (2024)
HTML Structure Exploration in 3D Software Cities
por: Hansen, Malte, et al.
Publicado: (2025)
por: Hansen, Malte, et al.
Publicado: (2025)
Understanding Self-Admitted Technical Debt in Test Code: An Empirical Study
por: Nakamura, Ibuki, et al.
Publicado: (2025)
por: Nakamura, Ibuki, et al.
Publicado: (2025)
An Empirical Study of the Non-determinism of ChatGPT in Code Generation
por: Ouyang, Shuyin, et al.
Publicado: (2023)
por: Ouyang, Shuyin, et al.
Publicado: (2023)
Guiding AI to Fix Its Own Flaws: An Empirical Study on LLM-Driven Secure Code Generation
por: Yan, Hao, et al.
Publicado: (2025)
por: Yan, Hao, et al.
Publicado: (2025)
An Empirical Study of Perceptions of General LLMs and Multimodal LLMs on Hugging Face
por: Liu, Yujian, et al.
Publicado: (2026)
por: Liu, Yujian, et al.
Publicado: (2026)
Understanding Chain-of-Thought Effectiveness in Code Generation: An Empirical and Information-Theoretic Analysis
por: Jin, Naizhu, et al.
Publicado: (2025)
por: Jin, Naizhu, et al.
Publicado: (2025)
Do Code LLMs Understand Design Patterns?
por: Pan, Zhenyu, et al.
Publicado: (2025)
por: Pan, Zhenyu, et al.
Publicado: (2025)
An Empirical Study of Knowledge Distillation for Code Understanding Tasks
por: Wang, Ruiqi, et al.
Publicado: (2025)
por: Wang, Ruiqi, et al.
Publicado: (2025)
Towards Identifying Code Proficiency through the Analysis of Python Textbooks
por: Rojpaisarnkit, Ruksit, et al.
Publicado: (2024)
por: Rojpaisarnkit, Ruksit, et al.
Publicado: (2024)
Analyzing Prominent LLMs: An Empirical Study of Performance and Complexity in Solving LeetCode Problems
por: Guimaraes, Everton, et al.
Publicado: (2025)
por: Guimaraes, Everton, et al.
Publicado: (2025)
Are Coding Agents Generating Over-Mocked Tests? An Empirical Study
por: Hora, Andre, et al.
Publicado: (2026)
por: Hora, Andre, et al.
Publicado: (2026)
An Empirical Study of Retrieval-Augmented Code Generation: Challenges and Opportunities
por: Yang, Zezhou, et al.
Publicado: (2025)
por: Yang, Zezhou, et al.
Publicado: (2025)
Understanding the Challenges and Opportunities of Generative AI Apps: An Empirical Study
por: AlMulla, Buthayna, et al.
Publicado: (2025)
por: AlMulla, Buthayna, et al.
Publicado: (2025)
Exploring the Effectiveness of LLMs in Automated Logging Generation: An Empirical Study
por: Li, Yichen, et al.
Publicado: (2023)
por: Li, Yichen, et al.
Publicado: (2023)
Contextual Fairness-Aware Practices in ML: A Cost-Effective Empirical Evaluation
por: Parziale, Alessandra, et al.
Publicado: (2025)
por: Parziale, Alessandra, et al.
Publicado: (2025)
Explaining Explanation: An Empirical Study on Explanation in Code Reviews
por: Widyasari, Ratnadira, et al.
Publicado: (2023)
por: Widyasari, Ratnadira, et al.
Publicado: (2023)
Assessing Small Language Models for Code Generation: An Empirical Study with Benchmarks
por: Hasan, Md Mahade, et al.
Publicado: (2025)
por: Hasan, Md Mahade, et al.
Publicado: (2025)
Prompt Engineering or Fine-Tuning: An Empirical Assessment of LLMs for Code
por: Shin, Jiho, et al.
Publicado: (2023)
por: Shin, Jiho, et al.
Publicado: (2023)
LLMs for Generation of Architectural Components: An Exploratory Empirical Study in the Serverless World
por: Arun, Shrikara, et al.
Publicado: (2025)
por: Arun, Shrikara, et al.
Publicado: (2025)
On the Use of Agentic Coding Manifests: An Empirical Study of Claude Code
por: Chatlatanagulchai, Worawalan, et al.
Publicado: (2025)
por: Chatlatanagulchai, Worawalan, et al.
Publicado: (2025)
DSL or Code? Evaluating the Quality of LLM-Generated Algebraic Specifications: A Case Study in Optimization at Kinaxis
por: Ayoughi, Negin, et al.
Publicado: (2026)
por: Ayoughi, Negin, et al.
Publicado: (2026)
Debt Behind the AI Boom: A Large-Scale Empirical Study of AI-Generated Code in the Wild
por: Liu, Yue, et al.
Publicado: (2026)
por: Liu, Yue, et al.
Publicado: (2026)
An Empirical Study on the Effects of System Prompts in Instruction-Tuned Models for Code Generation
por: Cheng, Zaiyu, et al.
Publicado: (2026)
por: Cheng, Zaiyu, et al.
Publicado: (2026)
On the Effectiveness of Training Data Optimization for LLM-based Code Generation: An Empirical Study
por: Kuang, Shiqi, et al.
Publicado: (2025)
por: Kuang, Shiqi, et al.
Publicado: (2025)
What to Retrieve for Effective Retrieval-Augmented Code Generation? An Empirical Study and Beyond
por: Gu, Wenchao, et al.
Publicado: (2025)
por: Gu, Wenchao, et al.
Publicado: (2025)
An Empirical Study on Commit Message Generation using LLMs via In-Context Learning
por: Wu, Yifan, et al.
Publicado: (2025)
por: Wu, Yifan, et al.
Publicado: (2025)
An Empirical Study on the Effectiveness of Large Language Models for Binary Code Understanding
por: Shang, Xiuwei, et al.
Publicado: (2025)
por: Shang, Xiuwei, et al.
Publicado: (2025)
Understanding Code Agent Behaviour: An Empirical Study of Success and Failure Trajectories
por: Majgaonkar, Oorja, et al.
Publicado: (2025)
por: Majgaonkar, Oorja, et al.
Publicado: (2025)
An Empirical Study on Capability of Large Language Models in Understanding Code Semantics
por: Nguyen, Thu-Trang, et al.
Publicado: (2024)
por: Nguyen, Thu-Trang, et al.
Publicado: (2024)
From Industry Claims to Empirical Reality: An Empirical Study of Code Review Agents in Pull Requests
por: Chowdhury, Kowshik, et al.
Publicado: (2026)
por: Chowdhury, Kowshik, et al.
Publicado: (2026)
Using LLMs in Software Requirements Specifications: An Empirical Evaluation
por: Krishna, Madhava, et al.
Publicado: (2024)
por: Krishna, Madhava, et al.
Publicado: (2024)
An Empirical Study of Interaction Smells in Multi-Turn Human-LLM Collaborative Code Generation
por: Zhang, Binquan, et al.
Publicado: (2026)
por: Zhang, Binquan, et al.
Publicado: (2026)
Ejemplares similares
-
The Role of Code Proficiency in the Era of Generative AI
por: Robles, Gregorio, et al.
Publicado: (2024) -
Not Only for Developers: Exploring Plugin Maintenance for Knowledge-Centric Communities
por: Rosa, Giovanni, et al.
Publicado: (2026) -
Software development in the age of LLMs and XR
por: Gonzalez-Barahona, Jesus M.
Publicado: (2024) -
How Do Code Smells Affect Skill Growth in Scratch Novice Programmers?
por: Aragón, Ricardo Hidalgo, et al.
Publicado: (2025) -
Scaling Test-Driven Code Generation from Functions to Classes: An Empirical Study
por: Liang, Yunhao, et al.
Publicado: (2026)