Scaling Test-Driven Code Generation from Functions to Classes: An Empirical Study
Fuente:
arXiv
Guardado en:
| Autores principales: | Liang, Yunhao, Ying, Ruixuan, Ni, Shiwen, Cui, Zhe |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HyClone: Bridging LLM Understanding and Dynamic Execution for Semantic Code Clone Detection
por: Liang, Yunhao, et al.
Publicado: (2025)
por: Liang, Yunhao, et al.
Publicado: (2025)
LLM-Based Test-Driven Interactive Code Generation: User Study and Empirical Evaluation
por: Fakhoury, Sarah, et al.
Publicado: (2024)
por: Fakhoury, Sarah, et al.
Publicado: (2024)
Are Coding Agents Generating Over-Mocked Tests? An Empirical Study
por: Hora, Andre, et al.
Publicado: (2026)
por: Hora, Andre, et al.
Publicado: (2026)
Understanding Specification-Driven Code Generation with LLMs: An Empirical Study Design
por: Rosa, Giovanni, et al.
Publicado: (2026)
por: Rosa, Giovanni, et al.
Publicado: (2026)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
por: Cui, Yi
Publicado: (2025)
por: Cui, Yi
Publicado: (2025)
An Empirical Study of Proactive Coding Assistants in Real-World Software Development
por: Li, Lehui, et al.
Publicado: (2026)
por: Li, Lehui, et al.
Publicado: (2026)
A Preference-Driven Methodology for High-Quality Solidity Code Generation
por: Peng, Zhiyuan, et al.
Publicado: (2025)
por: Peng, Zhiyuan, et al.
Publicado: (2025)
Large Language Models for Automated Web-Form-Test Generation: An Empirical Study
por: Li, Tao, et al.
Publicado: (2024)
por: Li, Tao, et al.
Publicado: (2024)
A Large-Scale Empirical Study of AI-Generated Code in Real-World Repositories
por: Mao, Tianhao, et al.
Publicado: (2026)
por: Mao, Tianhao, et al.
Publicado: (2026)
CodeChemist: Functional Knowledge Transfer for Low-Resource Code Generation via Test-Time Scaling
por: Wang, Kaixin, et al.
Publicado: (2025)
por: Wang, Kaixin, et al.
Publicado: (2025)
An Empirical Study on the Code Refactoring Capability of Large Language Models
por: Cordeiro, Jonathan, et al.
Publicado: (2024)
por: Cordeiro, Jonathan, et al.
Publicado: (2024)
Understanding Self-Admitted Technical Debt in Test Code: An Empirical Study
por: Nakamura, Ibuki, et al.
Publicado: (2025)
por: Nakamura, Ibuki, et al.
Publicado: (2025)
Guiding AI to Fix Its Own Flaws: An Empirical Study on LLM-Driven Secure Code Generation
por: Yan, Hao, et al.
Publicado: (2025)
por: Yan, Hao, et al.
Publicado: (2025)
Debt Behind the AI Boom: A Large-Scale Empirical Study of AI-Generated Code in the Wild
por: Liu, Yue, et al.
Publicado: (2026)
por: Liu, Yue, et al.
Publicado: (2026)
Can We Classify Flaky Tests Using Only Test Code? An LLM-Based Empirical Study
por: Berndt, Alexander, et al.
Publicado: (2026)
por: Berndt, Alexander, et al.
Publicado: (2026)
An Empirical Study of the Non-determinism of ChatGPT in Code Generation
por: Ouyang, Shuyin, et al.
Publicado: (2023)
por: Ouyang, Shuyin, et al.
Publicado: (2023)
An Empirical Study of Retrieval-Augmented Code Generation: Challenges and Opportunities
por: Yang, Zezhou, et al.
Publicado: (2025)
por: Yang, Zezhou, et al.
Publicado: (2025)
Testing with AI Agents: An Empirical Study of Test Generation Frequency, Quality, and Coverage
por: Yoshimoto, Suzuka, et al.
Publicado: (2026)
por: Yoshimoto, Suzuka, et al.
Publicado: (2026)
An Exploratory Study of Bayesian Prompt Optimization for Test-Driven Code Generation with Large Language Models
por: Tomar, Shlok, et al.
Publicado: (2025)
por: Tomar, Shlok, et al.
Publicado: (2025)
Test-Driven Development for Code Generation
por: Mathews, Noble Saji, et al.
Publicado: (2024)
por: Mathews, Noble Saji, et al.
Publicado: (2024)
Assessing Small Language Models for Code Generation: An Empirical Study with Benchmarks
por: Hasan, Md Mahade, et al.
Publicado: (2025)
por: Hasan, Md Mahade, et al.
Publicado: (2025)
Understanding NPM Malicious Package Detection: A Benchmark-Driven Empirical Analysis
por: Guo, Wenbo, et al.
Publicado: (2026)
por: Guo, Wenbo, et al.
Publicado: (2026)
What Makes Software Bugs Escape Testing? Evidence from a Large-Scale Empirical Study
por: Cotroneo, Domenico, et al.
Publicado: (2026)
por: Cotroneo, Domenico, et al.
Publicado: (2026)
On the Use of Agentic Coding Manifests: An Empirical Study of Claude Code
por: Chatlatanagulchai, Worawalan, et al.
Publicado: (2025)
por: Chatlatanagulchai, Worawalan, et al.
Publicado: (2025)
An Empirical Study on the Effects of System Prompts in Instruction-Tuned Models for Code Generation
por: Cheng, Zaiyu, et al.
Publicado: (2026)
por: Cheng, Zaiyu, et al.
Publicado: (2026)
An Empirical Study of False Negatives and Positives of Static Code Analyzers From the Perspective of Historical Issues
por: Cui, Han, et al.
Publicado: (2024)
por: Cui, Han, et al.
Publicado: (2024)
On the Effectiveness of Training Data Optimization for LLM-based Code Generation: An Empirical Study
por: Kuang, Shiqi, et al.
Publicado: (2025)
por: Kuang, Shiqi, et al.
Publicado: (2025)
What to Retrieve for Effective Retrieval-Augmented Code Generation? An Empirical Study and Beyond
por: Gu, Wenchao, et al.
Publicado: (2025)
por: Gu, Wenchao, et al.
Publicado: (2025)
From Exploration to Specification: LLM-Based Property Generation for Mobile App Testing
por: Xiong, Yiheng, et al.
Publicado: (2026)
por: Xiong, Yiheng, et al.
Publicado: (2026)
Security Weaknesses of Copilot-Generated Code in GitHub Projects: An Empirical Study
por: Fu, Yujia, et al.
Publicado: (2023)
por: Fu, Yujia, et al.
Publicado: (2023)
An Empirical Study of Interaction Smells in Multi-Turn Human-LLM Collaborative Code Generation
por: Zhang, Binquan, et al.
Publicado: (2026)
por: Zhang, Binquan, et al.
Publicado: (2026)
AI builds, We Analyze: An Empirical Study of AI-Generated Build Code Quality
por: Ghammam, Anwar, et al.
Publicado: (2026)
por: Ghammam, Anwar, et al.
Publicado: (2026)
An Empirical Study on Automatically Detecting AI-Generated Source Code: How Far Are We?
por: Suh, Hyunjae, et al.
Publicado: (2024)
por: Suh, Hyunjae, et al.
Publicado: (2024)
Scaling Mobile Chaos Testing with AI-Driven Test Execution
por: Marcano, Juan, et al.
Publicado: (2026)
por: Marcano, Juan, et al.
Publicado: (2026)
An Empirical Study of LLM-Based Code Clone Detection
por: Zhu, Wenqing, et al.
Publicado: (2025)
por: Zhu, Wenqing, et al.
Publicado: (2025)
Explaining Explanation: An Empirical Study on Explanation in Code Reviews
por: Widyasari, Ratnadira, et al.
Publicado: (2023)
por: Widyasari, Ratnadira, et al.
Publicado: (2023)
ABTest: Behavior-Driven Testing for AI Coding Agents
por: Dai, Wuyang, et al.
Publicado: (2026)
por: Dai, Wuyang, et al.
Publicado: (2026)
Toward Functional and Non-Functional Evaluation of Application-Level Code Generation
por: Pan, Ruwei, et al.
Publicado: (2026)
por: Pan, Ruwei, et al.
Publicado: (2026)
Compact Constraint Encoding for LLM Code Generation: An Empirical Study of Token Economics and Constraint Compliance
por: Tang, Hanzhang
Publicado: (2026)
por: Tang, Hanzhang
Publicado: (2026)
Detect Repair Verify for Securing LLM Generated Code: A Multi-Language Empirical Study
por: Cheng, Cheng
Publicado: (2026)
por: Cheng, Cheng
Publicado: (2026)
Ejemplares similares
-
HyClone: Bridging LLM Understanding and Dynamic Execution for Semantic Code Clone Detection
por: Liang, Yunhao, et al.
Publicado: (2025) -
LLM-Based Test-Driven Interactive Code Generation: User Study and Empirical Evaluation
por: Fakhoury, Sarah, et al.
Publicado: (2024) -
Are Coding Agents Generating Over-Mocked Tests? An Empirical Study
por: Hora, Andre, et al.
Publicado: (2026) -
Understanding Specification-Driven Code Generation with LLMs: An Empirical Study Design
por: Rosa, Giovanni, et al.
Publicado: (2026) -
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
por: Cui, Yi
Publicado: (2025)