TENET: Leveraging Tests Beyond Validation for Code Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Yiran, Jiang, Nan, Liang, Shanchao, Wu, Yi, Tan, Lin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can Language Models Replace Programmers for Coding? REPOCOD Says 'Not Yet'
von: Liang, Shanchao, et al.
Veröffentlicht: (2024)
von: Liang, Shanchao, et al.
Veröffentlicht: (2024)
Unified Software Engineering Agent as AI Software Engineer
von: Applis, Leonhard, et al.
Veröffentlicht: (2025)
von: Applis, Leonhard, et al.
Veröffentlicht: (2025)
The SWE-Bench Illusion: When State-of-the-Art LLMs Remember Instead of Reason
von: Liang, Shanchao, et al.
Veröffentlicht: (2025)
von: Liang, Shanchao, et al.
Veröffentlicht: (2025)
Nova: Generative Language Models for Assembly Code with Hierarchical Attention and Contrastive Learning
von: Jiang, Nan, et al.
Veröffentlicht: (2023)
von: Jiang, Nan, et al.
Veröffentlicht: (2023)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
von: Cui, Yi
Veröffentlicht: (2025)
von: Cui, Yi
Veröffentlicht: (2025)
Leveraging Metamemory Mechanisms for Enhanced Data-Free Code Generation in LLMs
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
Collu-Bench: A Benchmark for Predicting Language Model Hallucinations in Code
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
von: Liu, Fang, et al.
Veröffentlicht: (2024)
von: Liu, Fang, et al.
Veröffentlicht: (2024)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
von: Fu, Jia, et al.
Veröffentlicht: (2025)
von: Fu, Jia, et al.
Veröffentlicht: (2025)
Beyond Trusting Trust: Multi-Model Validation for Robust Code Generation
von: McDanel, Bradley
Veröffentlicht: (2025)
von: McDanel, Bradley
Veröffentlicht: (2025)
The Future of Software Testing: AI-Powered Test Case Generation and Validation
von: Baqar, Mohammad, et al.
Veröffentlicht: (2024)
von: Baqar, Mohammad, et al.
Veröffentlicht: (2024)
Test-Driven Development for Code Generation
von: Mathews, Noble Saji, et al.
Veröffentlicht: (2024)
von: Mathews, Noble Saji, et al.
Veröffentlicht: (2024)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
von: Wang, Xiaoyin, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyin, et al.
Veröffentlicht: (2024)
Leveraging GPT-4 for Vulnerability-Witnessing Unit Test Generation
von: Antal, Gábor, et al.
Veröffentlicht: (2025)
von: Antal, Gábor, et al.
Veröffentlicht: (2025)
RedCode: Risky Code Execution and Generation Benchmark for Code Agents
von: Guo, Chengquan, et al.
Veröffentlicht: (2024)
von: Guo, Chengquan, et al.
Veröffentlicht: (2024)
GitTaskBench: A Benchmark for Code Agents Solving Real-World Tasks Through Code Repository Leveraging
von: Ni, Ziyi, et al.
Veröffentlicht: (2025)
von: Ni, Ziyi, et al.
Veröffentlicht: (2025)
Revisit Self-Debugging with Self-Generated Tests for Code Generation
von: Chen, Xiancai, et al.
Veröffentlicht: (2025)
von: Chen, Xiancai, et al.
Veröffentlicht: (2025)
Contract-Coding: Towards Repo-Level Generation via Structured Symbolic Paradigm
von: Lin, Yi, et al.
Veröffentlicht: (2026)
von: Lin, Yi, et al.
Veröffentlicht: (2026)
Leveraging Large Language Models for Enhancing the Understandability of Generated Unit Tests
von: Deljouyi, Amirhossein, et al.
Veröffentlicht: (2024)
von: Deljouyi, Amirhossein, et al.
Veröffentlicht: (2024)
SELF-REDRAFT: Eliciting Intrinsic Exploration-Exploitation Balance in Test-Time Scaling for Code Generation
von: Chen, Yixiang, et al.
Veröffentlicht: (2025)
von: Chen, Yixiang, et al.
Veröffentlicht: (2025)
Leveraging LLMs for Multi-File DSL Code Generation: An Industrial Case Study
von: Chand, Sivajeet, et al.
Veröffentlicht: (2026)
von: Chand, Sivajeet, et al.
Veröffentlicht: (2026)
Bias Testing and Mitigation in LLM-based Code Generation
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
Hallucinations in Code Change to Natural Language Generation: Prevalence and Evaluation of Detection Metrics
von: Liu, Chunhua, et al.
Veröffentlicht: (2025)
von: Liu, Chunhua, et al.
Veröffentlicht: (2025)
RepoRepair: Leveraging Code Documentation for Repository-Level Automated Program Repair
von: Pan, Zhongqiang, et al.
Veröffentlicht: (2026)
von: Pan, Zhongqiang, et al.
Veröffentlicht: (2026)
ReCatcher: Towards LLMs Regression Testing for Code Generation
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2025)
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2025)
DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs
von: Liu, Zhihan, et al.
Veröffentlicht: (2024)
von: Liu, Zhihan, et al.
Veröffentlicht: (2024)
Uncertainty Quantification for LLM-based Code Generation
von: Xu, Senrong, et al.
Veröffentlicht: (2026)
von: Xu, Senrong, et al.
Veröffentlicht: (2026)
I Can Find You in Seconds! Leveraging Large Language Models for Code Authorship Attribution
von: Choi, Soohyeon, et al.
Veröffentlicht: (2025)
von: Choi, Soohyeon, et al.
Veröffentlicht: (2025)
GeoCode-GPT: A Large Language Model for Geospatial Code Generation Tasks
von: Hou, Shuyang, et al.
Veröffentlicht: (2024)
von: Hou, Shuyang, et al.
Veröffentlicht: (2024)
Uncovering Code Insights: Leveraging GitHub Artifacts for Deeper Code Understanding
von: Nevo, Ziv, et al.
Veröffentlicht: (2025)
von: Nevo, Ziv, et al.
Veröffentlicht: (2025)
DomAgent: Leveraging Knowledge Graphs and Case-Based Reasoning for Domain-Specific Code Generation
von: Wang, Shuai, et al.
Veröffentlicht: (2026)
von: Wang, Shuai, et al.
Veröffentlicht: (2026)
What Makes Code Generation Ethically Sourced?
von: Xu, Zhuolin, et al.
Veröffentlicht: (2025)
von: Xu, Zhuolin, et al.
Veröffentlicht: (2025)
Beyond Autoregression: An Empirical Study of Diffusion Large Language Models for Code Generation
von: Li, Chengze, et al.
Veröffentlicht: (2025)
von: Li, Chengze, et al.
Veröffentlicht: (2025)
Domain Adaptation for Code Model-based Unit Test Case Generation
von: Shin, Jiho, et al.
Veröffentlicht: (2023)
von: Shin, Jiho, et al.
Veröffentlicht: (2023)
Unseen Horizons: Unveiling the Real Capability of LLM Code Generation Beyond the Familiar
von: Zhang, Yuanliang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuanliang, et al.
Veröffentlicht: (2024)
From Code Generation to Software Testing: AI Copilot with Context-Based RAG
von: Wang, Yuchen, et al.
Veröffentlicht: (2025)
von: Wang, Yuchen, et al.
Veröffentlicht: (2025)
CATCODER: Repository-Level Code Generation with Relevant Code and Type Context
von: Pan, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Pan, Zhiyuan, et al.
Veröffentlicht: (2024)
A Study on the Improvement of Code Generation Quality Using Large Language Models Leveraging Product Documentation
von: Morimoto, Takuro, et al.
Veröffentlicht: (2025)
von: Morimoto, Takuro, et al.
Veröffentlicht: (2025)
CodeSift: An LLM-Based Reference-Less Framework for Automatic Code Validation
von: Aggarwal, Pooja, et al.
Veröffentlicht: (2024)
von: Aggarwal, Pooja, et al.
Veröffentlicht: (2024)
Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation
von: Ouyang, Shuyin, et al.
Veröffentlicht: (2026)
von: Ouyang, Shuyin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Can Language Models Replace Programmers for Coding? REPOCOD Says 'Not Yet'
von: Liang, Shanchao, et al.
Veröffentlicht: (2024) -
Unified Software Engineering Agent as AI Software Engineer
von: Applis, Leonhard, et al.
Veröffentlicht: (2025) -
The SWE-Bench Illusion: When State-of-the-Art LLMs Remember Instead of Reason
von: Liang, Shanchao, et al.
Veröffentlicht: (2025) -
Nova: Generative Language Models for Assembly Code with Hierarchical Attention and Contrastive Learning
von: Jiang, Nan, et al.
Veröffentlicht: (2023) -
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
von: Cui, Yi
Veröffentlicht: (2025)