Rethinking the effects of data contamination in Code Intelligence
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Zhen, Lin, Hongyi, He, Yifan, Wang, Junqi, Sun, Zeyu, Liu, Shuo, Xu, Jie, Wang, Pengpeng, Yu, Zhongxing, Liang, Qingyuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Contextualized Code Pretraining for Code Generation
di: Liu, Chen, et al.
Pubblicazione: (2026)
di: Liu, Chen, et al.
Pubblicazione: (2026)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
di: Gong, Zhihao, et al.
Pubblicazione: (2026)
di: Gong, Zhihao, et al.
Pubblicazione: (2026)
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
di: Gong, Zhihao, et al.
Pubblicazione: (2025)
di: Gong, Zhihao, et al.
Pubblicazione: (2025)
Directional Diffusion-Style Code Editing Pre-training
di: Liang, Qingyuan, et al.
Pubblicazione: (2025)
di: Liang, Qingyuan, et al.
Pubblicazione: (2025)
Prompt Alchemy: Automatic Prompt Refinement for Enhancing Code Generation
di: Ye, Sixiang, et al.
Pubblicazione: (2025)
di: Ye, Sixiang, et al.
Pubblicazione: (2025)
GramTrans: A Better Code Representation Approach in Code Generation
di: Zhang, Zhao, et al.
Pubblicazione: (2025)
di: Zhang, Zhao, et al.
Pubblicazione: (2025)
NARRepair: Non-Autoregressive Code Generation Model for Automatic Program Repair
di: Yang, Zhenyu, et al.
Pubblicazione: (2024)
di: Yang, Zhenyu, et al.
Pubblicazione: (2024)
Do Advanced Language Models Eliminate the Need for Prompt Engineering in Software Engineering?
di: Wang, Guoqing, et al.
Pubblicazione: (2024)
di: Wang, Guoqing, et al.
Pubblicazione: (2024)
Automatically Learning a Precise Measurement for Fault Diagnosis Capability of Test Cases
di: Zhao, Yifan, et al.
Pubblicazione: (2025)
di: Zhao, Yifan, et al.
Pubblicazione: (2025)
R2ComSync: Improving Code-Comment Synchronization with In-Context Learning and Reranking
di: Yang, Zhen, et al.
Pubblicazione: (2025)
di: Yang, Zhen, et al.
Pubblicazione: (2025)
Lyra: A Benchmark for Turducken-Style Code Generation
di: Liang, Qingyuan, et al.
Pubblicazione: (2021)
di: Liang, Qingyuan, et al.
Pubblicazione: (2021)
CupCleaner: A Hybrid Data Cleaning Approach for Comment Updating
di: Liang, Qingyuan, et al.
Pubblicazione: (2023)
di: Liang, Qingyuan, et al.
Pubblicazione: (2023)
Exploring and Lifting the Robustness of LLM-powered Automated Program Repair with Metamorphic Testing
di: Xue, Pengyu, et al.
Pubblicazione: (2024)
di: Xue, Pengyu, et al.
Pubblicazione: (2024)
IDOL: Improved Different Optimization Levels Testing for Solidity Compilers
di: Li, Lantian, et al.
Pubblicazione: (2025)
di: Li, Lantian, et al.
Pubblicazione: (2025)
Condor: A Code Discriminator Integrating General Semantics with Code Details
di: Liang, Qingyuan, et al.
Pubblicazione: (2024)
di: Liang, Qingyuan, et al.
Pubblicazione: (2024)
Exploring and Unleashing the Power of Large Language Models in Automated Code Translation
di: Yang, Zhen, et al.
Pubblicazione: (2024)
di: Yang, Zhen, et al.
Pubblicazione: (2024)
Towards Speeding up Program Repair with Non-Autoregressive Model
di: Yang, Zhenyu, et al.
Pubblicazione: (2025)
di: Yang, Zhenyu, et al.
Pubblicazione: (2025)
Parameter-Efficient Fine-Tuning with Attributed Patch Semantic Graph for Automated Patch Correctness Assessment
di: Yang, Zhenyu, et al.
Pubblicazione: (2025)
di: Yang, Zhenyu, et al.
Pubblicazione: (2025)
Search-Induced Issues in Web-Augmented LLM Code Generation: Detecting and Repairing Error-Inducing Pages
di: Wang, Guoqing, et al.
Pubblicazione: (2026)
di: Wang, Guoqing, et al.
Pubblicazione: (2026)
Knowledge-Enhanced Program Repair for Data Science Code
di: Ouyang, Shuyin, et al.
Pubblicazione: (2025)
di: Ouyang, Shuyin, et al.
Pubblicazione: (2025)
Rethinking the Evaluation of Microservice RCA with a Fault Propagation-Aware Benchmark
di: Fang, Aoyang, et al.
Pubblicazione: (2025)
di: Fang, Aoyang, et al.
Pubblicazione: (2025)
Solsmith: Solidity Random Program Generator for Compiler Testing
di: Li, Lantian, et al.
Pubblicazione: (2025)
di: Li, Lantian, et al.
Pubblicazione: (2025)
Understanding Typing-Related Bugs in Solidity Compiler
di: Li, Lantian, et al.
Pubblicazione: (2025)
di: Li, Lantian, et al.
Pubblicazione: (2025)
Instructive Code Retriever: Learn from Large Language Model's Feedback for Code Intelligence Tasks
di: Lu, Jiawei, et al.
Pubblicazione: (2024)
di: Lu, Jiawei, et al.
Pubblicazione: (2024)
DSCodeBench: A Realistic Benchmark for Data Science Code Generation
di: Ouyang, Shuyin, et al.
Pubblicazione: (2025)
di: Ouyang, Shuyin, et al.
Pubblicazione: (2025)
Environment-in-the-Loop: Rethinking Code Migration with LLM-based Agents
di: Li, Xiang, et al.
Pubblicazione: (2026)
di: Li, Xiang, et al.
Pubblicazione: (2026)
Rethinking Technology Stack Selection with AI Coding Proficiency
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2025)
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2025)
Compiler Optimization Testing Based on Optimization-Guided Equivalence Transformations
di: Wu, Jingwen, et al.
Pubblicazione: (2025)
di: Wu, Jingwen, et al.
Pubblicazione: (2025)
Sema Code: Decoupling AI Coding Agents into Programmable, Embeddable Infrastructure
di: Wang, Huacan, et al.
Pubblicazione: (2026)
di: Wang, Huacan, et al.
Pubblicazione: (2026)
Toward Functional and Non-Functional Evaluation of Application-Level Code Generation
di: Pan, Ruwei, et al.
Pubblicazione: (2026)
di: Pan, Ruwei, et al.
Pubblicazione: (2026)
A Vulnerability Code Intent Summary Dataset
di: Huang, Yifan, et al.
Pubblicazione: (2025)
di: Huang, Yifan, et al.
Pubblicazione: (2025)
Hyperion: Unveiling DApp Inconsistencies using LLM and Dataflow-Guided Symbolic Execution
di: Yang, Shuo, et al.
Pubblicazione: (2024)
di: Yang, Shuo, et al.
Pubblicazione: (2024)
LLMs are Bug Replicators: An Empirical Study on LLMs' Capability in Completing Bug-prone Code
di: Guo, Liwei, et al.
Pubblicazione: (2025)
di: Guo, Liwei, et al.
Pubblicazione: (2025)
RAG-Enhanced Commit Message Generation
di: Zhang, Linghao, et al.
Pubblicazione: (2024)
di: Zhang, Linghao, et al.
Pubblicazione: (2024)
Interleaved Learning and Exploration: A Self-Adaptive Fuzz Testing Framework for MLIR
di: Sun, Zeyu, et al.
Pubblicazione: (2025)
di: Sun, Zeyu, et al.
Pubblicazione: (2025)
Evaluating and Achieving Controllable Code Completion in Code LLM
di: Zhang, Jiajun, et al.
Pubblicazione: (2026)
di: Zhang, Jiajun, et al.
Pubblicazione: (2026)
WebCompass: Towards Multimodal Web Coding Evaluation for Code Language Models
di: Lei, Xinping, et al.
Pubblicazione: (2026)
di: Lei, Xinping, et al.
Pubblicazione: (2026)
Automated Commit Message Generation with Large Language Models: An Empirical Study and Beyond
di: Xue, Pengyu, et al.
Pubblicazione: (2024)
di: Xue, Pengyu, et al.
Pubblicazione: (2024)
Understanding Inconsistent State Update Vulnerabilities in Smart Contracts
di: Li, Lantian, et al.
Pubblicazione: (2025)
di: Li, Lantian, et al.
Pubblicazione: (2025)
LAURA: Enhancing Code Review Generation with Context-Enriched Retrieval-Augmented LLM
di: Zhang, Yuxin, et al.
Pubblicazione: (2025)
di: Zhang, Yuxin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Contextualized Code Pretraining for Code Generation
di: Liu, Chen, et al.
Pubblicazione: (2026) -
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
di: Gong, Zhihao, et al.
Pubblicazione: (2026) -
TRACE: Evaluating Execution Efficiency of LLM-Based Code Translation
di: Gong, Zhihao, et al.
Pubblicazione: (2025) -
Directional Diffusion-Style Code Editing Pre-training
di: Liang, Qingyuan, et al.
Pubblicazione: (2025) -
Prompt Alchemy: Automatic Prompt Refinement for Enhancing Code Generation
di: Ye, Sixiang, et al.
Pubblicazione: (2025)