LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xia, Yunhui, Shen, Wei, Wang, Yan, Liu, Jason Klein, Sun, Huifeng, Wu, Siyue, Hu, Jian, Xu, Xiaolong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Analyzing Prominent LLMs: An Empirical Study of Performance and Complexity in Solving LeetCode Problems
von: Guimaraes, Everton, et al.
Veröffentlicht: (2025)
von: Guimaraes, Everton, et al.
Veröffentlicht: (2025)
A case study on the transformative potential of AI in software engineering on LeetCode and ChatGPT
von: Merkel, Manuel, et al.
Veröffentlicht: (2025)
von: Merkel, Manuel, et al.
Veröffentlicht: (2025)
PuzzleMark: Implicit Jigsaw Learning for Robust Code Dataset Watermarking in Neural Code Completion Models
von: Huang, Haocheng, et al.
Veröffentlicht: (2026)
von: Huang, Haocheng, et al.
Veröffentlicht: (2026)
A Vulnerability Code Intent Summary Dataset
von: Huang, Yifan, et al.
Veröffentlicht: (2025)
von: Huang, Yifan, et al.
Veröffentlicht: (2025)
Clean Code, Better Models: Enhancing LLM Performance with Smell-Cleaned Dataset
von: Xue, Zhipeng, et al.
Veröffentlicht: (2025)
von: Xue, Zhipeng, et al.
Veröffentlicht: (2025)
Optimizing Datasets for Code Summarization: Is Code-Comment Coherence Enough?
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
von: Vitale, Antonio, et al.
Veröffentlicht: (2025)
CREME: Robustness Enhancement of Code LLMs via Layer-Aware Model Editing
von: Liu, Shuhan, et al.
Veröffentlicht: (2025)
von: Liu, Shuhan, et al.
Veröffentlicht: (2025)
NLPerturbator: Studying the Robustness of Code LLMs to Natural Language Variations
von: Chen, Junkai, et al.
Veröffentlicht: (2024)
von: Chen, Junkai, et al.
Veröffentlicht: (2024)
OpenCodeInstruct: A Large-scale Instruction Tuning Dataset for Code LLMs
von: Ahmad, Wasi Uddin, et al.
Veröffentlicht: (2025)
von: Ahmad, Wasi Uddin, et al.
Veröffentlicht: (2025)
Train in Vain: Functionality-Preserving Poisoning to Prevent Unauthorized Use of Code Datasets
von: Xiao, Yuan, et al.
Veröffentlicht: (2026)
von: Xiao, Yuan, et al.
Veröffentlicht: (2026)
Teaching Code LLMs to Use Autocompletion Tools in Repository-Level Code Generation
von: Wang, Chong, et al.
Veröffentlicht: (2024)
von: Wang, Chong, et al.
Veröffentlicht: (2024)
MegaVul: A C/C++ Vulnerability Dataset with Comprehensive Code Representation
von: Ni, Chao, et al.
Veröffentlicht: (2024)
von: Ni, Chao, et al.
Veröffentlicht: (2024)
Improving Automated Secure Code Reviews: A Synthetic Dataset for Code Vulnerability Flaws
von: Centellas-Claros, Leonardo, et al.
Veröffentlicht: (2025)
von: Centellas-Claros, Leonardo, et al.
Veröffentlicht: (2025)
A Dataset of Agentic AI Coding Tool Configurations
von: Galster, Matthias, et al.
Veröffentlicht: (2026)
von: Galster, Matthias, et al.
Veröffentlicht: (2026)
Secure-Instruct: An Automated Pipeline for Synthesizing Instruction-Tuning Datasets Using LLMs for Secure Code Generation
von: Li, Junjie, et al.
Veröffentlicht: (2025)
von: Li, Junjie, et al.
Veröffentlicht: (2025)
Open the Oyster: Empirical Evaluation and Improvement of Code Reasoning Confidence in LLMs
von: Wang, Shufan, et al.
Veröffentlicht: (2025)
von: Wang, Shufan, et al.
Veröffentlicht: (2025)
CodeSense: a Real-World Benchmark and Dataset for Code Semantic Reasoning
von: Roy, Monoshi Kumar, et al.
Veröffentlicht: (2025)
von: Roy, Monoshi Kumar, et al.
Veröffentlicht: (2025)
Evaluation of Code LLMs on Geospatial Code Generation
von: Gramacki, Piotr, et al.
Veröffentlicht: (2024)
von: Gramacki, Piotr, et al.
Veröffentlicht: (2024)
MADE-WIC: Multiple Annotated Datasets for Exploring Weaknesses In Code
von: Mock, Moritz, et al.
Veröffentlicht: (2024)
von: Mock, Moritz, et al.
Veröffentlicht: (2024)
An Exploratory Investigation into Code License Infringements in Large Language Model Training Datasets
von: Katzy, Jonathan, et al.
Veröffentlicht: (2024)
von: Katzy, Jonathan, et al.
Veröffentlicht: (2024)
CodeFort: Robust Training for Code Generation Models
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
On Evaluating the Efficiency of Source Code Generated by LLMs
von: Niu, Changan, et al.
Veröffentlicht: (2024)
von: Niu, Changan, et al.
Veröffentlicht: (2024)
An Approach to Detect Abnormal Submissions for CodeWorkout Dataset
von: Hicks, Alex, et al.
Veröffentlicht: (2024)
von: Hicks, Alex, et al.
Veröffentlicht: (2024)
Learned or Memorized ? Quantifying Memorization Advantage in Code LLMs
von: Euraste, Djiré Albérick, et al.
Veröffentlicht: (2026)
von: Euraste, Djiré Albérick, et al.
Veröffentlicht: (2026)
SelfPiCo: Self-Guided Partial Code Execution with LLMs
von: Xue, Zhipeng, et al.
Veröffentlicht: (2024)
von: Xue, Zhipeng, et al.
Veröffentlicht: (2024)
SR-Eval: Evaluating LLMs on Code Generation under Stepwise Requirement Refinement
von: Zhan, Zexun, et al.
Veröffentlicht: (2025)
von: Zhan, Zexun, et al.
Veröffentlicht: (2025)
Enhancing and Reporting Robustness Boundary of Neural Code Models for Intelligent Code Understanding
von: Han, Tingxu, et al.
Veröffentlicht: (2026)
von: Han, Tingxu, et al.
Veröffentlicht: (2026)
ADC: Enhancing Function Calling Via Adversarial Datasets and Code Line-Level Feedback
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
ComplexCodeEval: A Benchmark for Evaluating Large Code Models on More Complex Code
von: Feng, Jia, et al.
Veröffentlicht: (2024)
von: Feng, Jia, et al.
Veröffentlicht: (2024)
SynConfRoute: Syntax-Aware Routing for Efficient Code Completion with Small CodeLLMs
von: Thangarajah, Kishanthan, et al.
Veröffentlicht: (2026)
von: Thangarajah, Kishanthan, et al.
Veröffentlicht: (2026)
Write Your Own CodeChecker: An Automated Test-Driven Checker Development Approach with LLMs
von: Liu, Jun, et al.
Veröffentlicht: (2024)
von: Liu, Jun, et al.
Veröffentlicht: (2024)
DeCoMa: Detecting and Purifying Code Dataset Watermarks through Dual Channel Code Abstraction
von: Xiao, Yuan, et al.
Veröffentlicht: (2025)
von: Xiao, Yuan, et al.
Veröffentlicht: (2025)
Optimizing Token Consumption in LLMs: A Nano Surge Approach for Code Reasoning Efficiency
von: Hu, Junwei, et al.
Veröffentlicht: (2025)
von: Hu, Junwei, et al.
Veröffentlicht: (2025)
Exploring the Capabilities of LLMs for Code Change Related Tasks
von: Fan, Lishui, et al.
Veröffentlicht: (2024)
von: Fan, Lishui, et al.
Veröffentlicht: (2024)
SACS: A Code Smell Dataset using Semi-automatic Generation Approach
von: Zhang, Hanyu, et al.
Veröffentlicht: (2026)
von: Zhang, Hanyu, et al.
Veröffentlicht: (2026)
OSS License Identification at Scale: A Comprehensive Dataset Using World of Code
von: Jahanshahi, Mahmoud, et al.
Veröffentlicht: (2024)
von: Jahanshahi, Mahmoud, et al.
Veröffentlicht: (2024)
Token Sugar: Making Source Code Sweeter for LLMs through Token-Efficient Shorthand
von: Sun, Zhensu, et al.
Veröffentlicht: (2025)
von: Sun, Zhensu, et al.
Veröffentlicht: (2025)
LLMs Meet Library Evolution: Evaluating Deprecated API Usage in LLM-based Code Completion
von: Wang, Chong, et al.
Veröffentlicht: (2024)
von: Wang, Chong, et al.
Veröffentlicht: (2024)
CodeInsight: A Curated Dataset of Practical Coding Solutions from Stack Overflow
von: Beau, Nathanaël, et al.
Veröffentlicht: (2024)
von: Beau, Nathanaël, et al.
Veröffentlicht: (2024)
Evaluating LLMs Code Reasoning Under Real-World Context
von: Liu, Changshu
Veröffentlicht: (2026)
von: Liu, Changshu
Veröffentlicht: (2026)
Ähnliche Einträge
-
Analyzing Prominent LLMs: An Empirical Study of Performance and Complexity in Solving LeetCode Problems
von: Guimaraes, Everton, et al.
Veröffentlicht: (2025) -
A case study on the transformative potential of AI in software engineering on LeetCode and ChatGPT
von: Merkel, Manuel, et al.
Veröffentlicht: (2025) -
PuzzleMark: Implicit Jigsaw Learning for Robust Code Dataset Watermarking in Neural Code Completion Models
von: Huang, Haocheng, et al.
Veröffentlicht: (2026) -
A Vulnerability Code Intent Summary Dataset
von: Huang, Yifan, et al.
Veröffentlicht: (2025) -
Clean Code, Better Models: Enhancing LLM Performance with Smell-Cleaned Dataset
von: Xue, Zhipeng, et al.
Veröffentlicht: (2025)