Enregistré dans:
| Auteurs principaux: | Martins, Marcelo de Rezende, Gerosa, Marco A. |
|---|---|
| Format: | Preprint |
| Publié: |
2020
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2009.01959 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Applying Large Language Models API to Issue Classification Problem
par: Aracena, Gabriel, et autres
Publié: (2024)
par: Aracena, Gabriel, et autres
Publié: (2024)
CoCoST: Automatic Complex Code Generation with Online Searching and Correctness Testing
par: He, Xinyi, et autres
Publié: (2024)
par: He, Xinyi, et autres
Publié: (2024)
Neural Models for Source Code Synthesis and Completion
par: Niyogi, Mitodru
Publié: (2024)
par: Niyogi, Mitodru
Publié: (2024)
A Semantic-based Optimization Approach for Repairing LLMs: Case Study on Code Generation
par: Gu, Jian, et autres
Publié: (2025)
par: Gu, Jian, et autres
Publié: (2025)
SelfCodeAlign: Self-Alignment for Code Generation
par: Wei, Yuxiang, et autres
Publié: (2024)
par: Wei, Yuxiang, et autres
Publié: (2024)
IntentCoding: Amplifying User Intent in Code Generation
par: Fang, Zheng, et autres
Publié: (2026)
par: Fang, Zheng, et autres
Publié: (2026)
LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
par: Xia, Yunhui, et autres
Publié: (2025)
par: Xia, Yunhui, et autres
Publié: (2025)
CodeJudge: Evaluating Code Generation with Large Language Models
par: Tong, Weixi, et autres
Publié: (2024)
par: Tong, Weixi, et autres
Publié: (2024)
LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
par: Jain, Naman, et autres
Publié: (2024)
par: Jain, Naman, et autres
Publié: (2024)
NaturalCodeBench: Examining Coding Performance Mismatch on HumanEval and Natural User Prompts
par: Zhang, Shudan, et autres
Publié: (2024)
par: Zhang, Shudan, et autres
Publié: (2024)
CodeUltraFeedback: An LLM-as-a-Judge Dataset for Aligning Large Language Models to Coding Preferences
par: Weyssow, Martin, et autres
Publié: (2024)
par: Weyssow, Martin, et autres
Publié: (2024)
CONCUR: Benchmarking LLMs for Concurrent Code Generation
par: Huang, Jue, et autres
Publié: (2026)
par: Huang, Jue, et autres
Publié: (2026)
Learning Code Preference via Synthetic Evolution
par: Liu, Jiawei, et autres
Publié: (2024)
par: Liu, Jiawei, et autres
Publié: (2024)
Evaluating Language Models for Efficient Code Generation
par: Liu, Jiawei, et autres
Publié: (2024)
par: Liu, Jiawei, et autres
Publié: (2024)
StackEval: Benchmarking LLMs in Coding Assistance
par: Shah, Nidhish, et autres
Publié: (2024)
par: Shah, Nidhish, et autres
Publié: (2024)
A Multi-Perspective Architecture for Semantic Code Search
par: Haldar, Rajarshi, et autres
Publié: (2020)
par: Haldar, Rajarshi, et autres
Publié: (2020)
Towards Understanding What Code Language Models Learned
par: Ahmed, Toufique, et autres
Publié: (2023)
par: Ahmed, Toufique, et autres
Publié: (2023)
Wisdom and Delusion of LLM Ensembles for Code Generation and Repair
par: Vallecillos-Ruiz, Fernando, et autres
Publié: (2025)
par: Vallecillos-Ruiz, Fernando, et autres
Publié: (2025)
GiFT: Gibbs Fine-Tuning for Code Generation
par: Li, Haochen, et autres
Publié: (2025)
par: Li, Haochen, et autres
Publié: (2025)
MetaLint: Easy-to-Hard Generalization for Code Linting
par: Naik, Atharva, et autres
Publié: (2025)
par: Naik, Atharva, et autres
Publié: (2025)
RepoQA: Evaluating Long Context Code Understanding
par: Liu, Jiawei, et autres
Publié: (2024)
par: Liu, Jiawei, et autres
Publié: (2024)
Uncertainty Awareness of Large Language Models Under Code Distribution Shifts: A Benchmark Study
par: Li, Yufei, et autres
Publié: (2024)
par: Li, Yufei, et autres
Publié: (2024)
Hybrid-Gym: Training Coding Agents to Generalize Across Tasks
par: Xie, Yiqing, et autres
Publié: (2026)
par: Xie, Yiqing, et autres
Publié: (2026)
Data Augmentation for Code Translation with Comparable Corpora and Multiple References
par: Xie, Yiqing, et autres
Publié: (2023)
par: Xie, Yiqing, et autres
Publié: (2023)
AST-T5: Structure-Aware Pretraining for Code Generation and Understanding
par: Gong, Linyuan, et autres
Publié: (2024)
par: Gong, Linyuan, et autres
Publié: (2024)
Position: Intelligent Coding Systems Should Write Programs with Justifications
par: Xu, Xiangzhe, et autres
Publié: (2025)
par: Xu, Xiangzhe, et autres
Publié: (2025)
Planning-Aware Code Infilling via Horizon-Length Prediction
par: Ding, Yifeng, et autres
Publié: (2024)
par: Ding, Yifeng, et autres
Publié: (2024)
Evaluating the Generalization Capabilities of Large Language Models on Code Reasoning
par: Yang, Rem, et autres
Publié: (2025)
par: Yang, Rem, et autres
Publié: (2025)
Quantifying Contamination in Evaluating Code Generation Capabilities of Language Models
par: Riddell, Martin, et autres
Publié: (2024)
par: Riddell, Martin, et autres
Publié: (2024)
CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation
par: Peng, Jinjun, et autres
Publié: (2025)
par: Peng, Jinjun, et autres
Publié: (2025)
Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering
par: Ridnik, Tal, et autres
Publié: (2024)
par: Ridnik, Tal, et autres
Publié: (2024)
SceneGenAgent: Precise Industrial Scene Generation with Coding Agent
par: Xia, Xiao, et autres
Publié: (2024)
par: Xia, Xiao, et autres
Publié: (2024)
Top Leaderboard Ranking = Top Coding Proficiency, Always? EvoEval: Evolving Coding Benchmarks via LLM
par: Xia, Chunqiu Steven, et autres
Publié: (2024)
par: Xia, Chunqiu Steven, et autres
Publié: (2024)
Sense and Sensitivity: Examining the Influence of Semantic Recall on Long Context Code Reasoning
par: Štorek, Adam, et autres
Publié: (2025)
par: Štorek, Adam, et autres
Publié: (2025)
Can OpenSource beat ChatGPT? -- A Comparative Study of Large Language Models for Text-to-Code Generation
par: Mayer, Luis, et autres
Publié: (2024)
par: Mayer, Luis, et autres
Publié: (2024)
SODBench: A Large Language Model Approach to Documenting Spreadsheet Operations
par: Indika, Amila, et autres
Publié: (2025)
par: Indika, Amila, et autres
Publié: (2025)
Code Review Without Borders: Evaluating Synthetic vs. Real Data for Review Recommendation
par: Cohen, Yogev, et autres
Publié: (2025)
par: Cohen, Yogev, et autres
Publié: (2025)
Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation
par: Haider, Md. Asif, et autres
Publié: (2024)
par: Haider, Md. Asif, et autres
Publié: (2024)
Test Code Generation for Telecom Software Systems using Two-Stage Generative Model
par: Nabeel, Mohamad, et autres
Publié: (2024)
par: Nabeel, Mohamad, et autres
Publié: (2024)
Neuron Patching: Semantic-based Neuron-level Language Model Repair for Code Generation
par: Gu, Jian, et autres
Publié: (2023)
par: Gu, Jian, et autres
Publié: (2023)
Documents similaires
-
Applying Large Language Models API to Issue Classification Problem
par: Aracena, Gabriel, et autres
Publié: (2024) -
CoCoST: Automatic Complex Code Generation with Online Searching and Correctness Testing
par: He, Xinyi, et autres
Publié: (2024) -
Neural Models for Source Code Synthesis and Completion
par: Niyogi, Mitodru
Publié: (2024) -
A Semantic-based Optimization Approach for Repairing LLMs: Case Study on Code Generation
par: Gu, Jian, et autres
Publié: (2025) -
SelfCodeAlign: Self-Alignment for Code Generation
par: Wei, Yuxiang, et autres
Publié: (2024)