Language Models for Code Completion: A Practical Evaluation
Fuente:
arXiv
Guardado en:
| Autores principales: | Izadi, Maliheh, Katzy, Jonathan, van Dam, Tim, Otten, Marc, Popescu, Razvan Mihai, van Deursen, Arie |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
An Exploratory Investigation into Code License Infringements in Large Language Model Training Datasets
por: Katzy, Jonathan, et al.
Publicado: (2024)
por: Katzy, Jonathan, et al.
Publicado: (2024)
The Heap: A Contamination-Free Multilingual Code Dataset for Evaluating Large Language Models
por: Katzy, Jonathan, et al.
Publicado: (2025)
por: Katzy, Jonathan, et al.
Publicado: (2025)
A Transformer-Based Approach for Smart Invocation of Automatic Code Completion
por: de Moor, Aral, et al.
Publicado: (2024)
por: de Moor, Aral, et al.
Publicado: (2024)
Traces of Memorisation in Large Language Models for Code
por: Al-Kaswan, Ali, et al.
Publicado: (2023)
por: Al-Kaswan, Ali, et al.
Publicado: (2023)
Automated Attention Pattern Discovery at Scale in Large Language Models
por: Katzy, Jonathan, et al.
Publicado: (2026)
por: Katzy, Jonathan, et al.
Publicado: (2026)
A Qualitative Investigation into LLM-Generated Multilingual Code Comments and Automatic Evaluation Metrics
por: Katzy, Jonathan, et al.
Publicado: (2025)
por: Katzy, Jonathan, et al.
Publicado: (2025)
TreeRanker: Fast and Model-agnostic Ranking System for Code Suggestions in IDEs
por: Cipollone, Daniele, et al.
Publicado: (2025)
por: Cipollone, Daniele, et al.
Publicado: (2025)
Evaluating Non-English Developer Support in Machine Learning for Software Engineering
por: Katzy, Jonathan, et al.
Publicado: (2026)
por: Katzy, Jonathan, et al.
Publicado: (2026)
Model See, Model Do? Exposure-Aware Evaluation of Bug-vs-Fix Preference in Code LLMs
por: Al-Kaswan, Ali, et al.
Publicado: (2026)
por: Al-Kaswan, Ali, et al.
Publicado: (2026)
Code Red! On the Harmfulness of Applying Off-the-shelf Large Language Models to Programming Tasks
por: Al-Kaswan, Ali, et al.
Publicado: (2025)
por: Al-Kaswan, Ali, et al.
Publicado: (2025)
Investigating Autonomous Agent Contributions in the Wild: Activity Patterns and Code Change over Time
por: Popescu, Razvan Mihai, et al.
Publicado: (2026)
por: Popescu, Razvan Mihai, et al.
Publicado: (2026)
AST-PAC: AST-guided Membership Inference for Code
por: Koohestani, Roham, et al.
Publicado: (2026)
por: Koohestani, Roham, et al.
Publicado: (2026)
Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture the Flag Challenges
por: Al-Kaswan, Ali, et al.
Publicado: (2026)
por: Al-Kaswan, Ali, et al.
Publicado: (2026)
Long Code Arena: a Set of Benchmarks for Long-Context Code Models
por: Bogomolov, Egor, et al.
Publicado: (2024)
por: Bogomolov, Egor, et al.
Publicado: (2024)
Neural Models for Source Code Synthesis and Completion
por: Niyogi, Mitodru
Publicado: (2024)
por: Niyogi, Mitodru
Publicado: (2024)
Data vs. Model Machine Learning Fairness Testing: An Empirical Study
por: Shome, Arumoy, et al.
Publicado: (2024)
por: Shome, Arumoy, et al.
Publicado: (2024)
SIMCOPILOT: Evaluating Large Language Models for Copilot-Style Code Generation
por: Jiang, Mingchao, et al.
Publicado: (2025)
por: Jiang, Mingchao, et al.
Publicado: (2025)
McUDI: Model-Centric Unsupervised Degradation Indicator for Failure Prediction AIOps Solutions
por: Poenaru-Olaru, Lorena, et al.
Publicado: (2024)
por: Poenaru-Olaru, Lorena, et al.
Publicado: (2024)
Evaluating Large Language Models for Functional and Maintainable Code in Industrial Settings: A Case Study at ASML
por: Mundhra, Yash, et al.
Publicado: (2025)
por: Mundhra, Yash, et al.
Publicado: (2025)
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks
por: Kadosh, Tal, et al.
Publicado: (2023)
por: Kadosh, Tal, et al.
Publicado: (2023)
Large Language Models for Multilingual Code Intelligence: A Survey
por: Jiang, Chao, et al.
Publicado: (2026)
por: Jiang, Chao, et al.
Publicado: (2026)
Do Large Code Models Understand Programming Concepts? Counterfactual Analysis for Code Predicates
por: Hooda, Ashish, et al.
Publicado: (2024)
por: Hooda, Ashish, et al.
Publicado: (2024)
PerfRL: A Small Language Model Framework for Efficient Code Optimization
por: Duan, Shukai, et al.
Publicado: (2023)
por: Duan, Shukai, et al.
Publicado: (2023)
SwiftEval: Developing a Language-Specific Benchmark for LLM-generated Code Evaluation
por: Petrukha, Ivan, et al.
Publicado: (2025)
por: Petrukha, Ivan, et al.
Publicado: (2025)
Towards Automatic Translation of Machine Learning Visual Insights to Analytical Assertions
por: Shome, Arumoy, et al.
Publicado: (2024)
por: Shome, Arumoy, et al.
Publicado: (2024)
NExT: Teaching Large Language Models to Reason about Code Execution
por: Ni, Ansong, et al.
Publicado: (2024)
por: Ni, Ansong, et al.
Publicado: (2024)
Investigating the Performance of Language Models for Completing Code in Functional Programming Languages: a Haskell Case Study
por: van Dam, Tim, et al.
Publicado: (2024)
por: van Dam, Tim, et al.
Publicado: (2024)
Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions
por: Cassano, Federico, et al.
Publicado: (2023)
por: Cassano, Federico, et al.
Publicado: (2023)
Constrained Decoding for Fill-in-the-Middle Code Language Models via Efficient Left and Right Quotienting of Context-Sensitive Grammars
por: Melcer, Daniel, et al.
Publicado: (2024)
por: Melcer, Daniel, et al.
Publicado: (2024)
Black-Box Adversarial Attacks on LLM-Based Code Completion
por: Jenko, Slobodan, et al.
Publicado: (2024)
por: Jenko, Slobodan, et al.
Publicado: (2024)
JavaBench: A Benchmark of Object-Oriented Code Generation for Evaluating Large Language Models
por: Cao, Jialun, et al.
Publicado: (2024)
por: Cao, Jialun, et al.
Publicado: (2024)
Large Language Models for Code Summarization
por: Szalontai, Balázs, et al.
Publicado: (2024)
por: Szalontai, Balázs, et al.
Publicado: (2024)
Automatically Testing Functional Properties of Code Translation Models
por: Eniser, Hasan Ferit, et al.
Publicado: (2023)
por: Eniser, Hasan Ferit, et al.
Publicado: (2023)
MLCPD: A Unified Multi-Language Code Parsing Dataset with Universal AST Schema
por: Gajjar, Jugal, et al.
Publicado: (2025)
por: Gajjar, Jugal, et al.
Publicado: (2025)
Unlocking the Power of Environment Assumptions for Unit Proofs
por: Priya, Siddharth, et al.
Publicado: (2024)
por: Priya, Siddharth, et al.
Publicado: (2024)
Is Your Anomaly Detector Ready for Change? Adapting AIOps Solutions to the Real World
por: Poenaru-Olaru, Lorena, et al.
Publicado: (2023)
por: Poenaru-Olaru, Lorena, et al.
Publicado: (2023)
Reflections on the design, applications and implementations of the normative specification language eFLINT
por: van Binsbergen, L. Thomas, et al.
Publicado: (2025)
por: van Binsbergen, L. Thomas, et al.
Publicado: (2025)
ChatDBG: Augmenting Debugging with Large Language Models
por: Levin, Kyla H., et al.
Publicado: (2024)
por: Levin, Kyla H., et al.
Publicado: (2024)
CodeMind: Evaluating Large Language Models for Code Reasoning
por: Liu, Changshu, et al.
Publicado: (2024)
por: Liu, Changshu, et al.
Publicado: (2024)
Large Language Models for Code: Security Hardening and Adversarial Testing
por: He, Jingxuan, et al.
Publicado: (2023)
por: He, Jingxuan, et al.
Publicado: (2023)
Ejemplares similares
-
An Exploratory Investigation into Code License Infringements in Large Language Model Training Datasets
por: Katzy, Jonathan, et al.
Publicado: (2024) -
The Heap: A Contamination-Free Multilingual Code Dataset for Evaluating Large Language Models
por: Katzy, Jonathan, et al.
Publicado: (2025) -
A Transformer-Based Approach for Smart Invocation of Automatic Code Completion
por: de Moor, Aral, et al.
Publicado: (2024) -
Traces of Memorisation in Large Language Models for Code
por: Al-Kaswan, Ali, et al.
Publicado: (2023) -
Automated Attention Pattern Discovery at Scale in Large Language Models
por: Katzy, Jonathan, et al.
Publicado: (2026)