Evaluation of large language models for assessing code maintainability
Fuente:
arXiv
Saved in:
| Main Authors: | Dillmann, Marc, Siebert, Julien, Trendowicz, Adam |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ESALE: Enhancing Code-Summary Alignment Learning for Source Code Summarization
by: Fang, Chunrong, et al.
Published: (2024)
by: Fang, Chunrong, et al.
Published: (2024)
Bug In the Code Stack: Can LLMs Find Bugs in Large Python Code Stacks
by: Lee, Hokyung, et al.
Published: (2024)
by: Lee, Hokyung, et al.
Published: (2024)
Source Code Summarization in the Era of Large Language Models
by: Sun, Weisong, et al.
Published: (2024)
by: Sun, Weisong, et al.
Published: (2024)
Knowledge-Guided Multi-Agent Framework for Automated Requirements Development: A Vision
by: Huang, Jiangping, et al.
Published: (2025)
by: Huang, Jiangping, et al.
Published: (2025)
Commenting Higher-level Code Unit: Full Code, Reduced Code, or Hierarchical Code Summarization
by: Sun, Weisong, et al.
Published: (2025)
by: Sun, Weisong, et al.
Published: (2025)
Learning Software Bug Reports: A Systematic Literature Review
by: Long, Guoming, et al.
Published: (2025)
by: Long, Guoming, et al.
Published: (2025)
ContractEval: A Benchmark for Evaluating Contract-Satisfying Assertions in Code Generation
by: Lim, Soohan, et al.
Published: (2025)
by: Lim, Soohan, et al.
Published: (2025)
GEML: A Grammar-based Evolutionary Machine Learning Approach for Design-Pattern Detection
by: Barbudo, Rafael, et al.
Published: (2024)
by: Barbudo, Rafael, et al.
Published: (2024)
Lost in Translation? Converting RegExes for Log Parsing into Dynatrace Pattern Language
by: Fragner, Julian, et al.
Published: (2025)
by: Fragner, Julian, et al.
Published: (2025)
Human-Aligned Enhancement of Programming Answers with LLMs Guided by User Feedback
by: Bappon, Suborno Deb, et al.
Published: (2026)
by: Bappon, Suborno Deb, et al.
Published: (2026)
Automated Code Review Using Large Language Models at Ericsson: An Experience Report
by: Ramesh, Shweta, et al.
Published: (2025)
by: Ramesh, Shweta, et al.
Published: (2025)
Beyond Greenfield: The D3 Framework for AI-Driven Productivity in Brownfield Engineering
by: Sharma, Krishna Kumaar
Published: (2025)
by: Sharma, Krishna Kumaar
Published: (2025)
TCProF: Time-Complexity Prediction SSL Framework
by: Hahn, Joonghyuk, et al.
Published: (2025)
by: Hahn, Joonghyuk, et al.
Published: (2025)
MEC$^3$O: Multi-Expert Consensus for Code Time Complexity Prediction
by: Hahn, Joonghyuk, et al.
Published: (2025)
by: Hahn, Joonghyuk, et al.
Published: (2025)
Benchmarking Energy Efficiency of Large Language Models Using vLLM
by: Pronk, K., et al.
Published: (2025)
by: Pronk, K., et al.
Published: (2025)
Show Me Your Code! Kill Code Poisoning: A Lightweight Method Based on Code Naturalness
by: Sun, Weisong, et al.
Published: (2025)
by: Sun, Weisong, et al.
Published: (2025)
Simple and Effective Baselines for Code Summarisation Evaluation
by: Robinson, Jade, et al.
Published: (2025)
by: Robinson, Jade, et al.
Published: (2025)
When Retrieval Hurts Code Completion: A Diagnostic Study of Stale Repository Context
by: Weng, Haojun, et al.
Published: (2026)
by: Weng, Haojun, et al.
Published: (2026)
Finding a Crab in the C: Assured Translation via Comparative Symbolic Execution
by: Helbling, Caleb, et al.
Published: (2026)
by: Helbling, Caleb, et al.
Published: (2026)
cozy: Comparative Symbolic Execution for Binary Programs
by: Helbling, Caleb, et al.
Published: (2025)
by: Helbling, Caleb, et al.
Published: (2025)
From Understanding to Excelling: Template-Free Algorithm Design through Structural-Functional Co-Evolution
by: Zhao, Zhe, et al.
Published: (2025)
by: Zhao, Zhe, et al.
Published: (2025)
The Impact of Large Language Models on Open-source Innovation: Evidence from GitHub Copilot
by: Yeverechyahu, Doron, et al.
Published: (2024)
by: Yeverechyahu, Doron, et al.
Published: (2024)
When Many-Shot Prompting Fails: An Empirical Study of LLM Code Translation
by: Oskooei, Amirkia Rafiei, et al.
Published: (2025)
by: Oskooei, Amirkia Rafiei, et al.
Published: (2025)
Leveraging Large Language Models for Use Case Model Generation from Software Requirements
by: Eisenreich, Tobias, et al.
Published: (2025)
by: Eisenreich, Tobias, et al.
Published: (2025)
Automating Domain-Driven Design: Experience with a Prompting Framework
by: Eisenreich, Tobias, et al.
Published: (2026)
by: Eisenreich, Tobias, et al.
Published: (2026)
Lore: Repurposing Git Commit Messages as a Structured Knowledge Protocol for AI Coding Agents
by: Stetsenko, Ivan
Published: (2026)
by: Stetsenko, Ivan
Published: (2026)
Variational Prefix Tuning for Diverse and Accurate Code Summarization Using Pre-trained Language Models
by: Zhao, Junda, et al.
Published: (2025)
by: Zhao, Junda, et al.
Published: (2025)
SLEAN: Simple Lightweight Ensemble Analysis Network for Multi-Provider LLM Coordination: Design, Implementation, and Vibe Coding Bug Investigation Case Study
by: Vargas, Matheus J. T.
Published: (2025)
by: Vargas, Matheus J. T.
Published: (2025)
Migrating Esope to Fortran 2008 using model transformations
by: Sow, Younoussa, et al.
Published: (2026)
by: Sow, Younoussa, et al.
Published: (2026)
A Story About Cohesion and Separation: Label-Free Metric for Log Parser Evaluation
by: Qin, Qiaolin, et al.
Published: (2025)
by: Qin, Qiaolin, et al.
Published: (2025)
Distilling Desired Comments for Enhanced Code Review with Large Language Models
by: Yu, Yongda, et al.
Published: (2024)
by: Yu, Yongda, et al.
Published: (2024)
From Untestable to Testable: Metamorphic Testing in the Age of LLMs
by: Terragni, Valerio
Published: (2026)
by: Terragni, Valerio
Published: (2026)
Assessing Data Augmentation-Induced Bias in Training and Testing of Machine Learning Models
by: More, Riddhi, et al.
Published: (2025)
by: More, Riddhi, et al.
Published: (2025)
An Analysis of LLM Fine-Tuning and Few-Shot Learning for Flaky Test Detection and Classification
by: More, Riddhi, et al.
Published: (2025)
by: More, Riddhi, et al.
Published: (2025)
Fine-Tuning LLMs to Analyze Multiple Dimensions of Code Review: A Maximum Entropy Regulated Long Chain-of-Thought Approach
by: Yu, Yongda, et al.
Published: (2025)
by: Yu, Yongda, et al.
Published: (2025)
Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study
by: Alshaikh, Moaath, et al.
Published: (2026)
by: Alshaikh, Moaath, et al.
Published: (2026)
A Systematic Approach for Assessing Large Language Models' Test Case Generation Capability
by: Chang, Hung-Fu, et al.
Published: (2025)
by: Chang, Hung-Fu, et al.
Published: (2025)
Is It Time To Treat Prompts As Code? A Multi-Use Case Study For Prompt Optimization Using DSPy
by: Lemos, Francisca, et al.
Published: (2025)
by: Lemos, Francisca, et al.
Published: (2025)
XPath Agent: An Efficient XPath Programming Agent Based on LLM for Web Crawler
by: Li, Yu, et al.
Published: (2024)
by: Li, Yu, et al.
Published: (2024)
CWM: An Open-Weights LLM for Research on Code Generation with World Models
by: FAIR CodeGen team, et al.
Published: (2025)
by: FAIR CodeGen team, et al.
Published: (2025)
Similar Items
-
ESALE: Enhancing Code-Summary Alignment Learning for Source Code Summarization
by: Fang, Chunrong, et al.
Published: (2024) -
Bug In the Code Stack: Can LLMs Find Bugs in Large Python Code Stacks
by: Lee, Hokyung, et al.
Published: (2024) -
Source Code Summarization in the Era of Large Language Models
by: Sun, Weisong, et al.
Published: (2024) -
Knowledge-Guided Multi-Agent Framework for Automated Requirements Development: A Vision
by: Huang, Jiangping, et al.
Published: (2025) -
Commenting Higher-level Code Unit: Full Code, Reduced Code, or Hierarchical Code Summarization
by: Sun, Weisong, et al.
Published: (2025)