How Propense Are Large Language Models at Producing Code Smells? A Benchmarking Study
Fuente:
arXiv
Saved in:
| Main Authors: | Velasco, Alejandro, Rodriguez-Cardenas, Daniel, Alif, Luftar Rahman, Palacio, David N., Poshyvanyk, Denys |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Causal Perspective on Measuring, Explaining and Mitigating Smells in LLM-Generated Code
by: Velasco, Alejandro, et al.
Published: (2025)
by: Velasco, Alejandro, et al.
Published: (2025)
Which Syntactic Capabilities Are Statistically Learned by Masked Language Models for Code?
by: Velasco, Alejandro, et al.
Published: (2024)
by: Velasco, Alejandro, et al.
Published: (2024)
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We?
by: O'Brien, Conor, et al.
Published: (2024)
by: O'Brien, Conor, et al.
Published: (2024)
SnipGen: A Mining Repository Framework for Evaluating LLMs for Code
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2025)
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2025)
Towards Enabling An Artificial Self-Construction Software Life-cycle via Autopoietic Architectures
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation
by: Khati, Dipin, et al.
Published: (2025)
by: Khati, Dipin, et al.
Published: (2025)
Towards More Trustworthy and Interpretable LLMs for Code through Syntax-Grounded Explanations
by: Palacio, David N., et al.
Published: (2024)
by: Palacio, David N., et al.
Published: (2024)
On Interpreting the Effectiveness of Unsupervised Software Traceability with Information Theory
by: Palacio, David N., et al.
Published: (2024)
by: Palacio, David N., et al.
Published: (2024)
Toward a Theory of Causation for Interpreting Neural Code Models
by: Palacio, David N., et al.
Published: (2023)
by: Palacio, David N., et al.
Published: (2023)
Toward Neurosymbolic Program Comprehension
by: Velasco, Alejandro, et al.
Published: (2025)
by: Velasco, Alejandro, et al.
Published: (2025)
Rethinking Software Empirical Studies with Structural Causal Models
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
Detecting and Correcting Hallucinations in LLM-Generated Code via Deterministic AST Analysis
by: Khati, Dipin, et al.
Published: (2026)
by: Khati, Dipin, et al.
Published: (2026)
Tricky$^2$: Towards a Benchmark for Evaluating Human and LLM Error Interactions
by: Granger, Cole, et al.
Published: (2026)
by: Granger, Cole, et al.
Published: (2026)
On the Effectiveness of LLM-as-a-judge for Code Generation and Summarization
by: Crupi, Giuseppe, et al.
Published: (2025)
by: Crupi, Giuseppe, et al.
Published: (2025)
Understanding Privacy Risks in Code Models Through Training Dynamics: A Causal Approach
by: Yang, Hua, et al.
Published: (2025)
by: Yang, Hua, et al.
Published: (2025)
On the Generalizability of Transformer Models to Code Completions of Different Lengths
by: Cooper, Nathan, et al.
Published: (2025)
by: Cooper, Nathan, et al.
Published: (2025)
Towards More Trustworthy Deep Code Models by Enabling Out-of-Distribution Detection
by: Yan, Yanfu, et al.
Published: (2025)
by: Yan, Yanfu, et al.
Published: (2025)
A Path Less Traveled: Reimagining Software Engineering Automation via a Neurosymbolic Paradigm
by: Mastropaolo, Antonio, et al.
Published: (2025)
by: Mastropaolo, Antonio, et al.
Published: (2025)
How Do Semantically Equivalent Code Transformations Impact Membership Inference on LLMs for Code?
by: Yang, Hua, et al.
Published: (2025)
by: Yang, Hua, et al.
Published: (2025)
Mapping the Trust Terrain: LLMs in Software Engineering -- Insights and Perspectives
by: Khati, Dipin, et al.
Published: (2025)
by: Khati, Dipin, et al.
Published: (2025)
On the Prevalence, Evolution, and Impact of Code Smells in Simulation Modelling Software
by: Mahbub, Riasat, et al.
Published: (2024)
by: Mahbub, Riasat, et al.
Published: (2024)
Testing Practices, Challenges, and Developer Perspectives in Open-Source IoT Platforms
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2025)
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2025)
Towards Comprehensive Benchmarking Infrastructure for LLMs In Software Engineering
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
Beyond Strict Rules: Assessing the Effectiveness of Large Language Models for Code Smell Detection
by: Souza, Saymon, et al.
Published: (2026)
by: Souza, Saymon, et al.
Published: (2026)
Challenges and Practices in Quantum Software Testing and Debugging: Insights from Practitioners
by: Zappin, Jake, et al.
Published: (2025)
by: Zappin, Jake, et al.
Published: (2025)
When Quantum Meets Classical: Characterizing Hybrid Quantum-Classical Issues Discussed in Developer Forums
by: Zappin, Jake, et al.
Published: (2024)
by: Zappin, Jake, et al.
Published: (2024)
Bridging the Quantum Divide: Aligning Academic and Industry Goals in Software Engineering
by: Zappin, Jake, et al.
Published: (2025)
by: Zappin, Jake, et al.
Published: (2025)
Perspective of Software Engineering Researchers on Machine Learning Practices Regarding Research, Review, and Education
by: Mojica-Hanke, Anamaria, et al.
Published: (2024)
by: Mojica-Hanke, Anamaria, et al.
Published: (2024)
Prompting in Practice: Investigating Software Practitioners' Use of Generative AI Tools
by: Otten, Daniel, et al.
Published: (2025)
by: Otten, Daniel, et al.
Published: (2025)
Evaluating Large Language Models in Detecting Test Smells
by: Lucas, Keila, et al.
Published: (2024)
by: Lucas, Keila, et al.
Published: (2024)
BOMs Away! Inside the Minds of Stakeholders: A Comprehensive Study of Bills of Materials for Software Systems
by: Stalnaker, Trevor, et al.
Published: (2023)
by: Stalnaker, Trevor, et al.
Published: (2023)
Toward Explaining Large Language Models in Software Engineering Tasks
by: Vitale, Antonio, et al.
Published: (2025)
by: Vitale, Antonio, et al.
Published: (2025)
How Do Communities of ML-Enabled Systems Smell? A Cross-Sectional Study on the Prevalence of Community Smells
by: Annunziata, Giusy, et al.
Published: (2025)
by: Annunziata, Giusy, et al.
Published: (2025)
What Breaks When LLMs Code? Characterizing Operational Safety Failures of Agentic Code Assistants
by: Hasan, Alif Al, et al.
Published: (2026)
by: Hasan, Alif Al, et al.
Published: (2026)
An Event-Driven Tool for Context-Aware Code Smell Detection Using SmellDSL
by: Viegas, Matheus dos Santos, et al.
Published: (2026)
by: Viegas, Matheus dos Santos, et al.
Published: (2026)
"The Law Doesn't Work Like a Computer": Exploring Software Licensing Issues Faced by Legal Practitioners
by: Wintersgill, Nathan, et al.
Published: (2024)
by: Wintersgill, Nathan, et al.
Published: (2024)
"False negative -- that one is going to kill you": Understanding Industry Perspectives of Static Analysis based Security Testing
by: Ami, Amit Seal, et al.
Published: (2023)
by: Ami, Amit Seal, et al.
Published: (2023)
Smells Depend on the Context: An Interview Study of Issue Tracking Problems and Smells in Practice
by: Montgomery, Lloyd, et al.
Published: (2026)
by: Montgomery, Lloyd, et al.
Published: (2026)
Empirical Characterization of Logging Smells in Machine Learning Code
by: Foalem, Patrick Loic, et al.
Published: (2026)
by: Foalem, Patrick Loic, et al.
Published: (2026)
Towards Automated Detection of Inline Code Comment Smells
by: Oztas, Ipek, et al.
Published: (2025)
by: Oztas, Ipek, et al.
Published: (2025)
Similar Items
-
A Causal Perspective on Measuring, Explaining and Mitigating Smells in LLM-Generated Code
by: Velasco, Alejandro, et al.
Published: (2025) -
Which Syntactic Capabilities Are Statistically Learned by Masked Language Models for Code?
by: Velasco, Alejandro, et al.
Published: (2024) -
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We?
by: O'Brien, Conor, et al.
Published: (2024) -
SnipGen: A Mining Repository Framework for Evaluating LLMs for Code
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2025) -
Towards Enabling An Artificial Self-Construction Software Life-cycle via Autopoietic Architectures
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)