Analyzing Prominent LLMs: An Empirical Study of Performance and Complexity in Solving LeetCode Problems
Fuente:
arXiv
Saved in:
| Main Authors: | Guimaraes, Everton, Nascimento, Nathalia, Shivalingaiah, Chandan, Nelapati, Asish |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A case study on the transformative potential of AI in software engineering on LeetCode and ChatGPT
by: Merkel, Manuel, et al.
Published: (2025)
by: Merkel, Manuel, et al.
Published: (2025)
Designing Empirical Studies on LLM-Based Code Generation: Towards a Reference Framework
by: Nascimento, Nathalia, et al.
Published: (2025)
by: Nascimento, Nathalia, et al.
Published: (2025)
LLM4DS: Evaluating Large Language Models for Data Science Code Generation
by: Nascimento, Nathalia, et al.
Published: (2024)
by: Nascimento, Nathalia, et al.
Published: (2024)
LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
by: Xia, Yunhui, et al.
Published: (2025)
by: Xia, Yunhui, et al.
Published: (2025)
Evaluating ChatGPT-3.5 Efficiency in Solving Coding Problems of Different Complexity Levels: An Empirical Analysis
by: Li, Minda, et al.
Published: (2024)
by: Li, Minda, et al.
Published: (2024)
AI builds, We Analyze: An Empirical Study of AI-Generated Build Code Quality
by: Ghammam, Anwar, et al.
Published: (2026)
by: Ghammam, Anwar, et al.
Published: (2026)
LLMs are Bug Replicators: An Empirical Study on LLMs' Capability in Completing Bug-prone Code
by: Guo, Liwei, et al.
Published: (2025)
by: Guo, Liwei, et al.
Published: (2025)
An Empirical Study of False Negatives and Positives of Static Code Analyzers From the Perspective of Historical Issues
by: Cui, Han, et al.
Published: (2024)
by: Cui, Han, et al.
Published: (2024)
Understanding Specification-Driven Code Generation with LLMs: An Empirical Study Design
by: Rosa, Giovanni, et al.
Published: (2026)
by: Rosa, Giovanni, et al.
Published: (2026)
Are "Solved Issues" in SWE-bench Really Solved Correctly? An Empirical Study
by: Wang, You, et al.
Published: (2025)
by: Wang, You, et al.
Published: (2025)
Isolating Language-Coding from Problem-Solving: Benchmarking LLMs with PseudoEval
by: Wu, Jiarong, et al.
Published: (2025)
by: Wu, Jiarong, et al.
Published: (2025)
Toward Interactive Optimization of Source Code Differences: An Empirical Study of Its Performance
by: Yagi, Tsukasa, et al.
Published: (2024)
by: Yagi, Tsukasa, et al.
Published: (2024)
Prompt Engineering or Fine-Tuning: An Empirical Assessment of LLMs for Code
by: Shin, Jiho, et al.
Published: (2023)
by: Shin, Jiho, et al.
Published: (2023)
On the Use of Agentic Coding Manifests: An Empirical Study of Claude Code
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
Variability-Aware Machine Learning Model Selection: Feature Modeling, Instantiation, and Experimental Case Study
by: Tavares, Cristina, et al.
Published: (2024)
by: Tavares, Cristina, et al.
Published: (2024)
How Do Agents Perform Code Optimization? An Empirical Study
by: Peng, Huiyun, et al.
Published: (2025)
by: Peng, Huiyun, et al.
Published: (2025)
An Empirical Study of Perceptions of General LLMs and Multimodal LLMs on Hugging Face
by: Liu, Yujian, et al.
Published: (2026)
by: Liu, Yujian, et al.
Published: (2026)
An Empirical Study of LLM-Based Code Clone Detection
by: Zhu, Wenqing, et al.
Published: (2025)
by: Zhu, Wenqing, et al.
Published: (2025)
Explaining Explanation: An Empirical Study on Explanation in Code Reviews
by: Widyasari, Ratnadira, et al.
Published: (2023)
by: Widyasari, Ratnadira, et al.
Published: (2023)
How Does Cognitive Capability and Personality Influence Problem Solving in Coding Interview Puzzles?
by: Hidellaarachchi, Dulaji, et al.
Published: (2025)
by: Hidellaarachchi, Dulaji, et al.
Published: (2025)
An Empirical Study on the Capability of LLMs in Decomposing Bug Reports
by: Chen, Zhiyuan, et al.
Published: (2025)
by: Chen, Zhiyuan, et al.
Published: (2025)
Towards Evaluation Guidelines for Empirical Studies involving LLMs
by: Wagner, Stefan, et al.
Published: (2024)
by: Wagner, Stefan, et al.
Published: (2024)
An Empirical Study on the Potential of LLMs in Automated Software Refactoring
by: Liu, Bo, et al.
Published: (2024)
by: Liu, Bo, et al.
Published: (2024)
Can LLMs Solve Science or Just Write Code? Evaluating Quantum Solver Generation
by: Baresi, Luciano, et al.
Published: (2026)
by: Baresi, Luciano, et al.
Published: (2026)
From Industry Claims to Empirical Reality: An Empirical Study of Code Review Agents in Pull Requests
by: Chowdhury, Kowshik, et al.
Published: (2026)
by: Chowdhury, Kowshik, et al.
Published: (2026)
Code vs Serialized AST Inputs for LLM-Based Code Summarization: An Empirical Study
by: Dong, Shijia, et al.
Published: (2026)
by: Dong, Shijia, et al.
Published: (2026)
Toward Effective Secure Code Reviews: An Empirical Study of Security-Related Coding Weaknesses
by: Charoenwet, Wachiraphan, et al.
Published: (2023)
by: Charoenwet, Wachiraphan, et al.
Published: (2023)
Rethinking Code Review Workflows with LLM Assistance: An Empirical Study
by: Aðalsteinsson, Fannar Steinn, et al.
Published: (2025)
by: Aðalsteinsson, Fannar Steinn, et al.
Published: (2025)
Agent READMEs: An Empirical Study of Context Files for Agentic Coding
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
An Empirical Study of Retrieval-Augmented Code Generation: Challenges and Opportunities
by: Yang, Zezhou, et al.
Published: (2025)
by: Yang, Zezhou, et al.
Published: (2025)
An Empirical Study of the Non-determinism of ChatGPT in Code Generation
by: Ouyang, Shuyin, et al.
Published: (2023)
by: Ouyang, Shuyin, et al.
Published: (2023)
Are Coding Agents Generating Over-Mocked Tests? An Empirical Study
by: Hora, Andre, et al.
Published: (2026)
by: Hora, Andre, et al.
Published: (2026)
An Empirical Study of Static Analysis Tools for Secure Code Review
by: Charoenwet, Wachiraphan, et al.
Published: (2024)
by: Charoenwet, Wachiraphan, et al.
Published: (2024)
An Empirical Study on the Code Refactoring Capability of Large Language Models
by: Cordeiro, Jonathan, et al.
Published: (2024)
by: Cordeiro, Jonathan, et al.
Published: (2024)
Changes in Coding Behavior and Performance Since the Introduction of LLMs
by: Zhang, Yufan, et al.
Published: (2026)
by: Zhang, Yufan, et al.
Published: (2026)
Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs
by: Kim, Jinhan, et al.
Published: (2025)
by: Kim, Jinhan, et al.
Published: (2025)
Exploring the Effectiveness of LLMs in Automated Logging Generation: An Empirical Study
by: Li, Yichen, et al.
Published: (2023)
by: Li, Yichen, et al.
Published: (2023)
An Empirical Study on the Performance and Energy Usage of Compiled Python Code
by: Stoico, Vincenzo, et al.
Published: (2025)
by: Stoico, Vincenzo, et al.
Published: (2025)
An Empirical Study: MEMS as a Static Performance Metric
by: Zhang, Liwei, et al.
Published: (2025)
by: Zhang, Liwei, et al.
Published: (2025)
On the Use of Agentic Coding: An Empirical Study of Pull Requests on GitHub
by: Watanabe, Miku, et al.
Published: (2025)
by: Watanabe, Miku, et al.
Published: (2025)
Similar Items
-
A case study on the transformative potential of AI in software engineering on LeetCode and ChatGPT
by: Merkel, Manuel, et al.
Published: (2025) -
Designing Empirical Studies on LLM-Based Code Generation: Towards a Reference Framework
by: Nascimento, Nathalia, et al.
Published: (2025) -
LLM4DS: Evaluating Large Language Models for Data Science Code Generation
by: Nascimento, Nathalia, et al.
Published: (2024) -
LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
by: Xia, Yunhui, et al.
Published: (2025) -
Evaluating ChatGPT-3.5 Efficiency in Solving Coding Problems of Different Complexity Levels: An Empirical Analysis
by: Li, Minda, et al.
Published: (2024)