Can ChatGPT replace StackOverflow? A Study on Robustness and Reliability of Large Language Model Code Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhong, Li, Wang, Zilong |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Privacy Questions From Stack Overflow: Can ChatGPT Compete?
by: Delile, Zack, et al.
Published: (2023)
by: Delile, Zack, et al.
Published: (2023)
Just another copy and paste? Comparing the security vulnerabilities of ChatGPT generated code and StackOverflow answers
by: Hamer, Sivana, et al.
Published: (2024)
by: Hamer, Sivana, et al.
Published: (2024)
Is Stack Overflow Obsolete? An Empirical Study of the Characteristics of ChatGPT Answers to Stack Overflow Questions
by: Kabir, Samia, et al.
Published: (2023)
by: Kabir, Samia, et al.
Published: (2023)
Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
by: Liu, Yi, et al.
Published: (2023)
by: Liu, Yi, et al.
Published: (2023)
SOGPTSpotter: Detecting ChatGPT-Generated Answers on Stack Overflow
by: Ma, Suyu, et al.
Published: (2026)
by: Ma, Suyu, et al.
Published: (2026)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Safety Analysis in the Era of Large Language Models: A Case Study of STPA using ChatGPT
by: Qi, Yi, et al.
Published: (2023)
by: Qi, Yi, et al.
Published: (2023)
SOSecure: Safer Code Generation with RAG and StackOverflow Discussions
by: Mukherjee, Manisha, et al.
Published: (2025)
by: Mukherjee, Manisha, et al.
Published: (2025)
Can ChatGPT Support Developers? An Empirical Evaluation of Large Language Models for Code Generation
by: Jin, Kailun, et al.
Published: (2024)
by: Jin, Kailun, et al.
Published: (2024)
LLMs and Stack Overflow Discussions: Reliability, Impact, and Challenges
by: Da Silva, Leuson, et al.
Published: (2024)
by: Da Silva, Leuson, et al.
Published: (2024)
ChatGPT vs. DeepSeek: A Comparative Study on AI-Based Code Generation
by: Manik, Md Motaleb Hossen
Published: (2025)
by: Manik, Md Motaleb Hossen
Published: (2025)
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
by: Zhong, Li, et al.
Published: (2024)
by: Zhong, Li, et al.
Published: (2024)
Exploring ChatGPT's Capabilities on Vulnerability Management
by: Liu, Peiyu, et al.
Published: (2023)
by: Liu, Peiyu, et al.
Published: (2023)
Beyond Code Generation: An Observational Study of ChatGPT Usage in Software Engineering Practice
by: Khojah, Ranim, et al.
Published: (2024)
by: Khojah, Ranim, et al.
Published: (2024)
Evaluation of the Code Generation Capabilities of ChatGPT 4: A Comparative Analysis in 19 Programming Languages
by: Gilbert, L. C.
Published: (2025)
by: Gilbert, L. C.
Published: (2025)
Can OpenSource beat ChatGPT? -- A Comparative Study of Large Language Models for Text-to-Code Generation
by: Mayer, Luis, et al.
Published: (2024)
by: Mayer, Luis, et al.
Published: (2024)
Unmasking the giant: A comprehensive evaluation of ChatGPT's proficiency in coding algorithms and data structures
by: Arefin, Sayed Erfan, et al.
Published: (2023)
by: Arefin, Sayed Erfan, et al.
Published: (2023)
Can ChatGPT support software verification?
by: Janßen, Christian, et al.
Published: (2023)
by: Janßen, Christian, et al.
Published: (2023)
Do Prompt Patterns Affect Code Quality? A First Empirical Assessment of ChatGPT-Generated Code
by: Della Porta, Antonio, et al.
Published: (2025)
by: Della Porta, Antonio, et al.
Published: (2025)
Experimental Analysis of Productive Interaction Strategy with ChatGPT: User Study on Function and Project-level Code Generation Tasks
by: Hyun, Sangwon, et al.
Published: (2025)
by: Hyun, Sangwon, et al.
Published: (2025)
Is ChatGPT a Good Software Librarian? An Exploratory Study on the Use of ChatGPT for Software Library Recommendations
by: Latendresse, Jasmine, et al.
Published: (2024)
by: Latendresse, Jasmine, et al.
Published: (2024)
An Empirical Study of OpenAI API Discussions on Stack Overflow
by: Chen, Xiang, et al.
Published: (2025)
by: Chen, Xiang, et al.
Published: (2025)
How Robust are LLM-Generated Library Imports? An Empirical Study using Stack Overflow
by: Latendresse, Jasmine, et al.
Published: (2025)
by: Latendresse, Jasmine, et al.
Published: (2025)
CodeMirage: Hallucinations in Code Generated by Large Language Models
by: Agarwal, Vibhor, et al.
Published: (2024)
by: Agarwal, Vibhor, et al.
Published: (2024)
An Empirical Exploration of ChatGPT's Ability to Support Problem Formulation Tasks for Mission Engineering and a Documentation of its Performance Variability
by: Ofsa, Max, et al.
Published: (2025)
by: Ofsa, Max, et al.
Published: (2025)
Evaluating ChatGPT-3.5 Efficiency in Solving Coding Problems of Different Complexity Levels: An Empirical Analysis
by: Li, Minda, et al.
Published: (2024)
by: Li, Minda, et al.
Published: (2024)
Programming with AI: Evaluating ChatGPT, Gemini, AlphaCode, and GitHub Copilot for Programmers
by: Siam, Md Kamrul, et al.
Published: (2024)
by: Siam, Md Kamrul, et al.
Published: (2024)
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
by: Wang, Peiding, et al.
Published: (2025)
by: Wang, Peiding, et al.
Published: (2025)
A Survey on Large Language Models for Code Generation
by: Jiang, Juyong, et al.
Published: (2024)
by: Jiang, Juyong, et al.
Published: (2024)
Exploring Challenges in Test Mocking: Developer Questions and Insights from StackOverflow
by: Ahmed, Mumtahina, et al.
Published: (2025)
by: Ahmed, Mumtahina, et al.
Published: (2025)
Exploring Data-Efficient Adaptation of Large Language Models for Code Generation
by: Jiang, Xue, et al.
Published: (2024)
by: Jiang, Xue, et al.
Published: (2024)
ArchCode: Incorporating Software Requirements in Code Generation with Large Language Models
by: Han, Hojae, et al.
Published: (2024)
by: Han, Hojae, et al.
Published: (2024)
WIP: Assessing the Effectiveness of ChatGPT in Preparatory Testing Activities
by: Haldar, Susmita, et al.
Published: (2025)
by: Haldar, Susmita, et al.
Published: (2025)
GeoCode-GPT: A Large Language Model for Geospatial Code Generation Tasks
by: Hou, Shuyang, et al.
Published: (2024)
by: Hou, Shuyang, et al.
Published: (2024)
Can LLMs Generate Reliable Test Case Generators? A Study on Competition-Level Programming Problems
by: Cao, Yuhan, et al.
Published: (2025)
by: Cao, Yuhan, et al.
Published: (2025)
How Diversely Can Language Models Solve Problems? Exploring the Algorithmic Diversity of Model-Generated Code
by: Lee, Seonghyeon, et al.
Published: (2025)
by: Lee, Seonghyeon, et al.
Published: (2025)
Benchmarking Large Language Models for ABAP Code Generation: An Empirical Study on Iterative Improvement by Compiler Feedback
by: Wallraven, Stephan, et al.
Published: (2026)
by: Wallraven, Stephan, et al.
Published: (2026)
Revisiting the Reliability of Language Models in Instruction-Following
by: Dong, Jianshuo, et al.
Published: (2025)
by: Dong, Jianshuo, et al.
Published: (2025)
Self-Explained Keywords Empower Large Language Models for Code Generation
by: Fan, Lishui, et al.
Published: (2024)
by: Fan, Lishui, et al.
Published: (2024)
How Do Java Developers Reuse StackOverflow Answers in Their GitHub Projects?
by: Chen, Juntong, et al.
Published: (2023)
by: Chen, Juntong, et al.
Published: (2023)
Similar Items
-
Evaluating Privacy Questions From Stack Overflow: Can ChatGPT Compete?
by: Delile, Zack, et al.
Published: (2023) -
Just another copy and paste? Comparing the security vulnerabilities of ChatGPT generated code and StackOverflow answers
by: Hamer, Sivana, et al.
Published: (2024) -
Is Stack Overflow Obsolete? An Empirical Study of the Characteristics of ChatGPT Answers to Stack Overflow Questions
by: Kabir, Samia, et al.
Published: (2023) -
Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
by: Liu, Yi, et al.
Published: (2023) -
SOGPTSpotter: Detecting ChatGPT-Generated Answers on Stack Overflow
by: Ma, Suyu, et al.
Published: (2026)