How Robust are LLM-Generated Library Imports? An Empirical Study using Stack Overflow
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Latendresse, Jasmine, Khatoonabadi, SayedHassan, Shihab, Emad |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Is ChatGPT a Good Software Librarian? An Exploratory Study on the Use of ChatGPT for Software Library Recommendations
par: Latendresse, Jasmine, et autres
Publié: (2024)
par: Latendresse, Jasmine, et autres
Publié: (2024)
Beyond Synthetic Benchmarks: Evaluating LLM Performance on Real-World Class-Level Code Generation
par: Rahman, Musfiqur, et autres
Publié: (2025)
par: Rahman, Musfiqur, et autres
Publié: (2025)
OpenClassGen: A Large-Scale Corpus of Real-World Python Classes for LLM Research
par: Rahman, Musfiqur, et autres
Publié: (2025)
par: Rahman, Musfiqur, et autres
Publié: (2025)
Tokenomics: Quantifying Where Tokens Are Used in Agentic Software Engineering
par: Salim, Mohamad, et autres
Publié: (2026)
par: Salim, Mohamad, et autres
Publié: (2026)
Automatic Detection of LLM-Generated Code: A Comparative Case Study of Contemporary Models Across Function and Class Granularities
par: Rahman, Musfiqur, et autres
Publié: (2024)
par: Rahman, Musfiqur, et autres
Publié: (2024)
Automated File-Level Logging Generation for Machine Learning Applications using LLMs: A Case Study using GPT-4o Mini
par: Rodriguez, Mayra Sofia Ruiz, et autres
Publié: (2025)
par: Rodriguez, Mayra Sofia Ruiz, et autres
Publié: (2025)
Evaluating the Use of LLMs for Documentation to Code Traceability
par: Alor, Ebube, et autres
Publié: (2025)
par: Alor, Ebube, et autres
Publié: (2025)
Understanding the Helpfulness of Stale Bot for Pull-based Development: An Empirical Study of 20 Large Open-Source Projects
par: Khatoonabadi, SayedHassan, et autres
Publié: (2023)
par: Khatoonabadi, SayedHassan, et autres
Publié: (2023)
Synergizing LLMs and Knowledge Graphs: A Novel Approach to Software Repository-Related Question Answering
par: Abedu, Samuel, et autres
Publié: (2024)
par: Abedu, Samuel, et autres
Publié: (2024)
An Approach for Auto Generation of Labeling Functions for Software Engineering Chatbots
par: Alor, Ebube, et autres
Publié: (2024)
par: Alor, Ebube, et autres
Publié: (2024)
On Wasted Contributions: Understanding the Dynamics of Contributor-Abandoned Pull Requests
par: Khatoonabadi, SayedHassan, et autres
Publié: (2021)
par: Khatoonabadi, SayedHassan, et autres
Publié: (2021)
Predicting the First Response Latency of Maintainers and Contributors in Pull Requests
par: Khatoonabadi, SayedHassan, et autres
Publié: (2023)
par: Khatoonabadi, SayedHassan, et autres
Publié: (2023)
Evaluating the Use of LLMs for Automated DOM-Level Resolution of Web Performance Issues
par: Peters, Gideon, et autres
Publié: (2026)
par: Peters, Gideon, et autres
Publié: (2026)
The Impact of Environment Configurations on the Stability of AI-Enabled Systems
par: Rahman, Musfiqur, et autres
Publié: (2024)
par: Rahman, Musfiqur, et autres
Publié: (2024)
The Impact of Large Language Models (LLMs) on Code Review Process
par: Collante, Antonio, et autres
Publié: (2025)
par: Collante, Antonio, et autres
Publié: (2025)
Is Stack Overflow Obsolete? An Empirical Study of the Characteristics of ChatGPT Answers to Stack Overflow Questions
par: Kabir, Samia, et autres
Publié: (2023)
par: Kabir, Samia, et autres
Publié: (2023)
An Empirical Study of OpenAI API Discussions on Stack Overflow
par: Chen, Xiang, et autres
Publié: (2025)
par: Chen, Xiang, et autres
Publié: (2025)
Will It Survive? Deciphering the Fate of AI-Generated Code in Open Source
par: Rahman, Musfiqur, et autres
Publié: (2026)
par: Rahman, Musfiqur, et autres
Publié: (2026)
Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild
par: Zhao, Zhimin, et autres
Publié: (2026)
par: Zhao, Zhimin, et autres
Publié: (2026)
On the Impact of Black-box Deployment Strategies for Edge AI on Latency and Model Performance
par: Singh, Jaskirat, et autres
Publié: (2024)
par: Singh, Jaskirat, et autres
Publié: (2024)
PCBSchemaGen: Constraint-Guided Schematic Design via LLM for Printed Circuit Boards (PCB)
par: Zou, Huanghaohe, et autres
Publié: (2026)
par: Zou, Huanghaohe, et autres
Publié: (2026)
LLMs and Stack Overflow Discussions: Reliability, Impact, and Challenges
par: Da Silva, Leuson, et autres
Publié: (2024)
par: Da Silva, Leuson, et autres
Publié: (2024)
How Efficient is LLM-Generated Code? A Rigorous & High-Standard Benchmark
par: Qiu, Ruizhong, et autres
Publié: (2024)
par: Qiu, Ruizhong, et autres
Publié: (2024)
Parameter-Efficient Fine-Tuning of Large Language Models for Unit Test Generation: An Empirical Study
par: Storhaug, André, et autres
Publié: (2024)
par: Storhaug, André, et autres
Publié: (2024)
An Empirical Study of Fault Localisation Techniques for Deep Learning
par: Humbatova, Nargiz, et autres
Publié: (2024)
par: Humbatova, Nargiz, et autres
Publié: (2024)
On the Impact of Code Comments for Automated Bug-Fixing: An Empirical Study
par: Vitale, Antonio, et autres
Publié: (2026)
par: Vitale, Antonio, et autres
Publié: (2026)
How Robustly do LLMs Understand Execution Semantics?
par: Spiess, Claudio, et autres
Publié: (2026)
par: Spiess, Claudio, et autres
Publié: (2026)
Can ChatGPT replace StackOverflow? A Study on Robustness and Reliability of Large Language Model Code Generation
par: Zhong, Li, et autres
Publié: (2023)
par: Zhong, Li, et autres
Publié: (2023)
Can LLMs Generate Architectural Design Decisions? -An Exploratory Empirical study
par: Dhar, Rudra, et autres
Publié: (2024)
par: Dhar, Rudra, et autres
Publié: (2024)
More with Less: An Empirical Study of Turn-Control Strategies for Efficient Coding Agents
par: Gao, Pengfei, et autres
Publié: (2025)
par: Gao, Pengfei, et autres
Publié: (2025)
Operational Robustness of LLMs on Code Generation
par: Paul, Debalina Ghosh, et autres
Publié: (2026)
par: Paul, Debalina Ghosh, et autres
Publié: (2026)
Stack Trace Deduplication: Faster, More Accurately, and in More Realistic Scenarios
par: Shibaev, Egor, et autres
Publié: (2024)
par: Shibaev, Egor, et autres
Publié: (2024)
Order Matters! An Empirical Study on Large Language Models' Input Order Bias in Software Fault Localization
par: Rafi, Md Nakhla, et autres
Publié: (2024)
par: Rafi, Md Nakhla, et autres
Publié: (2024)
Understanding LLM-Driven Test Oracle Generation
par: Bodicoat, Adam, et autres
Publié: (2026)
par: Bodicoat, Adam, et autres
Publié: (2026)
Mutation-Guided LLM-based Test Generation at Meta
par: Foster, Christopher, et autres
Publié: (2025)
par: Foster, Christopher, et autres
Publié: (2025)
Enhancing LLM-Based Test Generation by Eliminating Covered Code
par: Xu, WeiZhe, et autres
Publié: (2026)
par: Xu, WeiZhe, et autres
Publié: (2026)
Semantic Voting: Execution-Grounded Consensus for LLM Code Generation
par: Jiang, Shan, et autres
Publié: (2026)
par: Jiang, Shan, et autres
Publié: (2026)
StackSight: Unveiling WebAssembly through Large Language Models and Neurosymbolic Chain-of-Thought Decompilation
par: Fang, Weike, et autres
Publié: (2024)
par: Fang, Weike, et autres
Publié: (2024)
Insights Generator: Systematic Corpus-Level Trace Diagnostics for LLM Agents
par: Manglik, Akshay, et autres
Publié: (2026)
par: Manglik, Akshay, et autres
Publié: (2026)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
par: Vulićević, Jelena Ilić
Publié: (2026)
par: Vulićević, Jelena Ilić
Publié: (2026)
Documents similaires
-
Is ChatGPT a Good Software Librarian? An Exploratory Study on the Use of ChatGPT for Software Library Recommendations
par: Latendresse, Jasmine, et autres
Publié: (2024) -
Beyond Synthetic Benchmarks: Evaluating LLM Performance on Real-World Class-Level Code Generation
par: Rahman, Musfiqur, et autres
Publié: (2025) -
OpenClassGen: A Large-Scale Corpus of Real-World Python Classes for LLM Research
par: Rahman, Musfiqur, et autres
Publié: (2025) -
Tokenomics: Quantifying Where Tokens Are Used in Agentic Software Engineering
par: Salim, Mohamad, et autres
Publié: (2026) -
Automatic Detection of LLM-Generated Code: A Comparative Case Study of Contemporary Models Across Function and Class Granularities
par: Rahman, Musfiqur, et autres
Publié: (2024)