Saved in:
| Main Authors: | Cheng, Ti-Chung, Badea, Carmen, Bird, Christian, Zimmermann, Thomas, DeLine, Robert, Forsgren, Nicole, Ford, Denae |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.00880 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can GPT-4 Replicate Empirical Software Engineering Research?
by: Liang, Jenny T., et al.
Published: (2023)
by: Liang, Jenny T., et al.
Published: (2023)
AI Where It Matters: Where, Why, and How Developers Want AI Support in Daily Work
by: Choudhuri, Rudrajit, et al.
Published: (2025)
by: Choudhuri, Rudrajit, et al.
Published: (2025)
Investigating and Designing for Trust in AI-powered Code Generation Tools
by: Wang, Ruotong, et al.
Published: (2023)
by: Wang, Ruotong, et al.
Published: (2023)
To Copilot and Beyond: 22 AI Systems Developers Want Built
by: Choudhuri, Rudrajit, et al.
Published: (2026)
by: Choudhuri, Rudrajit, et al.
Published: (2026)
"Maybe We Need Some More Examples:" Individual and Team Drivers of Developer GenAI Tool Use
by: Miller, Courtney, et al.
Published: (2025)
by: Miller, Courtney, et al.
Published: (2025)
Beyond the Comfort Zone: Emerging Solutions to Overcome Challenges in Integrating LLMs into Software Products
by: Nahar, Nadia, et al.
Published: (2024)
by: Nahar, Nadia, et al.
Published: (2024)
Prompting in the Wild: An Empirical Study of Prompt Evolution in Software Repositories
by: Tafreshipour, Mahan, et al.
Published: (2024)
by: Tafreshipour, Mahan, et al.
Published: (2024)
Reproducibility of Build Environments through Space and Time
by: Malka, Julien, et al.
Published: (2024)
by: Malka, Julien, et al.
Published: (2024)
Selection of Prompt Engineering Techniques for Code Generation through Predicting Code Complexity
by: Wang, Chung-Yu, et al.
Published: (2024)
by: Wang, Chung-Yu, et al.
Published: (2024)
Efficient Prime Paths Generation
by: Zelek, Jakub, et al.
Published: (2026)
by: Zelek, Jakub, et al.
Published: (2026)
Good Vibrations? A Qualitative Study of Co-Creation, Communication, Flow, and Trust in Vibe Coding
by: Pimenova, Veronica, et al.
Published: (2025)
by: Pimenova, Veronica, et al.
Published: (2025)
The First Prompt Counts the Most! An Evaluation of Large Language Models on Iterative Example-Based Code Generation
by: Fu, Yingjie, et al.
Published: (2024)
by: Fu, Yingjie, et al.
Published: (2024)
Studying LLM Performance on Closed- and Open-source Data
by: Ahmed, Toufique, et al.
Published: (2024)
by: Ahmed, Toufique, et al.
Published: (2024)
MOS: Towards Effective Smart Contract Vulnerability Detection through Mixture-of-Experts Tuning of Large Language Models
by: Yuan, Hang, et al.
Published: (2025)
by: Yuan, Hang, et al.
Published: (2025)
Green Prompt Engineering: Investigating the Energy Impact of Prompt Design in Software Engineering
by: De Martino, Vincenzo, et al.
Published: (2025)
by: De Martino, Vincenzo, et al.
Published: (2025)
Will It Break in Production? Metric-Driven Prediction of Residual Defects in Python Systems
by: De Rosa, Giuseppe, et al.
Published: (2026)
by: De Rosa, Giuseppe, et al.
Published: (2026)
ARISE -- Adaptive Refinement and Iterative Scenario Engineering
by: Poddubnyy, Konstantin, et al.
Published: (2026)
by: Poddubnyy, Konstantin, et al.
Published: (2026)
Understanding Dominant Themes in Reviewing Agentic AI-authored Code
by: Haider, Md. Asif, et al.
Published: (2026)
by: Haider, Md. Asif, et al.
Published: (2026)
Uncovering Scientific Software Sustainability through Community Engagement and Software Quality Metrics
by: Ahmed, Sharif, et al.
Published: (2025)
by: Ahmed, Sharif, et al.
Published: (2025)
Exploring Sustainability in Scientific Software through Code Quality & Test Coverage Metrics
by: Rahman, Sheikh Md. Mushfiqur, et al.
Published: (2026)
by: Rahman, Sheikh Md. Mushfiqur, et al.
Published: (2026)
Evaluating the Performance of Large Language Models in Competitive Programming: A Multi-Year, Multi-Grade Analysis
by: Dumitran, Adrian Marius, et al.
Published: (2024)
by: Dumitran, Adrian Marius, et al.
Published: (2024)
How Do Microservice API Patterns Impact Understandability? A Controlled Experiment
by: Bogner, Justus, et al.
Published: (2024)
by: Bogner, Justus, et al.
Published: (2024)
Does Functional Package Management Enable Reproducible Builds at Scale? Yes
by: Malka, Julien, et al.
Published: (2025)
by: Malka, Julien, et al.
Published: (2025)
Docker Does Not Guarantee Reproducibility
by: Malka, Julien, et al.
Published: (2026)
by: Malka, Julien, et al.
Published: (2026)
Improving Assignment Submission in Higher Education through a Git-Enabled System: An Iterative Case Study
by: Babatunde, Ololade, et al.
Published: (2025)
by: Babatunde, Ololade, et al.
Published: (2025)
An Empirical Study on the Effects of System Prompts in Instruction-Tuned Models for Code Generation
by: Cheng, Zaiyu, et al.
Published: (2026)
by: Cheng, Zaiyu, et al.
Published: (2026)
Bias Ahead: Sensitive Prompts as Early Warnings for Fairness in Large Language Models
by: Voria, Gianmario, et al.
Published: (2026)
by: Voria, Gianmario, et al.
Published: (2026)
E-code: Mastering Efficient Code Generation through Pretrained Models and Expert Encoder Group
by: Pan, Yue, et al.
Published: (2024)
by: Pan, Yue, et al.
Published: (2024)
Task-oriented Prompt Enhancement via Script Generation
by: Wang, Chung-Yu, et al.
Published: (2024)
by: Wang, Chung-Yu, et al.
Published: (2024)
On Fixing Insecure AI-Generated Code through Model Fine-Tuning and Prompting Strategies
by: Jahromi, Ali Soltanian Fard, et al.
Published: (2026)
by: Jahromi, Ali Soltanian Fard, et al.
Published: (2026)
Statistical-Based Metric Threshold Setting Method for Software Fault Prediction in Firmware Projects: An Industrial Experience
by: De Luca, Marco, et al.
Published: (2026)
by: De Luca, Marco, et al.
Published: (2026)
Automatic Test-Case Reduction in Proof Assistants: A Case Study in Coq
by: Gross, Jason, et al.
Published: (2022)
by: Gross, Jason, et al.
Published: (2022)
BinMetric: A Comprehensive Binary Analysis Benchmark for Large Language Models
by: Shang, Xiuwei, et al.
Published: (2025)
by: Shang, Xiuwei, et al.
Published: (2025)
Predicting the Understandability of Computational Notebooks through Code Metrics Analysis
by: Ghahfarokhi, Mojtaba Mostafavi, et al.
Published: (2024)
by: Ghahfarokhi, Mojtaba Mostafavi, et al.
Published: (2024)
SCOPE: A Dataset of Stereotyped Prompts for Counterfactual Fairness Assessment of LLMs
by: Parziale, Alessandra, et al.
Published: (2026)
by: Parziale, Alessandra, et al.
Published: (2026)
MetricSynth: Framework for Aggregating DORA and KPI Metrics Across Multi-Platform Engineering
by: Jain, Pallav, et al.
Published: (2025)
by: Jain, Pallav, et al.
Published: (2025)
Towards a Responsible AI Metrics Catalogue: A Collection of Metrics for AI Accountability
by: Xia, Boming, et al.
Published: (2023)
by: Xia, Boming, et al.
Published: (2023)
PromptPex: Automatic Test Generation for Language Model Prompts
by: Sharma, Reshabh K, et al.
Published: (2025)
by: Sharma, Reshabh K, et al.
Published: (2025)
PromptSet: A Programmer's Prompting Dataset
by: Pister, Kaiser, et al.
Published: (2024)
by: Pister, Kaiser, et al.
Published: (2024)
SE Journals in 2036: Looking Back at the Future We Need to Have
by: Menzies, Tim, et al.
Published: (2026)
by: Menzies, Tim, et al.
Published: (2026)
Similar Items
-
Can GPT-4 Replicate Empirical Software Engineering Research?
by: Liang, Jenny T., et al.
Published: (2023) -
AI Where It Matters: Where, Why, and How Developers Want AI Support in Daily Work
by: Choudhuri, Rudrajit, et al.
Published: (2025) -
Investigating and Designing for Trust in AI-powered Code Generation Tools
by: Wang, Ruotong, et al.
Published: (2023) -
To Copilot and Beyond: 22 AI Systems Developers Want Built
by: Choudhuri, Rudrajit, et al.
Published: (2026) -
"Maybe We Need Some More Examples:" Individual and Team Drivers of Developer GenAI Tool Use
by: Miller, Courtney, et al.
Published: (2025)