Guidelines for Empirical Studies in Software Engineering involving Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Baltes, Sebastian, Angermeir, Florian, Arora, Chetan, Barón, Marvin Muñoz, Chen, Chunyang, Böhme, Lukas, Calefato, Fabio, Ernst, Neil, Falessi, Davide, Fitzgerald, Brian, Fucci, Davide, He, Junda, Treude, Christoph, Kalinowski, Marcos, Lambiase, Stefano, Russo, Daniel, Lungu, Mircea, Montes, Cristina Martinez, Prechelt, Lutz, Ralph, Paul, van Tonder, Rijnard, Wagner, Stefan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Evaluation Guidelines for Empirical Studies involving LLMs
von: Wagner, Stefan, et al.
Veröffentlicht: (2024)
von: Wagner, Stefan, et al.
Veröffentlicht: (2024)
User Misconceptions of LLM-Based Conversational Programming Assistants
von: O'Brien, Gabrielle, et al.
Veröffentlicht: (2025)
von: O'Brien, Gabrielle, et al.
Veröffentlicht: (2025)
Characterizing Requirements Smells
von: Gentili, Emanuele, et al.
Veröffentlicht: (2024)
von: Gentili, Emanuele, et al.
Veröffentlicht: (2024)
DRIVE-T: A Methodology for Discriminative and Representative Data Viz Item Selection for Literacy Construct and Assessment
von: Locoro, Angela, et al.
Veröffentlicht: (2025)
von: Locoro, Angela, et al.
Veröffentlicht: (2025)
An Extensive Comparison of Static Application Security Testing Tools
von: Esposito, Matteo, et al.
Veröffentlicht: (2024)
von: Esposito, Matteo, et al.
Veröffentlicht: (2024)
Beyond Literacy: Predicting Interpretation Correctness of Visualizations with User Traits, Item Difficulty, and Rasch Scores
von: Falessi, Davide, et al.
Veröffentlicht: (2026)
von: Falessi, Davide, et al.
Veröffentlicht: (2026)
Self-Admitted GenAI Usage in Open-Source Software
von: Xiao, Tao, et al.
Veröffentlicht: (2025)
von: Xiao, Tao, et al.
Veröffentlicht: (2025)
Anticipating Bugs: Ticket-Level Bug Prediction and Temporal Proximity Effects
von: La Prova, Daniele, et al.
Veröffentlicht: (2025)
von: La Prova, Daniele, et al.
Veröffentlicht: (2025)
Operationalizing Ethics for AI Agents: How Developers Encode Values into Repository Context Files
von: Treude, Christoph, et al.
Veröffentlicht: (2026)
von: Treude, Christoph, et al.
Veröffentlicht: (2026)
AI Slop and the Software Commons
von: Baltes, Sebastian, et al.
Veröffentlicht: (2026)
von: Baltes, Sebastian, et al.
Veröffentlicht: (2026)
"An Endless Stream of AI Slop": The Growing Burden of AI-Assisted Software Development
von: Baltes, Sebastian, et al.
Veröffentlicht: (2026)
von: Baltes, Sebastian, et al.
Veröffentlicht: (2026)
A Multivocal Literature Review on the Benefits and Limitations of Automated Machine Learning Tools
von: Azevedo, Kelly, et al.
Veröffentlicht: (2024)
von: Azevedo, Kelly, et al.
Veröffentlicht: (2024)
Professional Insights into Benefits and Limitations of Implementing MLOps Principles
von: Araujo, Gabriel, et al.
Veröffentlicht: (2024)
von: Araujo, Gabriel, et al.
Veröffentlicht: (2024)
Assessing the Use of AutoML for Data-Driven Software Engineering
von: Calefato, Fabio, et al.
Veröffentlicht: (2023)
von: Calefato, Fabio, et al.
Veröffentlicht: (2023)
Characterizing Data Visualization Literacy: a Systematic Literature Review
von: Beschi, Sara, et al.
Veröffentlicht: (2025)
von: Beschi, Sara, et al.
Veröffentlicht: (2025)
Dealing with SonarQube Cloud: Initial Results from a Mining Software Repository Study
von: Nocera, Sabato, et al.
Veröffentlicht: (2025)
von: Nocera, Sabato, et al.
Veröffentlicht: (2025)
Crossover Designs in Software Engineering Experiments: Review of the State of Analysis
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
A Dataset of Agentic AI Coding Tool Configurations
von: Galster, Matthias, et al.
Veröffentlicht: (2026)
von: Galster, Matthias, et al.
Veröffentlicht: (2026)
Reflections on the Reproducibility of Commercial LLM Performance in Empirical Software Engineering Studies
von: Angermeir, Florian, et al.
Veröffentlicht: (2025)
von: Angermeir, Florian, et al.
Veröffentlicht: (2025)
Information-Theoretic Detection of Unusual Source Code Changes
von: Torres, Adriano, et al.
Veröffentlicht: (2025)
von: Torres, Adriano, et al.
Veröffentlicht: (2025)
Context Engineering for AI Agents in Open-Source Software
von: Mohsenimofidi, Seyedmoein, et al.
Veröffentlicht: (2025)
von: Mohsenimofidi, Seyedmoein, et al.
Veröffentlicht: (2025)
ThreMoLIA: Threat Modeling of Large Language Model-Integrated Applications
von: Jedrzejewski, Felix Viktor, et al.
Veröffentlicht: (2025)
von: Jedrzejewski, Felix Viktor, et al.
Veröffentlicht: (2025)
Optimizing Large Language Model Hyperparameters for Code Generation
von: Arora, Chetan, et al.
Veröffentlicht: (2024)
von: Arora, Chetan, et al.
Veröffentlicht: (2024)
LLM-Based Multi-Agent Systems for Software Engineering: Literature Review, Vision and the Road Ahead
von: He, Junda, et al.
Veröffentlicht: (2024)
von: He, Junda, et al.
Veröffentlicht: (2024)
$Classi|Q\rangle$ Towards a Translation Framework To Bridge The Classical-Quantum Programming Gap
von: Esposito, Matteo, et al.
Veröffentlicht: (2024)
von: Esposito, Matteo, et al.
Veröffentlicht: (2024)
Presentación
von: Patrizia Calefato
Veröffentlicht: (2008)
von: Patrizia Calefato
Veröffentlicht: (2008)
Modas juveniles y nuevas identidades culturales
von: Patrizia Calefato
Veröffentlicht: (2020)
von: Patrizia Calefato
Veröffentlicht: (2020)
“Espacio continuo de transformación”: la mirada semiótica sobre la traducción cultural
von: Patrizia Calefato
Veröffentlicht: (2008)
von: Patrizia Calefato
Veröffentlicht: (2008)
A Second Look at the Impact of Passive Voice Requirements on Domain Modeling: Bayesian Reanalysis of an Experiment
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
Policy-driven Software Bill of Materials on GitHub: An Empirical Study
von: Novikov, Oleksii, et al.
Veröffentlicht: (2025)
von: Novikov, Oleksii, et al.
Veröffentlicht: (2025)
NLP4RE Tools: Classification, Overview, and Management
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
Augmenting Software Bills of Materials with Software Vulnerability Description: A Preliminary Study on GitHub
von: Fucci, Davide, et al.
Veröffentlicht: (2025)
von: Fucci, Davide, et al.
Veröffentlicht: (2025)
GenAI Is No Silver Bullet for Qualitative Research in Software Engineering
von: Ernst, Neil A., et al.
Veröffentlicht: (2026)
von: Ernst, Neil A., et al.
Veröffentlicht: (2026)
The Silent Scientist: When Software Research Fails to Reach Its Audience
von: Wyrich, Marvin, et al.
Veröffentlicht: (2025)
von: Wyrich, Marvin, et al.
Veröffentlicht: (2025)
Backdoor Attacks on Open Vocabulary Object Detectors via Multi-Modal Prompt Tuning
von: Raj, Ankita, et al.
Veröffentlicht: (2025)
von: Raj, Ankita, et al.
Veröffentlicht: (2025)
Practical Guidelines for the Selection and Evaluation of Natural Language Processing Techniques in Requirements Engineering
von: Sabetzadeh, Mehrdad, et al.
Veröffentlicht: (2024)
von: Sabetzadeh, Mehrdad, et al.
Veröffentlicht: (2024)
Replications, Revisions, and Reanalyses: Managing Variance Theories in Software Engineering
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
Privacy by Design: Aligning GDPR and Software Engineering Specifications with a Requirements Engineering Approach
von: Kosenkov, Oleksandr, et al.
Veröffentlicht: (2025)
von: Kosenkov, Oleksandr, et al.
Veröffentlicht: (2025)
Designing NLP-based solutions for requirements variability management: experiences from a design science study at Visma
von: Elahidoost, Parisa, et al.
Veröffentlicht: (2024)
von: Elahidoost, Parisa, et al.
Veröffentlicht: (2024)
Measuring the Fitness-for-Purpose of Requirements: An initial Model of Activities and Attributes
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
von: Frattini, Julian, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Evaluation Guidelines for Empirical Studies involving LLMs
von: Wagner, Stefan, et al.
Veröffentlicht: (2024) -
User Misconceptions of LLM-Based Conversational Programming Assistants
von: O'Brien, Gabrielle, et al.
Veröffentlicht: (2025) -
Characterizing Requirements Smells
von: Gentili, Emanuele, et al.
Veröffentlicht: (2024) -
DRIVE-T: A Methodology for Discriminative and Representative Data Viz Item Selection for Literacy Construct and Assessment
von: Locoro, Angela, et al.
Veröffentlicht: (2025) -
An Extensive Comparison of Static Application Security Testing Tools
von: Esposito, Matteo, et al.
Veröffentlicht: (2024)