Empirical and Sustainability Aspects of Software Engineering Research in the Era of Large Language Models: A Reflection
Fuente:
arXiv
Saved in:
| Main Authors: | Williams, David, Hort, Max, Kechagia, Maria, Aleti, Aldeida, Petke, Justyna, Sarro, Federica |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Test-based Patch Clustering for Automatically-Generated Patches Assessment
by: Martinez, Matias, et al.
Published: (2022)
by: Martinez, Matias, et al.
Published: (2022)
Unveiling Practical Shortcomings of Patch Overfitting Detection Techniques
by: Williams, David, et al.
Published: (2026)
by: Williams, David, et al.
Published: (2026)
Hot Fixing Software: A Comprehensive Review of Terminology, Techniques, and Applications
by: Hanna, Carol, et al.
Published: (2024)
by: Hanna, Carol, et al.
Published: (2024)
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
by: Hanna, Carol, et al.
Published: (2025)
by: Hanna, Carol, et al.
Published: (2025)
Trustworthy AI Software Engineers
by: Aleti, Aldeida, et al.
Published: (2026)
by: Aleti, Aldeida, et al.
Published: (2026)
FairRF: Multi-Objective Search for Single and Intersectional Software Fairness
by: d'Alosio, Giordano, et al.
Published: (2026)
by: d'Alosio, Giordano, et al.
Published: (2026)
A Comprehensive Survey of Benchmarks for Automated Improvement of Software's Non-Functional Properties
by: Blot, Aymeric, et al.
Published: (2022)
by: Blot, Aymeric, et al.
Published: (2022)
LLM-Guided Genetic Improvement: Envisioning Semantic Aware Automated Software Evolution
by: Even-Mendoza, Karine, et al.
Published: (2025)
by: Even-Mendoza, Karine, et al.
Published: (2025)
BayesInsights: Modelling Software Delivery and Developer Experience with Bayesian Networks at Bloomberg
by: Kirbas, Serkan, et al.
Published: (2026)
by: Kirbas, Serkan, et al.
Published: (2026)
Enhancing Large Language Models for Text-to-Testcase Generation
by: Alagarsamy, Saranya, et al.
Published: (2024)
by: Alagarsamy, Saranya, et al.
Published: (2024)
PAFOT: A Position-Based Approach for Finding Optimal Tests of Autonomous Vehicles
by: Crespo-Rodriguez, Victor, et al.
Published: (2024)
by: Crespo-Rodriguez, Victor, et al.
Published: (2024)
Experimental evaluation of architectural software performance design patterns in microservices
by: Meijer, Willem, et al.
Published: (2024)
by: Meijer, Willem, et al.
Published: (2024)
Hot Fixing in the Wild
by: Hanna, Carol, et al.
Published: (2026)
by: Hanna, Carol, et al.
Published: (2026)
Requirements-Driven Automated Software Testing: A Systematic Review
by: Wang, Fanyu, et al.
Published: (2025)
by: Wang, Fanyu, et al.
Published: (2025)
Fairness Testing: A Comprehensive Survey and Analysis of Trends
by: Chen, Zhenpeng, et al.
Published: (2022)
by: Chen, Zhenpeng, et al.
Published: (2022)
Charting The Evolution of Solidity Error Handling
by: Mitropoulos, Charalambos, et al.
Published: (2024)
by: Mitropoulos, Charalambos, et al.
Published: (2024)
Understanding Fairness in Software Engineering: Insights from Stack Exchange
by: Sesari, Emeralda, et al.
Published: (2024)
by: Sesari, Emeralda, et al.
Published: (2024)
UntrustVul: An Automated Approach for Identifying Untrustworthy Alerts in Vulnerability Detection Models
by: Tung, Lam Nguyen, et al.
Published: (2025)
by: Tung, Lam Nguyen, et al.
Published: (2025)
Reinforcement Learning for Mutation Operator Selection in Automated Program Repair
by: Hanna, Carol, et al.
Published: (2023)
by: Hanna, Carol, et al.
Published: (2023)
Neuron Patching: Semantic-based Neuron-level Language Model Repair for Code Generation
by: Gu, Jian, et al.
Published: (2023)
by: Gu, Jian, et al.
Published: (2023)
ViBR: Automated Bug Replay from Video-based Reports using Vision-Language Models
by: Feng, Sidong, et al.
Published: (2026)
by: Feng, Sidong, et al.
Published: (2026)
Revisiting Sentiment Analysis for Software Engineering in the Era of Large Language Models
by: Zhang, Ting, et al.
Published: (2023)
by: Zhang, Ting, et al.
Published: (2023)
It is Giving Major Satisfaction: Why Fairness Matters for Software Practitioners
by: Sesari, Emeralda, et al.
Published: (2024)
by: Sesari, Emeralda, et al.
Published: (2024)
SafeTune: Search-based Harmfulness Minimisation for Large Language Models
by: d'Aloisio, Giordano, et al.
Published: (2026)
by: d'Aloisio, Giordano, et al.
Published: (2026)
The Role of Road Features and Vehicle Dynamics in Cost-Effective Autonomous Vehicles Safety Testing: Insights from Instance Space Analysis
by: Crespo-Rodriguez, Victor, et al.
Published: (2026)
by: Crespo-Rodriguez, Victor, et al.
Published: (2026)
From Domain Documents to Requirements: Retrieval-Augmented Generation in the Space Industry
by: Arora, Chetan, et al.
Published: (2025)
by: Arora, Chetan, et al.
Published: (2025)
The Current Challenges of Software Engineering in the Era of Large Language Models
by: Gao, Cuiyun, et al.
Published: (2024)
by: Gao, Cuiyun, et al.
Published: (2024)
A Comparative Study on Large Language Models for Log Parsing
by: Astekin, Merve, et al.
Published: (2024)
by: Astekin, Merve, et al.
Published: (2024)
Enhancing Energy-Awareness in Deep Learning through Fine-Grained Energy Measurement
by: Rajput, Saurabhsingh, et al.
Published: (2023)
by: Rajput, Saurabhsingh, et al.
Published: (2023)
Comparing AI Coding Agents: A Task-Stratified Analysis of Pull Request Acceptance
by: Pinna, Giovanni, et al.
Published: (2026)
by: Pinna, Giovanni, et al.
Published: (2026)
A Semantic-based Optimization Approach for Repairing LLMs: Case Study on Code Generation
by: Gu, Jian, et al.
Published: (2025)
by: Gu, Jian, et al.
Published: (2025)
How Do Generative Models Draw a Software Engineer? A Case Study on Stable Diffusion Bias
by: Fadahunsi, Tosin, et al.
Published: (2025)
by: Fadahunsi, Tosin, et al.
Published: (2025)
On the Compression of Language Models for Code: An Empirical Study on CodeBERT
by: d'Aloisio, Giordano, et al.
Published: (2024)
by: d'Aloisio, Giordano, et al.
Published: (2024)
Guidelines for Empirical Studies in Software Engineering involving Large Language Models
by: Baltes, Sebastian, et al.
Published: (2025)
by: Baltes, Sebastian, et al.
Published: (2025)
An Empirical Study on Decision-Making Aspects in Responsible Software Engineering for AI
by: Rani, Lekshmi Murali, et al.
Published: (2025)
by: Rani, Lekshmi Murali, et al.
Published: (2025)
JMigBench: A Benchmark for Evaluating LLMs on Source Code Migration (Java 8 to Java 11)
by: Amin, Nishil, et al.
Published: (2026)
by: Amin, Nishil, et al.
Published: (2026)
Exploring the Power of Diffusion Large Language Models for Software Engineering: An Empirical Investigation
by: Zhang, Jingyao, et al.
Published: (2025)
by: Zhang, Jingyao, et al.
Published: (2025)
Multi-Modal Requirements Data-based Acceptance Criteria Generation using LLMs
by: Wang, Fanyu, et al.
Published: (2025)
by: Wang, Fanyu, et al.
Published: (2025)
Do Research Software Engineers and Software Engineering Researchers Speak the Same Language?
by: Kehrer, Timo, et al.
Published: (2025)
by: Kehrer, Timo, et al.
Published: (2025)
SustainDiffusion: Optimising the Social and Environmental Sustainability of Stable Diffusion Models
by: d'Aloisio, Giordano, et al.
Published: (2025)
by: d'Aloisio, Giordano, et al.
Published: (2025)
Similar Items
-
Test-based Patch Clustering for Automatically-Generated Patches Assessment
by: Martinez, Matias, et al.
Published: (2022) -
Unveiling Practical Shortcomings of Patch Overfitting Detection Techniques
by: Williams, David, et al.
Published: (2026) -
Hot Fixing Software: A Comprehensive Review of Terminology, Techniques, and Applications
by: Hanna, Carol, et al.
Published: (2024) -
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
by: Hanna, Carol, et al.
Published: (2025) -
Trustworthy AI Software Engineers
by: Aleti, Aldeida, et al.
Published: (2026)