Rethinking Code Review Workflows with LLM Assistance: An Empirical Study
Fuente:
arXiv
Saved in:
| Main Authors: | Aðalsteinsson, Fannar Steinn, Magnússon, Björn Borgar, Milicevic, Mislav, Davidsson, Adam Nirving, Cheng, Chih-Hong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Detect Repair Verify for Securing LLM Generated Code: A Multi-Language Empirical Study
by: Cheng, Cheng
Published: (2026)
by: Cheng, Cheng
Published: (2026)
Code Review Automation Via Multi-task Federated LLM -- An Empirical Study
by: Kumar, Jahnavi, et al.
Published: (2024)
by: Kumar, Jahnavi, et al.
Published: (2024)
Detect--Repair--Verify for LLM-Generated Code: A Multi-Language, Multi-Granularity Empirical Study
by: Cheng, Cheng
Published: (2026)
by: Cheng, Cheng
Published: (2026)
An Empirical Study of Developers' Challenges in Implementing Workflows as Code: A Case Study on Apache Airflow
by: Yasmin, Jerin, et al.
Published: (2024)
by: Yasmin, Jerin, et al.
Published: (2024)
Workflow-Level Design Principles for Trustworthy GenAI in Automotive System Engineering
by: Cheng, Chih-Hong, et al.
Published: (2026)
by: Cheng, Chih-Hong, et al.
Published: (2026)
An Empirical Study of LLM-Based Code Clone Detection
by: Zhu, Wenqing, et al.
Published: (2025)
by: Zhu, Wenqing, et al.
Published: (2025)
Explaining Explanation: An Empirical Study on Explanation in Code Reviews
by: Widyasari, Ratnadira, et al.
Published: (2023)
by: Widyasari, Ratnadira, et al.
Published: (2023)
Quality Assurance for LLM-RAG Systems: Empirical Insights from Tourism Application Testing
by: Ahmed, Bestoun S., et al.
Published: (2025)
by: Ahmed, Bestoun S., et al.
Published: (2025)
An Empirical Study of the Evolution of GitHub Actions Workflows
by: Mazrae, Pooya Rostami, et al.
Published: (2026)
by: Mazrae, Pooya Rostami, et al.
Published: (2026)
An Empirical Study of Static Analysis Tools for Secure Code Review
by: Charoenwet, Wachiraphan, et al.
Published: (2024)
by: Charoenwet, Wachiraphan, et al.
Published: (2024)
Code vs Serialized AST Inputs for LLM-Based Code Summarization: An Empirical Study
by: Dong, Shijia, et al.
Published: (2026)
by: Dong, Shijia, et al.
Published: (2026)
Toward Effective Secure Code Reviews: An Empirical Study of Security-Related Coding Weaknesses
by: Charoenwet, Wachiraphan, et al.
Published: (2023)
by: Charoenwet, Wachiraphan, et al.
Published: (2023)
From Industry Claims to Empirical Reality: An Empirical Study of Code Review Agents in Pull Requests
by: Chowdhury, Kowshik, et al.
Published: (2026)
by: Chowdhury, Kowshik, et al.
Published: (2026)
An Empirical Study on Code Review Activity Prediction and Its Impact in Practice
by: Olewicki, Doriane, et al.
Published: (2024)
by: Olewicki, Doriane, et al.
Published: (2024)
Benchmarking and Studying the LLM-based Code Review
by: Zeng, Zhengran, et al.
Published: (2025)
by: Zeng, Zhengran, et al.
Published: (2025)
On the Effectiveness of Training Data Optimization for LLM-based Code Generation: An Empirical Study
by: Kuang, Shiqi, et al.
Published: (2025)
by: Kuang, Shiqi, et al.
Published: (2025)
Demystifying Errors in LLM Reasoning Traces: An Empirical Study of Code Execution Simulation
by: Abdollahi, Mohammad, et al.
Published: (2025)
by: Abdollahi, Mohammad, et al.
Published: (2025)
An Empirical Study of Complexity, Heterogeneity, and Compliance of GitHub Actions Workflows
by: Abrokwah, Edward, et al.
Published: (2025)
by: Abrokwah, Edward, et al.
Published: (2025)
An Empirical Study on the Effects of System Prompts in Instruction-Tuned Models for Code Generation
by: Cheng, Zaiyu, et al.
Published: (2026)
by: Cheng, Zaiyu, et al.
Published: (2026)
Deciphering Refactoring Branch Dynamics in Modern Code Review: An Empirical Study on Qt
by: AlOmar, Eman Abdullah
Published: (2024)
by: AlOmar, Eman Abdullah
Published: (2024)
I Can't Share Code, but I need Translation -- An Empirical Study on Code Translation through Federated LLM
by: Kumar, Jahnavi, et al.
Published: (2025)
by: Kumar, Jahnavi, et al.
Published: (2025)
Decoding Human-LLM Collaboration in Coding: An Empirical Study of Multi-Turn Conversations in the Wild
by: Zhang, Binquan, et al.
Published: (2025)
by: Zhang, Binquan, et al.
Published: (2025)
LLM-Based Test-Driven Interactive Code Generation: User Study and Empirical Evaluation
by: Fakhoury, Sarah, et al.
Published: (2024)
by: Fakhoury, Sarah, et al.
Published: (2024)
An Empirical Study of Interaction Smells in Multi-Turn Human-LLM Collaborative Code Generation
by: Zhang, Binquan, et al.
Published: (2026)
by: Zhang, Binquan, et al.
Published: (2026)
The Hidden Costs of Automation: An Empirical Study on GitHub Actions Workflow Maintenance
by: Valenzuela-Toledo, Pablo, et al.
Published: (2024)
by: Valenzuela-Toledo, Pablo, et al.
Published: (2024)
Environment-in-the-Loop: Rethinking Code Migration with LLM-based Agents
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
Learn to Code Sustainably: An Empirical Study on LLM-based Green Code Generation
by: Vartziotis, Tina, et al.
Published: (2024)
by: Vartziotis, Tina, et al.
Published: (2024)
Compact Constraint Encoding for LLM Code Generation: An Empirical Study of Token Economics and Constraint Compliance
by: Tang, Hanzhang
Published: (2026)
by: Tang, Hanzhang
Published: (2026)
On the Use of Agentic Coding Manifests: An Empirical Study of Claude Code
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
Empirical Studies on Adversarial Reverse Engineering with Students
by: Tab, et al.
Published: (2026)
by: Tab, et al.
Published: (2026)
Guiding AI to Fix Its Own Flaws: An Empirical Study on LLM-Driven Secure Code Generation
by: Yan, Hao, et al.
Published: (2025)
by: Yan, Hao, et al.
Published: (2025)
Can We Classify Flaky Tests Using Only Test Code? An LLM-Based Empirical Study
by: Berndt, Alexander, et al.
Published: (2026)
by: Berndt, Alexander, et al.
Published: (2026)
Failure-Aware Enhancements for Large Language Model (LLM) Code Generation: An Empirical Study on Decision Framework
by: Shen, Jianru, et al.
Published: (2026)
by: Shen, Jianru, et al.
Published: (2026)
What Drives Issue Resolution Speed? An Empirical Study of Scientific Workflow Systems on GitHub
by: Alam, Khairul, et al.
Published: (2025)
by: Alam, Khairul, et al.
Published: (2025)
Modeling Sampling Workflows for Code Repositories
by: Lefeuvre, Romain, et al.
Published: (2026)
by: Lefeuvre, Romain, et al.
Published: (2026)
An Empirical Study on Challenges for LLM Application Developers
by: Chen, Xiang, et al.
Published: (2024)
by: Chen, Xiang, et al.
Published: (2024)
Engagement in Code Review: Emotional, Behavioral, and Cognitive Dimensions in Peer vs. LLM Interactions
by: Alami, Adam, et al.
Published: (2025)
by: Alami, Adam, et al.
Published: (2025)
Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review
by: Kamalı, Hüseyin Özgür, et al.
Published: (2026)
by: Kamalı, Hüseyin Özgür, et al.
Published: (2026)
Agent READMEs: An Empirical Study of Context Files for Agentic Coding
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
An Empirical Study of Retrieval-Augmented Code Generation: Challenges and Opportunities
by: Yang, Zezhou, et al.
Published: (2025)
by: Yang, Zezhou, et al.
Published: (2025)
Similar Items
-
Detect Repair Verify for Securing LLM Generated Code: A Multi-Language Empirical Study
by: Cheng, Cheng
Published: (2026) -
Code Review Automation Via Multi-task Federated LLM -- An Empirical Study
by: Kumar, Jahnavi, et al.
Published: (2024) -
Detect--Repair--Verify for LLM-Generated Code: A Multi-Language, Multi-Granularity Empirical Study
by: Cheng, Cheng
Published: (2026) -
An Empirical Study of Developers' Challenges in Implementing Workflows as Code: A Case Study on Apache Airflow
by: Yasmin, Jerin, et al.
Published: (2024) -
Workflow-Level Design Principles for Trustworthy GenAI in Automotive System Engineering
by: Cheng, Chih-Hong, et al.
Published: (2026)