Evaluating Software Process Models for Multi-Agent Class-Level Code Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shafin, Wasique Islam, Rafi, Md Nakhla, Li, Zhenhao, Chen, Tse-Hsun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Multi-Agent Approach to Fault Localization via Graph-Based Retrieval and Reflexion
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024)
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024)
CI-Repair-Bench: A Repository-Aware Benchmark for Automated Patch Validation via CI Workflows
von: Muna, Rabeya Khatun, et al.
Veröffentlicht: (2026)
von: Muna, Rabeya Khatun, et al.
Veröffentlicht: (2026)
Back to the Future! Studying Data Cleanness in Defects4J and its Impact on Fault Localization
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2023)
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2023)
Order Matters! An Empirical Study on Large Language Models' Input Order Bias in Software Fault Localization
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024)
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024)
Towards Better Graph Neural Network-based Fault Localization Through Enhanced Code Representation
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024)
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024)
SBEST: Spectrum-Based Fault Localization Without Fault-Triggering Tests
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024)
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024)
Crash Report Enhancement with Large Language Models: An Empirical Study
von: Fahim, S M Farah Al, et al.
Veröffentlicht: (2025)
von: Fahim, S M Farah Al, et al.
Veröffentlicht: (2025)
RobuNFR: Evaluating the Robustness of Large Language Models on Non-Functional Requirements Aware Code Generation
von: Lin, Feng, et al.
Veröffentlicht: (2025)
von: Lin, Feng, et al.
Veröffentlicht: (2025)
Towards Structured, State-Aware, and Execution-Grounded Reasoning for Software Engineering Agents
von: Tse-Hsun, et al.
Veröffentlicht: (2026)
von: Tse-Hsun, et al.
Veröffentlicht: (2026)
Discovery of Timeline and Crowd Reaction of Software Vulnerability Disclosures
von: Heng, Yi Wen, et al.
Veröffentlicht: (2024)
von: Heng, Yi Wen, et al.
Veröffentlicht: (2024)
Studying and Benchmarking Large Language Models For Log Level Suggestion
von: Heng, Yi Wen, et al.
Veröffentlicht: (2024)
von: Heng, Yi Wen, et al.
Veröffentlicht: (2024)
Identifying Performance-Sensitive Configurations in Software Systems through Code Analysis with LLM Agents
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
A Survey of Code Review Benchmarks and Evaluation Practices in Pre-LLM and LLM Era
von: Khan, Taufiqul Islam, et al.
Veröffentlicht: (2026)
von: Khan, Taufiqul Islam, et al.
Veröffentlicht: (2026)
SWE-Refactor: A Repository-Level Benchmark for Real-World LLM-Based Code Refactoring
von: Xu, Yisen, et al.
Veröffentlicht: (2026)
von: Xu, Yisen, et al.
Veröffentlicht: (2026)
CODEPROMPTZIP: Code-specific Prompt Compression for Retrieval-Augmented Generation in Coding Tasks with LMs
von: He, Pengfei, et al.
Veröffentlicht: (2025)
von: He, Pengfei, et al.
Veröffentlicht: (2025)
SOEN-101: Code Generation by Emulating Software Process Models Using Large Language Model Agents
von: Lin, Feng, et al.
Veröffentlicht: (2024)
von: Lin, Feng, et al.
Veröffentlicht: (2024)
MANTRA: Enhancing Automated Method-Level Refactoring with Contextual RAG and Multi-Agent LLM Collaboration
von: Xu, Yisen, et al.
Veröffentlicht: (2025)
von: Xu, Yisen, et al.
Veröffentlicht: (2025)
Evaluating the Effectiveness and Efficiency of Demonstration Retrievers in RAG for Coding Tasks
von: He, Pengfei, et al.
Veröffentlicht: (2024)
von: He, Pengfei, et al.
Veröffentlicht: (2024)
Empowering AIOps: Leveraging Large Language Models for IT Operations Management
von: Vitui, Arthur, et al.
Veröffentlicht: (2025)
von: Vitui, Arthur, et al.
Veröffentlicht: (2025)
Knowledge-Guided Multi-Agent Framework for Application-Level Software Code Generation
von: Xiong, Qian, et al.
Veröffentlicht: (2025)
von: Xiong, Qian, et al.
Veröffentlicht: (2025)
Studying the Impact of Early Test Termination Due to Assertion Failure on Code Coverage and Spectrum-based Fault Localization
von: Uddin, Md. Ashraf, et al.
Veröffentlicht: (2025)
von: Uddin, Md. Ashraf, et al.
Veröffentlicht: (2025)
On Rank Aggregating Test Prioritizations
von: Mondal, Shouvick, et al.
Veröffentlicht: (2024)
von: Mondal, Shouvick, et al.
Veröffentlicht: (2024)
ClassEval-T: Evaluating Large Language Models in Class-Level Code Translation
von: Xue, Pengyu, et al.
Veröffentlicht: (2024)
von: Xue, Pengyu, et al.
Veröffentlicht: (2024)
When LLMs Lag Behind: Knowledge Conflicts from Evolving APIs in Code Generation
von: Ashik, Ahmed Nusayer, et al.
Veröffentlicht: (2026)
von: Ashik, Ahmed Nusayer, et al.
Veröffentlicht: (2026)
SLICET5: Static Program Slicing using Language Models with Copy Mechanism and Constrained Decoding
von: He, Pengfei, et al.
Veröffentlicht: (2025)
von: He, Pengfei, et al.
Veröffentlicht: (2025)
A Comprehensive Framework for Evaluating API-oriented Code Generation in Large Language Models
von: Wu, Yixi, et al.
Veröffentlicht: (2024)
von: Wu, Yixi, et al.
Veröffentlicht: (2024)
From Historical Patches to Repair Plans: Outcome-Conditioned Reasoning for Repository-Level Program Repair
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
Peer Code Review in Research Software Development: The Research Software Engineer Perspective
von: Malik, Md Ariful Islam, et al.
Veröffentlicht: (2025)
von: Malik, Md Ariful Islam, et al.
Veröffentlicht: (2025)
HistoryFinder: Advancing Method-Level Source Code History Generation with Accurate Oracles and Enhanced Algorithm
von: Islam, Shahidul, et al.
Veröffentlicht: (2025)
von: Islam, Shahidul, et al.
Veröffentlicht: (2025)
A Combined Feature Embedding Tools for Multi-Class Software Defect and Identification
von: Sultan, Md. Fahim, et al.
Veröffentlicht: (2024)
von: Sultan, Md. Fahim, et al.
Veröffentlicht: (2024)
An Empirical Study on the Characteristics of Database Access Bugs in Java Applications
von: Liu, Wei, et al.
Veröffentlicht: (2024)
von: Liu, Wei, et al.
Veröffentlicht: (2024)
Towards Realistic Project-Level Code Generation via Multi-Agent Collaboration and Semantic Architecture Modeling
von: Zhao, Qianhui, et al.
Veröffentlicht: (2025)
von: Zhao, Qianhui, et al.
Veröffentlicht: (2025)
CODESIM: Multi-Agent Code Generation and Problem Solving through Simulation-Driven Planning and Debugging
von: Islam, Md. Ashraful, et al.
Veröffentlicht: (2025)
von: Islam, Md. Ashraful, et al.
Veröffentlicht: (2025)
Deep Learning Based Code Generation Methods: Literature Review
von: Yang, Zezhou, et al.
Veröffentlicht: (2023)
von: Yang, Zezhou, et al.
Veröffentlicht: (2023)
LLM-Based Detection of Tangled Code Changes for Higher-Quality Method-Level Bug Datasets
von: Opu, Md Nahidul Islam, et al.
Veröffentlicht: (2025)
von: Opu, Md Nahidul Islam, et al.
Veröffentlicht: (2025)
Evaluating Classical Software Process Models as Coordination Mechanisms for LLM-Based Software Generation
von: Ha, Duc Minh, et al.
Veröffentlicht: (2025)
von: Ha, Duc Minh, et al.
Veröffentlicht: (2025)
Process-Level Trajectory Evaluation for Environment Configuration in Software Engineering Agents
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025)
von: Kuang, Jiayi, et al.
Veröffentlicht: (2025)
LibreLog: Accurate and Efficient Unsupervised Log Parsing Using Open-Source Large Language Models
von: Ma, Zeyang, et al.
Veröffentlicht: (2024)
von: Ma, Zeyang, et al.
Veröffentlicht: (2024)
ZS4C: Zero-Shot Synthesis of Compilable Code for Incomplete Code Snippets using LLMs
von: Kabir, Azmain, et al.
Veröffentlicht: (2024)
von: Kabir, Azmain, et al.
Veröffentlicht: (2024)
Code Refactoring with LLM: A Comprehensive Evaluation With Few-Shot Settings
von: Tapader, Md. Raihan, et al.
Veröffentlicht: (2025)
von: Tapader, Md. Raihan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Multi-Agent Approach to Fault Localization via Graph-Based Retrieval and Reflexion
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024) -
CI-Repair-Bench: A Repository-Aware Benchmark for Automated Patch Validation via CI Workflows
von: Muna, Rabeya Khatun, et al.
Veröffentlicht: (2026) -
Back to the Future! Studying Data Cleanness in Defects4J and its Impact on Fault Localization
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2023) -
Order Matters! An Empirical Study on Large Language Models' Input Order Bias in Software Fault Localization
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024) -
Towards Better Graph Neural Network-based Fault Localization Through Enhanced Code Representation
von: Rafi, Md Nakhla, et al.
Veröffentlicht: (2024)