MANTRA: Enhancing Automated Method-Level Refactoring with Contextual RAG and Multi-Agent LLM Collaboration
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Yisen, Lin, Feng, Yang, Jinqiu, Tse-Hsun, Chen, Tsantalis, Nikolaos |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SWE-Refactor: A Repository-Level Benchmark for Real-World LLM-Based Code Refactoring
by: Xu, Yisen, et al.
Published: (2026)
by: Xu, Yisen, et al.
Published: (2026)
A Novel Refactoring and Semantic Aware Abstract Syntax Tree Differencing Tool and a Benchmark for Evaluating the Accuracy of Diff Tools
by: Alikhanifard, Pouria, et al.
Published: (2024)
by: Alikhanifard, Pouria, et al.
Published: (2024)
Refactoring-aware Block Tracking in Commit History
by: Hasan, Mohammed Tayeeb, et al.
Published: (2024)
by: Hasan, Mohammed Tayeeb, et al.
Published: (2024)
An Empirical Study of Refactoring Engine Bugs
by: Wang, Haibo, et al.
Published: (2024)
by: Wang, Haibo, et al.
Published: (2024)
AI-Generated Smells: An Analysis of Code and Architecture in LLM and Agent-Driven Development
by: Zhu, Yuecai, et al.
Published: (2026)
by: Zhu, Yuecai, et al.
Published: (2026)
From Historical Patches to Repair Plans: Outcome-Conditioned Reasoning for Repository-Level Program Repair
by: Li, Chenglin, et al.
Published: (2026)
by: Li, Chenglin, et al.
Published: (2026)
Leveraging LLMs, IDEs, and Semantic Embeddings for Automated Move Method Refactoring
by: Bellur, Abhiram, et al.
Published: (2025)
by: Bellur, Abhiram, et al.
Published: (2025)
RobuNFR: Evaluating the Robustness of Large Language Models on Non-Functional Requirements Aware Code Generation
by: Lin, Feng, et al.
Published: (2025)
by: Lin, Feng, et al.
Published: (2025)
Evaluating Software Process Models for Multi-Agent Class-Level Code Generation
by: Shafin, Wasique Islam, et al.
Published: (2025)
by: Shafin, Wasique Islam, et al.
Published: (2025)
Towards Structured, State-Aware, and Execution-Grounded Reasoning for Software Engineering Agents
by: Tse-Hsun, et al.
Published: (2026)
by: Tse-Hsun, et al.
Published: (2026)
Are Benchmark Tests Strong Enough? Mutation-Guided Diagnosis and Augmentation of Regression Suites
by: Li, Chenglin, et al.
Published: (2026)
by: Li, Chenglin, et al.
Published: (2026)
Evaluating the Effectiveness and Efficiency of Demonstration Retrievers in RAG for Coding Tasks
by: He, Pengfei, et al.
Published: (2024)
by: He, Pengfei, et al.
Published: (2024)
HEJ-Robust: A Robustness Benchmark for LLM-Based Automated Program Repair
by: Rabbi, Fazle, et al.
Published: (2026)
by: Rabbi, Fazle, et al.
Published: (2026)
SBEST: Spectrum-Based Fault Localization Without Fault-Triggering Tests
by: Rafi, Md Nakhla, et al.
Published: (2024)
by: Rafi, Md Nakhla, et al.
Published: (2024)
Automated Unit Test Refactoring
by: Gao, Yi, et al.
Published: (2024)
by: Gao, Yi, et al.
Published: (2024)
Identifying Performance-Sensitive Configurations in Software Systems through Code Analysis with LLM Agents
by: Wang, Zehao, et al.
Published: (2024)
by: Wang, Zehao, et al.
Published: (2024)
LLM-based Multi-Agent System for Intelligent Refactoring of Haskell Code
by: Siddeeq, Shahbaz, et al.
Published: (2025)
by: Siddeeq, Shahbaz, et al.
Published: (2025)
A Multi-Agent Approach to Fault Localization via Graph-Based Retrieval and Reflexion
by: Rafi, Md Nakhla, et al.
Published: (2024)
by: Rafi, Md Nakhla, et al.
Published: (2024)
Empowering AIOps: Leveraging Large Language Models for IT Operations Management
by: Vitui, Arthur, et al.
Published: (2025)
by: Vitui, Arthur, et al.
Published: (2025)
On Rank Aggregating Test Prioritizations
by: Mondal, Shouvick, et al.
Published: (2024)
by: Mondal, Shouvick, et al.
Published: (2024)
CI-Repair-Bench: A Repository-Aware Benchmark for Automated Patch Validation via CI Workflows
by: Muna, Rabeya Khatun, et al.
Published: (2026)
by: Muna, Rabeya Khatun, et al.
Published: (2026)
A Survey of Code Review Benchmarks and Evaluation Practices in Pre-LLM and LLM Era
by: Khan, Taufiqul Islam, et al.
Published: (2026)
by: Khan, Taufiqul Islam, et al.
Published: (2026)
A Multi-Language Perspective on the Robustness of LLM Code Generation
by: Rabbi, Fazle, et al.
Published: (2025)
by: Rabbi, Fazle, et al.
Published: (2025)
PopSweeper: Automatically Detecting and Resolving App-Blocking Pop-Ups to Assist Automated Mobile GUI Testing
by: Guo, Linqiang, et al.
Published: (2024)
by: Guo, Linqiang, et al.
Published: (2024)
"Refactoring Runaway": Understanding and Mitigating Tangled Refactorings in Coding Agents for Issue Resolution
by: Tian, Zhao, et al.
Published: (2026)
by: Tian, Zhao, et al.
Published: (2026)
Multi-Agent Coordinated Rename Refactoring
by: Bellur, Abhiram, et al.
Published: (2026)
by: Bellur, Abhiram, et al.
Published: (2026)
Automated Extract Method Refactoring with Open-Source LLMs: A Comparative Study
by: Chand, Sivajeet, et al.
Published: (2025)
by: Chand, Sivajeet, et al.
Published: (2025)
CODEPROMPTZIP: Code-specific Prompt Compression for Retrieval-Augmented Generation in Coding Tasks with LMs
by: He, Pengfei, et al.
Published: (2025)
by: He, Pengfei, et al.
Published: (2025)
SLICET5: Static Program Slicing using Language Models with Copy Mechanism and Constrained Decoding
by: He, Pengfei, et al.
Published: (2025)
by: He, Pengfei, et al.
Published: (2025)
GUIWatcher: Automatically Detecting GUI Lags by Analyzing Mobile Application Screencasts
by: Liu, Wei, et al.
Published: (2025)
by: Liu, Wei, et al.
Published: (2025)
Studying and Benchmarking Large Language Models For Log Level Suggestion
by: Heng, Yi Wen, et al.
Published: (2024)
by: Heng, Yi Wen, et al.
Published: (2024)
Bias Unveiled: Investigating Social Bias in LLM-Generated Code
by: Ling, Lin, et al.
Published: (2024)
by: Ling, Lin, et al.
Published: (2024)
Investigating Student Reasoning in Method-Level Code Refactoring: A Think-Aloud Study
by: Oliveira, Eduardo Carneiro, et al.
Published: (2024)
by: Oliveira, Eduardo Carneiro, et al.
Published: (2024)
MobileUPReg: Identifying User-Perceived Performance Regressions in Mobile OS Versions
by: Liu, Wei, et al.
Published: (2025)
by: Liu, Wei, et al.
Published: (2025)
An Empirical Study on the Potential of LLMs in Automated Software Refactoring
by: Liu, Bo, et al.
Published: (2024)
by: Liu, Bo, et al.
Published: (2024)
RefAgent: A Multi-agent LLM-based Framework for Automatic Software Refactoring
by: Oueslati, Khouloud, et al.
Published: (2025)
by: Oueslati, Khouloud, et al.
Published: (2025)
An Empirical Study on the Characteristics of Database Access Bugs in Java Applications
by: Liu, Wei, et al.
Published: (2024)
by: Liu, Wei, et al.
Published: (2024)
Testing Refactoring Engine via Historical Bug Report driven LLM
by: Wang, Haibo, et al.
Published: (2025)
by: Wang, Haibo, et al.
Published: (2025)
MANTRA: a Framework for Multi-stage Adaptive Noise TReAtment During Training
by: Zhao, Zixiao, et al.
Published: (2025)
by: Zhao, Zixiao, et al.
Published: (2025)
How do Agents Refactor: An Empirical Study
by: Ottenhof, Lukas, et al.
Published: (2026)
by: Ottenhof, Lukas, et al.
Published: (2026)
Similar Items
-
SWE-Refactor: A Repository-Level Benchmark for Real-World LLM-Based Code Refactoring
by: Xu, Yisen, et al.
Published: (2026) -
A Novel Refactoring and Semantic Aware Abstract Syntax Tree Differencing Tool and a Benchmark for Evaluating the Accuracy of Diff Tools
by: Alikhanifard, Pouria, et al.
Published: (2024) -
Refactoring-aware Block Tracking in Commit History
by: Hasan, Mohammed Tayeeb, et al.
Published: (2024) -
An Empirical Study of Refactoring Engine Bugs
by: Wang, Haibo, et al.
Published: (2024) -
AI-Generated Smells: An Analysis of Code and Architecture in LLM and Agent-Driven Development
by: Zhu, Yuecai, et al.
Published: (2026)