Is Multi-Agent Debate (MAD) the Silver Bullet? An Empirical Analysis of MAD in Code Summarization and Translation
Fuente:
arXiv
Saved in:
| Main Authors: | Chun, Jina, Chen, Qihong, Li, Jiawei, Ahmed, Iftekhar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Empirical Study on Automatically Detecting AI-Generated Source Code: How Far Are We?
by: Suh, Hyunjae, et al.
Published: (2024)
by: Suh, Hyunjae, et al.
Published: (2024)
Does the Order of Fine-tuning Matter and Why?
by: Chen, Qihong, et al.
Published: (2024)
by: Chen, Qihong, et al.
Published: (2024)
Enhancing Code Generation for Low-Resource Languages: No Silver Bullet
by: Giagnorio, Alessandro, et al.
Published: (2025)
by: Giagnorio, Alessandro, et al.
Published: (2025)
Which Prompting Technique Should I Use? An Empirical Investigation of Prompting Techniques for Software Engineering Tasks
by: Santana Jr, E. G., et al.
Published: (2025)
by: Santana Jr, E. G., et al.
Published: (2025)
Are Decoder-Only Large Language Models the Silver Bullet for Code Search?
by: Chen, Yuxuan, et al.
Published: (2024)
by: Chen, Yuxuan, et al.
Published: (2024)
Consider What Humans Consider: Optimizing Commit Message Leveraging Contexts Considered By Human
by: Li, Jiawei, et al.
Published: (2025)
by: Li, Jiawei, et al.
Published: (2025)
Optimization is Better than Generation: Optimizing Commit Message Leveraging Human-written Commit Message
by: Li, Jiawei, et al.
Published: (2025)
by: Li, Jiawei, et al.
Published: (2025)
LLMs Are Not a Silver Bullet: A Case Study on Software Fairness
by: Li, Xinyue, et al.
Published: (2026)
by: Li, Xinyue, et al.
Published: (2026)
Does Documentation Matter? An Empirical Study of Practitioners' Perspective on Open-Source Software Adoption
by: Imani, Aaron, et al.
Published: (2024)
by: Imani, Aaron, et al.
Published: (2024)
Investigating the Impact of Code Comment Inconsistency on Bug Introducing
by: Radmanesh, Shiva, et al.
Published: (2024)
by: Radmanesh, Shiva, et al.
Published: (2024)
GenAI Is No Silver Bullet for Qualitative Research in Software Engineering
by: Ernst, Neil A., et al.
Published: (2026)
by: Ernst, Neil A., et al.
Published: (2026)
Human or LLM? A Comparative Study on Accessible Code Generation Capability
by: Suh, Hyunjae, et al.
Published: (2025)
by: Suh, Hyunjae, et al.
Published: (2025)
No Silver Bullets: Why Understanding Software Cycle Time is Messy, Not Magic
by: Flournoy, John C., et al.
Published: (2025)
by: Flournoy, John C., et al.
Published: (2025)
A Deep Dive Into Large Language Model Code Generation Mistakes: What and Why?
by: Chen, QiHong, et al.
Published: (2024)
by: Chen, QiHong, et al.
Published: (2024)
Is Self-Repair a Silver Bullet for Code Generation?
by: Olausson, Theo X., et al.
Published: (2023)
by: Olausson, Theo X., et al.
Published: (2023)
Code vs Serialized AST Inputs for LLM-Based Code Summarization: An Empirical Study
by: Dong, Shijia, et al.
Published: (2026)
by: Dong, Shijia, et al.
Published: (2026)
Do AI Coding Agents Log Like Humans? An Empirical Study
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
Prompting in the Wild: An Empirical Study of Prompt Evolution in Software Repositories
by: Tafreshipour, Mahan, et al.
Published: (2024)
by: Tafreshipour, Mahan, et al.
Published: (2024)
LegacyTranslate: LLM-based Multi-Agent Method for Legacy Code Translation
by: Moti, Zahra, et al.
Published: (2026)
by: Moti, Zahra, et al.
Published: (2026)
From Bias To Improved Prompts: A Case Study of Bias Mitigation of Clone Detection Models
by: Chen, QiHong, et al.
Published: (2025)
by: Chen, QiHong, et al.
Published: (2025)
Agent READMEs: An Empirical Study of Context Files for Agentic Coding
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
by: Chatlatanagulchai, Worawalan, et al.
Published: (2025)
Inside Out: Uncovering How Comment Internalization Steers LLMs for Better or Worse
by: Imani, Aaron, et al.
Published: (2025)
by: Imani, Aaron, et al.
Published: (2025)
Context Conquers Parameters: Outperforming Proprietary LLM in Commit Message Generation
by: Imani, Aaron, et al.
Published: (2024)
by: Imani, Aaron, et al.
Published: (2024)
I Can't Share Code, but I need Translation -- An Empirical Study on Code Translation through Federated LLM
by: Kumar, Jahnavi, et al.
Published: (2025)
by: Kumar, Jahnavi, et al.
Published: (2025)
RepoTransAgent: Multi-Agent LLM Framework for Repository-Aware Code Translation
by: Guan, Ziqi, et al.
Published: (2025)
by: Guan, Ziqi, et al.
Published: (2025)
Beyond Self-learned Attention: Mitigating Attention Bias in Transformer-based Models Using Attention Guidance
by: Gesi, Jiri, et al.
Published: (2024)
by: Gesi, Jiri, et al.
Published: (2024)
Test Smell: A Parasitic Energy Consumer in Software Testing
by: Misu, Md Rakib Hossain, et al.
Published: (2023)
by: Misu, Md Rakib Hossain, et al.
Published: (2023)
Examining LLMs Ability to Summarize Code Through Mutation-Analysis
by: Khatib, Lara, et al.
Published: (2026)
by: Khatib, Lara, et al.
Published: (2026)
Calibration of Large Language Models on Code Summarization
by: Virk, Yuvraj, et al.
Published: (2024)
by: Virk, Yuvraj, et al.
Published: (2024)
Agentic Refactoring: An Empirical Study of AI Coding Agents
by: Horikawa, Kosei, et al.
Published: (2025)
by: Horikawa, Kosei, et al.
Published: (2025)
Analysis on LLMs Performance for Code Summarization
by: Akib, Md. Ahnaf, et al.
Published: (2024)
by: Akib, Md. Ahnaf, et al.
Published: (2024)
Resource-Efficient & Effective Code Summarization
by: Afrin, Saima, et al.
Published: (2025)
by: Afrin, Saima, et al.
Published: (2025)
Are Coding Agents Generating Over-Mocked Tests? An Empirical Study
by: Hora, Andre, et al.
Published: (2026)
by: Hora, Andre, et al.
Published: (2026)
From Industry Claims to Empirical Reality: An Empirical Study of Code Review Agents in Pull Requests
by: Chowdhury, Kowshik, et al.
Published: (2026)
by: Chowdhury, Kowshik, et al.
Published: (2026)
Optimizing Datasets for Code Summarization: Is Code-Comment Coherence Enough?
by: Vitale, Antonio, et al.
Published: (2025)
by: Vitale, Antonio, et al.
Published: (2025)
Using AI-Based Coding Assistants in Practice: State of Affairs, Perceptions, and Ways Forward
by: Sergeyuk, Agnia, et al.
Published: (2024)
by: Sergeyuk, Agnia, et al.
Published: (2024)
What Makes a Great Software Quality Assurance Engineer?
by: Farias, Roselane Silva, et al.
Published: (2024)
by: Farias, Roselane Silva, et al.
Published: (2024)
On the Use of Agentic Coding: An Empirical Study of Pull Requests on GitHub
by: Watanabe, Miku, et al.
Published: (2025)
by: Watanabe, Miku, et al.
Published: (2025)
Evaluating LLMs Effectiveness in Detecting and Correcting Test Smells: An Empirical Study
by: Santana Jr, E. G., et al.
Published: (2025)
by: Santana Jr, E. G., et al.
Published: (2025)
Understanding LLM-Centric Challenges for Deep Learning Frameworks: An Empirical Analysis
by: Mu, Yanzhou, et al.
Published: (2025)
by: Mu, Yanzhou, et al.
Published: (2025)
Similar Items
-
An Empirical Study on Automatically Detecting AI-Generated Source Code: How Far Are We?
by: Suh, Hyunjae, et al.
Published: (2024) -
Does the Order of Fine-tuning Matter and Why?
by: Chen, Qihong, et al.
Published: (2024) -
Enhancing Code Generation for Low-Resource Languages: No Silver Bullet
by: Giagnorio, Alessandro, et al.
Published: (2025) -
Which Prompting Technique Should I Use? An Empirical Investigation of Prompting Techniques for Software Engineering Tasks
by: Santana Jr, E. G., et al.
Published: (2025) -
Are Decoder-Only Large Language Models the Silver Bullet for Code Search?
by: Chen, Yuxuan, et al.
Published: (2024)