Multi-Agent Debate for LLM Judges with Adaptive Stability Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Tianyu, Tan, Zhen, Wang, Song, Qu, Huaizhi, Chen, Tianlong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DOGe: Defensive Output Generation for LLM Protection Against Knowledge Distillation
by: Li, Pingzhi, et al.
Published: (2025)
by: Li, Pingzhi, et al.
Published: (2025)
AnyMAC: Cascading Flexible Multi-Agent Collaboration via Next-Agent Prediction
by: Wang, Song, et al.
Published: (2025)
by: Wang, Song, et al.
Published: (2025)
Beyond Redundancy: Diverse and Specialized Multi-Expert Sparse Autoencoder
by: Xu, Zhen, et al.
Published: (2025)
by: Xu, Zhen, et al.
Published: (2025)
$\textit{Agents Under Siege}$: Breaking Pragmatic Multi-Agent LLM Systems with Optimized Prompt Attacks
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025)
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025)
Adaptive Theory of Mind for LLM-based Multi-Agent Coordination
by: Mu, Chunjiang, et al.
Published: (2026)
by: Mu, Chunjiang, et al.
Published: (2026)
Dynamic Mixed-Precision Routing for Efficient Multi-step LLM Interaction
by: Li, Yuanzhe, et al.
Published: (2026)
by: Li, Yuanzhe, et al.
Published: (2026)
Who's Your Judge? On the Detectability of LLM-Generated Judgments
by: Li, Dawei, et al.
Published: (2025)
by: Li, Dawei, et al.
Published: (2025)
Beyond Detection: Exploring Evidence-based Multi-Agent Debate for Misinformation Intervention and Persuasion
by: Han, Chen, et al.
Published: (2025)
by: Han, Chen, et al.
Published: (2025)
Tuning-Free Accountable Intervention for LLM Deployment -- A Metacognitive Approach
by: Tan, Zhen, et al.
Published: (2024)
by: Tan, Zhen, et al.
Published: (2024)
Detecting Unfaithful Chain-of-Thought via Circuit-Guided Internal-External Discrepancy
by: Shen, Xu, et al.
Published: (2026)
by: Shen, Xu, et al.
Published: (2026)
PETS: A Principled Framework Towards Optimal Trajectory Allocation for Efficient Test-Time Self-Consistency
by: Liu, Zhangyi, et al.
Published: (2026)
by: Liu, Zhangyi, et al.
Published: (2026)
Exploring Health Misinformation Detection with Multi-Agent Debate
by: Chen, Chih-Han, et al.
Published: (2025)
by: Chen, Chih-Han, et al.
Published: (2025)
Graph-of-Agents: A Graph-based Framework for Multi-Agent LLM Collaboration
by: Yun, Sukwon, et al.
Published: (2026)
by: Yun, Sukwon, et al.
Published: (2026)
Combating Adversarial Attacks with Multi-Agent Debate
by: Chern, Steffi, et al.
Published: (2024)
by: Chern, Steffi, et al.
Published: (2024)
TrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them
by: Wang, Yidong, et al.
Published: (2025)
by: Wang, Yidong, et al.
Published: (2025)
CortexDebate: Debating Sparsely and Equally for Multi-Agent Debate
by: Sun, Yiliu, et al.
Published: (2025)
by: Sun, Yiliu, et al.
Published: (2025)
DynaDebate: Breaking Homogeneity in Multi-Agent Debate with Dynamic Path Generation
by: Li, Zhenghao, et al.
Published: (2026)
by: Li, Zhenghao, et al.
Published: (2026)
Gradientsys: A Multi-Agent LLM Scheduler with ReAct Orchestration
by: Song, Xinyuan, et al.
Published: (2025)
by: Song, Xinyuan, et al.
Published: (2025)
Auditing Multi-Agent LLM Reasoning Trees Outperforms Majority Vote and LLM-as-Judge
by: Yang, Wei, et al.
Published: (2026)
by: Yang, Wei, et al.
Published: (2026)
MAD-Sherlock: Multi-Agent Debate for Visual Misinformation Detection
by: Lakara, Kumud, et al.
Published: (2024)
by: Lakara, Kumud, et al.
Published: (2024)
Metacognitive Self-Correction for Multi-Agent System via Prototype-Guided Next-Execution Reconstruction
by: Shen, Xu, et al.
Published: (2025)
by: Shen, Xu, et al.
Published: (2025)
OxyGent: Making Multi-Agent Systems Modular, Observable, and Evolvable via Oxy Abstraction
by: Hu, Junxing, et al.
Published: (2026)
by: Hu, Junxing, et al.
Published: (2026)
DebFlow: Automating Agent Creation via Agent Debate
by: Su, Jinwei, et al.
Published: (2025)
by: Su, Jinwei, et al.
Published: (2025)
CollabEval: Enhancing LLM-as-a-Judge via Multi-Agent Collaboration
by: Qian, Yiyue, et al.
Published: (2026)
by: Qian, Yiyue, et al.
Published: (2026)
RUMAD: Reinforcement-Unifying Multi-Agent Debate
by: Wang, Chao, et al.
Published: (2026)
by: Wang, Chao, et al.
Published: (2026)
SensingAgents: A Multi-Agent Collaborative Framework for Robust IMU Activity Recognition
by: Zheng, Naiyu, et al.
Published: (2026)
by: Zheng, Naiyu, et al.
Published: (2026)
GraphRCG: Self-Conditioned Graph Generation
by: Wang, Song, et al.
Published: (2024)
by: Wang, Song, et al.
Published: (2024)
Multi-Agent Debate: A Unified Agentic Framework for Tabular Anomaly Detection
by: Wang, Pinqiao, et al.
Published: (2026)
by: Wang, Pinqiao, et al.
Published: (2026)
Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning
by: Chen, Xuhang, et al.
Published: (2025)
by: Chen, Xuhang, et al.
Published: (2025)
OR-R1: Automating Modeling and Solving of Operations Research Optimization Problem via Test-Time Reinforcement Learning
by: Ding, Zezhen, et al.
Published: (2025)
by: Ding, Zezhen, et al.
Published: (2025)
Contextualization Distillation from Large Language Model for Knowledge Graph Completion
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
Leveraging LLMs as Meta-Judges: A Multi-Agent Framework for Evaluating LLM Judgments
by: Li, Yuran, et al.
Published: (2025)
by: Li, Yuran, et al.
Published: (2025)
CascadeDebate: Multi-Agent Deliberation for Cost-Aware LLM Cascades
by: Chang, Raeyoung, et al.
Published: (2026)
by: Chang, Raeyoung, et al.
Published: (2026)
MV-Debate: Multi-view Agent Debate with Dynamic Reflection Gating for Multimodal Harmful Content Detection in Social Media
by: Lu, Rui, et al.
Published: (2025)
by: Lu, Rui, et al.
Published: (2025)
From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents
by: Tan, Haoran, et al.
Published: (2026)
by: Tan, Haoran, et al.
Published: (2026)
Cross-Modal Memory Compression for Efficient Multi-Agent Debate
by: Wu, Jing, et al.
Published: (2026)
by: Wu, Jing, et al.
Published: (2026)
Judging with Many Minds: Do More Perspectives Mean Less Prejudice? On Bias Amplifications and Resistance in Multi-Agent Based LLM-as-Judge
by: Ma, Chiyu, et al.
Published: (2025)
by: Ma, Chiyu, et al.
Published: (2025)
REAL: Regression-Aware Reinforcement Learning for LLM-as-a-Judge
by: Zhang, Yasi, et al.
Published: (2026)
by: Zhang, Yasi, et al.
Published: (2026)
AgentRec: Next-Generation LLM-Powered Multi-Agent Collaborative Recommendation with Adaptive Intelligence
by: Ma, Bo, et al.
Published: (2025)
by: Ma, Bo, et al.
Published: (2025)
MAPLE: Many-Shot Adaptive Pseudo-Labeling for In-Context Learning
by: Chen, Zihan, et al.
Published: (2025)
by: Chen, Zihan, et al.
Published: (2025)
Similar Items
-
DOGe: Defensive Output Generation for LLM Protection Against Knowledge Distillation
by: Li, Pingzhi, et al.
Published: (2025) -
AnyMAC: Cascading Flexible Multi-Agent Collaboration via Next-Agent Prediction
by: Wang, Song, et al.
Published: (2025) -
Beyond Redundancy: Diverse and Specialized Multi-Expert Sparse Autoencoder
by: Xu, Zhen, et al.
Published: (2025) -
$\textit{Agents Under Siege}$: Breaking Pragmatic Multi-Agent LLM Systems with Optimized Prompt Attacks
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025) -
Adaptive Theory of Mind for LLM-based Multi-Agent Coordination
by: Mu, Chunjiang, et al.
Published: (2026)