CascadeDebate: Multi-Agent Deliberation for Cost-Aware LLM Cascades
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chang, Raeyoung, Kwon, Dongwook, Lee, Jisoo, Verma, Nikhil |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GEMMAS: Graph-based Evaluation Metrics for Multi Agent Systems
von: Lee, Jisoo, et al.
Veröffentlicht: (2025)
von: Lee, Jisoo, et al.
Veröffentlicht: (2025)
Collaborate, Deliberate, Evaluate: How LLM Alignment Affects Coordinated Multi-Agent Outcomes
von: Nath, Abhijnan, et al.
Veröffentlicht: (2025)
von: Nath, Abhijnan, et al.
Veröffentlicht: (2025)
Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to Deliberation
von: Zhang, Zhiwei, et al.
Veröffentlicht: (2025)
von: Zhang, Zhiwei, et al.
Veröffentlicht: (2025)
Nemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation
von: Yang, Zhuolin, et al.
Veröffentlicht: (2026)
von: Yang, Zhuolin, et al.
Veröffentlicht: (2026)
Cascaded Language Models for Cost-effective Human-AI Decision-Making
von: Fanconi, Claudio, et al.
Veröffentlicht: (2025)
von: Fanconi, Claudio, et al.
Veröffentlicht: (2025)
DeliberationBench: When Do More Voices Hurt? A Controlled Study of Multi-LLM Deliberation Protocols
von: Kaushal, Vaarunay, et al.
Veröffentlicht: (2025)
von: Kaushal, Vaarunay, et al.
Veröffentlicht: (2025)
Calibrate-Then-Act: Cost-Aware Exploration in LLM Agents
von: Ding, Wenxuan, et al.
Veröffentlicht: (2026)
von: Ding, Wenxuan, et al.
Veröffentlicht: (2026)
DuplexCascade: Full-Duplex Speech-to-Speech Dialogue with VAD-Free Cascaded ASR-LLM-TTS Pipeline and Micro-Turn Optimization
von: Yang, Jianing, et al.
Veröffentlicht: (2026)
von: Yang, Jianing, et al.
Veröffentlicht: (2026)
Quality-Aware Translation Tagging in Multilingual RAG system
von: Moon, Hoyeon, et al.
Veröffentlicht: (2025)
von: Moon, Hoyeon, et al.
Veröffentlicht: (2025)
Nemotron-Cascade: Scaling Cascaded Reinforcement Learning for General-Purpose Reasoning Models
von: Wang, Boxin, et al.
Veröffentlicht: (2025)
von: Wang, Boxin, et al.
Veröffentlicht: (2025)
Designing Memory-Augmented AR Agents for Spatiotemporal Reasoning in Personalized Task Assistance
von: Choi, Dongwook, et al.
Veröffentlicht: (2025)
von: Choi, Dongwook, et al.
Veröffentlicht: (2025)
From Deferral to Learning: Online In-Context Knowledge Distillation for LLM Cascades
von: Wu, Yu, et al.
Veröffentlicht: (2025)
von: Wu, Yu, et al.
Veröffentlicht: (2025)
Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs
von: Anantaprayoon, Panatchakorn, et al.
Veröffentlicht: (2026)
von: Anantaprayoon, Panatchakorn, et al.
Veröffentlicht: (2026)
ACIArena: Toward Unified Evaluation for Agent Cascading Injection
von: An, Hengyu, et al.
Veröffentlicht: (2026)
von: An, Hengyu, et al.
Veröffentlicht: (2026)
Achieving Unanimous Consensus Through Multi-Agent Deliberation
von: Pokharel, Apurba, et al.
Veröffentlicht: (2025)
von: Pokharel, Apurba, et al.
Veröffentlicht: (2025)
Bayesian Orchestration of Multi-LLM Agents for Cost-Aware Sequential Decision-Making
von: Amin, Danial
Veröffentlicht: (2026)
von: Amin, Danial
Veröffentlicht: (2026)
Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning
von: Yue, Murong, et al.
Veröffentlicht: (2023)
von: Yue, Murong, et al.
Veröffentlicht: (2023)
From Debate to Deliberation: Structured Collective Reasoning with Typed Epistemic Acts
von: Prakash, Sunil
Veröffentlicht: (2026)
von: Prakash, Sunil
Veröffentlicht: (2026)
Is Escalation Worth It? A Decision-Theoretic Characterization of LLM Cascades
von: Bouchard, Dylan
Veröffentlicht: (2026)
von: Bouchard, Dylan
Veröffentlicht: (2026)
DIAMOND: An LLM-Driven Agent for Context-Aware Baseball Highlight Summarization
von: Kang, Jeonghun, et al.
Veröffentlicht: (2025)
von: Kang, Jeonghun, et al.
Veröffentlicht: (2025)
QUIET: A Multi-Blank Cascaded Story Cloze Benchmark for LLM Creative Generation Capability
von: Zou, Bo, et al.
Veröffentlicht: (2026)
von: Zou, Bo, et al.
Veröffentlicht: (2026)
Multiple LLM Agents Debate for Equitable Cultural Alignment
von: Ki, Dayeon, et al.
Veröffentlicht: (2025)
von: Ki, Dayeon, et al.
Veröffentlicht: (2025)
GroupDebate: Enhancing the Efficiency of Multi-Agent Debate Using Group Discussion
von: Liu, Tongxuan, et al.
Veröffentlicht: (2024)
von: Liu, Tongxuan, et al.
Veröffentlicht: (2024)
Dynamic Role Assignment for Multi-Agent Debate
von: Zhang, Miao, et al.
Veröffentlicht: (2026)
von: Zhang, Miao, et al.
Veröffentlicht: (2026)
Combating Adversarial Attacks with Multi-Agent Debate
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
von: Chern, Steffi, et al.
Veröffentlicht: (2024)
Translate Smart, not Hard: Cascaded Translation Systems with Quality-Aware Deferral
von: Farinhas, António, et al.
Veröffentlicht: (2025)
von: Farinhas, António, et al.
Veröffentlicht: (2025)
Efficient Contextual LLM Cascades through Budget-Constrained Policy Learning
von: Zhang, Xuechen, et al.
Veröffentlicht: (2024)
von: Zhang, Xuechen, et al.
Veröffentlicht: (2024)
MEMOREPAIR: Barrier-First Cascade Repair in Agentic Memory
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
von: Zhao, Yang, et al.
Veröffentlicht: (2026)
Training-Free Exponential Context Extension via Cascading KV Cache
von: Willette, Jeffrey, et al.
Veröffentlicht: (2024)
von: Willette, Jeffrey, et al.
Veröffentlicht: (2024)
MTRouter: Cost-Aware Multi-Turn LLM Routing with History-Model Joint Embeddings
von: Zhang, Yiqun, et al.
Veröffentlicht: (2026)
von: Zhang, Yiqun, et al.
Veröffentlicht: (2026)
iMAD: Intelligent Multi-Agent Debate for Efficient and Accurate LLM Inference
von: Fan, Wei, et al.
Veröffentlicht: (2025)
von: Fan, Wei, et al.
Veröffentlicht: (2025)
Cascading Large Language Models for Salient Event Graph Generation
von: Tan, Xingwei, et al.
Veröffentlicht: (2024)
von: Tan, Xingwei, et al.
Veröffentlicht: (2024)
Cascaded Self-Evaluation Augmented Training for Lightweight Multimodal LLMs
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents
von: Liu, Jiayu, et al.
Veröffentlicht: (2025)
von: Liu, Jiayu, et al.
Veröffentlicht: (2025)
SMaRT: Select, Mix, and ReinvenT -- A Strategy Fusion Framework for LLM-Driven Reasoning and Planning
von: Verma, Nikhil, et al.
Veröffentlicht: (2025)
von: Verma, Nikhil, et al.
Veröffentlicht: (2025)
MADIAVE: Multi-Agent Debate for Implicit Attribute Value Extraction
von: Huang, Wei-Chieh, et al.
Veröffentlicht: (2025)
von: Huang, Wei-Chieh, et al.
Veröffentlicht: (2025)
Faster Cascades via Speculative Decoding
von: Narasimhan, Harikrishna, et al.
Veröffentlicht: (2024)
von: Narasimhan, Harikrishna, et al.
Veröffentlicht: (2024)
Completing Missing Annotation: Multi-Agent Debate for Accurate and Scalable Relevant Assessment for IR Benchmarks
von: Ban, Minjeong, et al.
Veröffentlicht: (2026)
von: Ban, Minjeong, et al.
Veröffentlicht: (2026)
Evaluating Behavioral Alignment in Conflict Dialogue: A Multi-Dimensional Comparison of LLM Agents and Humans
von: Kwon, Deuksin, et al.
Veröffentlicht: (2025)
von: Kwon, Deuksin, et al.
Veröffentlicht: (2025)
Case-Aware LLM-as-a-Judge Evaluation for Enterprise-Scale RAG Systems
von: Chhabra, Mukul, et al.
Veröffentlicht: (2026)
von: Chhabra, Mukul, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
GEMMAS: Graph-based Evaluation Metrics for Multi Agent Systems
von: Lee, Jisoo, et al.
Veröffentlicht: (2025) -
Collaborate, Deliberate, Evaluate: How LLM Alignment Affects Coordinated Multi-Agent Outcomes
von: Nath, Abhijnan, et al.
Veröffentlicht: (2025) -
Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to Deliberation
von: Zhang, Zhiwei, et al.
Veröffentlicht: (2025) -
Nemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation
von: Yang, Zhuolin, et al.
Veröffentlicht: (2026) -
Cascaded Language Models for Cost-effective Human-AI Decision-Making
von: Fanconi, Claudio, et al.
Veröffentlicht: (2025)