Why Do Multi-Agent LLM Systems Fail?
Fuente:
arXiv
Salvato in:
| Autori principali: | Cemri, Mert, Pan, Melissa Z., Yang, Shuyi, Agrawal, Lakshya A., Chopra, Bhavya, Tiwari, Rishabh, Keutzer, Kurt, Parameswaran, Aditya, Klein, Dan, Ramchandran, Kannan, Zaharia, Matei, Gonzalez, Joseph E., Stoica, Ion |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
$\texttt{SPECS}$: Faster Test-Time Scaling through Speculative Drafts
di: Cemri, Mert, et al.
Pubblicazione: (2025)
di: Cemri, Mert, et al.
Pubblicazione: (2025)
Learning, Fast and Slow: Towards LLMs That Adapt Continually
di: Tiwari, Rishabh, et al.
Pubblicazione: (2026)
di: Tiwari, Rishabh, et al.
Pubblicazione: (2026)
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
di: Liu, Shu, et al.
Pubblicazione: (2026)
di: Liu, Shu, et al.
Pubblicazione: (2026)
Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems
di: Agarwal, Shubham, et al.
Pubblicazione: (2026)
di: Agarwal, Shubham, et al.
Pubblicazione: (2026)
Barbarians at the Gate: How AI is Upending Systems Research
di: Cheng, Audrey, et al.
Pubblicazione: (2025)
di: Cheng, Audrey, et al.
Pubblicazione: (2025)
AdaEvolve: Adaptive LLM Driven Zeroth-Order Optimization
di: Cemri, Mert, et al.
Pubblicazione: (2026)
di: Cemri, Mert, et al.
Pubblicazione: (2026)
HashAttention: Semantic Sparsity for Faster Inference
di: Desai, Aditya, et al.
Pubblicazione: (2024)
di: Desai, Aditya, et al.
Pubblicazione: (2024)
EvoX: Meta-Evolution for Automated Discovery
di: Liu, Shu, et al.
Pubblicazione: (2026)
di: Liu, Shu, et al.
Pubblicazione: (2026)
Let the Barbarians In: How AI Can Accelerate Systems Performance Research
di: Cheng, Audrey, et al.
Pubblicazione: (2025)
di: Cheng, Audrey, et al.
Pubblicazione: (2025)
vAttention: Verified Sparse Attention
di: Desai, Aditya, et al.
Pubblicazione: (2025)
di: Desai, Aditya, et al.
Pubblicazione: (2025)
The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More
di: Chen, Lingjiao, et al.
Pubblicazione: (2026)
di: Chen, Lingjiao, et al.
Pubblicazione: (2026)
Networks of Networks: Complexity Class Principles Applied to Compound AI Systems Design
di: Davis, Jared Quincy, et al.
Pubblicazione: (2024)
di: Davis, Jared Quincy, et al.
Pubblicazione: (2024)
LangProBe: a Language Programs Benchmark
di: Tan, Shangyin, et al.
Pubblicazione: (2025)
di: Tan, Shangyin, et al.
Pubblicazione: (2025)
Semantic Data Processing with Holistic Data Understanding
di: Sun, Youran, et al.
Pubblicazione: (2026)
di: Sun, Youran, et al.
Pubblicazione: (2026)
optimize_anything: A Universal API for Optimizing any Text Parameter
di: Agrawal, Lakshya A, et al.
Pubblicazione: (2026)
di: Agrawal, Lakshya A, et al.
Pubblicazione: (2026)
DeepScholar-Bench: A Live Benchmark and Automated Evaluation for Generative Research Synthesis
di: Patel, Liana, et al.
Pubblicazione: (2025)
di: Patel, Liana, et al.
Pubblicazione: (2025)
Are More LLM Calls All You Need? Towards Scaling Laws of Compound Inference Systems
di: Chen, Lingjiao, et al.
Pubblicazione: (2024)
di: Chen, Lingjiao, et al.
Pubblicazione: (2024)
Optimizing Model Selection for Compound AI Systems
di: Chen, Lingjiao, et al.
Pubblicazione: (2025)
di: Chen, Lingjiao, et al.
Pubblicazione: (2025)
Rethinking Dataset Discovery with DataScout
di: Lin, Rachel, et al.
Pubblicazione: (2025)
di: Lin, Rachel, et al.
Pubblicazione: (2025)
RAFT: Adapting Language Model to Domain Specific RAG
di: Zhang, Tianjun, et al.
Pubblicazione: (2024)
di: Zhang, Tianjun, et al.
Pubblicazione: (2024)
AI-Driven Research for Databases
di: Cheng, Audrey, et al.
Pubblicazione: (2026)
di: Cheng, Audrey, et al.
Pubblicazione: (2026)
Some Present-Day Problems of Romanian Library Science
di: Stoica, Ion
Pubblicazione: (1973)
di: Stoica, Ion
Pubblicazione: (1973)
The Central University Library, Bucharest. Over Seventy-five Years in the History of a Collection
di: Stoica, Ion
Pubblicazione: (1972)
di: Stoica, Ion
Pubblicazione: (1972)
GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
di: Agrawal, Lakshya A, et al.
Pubblicazione: (2025)
di: Agrawal, Lakshya A, et al.
Pubblicazione: (2025)
Revisiting Cache Freshness for Emerging Real-Time Applications
di: Mao, Ziming, et al.
Pubblicazione: (2024)
di: Mao, Ziming, et al.
Pubblicazione: (2024)
PQC Validator: Validating Post-Quantum Readiness in Cloud-Native 5G Core Networks
di: Chopra, Lakshya, et al.
Pubblicazione: (2026)
di: Chopra, Lakshya, et al.
Pubblicazione: (2026)
Merkle Tree Certificate Post-Quantum PKI for Kubernetes and Cloud-Native 5G/B5G Core
di: Chopra, Lakshya, et al.
Pubblicazione: (2026)
di: Chopra, Lakshya, et al.
Pubblicazione: (2026)
Reward Under Attack: Analyzing the Robustness and Hackability of Process Reward Models
di: Tiwari, Rishabh, et al.
Pubblicazione: (2026)
di: Tiwari, Rishabh, et al.
Pubblicazione: (2026)
BARE: Leveraging Base Language Models for Few-Shot Synthetic Data Generation
di: Zhu, Alan, et al.
Pubblicazione: (2025)
di: Zhu, Alan, et al.
Pubblicazione: (2025)
Toward a Theory of Tokenization in LLMs
di: Rajaraman, Nived, et al.
Pubblicazione: (2024)
di: Rajaraman, Nived, et al.
Pubblicazione: (2024)
The Fair Value of Data Under Heterogeneous Privacy Constraints in Federated Learning
di: Kang, Justin, et al.
Pubblicazione: (2023)
di: Kang, Justin, et al.
Pubblicazione: (2023)
MoE-Lightning: High-Throughput MoE Inference on Memory-constrained GPUs
di: Cao, Shiyi, et al.
Pubblicazione: (2024)
di: Cao, Shiyi, et al.
Pubblicazione: (2024)
Composing Policy Gradients and Prompt Optimization for Language Model Programs
di: Ziems, Noah, et al.
Pubblicazione: (2025)
di: Ziems, Noah, et al.
Pubblicazione: (2025)
Supporting Our AI Overlords: Redesigning Data Systems to be Agent-First
di: Liu, Shu, et al.
Pubblicazione: (2025)
di: Liu, Shu, et al.
Pubblicazione: (2025)
SAGE: A Realistic Benchmark for Semantic Understanding
di: Goel, Samarth, et al.
Pubblicazione: (2025)
di: Goel, Samarth, et al.
Pubblicazione: (2025)
Quantifying Positional Biases in Text Embedding Models
di: Lee, Reagan J., et al.
Pubblicazione: (2024)
di: Lee, Reagan J., et al.
Pubblicazione: (2024)
Resilience Quantification and its Support for Operational Resilience
di: Matei, Ion, et al.
Pubblicazione: (2026)
di: Matei, Ion, et al.
Pubblicazione: (2026)
Steering Semantic Data Processing With DocWrangler
di: Shankar, Shreya, et al.
Pubblicazione: (2025)
di: Shankar, Shreya, et al.
Pubblicazione: (2025)
SIEVE: Sample-Efficient Parametric Learning from Natural Language
di: Asawa, Parth, et al.
Pubblicazione: (2026)
di: Asawa, Parth, et al.
Pubblicazione: (2026)
MITRA: A Large-Scale Parallel Corpus and Multilingual Pretrained Language Model for Machine Translation and Semantic Retrieval for Pāli, Sanskrit, Buddhist Chinese, and Tibetan
di: Nehrdich, Sebastian, et al.
Pubblicazione: (2026)
di: Nehrdich, Sebastian, et al.
Pubblicazione: (2026)
Documenti analoghi
-
$\texttt{SPECS}$: Faster Test-Time Scaling through Speculative Drafts
di: Cemri, Mert, et al.
Pubblicazione: (2025) -
Learning, Fast and Slow: Towards LLMs That Adapt Continually
di: Tiwari, Rishabh, et al.
Pubblicazione: (2026) -
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
di: Liu, Shu, et al.
Pubblicazione: (2026) -
Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems
di: Agarwal, Shubham, et al.
Pubblicazione: (2026) -
Barbarians at the Gate: How AI is Upending Systems Research
di: Cheng, Audrey, et al.
Pubblicazione: (2025)