POIROT: Interrogating Agents for Failure Detection in Multi-Agent Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Varela, Iñaki Dellibarda, Sendra-Arranz, R., Romero-Sorozabal, Pablo, Valverde-García, J. M., Laudanski, Annemarie F., Gutiérrez, Álvaro, Rocon, Eduardo, Cebrian, Manuel |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking the Illusion of Thinking
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
Sensorimotor Self-Recognition in Multimodal Large Language Model-Driven Robots
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
The Hessian correspondence of hypersurfaces of degree 3 and 4
by: Sendra-Arranz, Javier
Published: (2023)
by: Sendra-Arranz, Javier
Published: (2023)
Questões para o trabalho profissional do Assistente Social no processo transexualizador
by: Pablo Cardozo Rocon
Published: (2018)
by: Pablo Cardozo Rocon
Published: (2018)
ACESSO À SAÚDE PELA POPULAÇÃO TRANS NO BRASIL: NAS ENTRELINHAS DA REVISÃO INTEGRATIVA
by: Pablo Cardozo Rocon
Published: (2020)
by: Pablo Cardozo Rocon
Published: (2020)
O que esperam pessoas trans do Sistema Único de Saúde?
by: Pablo Cardozo Rocon
Published: (2018)
by: Pablo Cardozo Rocon
Published: (2018)
Acesso de mulheres bissexuais e lésbicas em serviços públicos de saúde
by: Pablo Cardozo Rocon
Published: (2024)
by: Pablo Cardozo Rocon
Published: (2024)
Regulamentação da vida no processo transexualizador brasileiro: uma análise sobre a política pública
by: Pablo Cardozo Rocon
Published: (2016)
by: Pablo Cardozo Rocon
Published: (2016)
Aprendizagens com signos trans - uma transetopoiese disruptiva
by: Pablo Cardozo Rocon
Published: (2020)
by: Pablo Cardozo Rocon
Published: (2020)
Game theory of undirected graphical models
by: Portakal, Irem, et al.
Published: (2024)
by: Portakal, Irem, et al.
Published: (2024)
Desarrollo de un modelo de ajuste por el riesgo para el infarto agudo de miocardio en España: comparación con el modelo de Charlson y el modelo ices. Aplicaciones para medir resultados asistenciales
by: Juan Manuel Sendra Gutiérrez
Published: (2006)
by: Juan Manuel Sendra Gutiérrez
Published: (2006)
Tabaquismo y trastorno mental grave: conceptualización, abordaje teórico y estudios de intervención.
by: Juan Manuel Sendra-Gutiérrez
Published: (2012)
by: Juan Manuel Sendra-Gutiérrez
Published: (2012)
PICon: A Multi-Turn Interrogation Framework for Evaluating Persona Agent Consistency
by: Kim, Minseo, et al.
Published: (2026)
by: Kim, Minseo, et al.
Published: (2026)
Agents for Agents: An Interrogator-Based Secure Framework for Autonomous Internet of Underwater Things
by: Akarma, Ali, et al.
Published: (2026)
by: Akarma, Ali, et al.
Published: (2026)
Totally mixed conditional independence equilibria of generic games
by: Bouyer, Matthieu, et al.
Published: (2025)
by: Bouyer, Matthieu, et al.
Published: (2025)
Hilbert schemes of points on fold-like curves and their combinatorics
by: Ortiz, Ángel David Ríos, et al.
Published: (2025)
by: Ortiz, Ángel David Ríos, et al.
Published: (2025)
REVISITING PUBLIC SPACE IN POST-WAR SOCIAL HOUSING IN GREAT BRITAIN
by: Pablo Sendra
Published: (2013)
by: Pablo Sendra
Published: (2013)
Which Agent Causes Task Failures and When? On Automated Failure Attribution of LLM Multi-Agent Systems
by: Zhang, Shaokun, et al.
Published: (2025)
by: Zhang, Shaokun, et al.
Published: (2025)
A Communication Protocol Design aimed at a MultiAgent System Framework for Miniaturized Satellite Systems
by: Samantha Interiano-Valverde
Published: (2020)
by: Samantha Interiano-Valverde
Published: (2020)
AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems
by: Zhang, Boxuan, et al.
Published: (2026)
by: Zhang, Boxuan, et al.
Published: (2026)
POIROT: Investigating Direct Tangible vs. Digitally Mediated Interaction and Attitude Moderation in Multi-party Murder Mystery Games
by: Chen, Wen, et al.
Published: (2026)
by: Chen, Wen, et al.
Published: (2026)
Impact of Dietary Lipid to Carbohydrate Ratio on Elemental Stoichiometric Relationships in Growth Phenotypes of Ruditapes Decussatus
by: Kristina Arranz, et al.
Published: (2025)
by: Kristina Arranz, et al.
Published: (2025)
Interpretable Failure Analysis in Multi-Agent Reinforcement Learning Systems
by: Shefin, Risal Shahriar, et al.
Published: (2026)
by: Shefin, Risal Shahriar, et al.
Published: (2026)
Incentivized Network Dynamics in Digital Job Recruitment
by: Kolic, Blas, et al.
Published: (2024)
by: Kolic, Blas, et al.
Published: (2024)
Who is Introducing the Failure? Automatically Attributing Failures of Multi-Agent Systems via Spectrum Analysis
by: Ge, Yu, et al.
Published: (2025)
by: Ge, Yu, et al.
Published: (2025)
Efficient Failure Management for Multi-Agent Systems with Reasoning Trace Representation
by: Zhang, Lingzhe, et al.
Published: (2026)
by: Zhang, Lingzhe, et al.
Published: (2026)
Efficient Multi-Agent Collaboration with Tool Use for Online Planning in Complex Table Question Answering
by: Zhou, Wei, et al.
Published: (2024)
by: Zhou, Wei, et al.
Published: (2024)
A Sheaf Framework for Strategic Multi-Agent Systems: From Consensus to Nash Equilibria
by: Hernández, Manuel, et al.
Published: (2026)
by: Hernández, Manuel, et al.
Published: (2026)
Sheaf-Theoretic Planning: A Categorical Foundation for Resilient Multi-Agent Autonomous Systems
by: Hernández, Manuel, et al.
Published: (2026)
by: Hernández, Manuel, et al.
Published: (2026)
Rethinking Failure Attribution in Multi-Agent Systems: A Multi-Perspective Benchmark and Evaluation
by: In, Yeonjun, et al.
Published: (2026)
by: In, Yeonjun, et al.
Published: (2026)
AgentFixer: From Failure Detection to Fix Recommendations in LLM Agentic Systems
by: Mulian, Hadar, et al.
Published: (2026)
by: Mulian, Hadar, et al.
Published: (2026)
Community-Led Regeneration
by: Sendra, Pablo, et al.
Published: (2020)
by: Sendra, Pablo, et al.
Published: (2020)
VerifyMAS: Hypothesis Verification for Failure Attribution in LLM Multi-Agent Systems
by: Qiao, Hezhe, et al.
Published: (2026)
by: Qiao, Hezhe, et al.
Published: (2026)
Incremental Risk Assessment for Cascading Failures in Large-Scale Multi-Agent Systems
by: Liu, Guangyi, et al.
Published: (2026)
by: Liu, Guangyi, et al.
Published: (2026)
SentinelAgent: Graph-based Anomaly Detection in Multi-Agent Systems
by: He, Xu, et al.
Published: (2025)
by: He, Xu, et al.
Published: (2025)
Curalaba: cuando la política no entiende la guerra
by: Eduardo Cebrián
Published: (2008)
by: Eduardo Cebrián
Published: (2008)
Conditional Multi-Stage Failure Recovery for Embodied Agents
by: Farag, Youmna, et al.
Published: (2025)
by: Farag, Youmna, et al.
Published: (2025)
A favor del Plagio
by: Manuel Arranz
Published: (2007)
by: Manuel Arranz
Published: (2007)
When Stress Becomes Signal: Detecting Antifragility-Compatible Regimes in Multi-Agent LLM Systems
by: de la Chica, Jose Manuel, et al.
Published: (2026)
by: de la Chica, Jose Manuel, et al.
Published: (2026)
Founder effects shape the evolutionary dynamics of multimodality in open LLM families
by: Cebrian, Manuel
Published: (2026)
by: Cebrian, Manuel
Published: (2026)
Similar Items
-
Rethinking the Illusion of Thinking
by: Varela, Iñaki Dellibarda, et al.
Published: (2025) -
Sensorimotor Self-Recognition in Multimodal Large Language Model-Driven Robots
by: Varela, Iñaki Dellibarda, et al.
Published: (2025) -
The Hessian correspondence of hypersurfaces of degree 3 and 4
by: Sendra-Arranz, Javier
Published: (2023) -
Questões para o trabalho profissional do Assistente Social no processo transexualizador
by: Pablo Cardozo Rocon
Published: (2018) -
ACESSO À SAÚDE PELA POPULAÇÃO TRANS NO BRASIL: NAS ENTRELINHAS DA REVISÃO INTEGRATIVA
by: Pablo Cardozo Rocon
Published: (2020)