Diagnostics of cognitive failures in multi-agent expert systems using dynamic evaluation protocols and subsequent mutation of the processing context
Fuente:
arXiv
Saved in:
| Main Authors: | Sorstkins, Andrejs, Bailey, Josh, Baron, Dr Alistair |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
by: Sorstkins, Andrejs
Published: (2025)
by: Sorstkins, Andrejs
Published: (2025)
Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals
by: Sorstkins, Andrejs, et al.
Published: (2025)
by: Sorstkins, Andrejs, et al.
Published: (2025)
A process algebraic framework for multi-agent dynamic epistemic systems
by: Aldini, Alessandro
Published: (2024)
by: Aldini, Alessandro
Published: (2024)
Adaptive routing protocols for determining optimal paths in AI multi-agent systems: a priority- and learning-enhanced approach
by: Panayotov, Theodor, et al.
Published: (2025)
by: Panayotov, Theodor, et al.
Published: (2025)
The impact of multi-agent debate protocols on debate quality: a controlled case study
by: Marandi, Ramtin Zargari
Published: (2026)
by: Marandi, Ramtin Zargari
Published: (2026)
Fuzzy expert system for the process of collecting and purifying acidic water: a digital twin approach
by: Maratuly, Temirbolat, et al.
Published: (2026)
by: Maratuly, Temirbolat, et al.
Published: (2026)
Metric assessment protocol in the context of answer fluctuation on MCQ tasks
by: Goliakova, Ekaterina, et al.
Published: (2025)
by: Goliakova, Ekaterina, et al.
Published: (2025)
RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts
by: Wijk, Hjalmar, et al.
Published: (2024)
by: Wijk, Hjalmar, et al.
Published: (2024)
Context-Picker: Dynamic context selection using multi-stage reinforcement learning
by: Zhu, Siyuan, et al.
Published: (2025)
by: Zhu, Siyuan, et al.
Published: (2025)
System 0/1/2/3: Quad-process theory for multi-timescale embodied collective cognitive systems
by: Taniguchi, Tadahiro, et al.
Published: (2025)
by: Taniguchi, Tadahiro, et al.
Published: (2025)
Legal interpretation and AI: from expert systems to argumentation and LLMs
by: Janeček, Václav, et al.
Published: (2026)
by: Janeček, Václav, et al.
Published: (2026)
Diverse And Private Synthetic Datasets Generation for RAG evaluation: A multi-agent framework
by: Driouich, Ilias, et al.
Published: (2025)
by: Driouich, Ilias, et al.
Published: (2025)
Discovering mathematical concepts through a multi-agent system
by: Aggarwal, Daattavya, et al.
Published: (2026)
by: Aggarwal, Daattavya, et al.
Published: (2026)
Multi-agent cooperation through in-context co-player inference
by: Weis, Marissa A., et al.
Published: (2026)
by: Weis, Marissa A., et al.
Published: (2026)
Graders should cheat: privileged information enables expert-level automated evaluations
by: Zhou, Jin Peng, et al.
Published: (2025)
by: Zhou, Jin Peng, et al.
Published: (2025)
Memory poisoning and secure multi-agent systems
by: Torra, Vicenç, et al.
Published: (2026)
by: Torra, Vicenç, et al.
Published: (2026)
Enhanced Transformer architecture for in-context learning of dynamical systems
by: Rufolo, Matteo, et al.
Published: (2024)
by: Rufolo, Matteo, et al.
Published: (2024)
FREIDA: A Framework for developing quantitative agent based models based on qualitative expert knowledge
by: Oetker, Frederike, et al.
Published: (2023)
by: Oetker, Frederike, et al.
Published: (2023)
Setting up for failure: automatic discovery of the neural mechanisms of cognitive errors
by: Radmard, Puria, et al.
Published: (2025)
by: Radmard, Puria, et al.
Published: (2025)
Configurable multi-agent framework for scalable and realistic testing of llm-based agents
by: Wang, Sai, et al.
Published: (2025)
by: Wang, Sai, et al.
Published: (2025)
Log analysis is necessary for credible evaluation of AI agents
by: Kirgis, Peter, et al.
Published: (2026)
by: Kirgis, Peter, et al.
Published: (2026)
Large Language Models, scientific knowledge and factuality: A framework to streamline human expert evaluation
by: Wysocka, Magdalena, et al.
Published: (2023)
by: Wysocka, Magdalena, et al.
Published: (2023)
Robin: A multi-agent system for automating scientific discovery
by: Ghareeb, Ali Essam, et al.
Published: (2025)
by: Ghareeb, Ali Essam, et al.
Published: (2025)
From reactive to cognitive: brain-inspired spatial intelligence for embodied agents
by: Ruan, Shouwei, et al.
Published: (2025)
by: Ruan, Shouwei, et al.
Published: (2025)
Adaptive parameter sharing for multi-agent reinforcement learning
by: Li, Dapeng, et al.
Published: (2023)
by: Li, Dapeng, et al.
Published: (2023)
Realistic pedestrian-driver interaction modelling using multi-agent RL with human perceptual-motor constraints
by: Wang, Yueyang, et al.
Published: (2025)
by: Wang, Yueyang, et al.
Published: (2025)
LLMs learn governing principles of dynamical systems, revealing an in-context neural scaling law
by: Liu, Toni J. B., et al.
Published: (2024)
by: Liu, Toni J. B., et al.
Published: (2024)
An AI system to help scientists write expert-level empirical software
by: Aygün, Eser, et al.
Published: (2025)
by: Aygün, Eser, et al.
Published: (2025)
Stream-based perception for cognitive agents in mobile ecosystems
by: Dötterl, Jeremias, et al.
Published: (2024)
by: Dötterl, Jeremias, et al.
Published: (2024)
Agentic clinical reasoning over longitudinal myeloma records: a retrospective evaluation against expert consensus
by: Moll, Johannes, et al.
Published: (2026)
by: Moll, Johannes, et al.
Published: (2026)
Using multi-agent architecture to mitigate the risk of LLM hallucinations
by: Amer, Abd Elrahman, et al.
Published: (2025)
by: Amer, Abd Elrahman, et al.
Published: (2025)
Plancraft: an evaluation dataset for planning with LLM agents
by: Dagan, Gautier, et al.
Published: (2024)
by: Dagan, Gautier, et al.
Published: (2024)
A systematic review on expert systems for improving energy efficiency in the manufacturing industry
by: Ioshchikhes, Borys, et al.
Published: (2024)
by: Ioshchikhes, Borys, et al.
Published: (2024)
Acceleration method for generating perception failure scenarios based on editing Markov process
by: Cai, Canjie
Published: (2024)
by: Cai, Canjie
Published: (2024)
A Large Language Model-based multi-agent manufacturing system for intelligent shopfloor
by: Zhao, Zhen, et al.
Published: (2024)
by: Zhao, Zhen, et al.
Published: (2024)
WebExpert: domain-aware web agents with critic-guided expert experience for high-precision search
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
MAFA: A multi-agent framework for annotation
by: Hegazy, Mahmood, et al.
Published: (2025)
by: Hegazy, Mahmood, et al.
Published: (2025)
Beyond the high score: Prosocial ability profiles of multi-agent populations
by: Tesic, Marko, et al.
Published: (2025)
by: Tesic, Marko, et al.
Published: (2025)
Dynamic fairness-aware recommendation through multi-agent social choice
by: Aird, Amanda, et al.
Published: (2023)
by: Aird, Amanda, et al.
Published: (2023)
Reshaping MOFs text mining with a dynamic multi-agents framework of large language model
by: Lin, Zuhong, et al.
Published: (2025)
by: Lin, Zuhong, et al.
Published: (2025)
Similar Items
-
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
by: Sorstkins, Andrejs
Published: (2025) -
Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals
by: Sorstkins, Andrejs, et al.
Published: (2025) -
A process algebraic framework for multi-agent dynamic epistemic systems
by: Aldini, Alessandro
Published: (2024) -
Adaptive routing protocols for determining optimal paths in AI multi-agent systems: a priority- and learning-enhanced approach
by: Panayotov, Theodor, et al.
Published: (2025) -
The impact of multi-agent debate protocols on debate quality: a controlled case study
by: Marandi, Ramtin Zargari
Published: (2026)