Saved in:
| Main Authors: | Sorstkins, Andrejs, Bailey, Josh, Baron, Dr Alistair |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.15366 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
by: Sorstkins, Andrejs
Published: (2025)
by: Sorstkins, Andrejs
Published: (2025)
Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals
by: Sorstkins, Andrejs, et al.
Published: (2025)
by: Sorstkins, Andrejs, et al.
Published: (2025)
A process algebraic framework for multi-agent dynamic epistemic systems
by: Aldini, Alessandro
Published: (2024)
by: Aldini, Alessandro
Published: (2024)
Adaptive routing protocols for determining optimal paths in AI multi-agent systems: a priority- and learning-enhanced approach
by: Panayotov, Theodor, et al.
Published: (2025)
by: Panayotov, Theodor, et al.
Published: (2025)
The impact of multi-agent debate protocols on debate quality: a controlled case study
by: Marandi, Ramtin Zargari
Published: (2026)
by: Marandi, Ramtin Zargari
Published: (2026)
Fuzzy expert system for the process of collecting and purifying acidic water: a digital twin approach
by: Maratuly, Temirbolat, et al.
Published: (2026)
by: Maratuly, Temirbolat, et al.
Published: (2026)
Metric assessment protocol in the context of answer fluctuation on MCQ tasks
by: Goliakova, Ekaterina, et al.
Published: (2025)
by: Goliakova, Ekaterina, et al.
Published: (2025)
RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts
by: Wijk, Hjalmar, et al.
Published: (2024)
by: Wijk, Hjalmar, et al.
Published: (2024)
Diverse And Private Synthetic Datasets Generation for RAG evaluation: A multi-agent framework
by: Driouich, Ilias, et al.
Published: (2025)
by: Driouich, Ilias, et al.
Published: (2025)
System 0/1/2/3: Quad-process theory for multi-timescale embodied collective cognitive systems
by: Taniguchi, Tadahiro, et al.
Published: (2025)
by: Taniguchi, Tadahiro, et al.
Published: (2025)
Context-Picker: Dynamic context selection using multi-stage reinforcement learning
by: Zhu, Siyuan, et al.
Published: (2025)
by: Zhu, Siyuan, et al.
Published: (2025)
Discovering mathematical concepts through a multi-agent system
by: Aggarwal, Daattavya, et al.
Published: (2026)
by: Aggarwal, Daattavya, et al.
Published: (2026)
Multi-agent cooperation through in-context co-player inference
by: Weis, Marissa A., et al.
Published: (2026)
by: Weis, Marissa A., et al.
Published: (2026)
Memory poisoning and secure multi-agent systems
by: Torra, Vicenç, et al.
Published: (2026)
by: Torra, Vicenç, et al.
Published: (2026)
Graders should cheat: privileged information enables expert-level automated evaluations
by: Zhou, Jin Peng, et al.
Published: (2025)
by: Zhou, Jin Peng, et al.
Published: (2025)
Legal interpretation and AI: from expert systems to argumentation and LLMs
by: Janeček, Václav, et al.
Published: (2026)
by: Janeček, Václav, et al.
Published: (2026)
Enhanced Transformer architecture for in-context learning of dynamical systems
by: Rufolo, Matteo, et al.
Published: (2024)
by: Rufolo, Matteo, et al.
Published: (2024)
FREIDA: A Framework for developing quantitative agent based models based on qualitative expert knowledge
by: Oetker, Frederike, et al.
Published: (2023)
by: Oetker, Frederike, et al.
Published: (2023)
Setting up for failure: automatic discovery of the neural mechanisms of cognitive errors
by: Radmard, Puria, et al.
Published: (2025)
by: Radmard, Puria, et al.
Published: (2025)
Robin: A multi-agent system for automating scientific discovery
by: Ghareeb, Ali Essam, et al.
Published: (2025)
by: Ghareeb, Ali Essam, et al.
Published: (2025)
Configurable multi-agent framework for scalable and realistic testing of llm-based agents
by: Wang, Sai, et al.
Published: (2025)
by: Wang, Sai, et al.
Published: (2025)
Log analysis is necessary for credible evaluation of AI agents
by: Kirgis, Peter, et al.
Published: (2026)
by: Kirgis, Peter, et al.
Published: (2026)
Large Language Models, scientific knowledge and factuality: A framework to streamline human expert evaluation
by: Wysocka, Magdalena, et al.
Published: (2023)
by: Wysocka, Magdalena, et al.
Published: (2023)
Stream-based perception for cognitive agents in mobile ecosystems
by: Dötterl, Jeremias, et al.
Published: (2024)
by: Dötterl, Jeremias, et al.
Published: (2024)
From reactive to cognitive: brain-inspired spatial intelligence for embodied agents
by: Ruan, Shouwei, et al.
Published: (2025)
by: Ruan, Shouwei, et al.
Published: (2025)
Plancraft: an evaluation dataset for planning with LLM agents
by: Dagan, Gautier, et al.
Published: (2024)
by: Dagan, Gautier, et al.
Published: (2024)
Agentic clinical reasoning over longitudinal myeloma records: a retrospective evaluation against expert consensus
by: Moll, Johannes, et al.
Published: (2026)
by: Moll, Johannes, et al.
Published: (2026)
An AI system to help scientists write expert-level empirical software
by: Aygün, Eser, et al.
Published: (2025)
by: Aygün, Eser, et al.
Published: (2025)
LLMs learn governing principles of dynamical systems, revealing an in-context neural scaling law
by: Liu, Toni J. B., et al.
Published: (2024)
by: Liu, Toni J. B., et al.
Published: (2024)
Adaptive parameter sharing for multi-agent reinforcement learning
by: Li, Dapeng, et al.
Published: (2023)
by: Li, Dapeng, et al.
Published: (2023)
WebExpert: domain-aware web agents with critic-guided expert experience for high-precision search
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
Realistic pedestrian-driver interaction modelling using multi-agent RL with human perceptual-motor constraints
by: Wang, Yueyang, et al.
Published: (2025)
by: Wang, Yueyang, et al.
Published: (2025)
A Large Language Model-based multi-agent manufacturing system for intelligent shopfloor
by: Zhao, Zhen, et al.
Published: (2024)
by: Zhao, Zhen, et al.
Published: (2024)
A systematic review on expert systems for improving energy efficiency in the manufacturing industry
by: Ioshchikhes, Borys, et al.
Published: (2024)
by: Ioshchikhes, Borys, et al.
Published: (2024)
Using multi-agent architecture to mitigate the risk of LLM hallucinations
by: Amer, Abd Elrahman, et al.
Published: (2025)
by: Amer, Abd Elrahman, et al.
Published: (2025)
MAFA: A multi-agent framework for annotation
by: Hegazy, Mahmood, et al.
Published: (2025)
by: Hegazy, Mahmood, et al.
Published: (2025)
Reshaping MOFs text mining with a dynamic multi-agents framework of large language model
by: Lin, Zuhong, et al.
Published: (2025)
by: Lin, Zuhong, et al.
Published: (2025)
Group size effects and collective misalignment in LLM multi-agent systems
by: Flint, Ariel, et al.
Published: (2025)
by: Flint, Ariel, et al.
Published: (2025)
Acceleration method for generating perception failure scenarios based on editing Markov process
by: Cai, Canjie
Published: (2024)
by: Cai, Canjie
Published: (2024)
Large language models for post-publication research evaluation: Evidence from expert recommendations and citation indicators
by: Wu, Mengjia, et al.
Published: (2026)
by: Wu, Mengjia, et al.
Published: (2026)
Similar Items
-
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
by: Sorstkins, Andrejs
Published: (2025) -
Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals
by: Sorstkins, Andrejs, et al.
Published: (2025) -
A process algebraic framework for multi-agent dynamic epistemic systems
by: Aldini, Alessandro
Published: (2024) -
Adaptive routing protocols for determining optimal paths in AI multi-agent systems: a priority- and learning-enhanced approach
by: Panayotov, Theodor, et al.
Published: (2025) -
The impact of multi-agent debate protocols on debate quality: a controlled case study
by: Marandi, Ramtin Zargari
Published: (2026)