Salvato in:
| Autori principali: | Sorstkins, Andrejs, Bailey, Josh, Baron, Dr Alistair |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2509.15366 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
di: Sorstkins, Andrejs
Pubblicazione: (2025)
di: Sorstkins, Andrejs
Pubblicazione: (2025)
Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals
di: Sorstkins, Andrejs, et al.
Pubblicazione: (2025)
di: Sorstkins, Andrejs, et al.
Pubblicazione: (2025)
A process algebraic framework for multi-agent dynamic epistemic systems
di: Aldini, Alessandro
Pubblicazione: (2024)
di: Aldini, Alessandro
Pubblicazione: (2024)
Adaptive routing protocols for determining optimal paths in AI multi-agent systems: a priority- and learning-enhanced approach
di: Panayotov, Theodor, et al.
Pubblicazione: (2025)
di: Panayotov, Theodor, et al.
Pubblicazione: (2025)
The impact of multi-agent debate protocols on debate quality: a controlled case study
di: Marandi, Ramtin Zargari
Pubblicazione: (2026)
di: Marandi, Ramtin Zargari
Pubblicazione: (2026)
Fuzzy expert system for the process of collecting and purifying acidic water: a digital twin approach
di: Maratuly, Temirbolat, et al.
Pubblicazione: (2026)
di: Maratuly, Temirbolat, et al.
Pubblicazione: (2026)
Metric assessment protocol in the context of answer fluctuation on MCQ tasks
di: Goliakova, Ekaterina, et al.
Pubblicazione: (2025)
di: Goliakova, Ekaterina, et al.
Pubblicazione: (2025)
RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts
di: Wijk, Hjalmar, et al.
Pubblicazione: (2024)
di: Wijk, Hjalmar, et al.
Pubblicazione: (2024)
Diverse And Private Synthetic Datasets Generation for RAG evaluation: A multi-agent framework
di: Driouich, Ilias, et al.
Pubblicazione: (2025)
di: Driouich, Ilias, et al.
Pubblicazione: (2025)
System 0/1/2/3: Quad-process theory for multi-timescale embodied collective cognitive systems
di: Taniguchi, Tadahiro, et al.
Pubblicazione: (2025)
di: Taniguchi, Tadahiro, et al.
Pubblicazione: (2025)
Context-Picker: Dynamic context selection using multi-stage reinforcement learning
di: Zhu, Siyuan, et al.
Pubblicazione: (2025)
di: Zhu, Siyuan, et al.
Pubblicazione: (2025)
Discovering mathematical concepts through a multi-agent system
di: Aggarwal, Daattavya, et al.
Pubblicazione: (2026)
di: Aggarwal, Daattavya, et al.
Pubblicazione: (2026)
Multi-agent cooperation through in-context co-player inference
di: Weis, Marissa A., et al.
Pubblicazione: (2026)
di: Weis, Marissa A., et al.
Pubblicazione: (2026)
Memory poisoning and secure multi-agent systems
di: Torra, Vicenç, et al.
Pubblicazione: (2026)
di: Torra, Vicenç, et al.
Pubblicazione: (2026)
Graders should cheat: privileged information enables expert-level automated evaluations
di: Zhou, Jin Peng, et al.
Pubblicazione: (2025)
di: Zhou, Jin Peng, et al.
Pubblicazione: (2025)
Legal interpretation and AI: from expert systems to argumentation and LLMs
di: Janeček, Václav, et al.
Pubblicazione: (2026)
di: Janeček, Václav, et al.
Pubblicazione: (2026)
Enhanced Transformer architecture for in-context learning of dynamical systems
di: Rufolo, Matteo, et al.
Pubblicazione: (2024)
di: Rufolo, Matteo, et al.
Pubblicazione: (2024)
FREIDA: A Framework for developing quantitative agent based models based on qualitative expert knowledge
di: Oetker, Frederike, et al.
Pubblicazione: (2023)
di: Oetker, Frederike, et al.
Pubblicazione: (2023)
Setting up for failure: automatic discovery of the neural mechanisms of cognitive errors
di: Radmard, Puria, et al.
Pubblicazione: (2025)
di: Radmard, Puria, et al.
Pubblicazione: (2025)
Robin: A multi-agent system for automating scientific discovery
di: Ghareeb, Ali Essam, et al.
Pubblicazione: (2025)
di: Ghareeb, Ali Essam, et al.
Pubblicazione: (2025)
Configurable multi-agent framework for scalable and realistic testing of llm-based agents
di: Wang, Sai, et al.
Pubblicazione: (2025)
di: Wang, Sai, et al.
Pubblicazione: (2025)
Log analysis is necessary for credible evaluation of AI agents
di: Kirgis, Peter, et al.
Pubblicazione: (2026)
di: Kirgis, Peter, et al.
Pubblicazione: (2026)
Large Language Models, scientific knowledge and factuality: A framework to streamline human expert evaluation
di: Wysocka, Magdalena, et al.
Pubblicazione: (2023)
di: Wysocka, Magdalena, et al.
Pubblicazione: (2023)
Stream-based perception for cognitive agents in mobile ecosystems
di: Dötterl, Jeremias, et al.
Pubblicazione: (2024)
di: Dötterl, Jeremias, et al.
Pubblicazione: (2024)
From reactive to cognitive: brain-inspired spatial intelligence for embodied agents
di: Ruan, Shouwei, et al.
Pubblicazione: (2025)
di: Ruan, Shouwei, et al.
Pubblicazione: (2025)
Plancraft: an evaluation dataset for planning with LLM agents
di: Dagan, Gautier, et al.
Pubblicazione: (2024)
di: Dagan, Gautier, et al.
Pubblicazione: (2024)
Agentic clinical reasoning over longitudinal myeloma records: a retrospective evaluation against expert consensus
di: Moll, Johannes, et al.
Pubblicazione: (2026)
di: Moll, Johannes, et al.
Pubblicazione: (2026)
An AI system to help scientists write expert-level empirical software
di: Aygün, Eser, et al.
Pubblicazione: (2025)
di: Aygün, Eser, et al.
Pubblicazione: (2025)
LLMs learn governing principles of dynamical systems, revealing an in-context neural scaling law
di: Liu, Toni J. B., et al.
Pubblicazione: (2024)
di: Liu, Toni J. B., et al.
Pubblicazione: (2024)
Adaptive parameter sharing for multi-agent reinforcement learning
di: Li, Dapeng, et al.
Pubblicazione: (2023)
di: Li, Dapeng, et al.
Pubblicazione: (2023)
WebExpert: domain-aware web agents with critic-guided expert experience for high-precision search
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
Realistic pedestrian-driver interaction modelling using multi-agent RL with human perceptual-motor constraints
di: Wang, Yueyang, et al.
Pubblicazione: (2025)
di: Wang, Yueyang, et al.
Pubblicazione: (2025)
A Large Language Model-based multi-agent manufacturing system for intelligent shopfloor
di: Zhao, Zhen, et al.
Pubblicazione: (2024)
di: Zhao, Zhen, et al.
Pubblicazione: (2024)
A systematic review on expert systems for improving energy efficiency in the manufacturing industry
di: Ioshchikhes, Borys, et al.
Pubblicazione: (2024)
di: Ioshchikhes, Borys, et al.
Pubblicazione: (2024)
Using multi-agent architecture to mitigate the risk of LLM hallucinations
di: Amer, Abd Elrahman, et al.
Pubblicazione: (2025)
di: Amer, Abd Elrahman, et al.
Pubblicazione: (2025)
MAFA: A multi-agent framework for annotation
di: Hegazy, Mahmood, et al.
Pubblicazione: (2025)
di: Hegazy, Mahmood, et al.
Pubblicazione: (2025)
Reshaping MOFs text mining with a dynamic multi-agents framework of large language model
di: Lin, Zuhong, et al.
Pubblicazione: (2025)
di: Lin, Zuhong, et al.
Pubblicazione: (2025)
Group size effects and collective misalignment in LLM multi-agent systems
di: Flint, Ariel, et al.
Pubblicazione: (2025)
di: Flint, Ariel, et al.
Pubblicazione: (2025)
Acceleration method for generating perception failure scenarios based on editing Markov process
di: Cai, Canjie
Pubblicazione: (2024)
di: Cai, Canjie
Pubblicazione: (2024)
Large language models for post-publication research evaluation: Evidence from expert recommendations and citation indicators
di: Wu, Mengjia, et al.
Pubblicazione: (2026)
di: Wu, Mengjia, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Assessing RAG and HyDE on 1B vs. 4B-Parameter Gemma LLMs for Personal Assistants Integretion
di: Sorstkins, Andrejs
Pubblicazione: (2025) -
Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals
di: Sorstkins, Andrejs, et al.
Pubblicazione: (2025) -
A process algebraic framework for multi-agent dynamic epistemic systems
di: Aldini, Alessandro
Pubblicazione: (2024) -
Adaptive routing protocols for determining optimal paths in AI multi-agent systems: a priority- and learning-enhanced approach
di: Panayotov, Theodor, et al.
Pubblicazione: (2025) -
The impact of multi-agent debate protocols on debate quality: a controlled case study
di: Marandi, Ramtin Zargari
Pubblicazione: (2026)