Stay Focused: Problem Drift in Multi-Agent Debate
Fuente:
arXiv
Salvato in:
| Autori principali: | Becker, Jonas, Kaesberg, Lars Benedikt, Stephan, Andreas, Wahle, Jan Philip, Ruas, Terry, Gipp, Bela |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MALLM: Multi-Agent Large Language Models Framework
di: Becker, Jonas, et al.
Pubblicazione: (2025)
di: Becker, Jonas, et al.
Pubblicazione: (2025)
Text Generation: A Systematic Literature Review of Tasks, Evaluation, and Challenges
di: Becker, Jonas, et al.
Pubblicazione: (2024)
di: Becker, Jonas, et al.
Pubblicazione: (2024)
Towards Human Understanding of Paraphrase Types in Large Language Models
di: Meier, Dominik, et al.
Pubblicazione: (2024)
di: Meier, Dominik, et al.
Pubblicazione: (2024)
Voting or Consensus? Decision-Making in Multi-Agent Debate
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2025)
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2025)
CiteAssist: A System for Automated Preprint Citation and BibTeX Generation
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2024)
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2024)
SPaRC: A Spatial Pathfinding Reasoning Challenge
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2025)
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2025)
SemEval-2026 Task 3: Dimensional Aspect-Based Sentiment Analysis (DimABSA)
di: Yu, Liang-Chih, et al.
Pubblicazione: (2026)
di: Yu, Liang-Chih, et al.
Pubblicazione: (2026)
DimStance: Multilingual Datasets for Dimensional Stance Analysis
di: Becker, Jonas, et al.
Pubblicazione: (2026)
di: Becker, Jonas, et al.
Pubblicazione: (2026)
DimABSA: Building Multilingual and Multidomain Datasets for Dimensional Aspect-Based Sentiment Analysis
di: Lee, Lung-Hao, et al.
Pubblicazione: (2026)
di: Lee, Lung-Hao, et al.
Pubblicazione: (2026)
Mind the Gap Between Spatial Reasoning and Acting! Step-by-Step Evaluation of Agents With Spatial-Gym
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2026)
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2026)
Multi-Agent Reasoning Improves Compute Efficiency: Pareto-Optimal Test-Time Scaling
di: Wunderlich, Florian Valentin, et al.
Pubblicazione: (2026)
di: Wunderlich, Florian Valentin, et al.
Pubblicazione: (2026)
Paraphrase Types for Generation and Detection
di: Wahle, Jan Philip, et al.
Pubblicazione: (2023)
di: Wahle, Jan Philip, et al.
Pubblicazione: (2023)
propella-1: Multi-Property Document Annotation for LLM Data Curation at Scale
di: Idahl, Maximilian, et al.
Pubblicazione: (2026)
di: Idahl, Maximilian, et al.
Pubblicazione: (2026)
German Text Simplification: Finetuning Large Language Models with Semi-Synthetic Data
di: Klöser, Lars, et al.
Pubblicazione: (2024)
di: Klöser, Lars, et al.
Pubblicazione: (2024)
Residual Drift Dominates Contradiction in Multi-Turn Constraint Reasoning
di: Kawada, Sebastien
Pubblicazione: (2026)
di: Kawada, Sebastien
Pubblicazione: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
di: Ashuach, Tomer, et al.
Pubblicazione: (2025)
di: Ashuach, Tomer, et al.
Pubblicazione: (2025)
Literature Review Of Multi-Agent Debate For Problem-Solving
di: Tillmann, Arne
Pubblicazione: (2025)
di: Tillmann, Arne
Pubblicazione: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025)
di: Fadli, Samih
Pubblicazione: (2025)
When Persuasion Overrides Truth in Multi-Agent LLM Debates: Introducing a Confidence-Weighted Persuasion Override Rate (CW-POR)
di: Agarwal, Mahak, et al.
Pubblicazione: (2025)
di: Agarwal, Mahak, et al.
Pubblicazione: (2025)
Clinical Document Corpora -- Real Ones, Translated and Synthetic Substitutes, and Assorted Domain Proxies: A Survey of Diversity in Corpus Design, with Focus on German Text Data
di: Hahn, Udo
Pubblicazione: (2024)
di: Hahn, Udo
Pubblicazione: (2024)
Strategy Adaptation in Large Language Model Werewolf Agents
di: Nakamori, Fuya, et al.
Pubblicazione: (2025)
di: Nakamori, Fuya, et al.
Pubblicazione: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
di: Saji, Alan, et al.
Pubblicazione: (2025)
di: Saji, Alan, et al.
Pubblicazione: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)
di: Peters, Sydney, et al.
Pubblicazione: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
di: Collado-Montañez, Jaime, et al.
Pubblicazione: (2025)
di: Collado-Montañez, Jaime, et al.
Pubblicazione: (2025)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
di: Smădu, Răzvan-Alexandru, et al.
Pubblicazione: (2025)
Improving Retrospective Language Agents via Joint Policy Gradient Optimization
di: Feng, Xueyang, et al.
Pubblicazione: (2025)
di: Feng, Xueyang, et al.
Pubblicazione: (2025)
HACHIMI: Scalable and Controllable Student Persona Generation via Orchestrated Agents
di: Jiang, Yilin, et al.
Pubblicazione: (2026)
di: Jiang, Yilin, et al.
Pubblicazione: (2026)
Adaptive Focus Memory for Language Models
di: Cruz, Christopher
Pubblicazione: (2025)
di: Cruz, Christopher
Pubblicazione: (2025)
Are Generative Models Underconfident? Better Quality Estimation with Boosted Model Probability
di: Dinh, Tu Anh, et al.
Pubblicazione: (2025)
di: Dinh, Tu Anh, et al.
Pubblicazione: (2025)
Sigmoid Head for Quality Estimation under Language Ambiguity
di: Dinh, Tu Anh, et al.
Pubblicazione: (2026)
di: Dinh, Tu Anh, et al.
Pubblicazione: (2026)
Who Watches the Watchmen? Humans Disagree With Translation Metrics on Unseen Domains
di: Schmidt, Finn, et al.
Pubblicazione: (2026)
di: Schmidt, Finn, et al.
Pubblicazione: (2026)
Paraphrase Types Elicit Prompt Engineering Capabilities
di: Wahle, Jan Philip, et al.
Pubblicazione: (2024)
di: Wahle, Jan Philip, et al.
Pubblicazione: (2024)
Piecing Together Cross-Document Coreference Resolution Datasets: Systematic Dataset Analysis and Unification
di: Zhukova, Anastasia, et al.
Pubblicazione: (2026)
di: Zhukova, Anastasia, et al.
Pubblicazione: (2026)
You need to MIMIC to get FAME: Solving Meeting Transcript Scarcity with a Multi-Agent Conversations
di: Kirstein, Frederic, et al.
Pubblicazione: (2025)
di: Kirstein, Frederic, et al.
Pubblicazione: (2025)
Instructional Agents: Reducing Teaching Faculty Workload through Multi-Agent Instructional Design
di: Yao, Huaiyuan, et al.
Pubblicazione: (2025)
di: Yao, Huaiyuan, et al.
Pubblicazione: (2025)
Quality Estimation with $k$-nearest Neighbors and Automatic Evaluation for Model-specific Quality Estimation
di: Dinh, Tu Anh, et al.
Pubblicazione: (2024)
di: Dinh, Tu Anh, et al.
Pubblicazione: (2024)
Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents
di: Wang, Renxi, et al.
Pubblicazione: (2024)
di: Wang, Renxi, et al.
Pubblicazione: (2024)
Statistical Scouting Finds Debate-Safe but Not Debate-Useful Cases: A Matched-Ceiling Study of Open-Weight LLM Reasoning Protocols
di: Hu, Julia, et al.
Pubblicazione: (2026)
di: Hu, Julia, et al.
Pubblicazione: (2026)
What's under the hood: Investigating Automatic Metrics on Meeting Summarization
di: Kirstein, Frederic, et al.
Pubblicazione: (2024)
di: Kirstein, Frederic, et al.
Pubblicazione: (2024)
CADS: A Systematic Literature Review on the Challenges of Abstractive Dialogue Summarization
di: Kirstein, Frederic, et al.
Pubblicazione: (2024)
di: Kirstein, Frederic, et al.
Pubblicazione: (2024)
Documenti analoghi
-
MALLM: Multi-Agent Large Language Models Framework
di: Becker, Jonas, et al.
Pubblicazione: (2025) -
Text Generation: A Systematic Literature Review of Tasks, Evaluation, and Challenges
di: Becker, Jonas, et al.
Pubblicazione: (2024) -
Towards Human Understanding of Paraphrase Types in Large Language Models
di: Meier, Dominik, et al.
Pubblicazione: (2024) -
Voting or Consensus? Decision-Making in Multi-Agent Debate
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2025) -
CiteAssist: A System for Automated Preprint Citation and BibTeX Generation
di: Kaesberg, Lars Benedikt, et al.
Pubblicazione: (2024)