Disagreement as Data: Reasoning Trace Analytics in Multi-Agent Systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tajik, Elham, Borchers, Conrad, Shahrokhian, Bahar, Simon, Sebastian, Keramati, Ali, Pal, Sonika, Sankaranarayanan, Sreecharan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Temperature and Persona Shape LLM Agent Consensus With Minimal Accuracy Gains in Qualitative Coding
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
Mitigating "Epistemic Debt" in Generative AI-Scaffolded Novice Programming using Metacognitive Scripts
von: Sankaranarayanan, Sreecharan
Veröffentlicht: (2026)
von: Sankaranarayanan, Sreecharan
Veröffentlicht: (2026)
Can Large Language Models Match Tutoring System Adaptivity? A Benchmarking Study
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
Parametric Constraints for Bayesian Knowledge Tracing from First Principles
von: Shchepakin, Denis, et al.
Veröffentlicht: (2023)
von: Shchepakin, Denis, et al.
Veröffentlicht: (2023)
Disentangling Learning from Judgment: Representation Learning for Open Response Analytics
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
Disagreements in Reasoning: How a Model's Thinking Process Dictates Persuasion in Multi-Agent Systems
von: Zhao, Haodong, et al.
Veröffentlicht: (2025)
von: Zhao, Haodong, et al.
Veröffentlicht: (2025)
Benchmarking Educational LLMs with Analytics: A Case Study on Gender Bias in Feedback
von: Du, Yishan, et al.
Veröffentlicht: (2025)
von: Du, Yishan, et al.
Veröffentlicht: (2025)
DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning
von: Sivakumaran, Nithin, et al.
Veröffentlicht: (2025)
von: Sivakumaran, Nithin, et al.
Veröffentlicht: (2025)
Representation Learning to Study Temporal Dynamics in Tutorial Scaffolding
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
When Disagreements Elicit Robustness: Investigating Self-Repair Capabilities under LLM Multi-Agent Disagreements
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
Large Language Models as Students Who Think Aloud: Overly Coherent, Verbose, and Confident
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
Physiological and Semantic Patterns in Medical Teams Using an Intelligent Tutoring System
von: Huang, Xiaoshan, et al.
Veröffentlicht: (2026)
von: Huang, Xiaoshan, et al.
Veröffentlicht: (2026)
The Hidden Strength of Disagreement: Unraveling the Consensus-Diversity Tradeoff in Adaptive Multi-Agent Systems
von: Wu, Zengqing, et al.
Veröffentlicht: (2025)
von: Wu, Zengqing, et al.
Veröffentlicht: (2025)
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)
AgentAda: Skill-Adaptive Data Analytics for Tailored Insight Discovery
von: Abaskohi, Amirhossein, et al.
Veröffentlicht: (2025)
von: Abaskohi, Amirhossein, et al.
Veröffentlicht: (2025)
Toward Trait-Aware Learning Analytics
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
MSA at BEA 2025 Shared Task: Disagreement-Aware Instruction Tuning for Multi-Dimensional Evaluation of LLMs as Math Tutors
von: Hikal, Baraa, et al.
Veröffentlicht: (2025)
von: Hikal, Baraa, et al.
Veröffentlicht: (2025)
Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces
von: Zhang, Chenchen
Veröffentlicht: (2026)
von: Zhang, Chenchen
Veröffentlicht: (2026)
Understanding Teacher Revisions of Large Language Model-Generated Feedback
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
von: Borchers, Conrad, et al.
Veröffentlicht: (2026)
Can Reasoning Help Large Language Models Capture Human Annotator Disagreement?
von: Ni, Jingwei, et al.
Veröffentlicht: (2025)
von: Ni, Jingwei, et al.
Veröffentlicht: (2025)
Leveraging Annotator Disagreement for Text Classification
von: Xu, Jin, et al.
Veröffentlicht: (2024)
von: Xu, Jin, et al.
Veröffentlicht: (2024)
LLM-Generated Feedback Supports Learning If Learners Choose to Use It
von: Thomas, Danielle R., et al.
Veröffentlicht: (2025)
von: Thomas, Danielle R., et al.
Veröffentlicht: (2025)
G-Memory: Tracing Hierarchical Memory for Multi-Agent Systems
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
DiscoUQ: Structured Disagreement Analysis for Uncertainty Quantification in LLM Agent Ensembles
von: Jiang, Bo
Veröffentlicht: (2026)
von: Jiang, Bo
Veröffentlicht: (2026)
D3CODE: Disentangling Disagreements in Data across Cultures on Offensiveness Detection and Evaluation
von: Davani, Aida Mostafazadeh, et al.
Veröffentlicht: (2024)
von: Davani, Aida Mostafazadeh, et al.
Veröffentlicht: (2024)
Quantifying and Predicting Disagreement in Graded Human Ratings
von: Zhang, Leixin, et al.
Veröffentlicht: (2026)
von: Zhang, Leixin, et al.
Veröffentlicht: (2026)
From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning
von: Purohit, Kiran, et al.
Veröffentlicht: (2026)
von: Purohit, Kiran, et al.
Veröffentlicht: (2026)
TraceBack: Multi-Agent Decomposition for Fine-Grained Table Attribution
von: Anvekar, Tejas, et al.
Veröffentlicht: (2026)
von: Anvekar, Tejas, et al.
Veröffentlicht: (2026)
Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis
von: Lu, Junyu, et al.
Veröffentlicht: (2026)
von: Lu, Junyu, et al.
Veröffentlicht: (2026)
NUTMEG: Separating Signal From Noise in Annotator Disagreement
von: Ivey, Jonathan, et al.
Veröffentlicht: (2025)
von: Ivey, Jonathan, et al.
Veröffentlicht: (2025)
Do Differences in Values Influence Disagreements in Online Discussions?
von: van der Meer, Michiel, et al.
Veröffentlicht: (2023)
von: van der Meer, Michiel, et al.
Veröffentlicht: (2023)
Bridging the Gap: In-Context Learning for Modeling Human Disagreement
von: Muscato, Benedetta, et al.
Veröffentlicht: (2025)
von: Muscato, Benedetta, et al.
Veröffentlicht: (2025)
From Disagreement to Understanding: The Case for Ambiguity Detection in NLI
von: Jayaweera, Chathuri, et al.
Veröffentlicht: (2025)
von: Jayaweera, Chathuri, et al.
Veröffentlicht: (2025)
CogniAlign: Survivability-Grounded Multi-Agent Moral Reasoning for Safe and Transparent AI
von: Ali, Hasin Jawad, et al.
Veröffentlicht: (2025)
von: Ali, Hasin Jawad, et al.
Veröffentlicht: (2025)
ReasoningFlow: Semantic Structure of Complex Reasoning Traces
von: Lee, Jinu, et al.
Veröffentlicht: (2025)
von: Lee, Jinu, et al.
Veröffentlicht: (2025)
MASLegalBench: Benchmarking Multi-Agent Systems in Deductive Legal Reasoning
von: Jing, Huihao, et al.
Veröffentlicht: (2025)
von: Jing, Huihao, et al.
Veröffentlicht: (2025)
Learning to Break: Knowledge-Enhanced Reasoning in Multi-Agent Debate System
von: Wang, Haotian, et al.
Veröffentlicht: (2023)
von: Wang, Haotian, et al.
Veröffentlicht: (2023)
TraceSIR: A Multi-Agent Framework for Structured Analysis and Reporting of Agentic Execution Traces
von: Yang, Shu-Xun, et al.
Veröffentlicht: (2026)
von: Yang, Shu-Xun, et al.
Veröffentlicht: (2026)
Insight Agents: An LLM-Based Multi-Agent System for Data Insights
von: Bai, Jincheng, et al.
Veröffentlicht: (2026)
von: Bai, Jincheng, et al.
Veröffentlicht: (2026)
Beyond Consensus: Perspectivist Modeling and Evaluation of Annotator Disagreement in NLP
von: Xu, Yinuo, et al.
Veröffentlicht: (2026)
von: Xu, Yinuo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Temperature and Persona Shape LLM Agent Consensus With Minimal Accuracy Gains in Qualitative Coding
von: Borchers, Conrad, et al.
Veröffentlicht: (2025) -
Mitigating "Epistemic Debt" in Generative AI-Scaffolded Novice Programming using Metacognitive Scripts
von: Sankaranarayanan, Sreecharan
Veröffentlicht: (2026) -
Can Large Language Models Match Tutoring System Adaptivity? A Benchmarking Study
von: Borchers, Conrad, et al.
Veröffentlicht: (2025) -
Parametric Constraints for Bayesian Knowledge Tracing from First Principles
von: Shchepakin, Denis, et al.
Veröffentlicht: (2023) -
Disentangling Learning from Judgment: Representation Learning for Open Response Analytics
von: Borchers, Conrad, et al.
Veröffentlicht: (2025)