Saved in:
| Main Authors: | West, Alva, Weng, Yixuan, Zhu, Minjun, Lin, Zhen, Ning, Zhiyuan, Zhang, Yue |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.10401 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI-Generated Text is Non-Stationary: Detection via Temporal Tomography
by: West, Alva, et al.
Published: (2025)
by: West, Alva, et al.
Published: (2025)
T-Detect: Tail-Aware Statistical Normalization for Robust Detection of Adversarial Machine-Generated Text
by: West, Alva, et al.
Published: (2025)
by: West, Alva, et al.
Published: (2025)
DeepReviewer 2.0: A Traceable Agentic System for Auditable Scientific Peer Review
by: Weng, Yixuan, et al.
Published: (2026)
by: Weng, Yixuan, et al.
Published: (2026)
CycleResearcher: Improving Automated Research via Automated Review
by: Weng, Yixuan, et al.
Published: (2024)
by: Weng, Yixuan, et al.
Published: (2024)
AI Scientists Fail Without Strong Implementation Capability
by: Zhu, Minjun, et al.
Published: (2025)
by: Zhu, Minjun, et al.
Published: (2025)
Automatic Failure Attribution and Critical Step Prediction Method for Multi-Agent Systems Based on Causal Inference
by: Ma, Guoqing, et al.
Published: (2025)
by: Ma, Guoqing, et al.
Published: (2025)
Personality Alignment of Large Language Models
by: Zhu, Minjun, et al.
Published: (2024)
by: Zhu, Minjun, et al.
Published: (2024)
AutoFigure: Generating and Refining Publication-Ready Scientific Illustrations
by: Zhu, Minjun, et al.
Published: (2026)
by: Zhu, Minjun, et al.
Published: (2026)
DeepReview: Improving LLM-based Paper Review with Human-like Deep Thinking Process
by: Zhu, Minjun, et al.
Published: (2025)
by: Zhu, Minjun, et al.
Published: (2025)
AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems
by: Zhang, Boxuan, et al.
Published: (2026)
by: Zhang, Boxuan, et al.
Published: (2026)
DeepScientist: Advancing Frontier-Pushing Scientific Findings Progressively
by: Weng, Yixuan, et al.
Published: (2025)
by: Weng, Yixuan, et al.
Published: (2025)
Which Agent Causes Task Failures and When? On Automated Failure Attribution of LLM Multi-Agent Systems
by: Zhang, Shaokun, et al.
Published: (2025)
by: Zhang, Shaokun, et al.
Published: (2025)
When Only the Final Text Survives: Implicit Execution Tracing for Multi-Agent Attribution
by: Nian, Yi, et al.
Published: (2026)
by: Nian, Yi, et al.
Published: (2026)
PreAct: Prediction Enhances Agent's Planning Ability
by: Fu, Dayuan, et al.
Published: (2024)
by: Fu, Dayuan, et al.
Published: (2024)
SafeReview: Defending LLM-based Review Systems Against Adversarial Hidden Prompts
by: Xin, Yuan, et al.
Published: (2026)
by: Xin, Yuan, et al.
Published: (2026)
Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks
by: Tan, Rongyuan, et al.
Published: (2026)
by: Tan, Rongyuan, et al.
Published: (2026)
Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge
by: Li, Bin, et al.
Published: (2025)
by: Li, Bin, et al.
Published: (2025)
Bi-directional Bias Attribution: Debiasing Large Language Models without Modifying Prompts
by: Lin, Yujie, et al.
Published: (2026)
by: Lin, Yujie, et al.
Published: (2026)
MARS: Multi-Agent Adaptive Reasoning with Socratic Guidance for Automated Prompt Optimization
by: Zhang, Jian, et al.
Published: (2025)
by: Zhang, Jian, et al.
Published: (2025)
LingYi: Medical Conversational Question Answering System based on Multi-modal Knowledge Graphs
by: Xia, Fei, et al.
Published: (2022)
by: Xia, Fei, et al.
Published: (2022)
Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System
by: Chen, Weize, et al.
Published: (2024)
by: Chen, Weize, et al.
Published: (2024)
Attention with Dependency Parsing Augmentation for Fine-Grained Attribution
by: Ding, Qiang, et al.
Published: (2024)
by: Ding, Qiang, et al.
Published: (2024)
Code Broker: A Multi-Agent System for Automated Code Quality Assessment
by: Attrah, Samer
Published: (2026)
by: Attrah, Samer
Published: (2026)
AutoFigure-Edit: Generating Editable Scientific Illustration
by: Lin, Zhen, et al.
Published: (2026)
by: Lin, Zhen, et al.
Published: (2026)
An Empirical Study on Failures in Automated Issue Solving
by: Liu, Simiao, et al.
Published: (2025)
by: Liu, Simiao, et al.
Published: (2025)
MADIAVE: Multi-Agent Debate for Implicit Attribute Value Extraction
by: Huang, Wei-Chieh, et al.
Published: (2025)
by: Huang, Wei-Chieh, et al.
Published: (2025)
Predicting vs. Acting: A Trade-off Between World Modeling & Agent Modeling
by: Li, Margaret, et al.
Published: (2024)
by: Li, Margaret, et al.
Published: (2024)
CausalAgent: A Conversational Multi-Agent System for End-to-End Causal Inference
by: Zhu, Jiawei, et al.
Published: (2026)
by: Zhu, Jiawei, et al.
Published: (2026)
VerifyMAS: Hypothesis Verification for Failure Attribution in LLM Multi-Agent Systems
by: Qiao, Hezhe, et al.
Published: (2026)
by: Qiao, Hezhe, et al.
Published: (2026)
Is Your LLM Really Mastering the Concept? A Multi-Agent Benchmark
by: Xu, Shuhang, et al.
Published: (2025)
by: Xu, Shuhang, et al.
Published: (2025)
GEM: Graph-Enhanced Mixture-of-Experts with ReAct Agents for Dialogue State Tracking
by: Zhu, Ziqi, et al.
Published: (2026)
by: Zhu, Ziqi, et al.
Published: (2026)
CausalFlow: Causal Attribution and Counterfactual Repair for LLM Agent Failures
by: Bonagiri, Akash, et al.
Published: (2026)
by: Bonagiri, Akash, et al.
Published: (2026)
Small but Mighty: Enhancing Time Series Forecasting with Lightweight LLMs
by: Fan, Haoran, et al.
Published: (2025)
by: Fan, Haoran, et al.
Published: (2025)
From Flat Logs to Causal Graphs: Hierarchical Failure Attribution for LLM-based Multi-Agent Systems
by: Wang, Yawen, et al.
Published: (2026)
by: Wang, Yawen, et al.
Published: (2026)
See, Think, Act: Teaching Multimodal Agents to Effectively Interact with GUI by Identifying Toggles
by: Wu, Zongru, et al.
Published: (2025)
by: Wu, Zongru, et al.
Published: (2025)
Language-Specific Representation of Emotion-Concept Knowledge Causally Supports Emotion Inference
by: Li, Ming, et al.
Published: (2023)
by: Li, Ming, et al.
Published: (2023)
Insight Agents: An LLM-Based Multi-Agent System for Data Insights
by: Bai, Jincheng, et al.
Published: (2026)
by: Bai, Jincheng, et al.
Published: (2026)
MAD-Fact: A Multi-Agent Debate Framework for Long-Form Factuality Evaluation in LLMs
by: Ning, Yucheng, et al.
Published: (2025)
by: Ning, Yucheng, et al.
Published: (2025)
DocAgent: A Multi-Agent System for Automated Code Documentation Generation
by: Yang, Dayu, et al.
Published: (2025)
by: Yang, Dayu, et al.
Published: (2025)
Hidden Error Awareness in Chain-of-Thought Reasoning: The Signal Is Diagnostic, Not Causal
by: Yuan, Aojie, et al.
Published: (2026)
by: Yuan, Aojie, et al.
Published: (2026)
Similar Items
-
AI-Generated Text is Non-Stationary: Detection via Temporal Tomography
by: West, Alva, et al.
Published: (2025) -
T-Detect: Tail-Aware Statistical Normalization for Robust Detection of Adversarial Machine-Generated Text
by: West, Alva, et al.
Published: (2025) -
DeepReviewer 2.0: A Traceable Agentic System for Auditable Scientific Peer Review
by: Weng, Yixuan, et al.
Published: (2026) -
CycleResearcher: Improving Automated Research via Automated Review
by: Weng, Yixuan, et al.
Published: (2024) -
AI Scientists Fail Without Strong Implementation Capability
by: Zhu, Minjun, et al.
Published: (2025)