Where Do Deep-Research Agents Go Wrong? Span-Level Error Localization in Agent Trajectories
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Jiaming, Feng, Ziteng, Wu, Jiangtao, Li, Ruihao, Xie, Qianqian, Ren, Yuxiang, Zhu, He, Han, Xueming, Meng, Fanyu, Feng, Junlan, Liu, Jiaheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Where Did It All Go Wrong? A Hierarchical Look into Multi-Agent Error Attribution
di: Banerjee, Adi, et al.
Pubblicazione: (2025)
di: Banerjee, Adi, et al.
Pubblicazione: (2025)
Where Did It Go Wrong? Capability-Oriented Failure Attribution for Vision-and-Language Navigation Agents
di: Chen, Jianming, et al.
Pubblicazione: (2026)
di: Chen, Jianming, et al.
Pubblicazione: (2026)
DR$^{3}$-Eval: Towards Realistic and Reproducible Deep Research Evaluation
di: Xie, Qianqian, et al.
Pubblicazione: (2026)
di: Xie, Qianqian, et al.
Pubblicazione: (2026)
ExeChecker: Where Did I Go Wrong?
di: Gu, Yiwen, et al.
Pubblicazione: (2024)
di: Gu, Yiwen, et al.
Pubblicazione: (2024)
What's Wrong with the Absolute Trajectory Error?
di: Lee, Seong Hun, et al.
Pubblicazione: (2022)
di: Lee, Seong Hun, et al.
Pubblicazione: (2022)
Where Do You Go? Pedestrian Trajectory Prediction using Scene Features
di: Rezaei, Mohammad Ali, et al.
Pubblicazione: (2025)
di: Rezaei, Mohammad Ali, et al.
Pubblicazione: (2025)
The Wrong Way to Go.
di: Wright, H. Curtis
Pubblicazione: (1979)
di: Wright, H. Curtis
Pubblicazione: (1979)
Thinking-Based Non-Thinking: Solving the Reward Hacking Problem in Training Hybrid Reasoning Models via Reinforcement Learning
di: Gan, Siyuan, et al.
Pubblicazione: (2026)
di: Gan, Siyuan, et al.
Pubblicazione: (2026)
Where Did It Go Wrong? Attributing Undesirable LLM Behaviors via Representation Gradient Tracing
di: Li, Zhe, et al.
Pubblicazione: (2025)
di: Li, Zhe, et al.
Pubblicazione: (2025)
Where AI Assurance Might Go Wrong: Initial lessons from engineering of critical systems
di: Bloomfield, Robin, et al.
Pubblicazione: (2025)
di: Bloomfield, Robin, et al.
Pubblicazione: (2025)
Post‐Artesunate Delayed Hemolysis: Anything That Can Go Wrong Will Go Wrong—Murphy's Law
di: Beliza Chemutai, et al.
Pubblicazione: (2024)
di: Beliza Chemutai, et al.
Pubblicazione: (2024)
Knowledge Augmented Finetuning Matters in both RAG and Agent Based Dialog Systems
di: Cai, Yucheng, et al.
Pubblicazione: (2025)
di: Cai, Yucheng, et al.
Pubblicazione: (2025)
Harnessing Diverse Perspectives: A Multi-Agent Framework for Enhanced Error Detection in Knowledge Graphs
di: Li, Yu, et al.
Pubblicazione: (2025)
di: Li, Yu, et al.
Pubblicazione: (2025)
Strategy-Aware Optimization Modeling with Reasoning LLMs
di: Zhao, Ruiqing, et al.
Pubblicazione: (2026)
di: Zhao, Ruiqing, et al.
Pubblicazione: (2026)
Conformal Agent Error Attribution
di: Feng, Naihe, et al.
Pubblicazione: (2026)
di: Feng, Naihe, et al.
Pubblicazione: (2026)
Do Agent Societies Develop Intellectual Elites? The Hidden Power Laws of Collective Cognition in LLM Multi-Agent Systems
di: Venkatesh, Kavana, et al.
Pubblicazione: (2026)
di: Venkatesh, Kavana, et al.
Pubblicazione: (2026)
Where Do the Joules Go? Diagnosing Inference Energy Consumption
di: Chung, Jae-Won, et al.
Pubblicazione: (2026)
di: Chung, Jae-Won, et al.
Pubblicazione: (2026)
Fat replacers: Where Do We Go From Here?
di: Pszczola, D. E
Pubblicazione: (1997)
di: Pszczola, D. E
Pubblicazione: (1997)
Fat Replacers: Where Do We Go From Here?
di: Pszczola, Donald E
Pubblicazione: (1997)
di: Pszczola, Donald E
Pubblicazione: (1997)
Information Literacy--Where Do We Go from Here?
di: Koch, Melissa
Pubblicazione: (2001)
di: Koch, Melissa
Pubblicazione: (2001)
Beyond One-Size-Fits-All: Adaptive Subgraph Denoising for Zero-Shot Graph Learning with Large Language Models
di: Li, Fengzhi, et al.
Pubblicazione: (2026)
di: Li, Fengzhi, et al.
Pubblicazione: (2026)
Bibliographical Instruction: Where It's At and Where It's Going.
di: McNallie, Bruce
Pubblicazione: (1982)
di: McNallie, Bruce
Pubblicazione: (1982)
Modeling of Moving Sound Sources Based on Array Measurements
di: Meng, Fanyu
Pubblicazione: (2022)
di: Meng, Fanyu
Pubblicazione: (2022)
Wrong Face, Wrong Move: The Social Dynamics of Emotion Misperception in Agent-Based Models
di: Freire-Obregón, David
Pubblicazione: (2025)
di: Freire-Obregón, David
Pubblicazione: (2025)
Rethinking the Design of Reinforcement Learning-Based Deep Research Agents
di: Wan, Yi, et al.
Pubblicazione: (2025)
di: Wan, Yi, et al.
Pubblicazione: (2025)
Where Are We Going?
di: Franklin, Hardy R.
Pubblicazione: (1976)
di: Franklin, Hardy R.
Pubblicazione: (1976)
Go Where the Grants Are
di: Anderson, Cynthia, et al.
Pubblicazione: (2008)
di: Anderson, Cynthia, et al.
Pubblicazione: (2008)
Where Are We Going?
di: Laughlin, Mildred Knight
Pubblicazione: (1988)
di: Laughlin, Mildred Knight
Pubblicazione: (1988)
TrajAgent: An LLM-Agent Framework for Trajectory Modeling via Large-and-Small Model Collaboration
di: Du, Yuwei, et al.
Pubblicazione: (2024)
di: Du, Yuwei, et al.
Pubblicazione: (2024)
MobiFlow: Real-World Mobile Agent Benchmarking through Trajectory Fusion
di: Feng, Yunfei, et al.
Pubblicazione: (2026)
di: Feng, Yunfei, et al.
Pubblicazione: (2026)
Can Competition Enhance the Proficiency of Agents Powered by Large Language Models in the Realm of News-driven Time Series Forecasting?
di: Zhang, Yuxuan, et al.
Pubblicazione: (2025)
di: Zhang, Yuxuan, et al.
Pubblicazione: (2025)
Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
di: Song, Yifan, et al.
Pubblicazione: (2024)
di: Song, Yifan, et al.
Pubblicazione: (2024)
Levels of Autonomy for AI Agents
di: Feng, K. J. Kevin, et al.
Pubblicazione: (2025)
di: Feng, K. J. Kevin, et al.
Pubblicazione: (2025)
Explaining Why Things Go Where They Go: Interpretable Constructs of Human Organizational Preferences
di: Fashae, Emmanuel, et al.
Pubblicazione: (2025)
di: Fashae, Emmanuel, et al.
Pubblicazione: (2025)
What Can Go Wrong During Caplet Stripping ?
di: Floc'h, Fabien Le
Pubblicazione: (2026)
di: Floc'h, Fabien Le
Pubblicazione: (2026)
Where Do Tokens Go? Understanding Pruning Behaviors in STEP at High Resolutions
di: Szczepanski, Michal, et al.
Pubblicazione: (2025)
di: Szczepanski, Michal, et al.
Pubblicazione: (2025)
Where Do We Go From Here? The Future of Gender and Negotiation Research
di: Allison Elias
Pubblicazione: (2025)
di: Allison Elias
Pubblicazione: (2025)
Continuing Education for Special Librarianship; Where Do We Go From Here?
di: Sloane, Margaret N.
Pubblicazione: (1968)
di: Sloane, Margaret N.
Pubblicazione: (1968)
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
di: Chang, Jiale, et al.
Pubblicazione: (2026)
di: Chang, Jiale, et al.
Pubblicazione: (2026)
How Far Are We from Genuinely Useful Deep Research Agents?
di: Zhang, Dingling, et al.
Pubblicazione: (2025)
di: Zhang, Dingling, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Where Did It All Go Wrong? A Hierarchical Look into Multi-Agent Error Attribution
di: Banerjee, Adi, et al.
Pubblicazione: (2025) -
Where Did It Go Wrong? Capability-Oriented Failure Attribution for Vision-and-Language Navigation Agents
di: Chen, Jianming, et al.
Pubblicazione: (2026) -
DR$^{3}$-Eval: Towards Realistic and Reproducible Deep Research Evaluation
di: Xie, Qianqian, et al.
Pubblicazione: (2026) -
ExeChecker: Where Did I Go Wrong?
di: Gu, Yiwen, et al.
Pubblicazione: (2024) -
What's Wrong with the Absolute Trajectory Error?
di: Lee, Seong Hun, et al.
Pubblicazione: (2022)