Saved in:
| Main Authors: | Nian, Junjie, Chen, Kang, Zhang, Ge, Cao, Yixin, Jiang, Yugang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.31308 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SliceGraph: Mapping Process Isomers in Multi-Run Chain-of-Thought Reasoning
by: Chen, Kang, et al.
Published: (2026)
by: Chen, Kang, et al.
Published: (2026)
ARM: Role-Conditioned Neuron Transplantation for Training-Free Generalist LLM Agent Merging
by: Feng, Zhuoka, et al.
Published: (2026)
by: Feng, Zhuoka, et al.
Published: (2026)
NEX: Neuron Explore-Exploit Scoring for Label-Free Chain-of-Thought Selection and Model Ranking
by: Chen, Kang, et al.
Published: (2026)
by: Chen, Kang, et al.
Published: (2026)
Thinking Traps in Long Chain-of-Thought: A Measurable Study and Trap-Aware Adaptive Restart
by: Chen, Kang, et al.
Published: (2026)
by: Chen, Kang, et al.
Published: (2026)
AgentRx: Diagnosing AI Agent Failures from Execution Trajectories
by: Barke, Shraddha, et al.
Published: (2026)
by: Barke, Shraddha, et al.
Published: (2026)
Fixing the Broken Compass: Diagnosing and Improving Inference-Time Reward Modeling
by: Li, Jiachun, et al.
Published: (2025)
by: Li, Jiachun, et al.
Published: (2025)
Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
by: Ni, Jingwei, et al.
Published: (2026)
by: Ni, Jingwei, et al.
Published: (2026)
VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing
by: Shi, Haoyuan, et al.
Published: (2026)
by: Shi, Haoyuan, et al.
Published: (2026)
When Only the Final Text Survives: Implicit Execution Tracing for Multi-Agent Attribution
by: Nian, Yi, et al.
Published: (2026)
by: Nian, Yi, et al.
Published: (2026)
AIR: Improving Agent Safety through Incident Response
by: Xiao, Zibo, et al.
Published: (2026)
by: Xiao, Zibo, et al.
Published: (2026)
Diagnosing Harmful Continuation in Answer-Correct Long-CoT Training Traces
by: He, Chen, et al.
Published: (2026)
by: He, Chen, et al.
Published: (2026)
HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents
by: Jin, Chao, et al.
Published: (2026)
by: Jin, Chao, et al.
Published: (2026)
Knowledge Graph Completion by Intermediate Variables Regularization
by: Xiao, Changyi, et al.
Published: (2025)
by: Xiao, Changyi, et al.
Published: (2025)
Diagnosing Failure Modes of Shared-State Collaboration in Resource-Constrained Visual Agents
by: Zhou, Yunpeng
Published: (2026)
by: Zhou, Yunpeng
Published: (2026)
WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory Pruning
by: Wang, Junjie, et al.
Published: (2026)
by: Wang, Junjie, et al.
Published: (2026)
Complex Logical Query Answering by Calibrating Knowledge Graph Completion Models
by: Xiao, Changyi, et al.
Published: (2024)
by: Xiao, Changyi, et al.
Published: (2024)
Diagnosing Korean-Language LLM Political Bias via Census-Grounded Agent Simulation
by: Kang, Sungwoo
Published: (2026)
by: Kang, Sungwoo
Published: (2026)
Improving Large Language Models in Event Relation Logical Prediction
by: Chen, Meiqi, et al.
Published: (2023)
by: Chen, Meiqi, et al.
Published: (2023)
AgentsCourt: Building Judicial Decision-Making Agents with Court Debate Simulation and Legal Knowledge Augmentation
by: He, Zhitao, et al.
Published: (2024)
by: He, Zhitao, et al.
Published: (2024)
WebGraphEval: Multi-Turn Trajectory Evaluation for Web Agents using Graph Representation
by: Qian, Yaoyao, et al.
Published: (2025)
by: Qian, Yaoyao, et al.
Published: (2025)
Knowledge Graph Embedding by Normalizing Flows
by: Xiao, Changyi, et al.
Published: (2024)
by: Xiao, Changyi, et al.
Published: (2024)
Explaining RL Decisions with Trajectories
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2023)
by: Deshmukh, Shripad Vilasrao, et al.
Published: (2023)
Graph-Augmented Large Language Model Agents: Current Progress and Future Prospects
by: Liu, Yixin, et al.
Published: (2025)
by: Liu, Yixin, et al.
Published: (2025)
Diagnosing and Remedying Knowledge Deficiencies in LLMs via Label-free Curricular Meaningful Learning
by: Xiong, Kai, et al.
Published: (2024)
by: Xiong, Kai, et al.
Published: (2024)
CivRealm: A Learning and Reasoning Odyssey in Civilization for Decision-Making Agents
by: Qi, Siyuan, et al.
Published: (2024)
by: Qi, Siyuan, et al.
Published: (2024)
Agent Audit: A Security Analysis System for LLM Agent Applications
by: Zhang, Haiyue, et al.
Published: (2026)
by: Zhang, Haiyue, et al.
Published: (2026)
SciToolAgent: A Knowledge Graph-Driven Scientific Agent for Multi-Tool Integration
by: Ding, Keyan, et al.
Published: (2025)
by: Ding, Keyan, et al.
Published: (2025)
Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resolution Profiles
by: Yan, Lu, et al.
Published: (2026)
by: Yan, Lu, et al.
Published: (2026)
CORE-Acu: Structured Reasoning Traces and Knowledge Graph Safety Verification for Acupuncture Clinical Decision Support
by: Xu, Liuyi, et al.
Published: (2026)
by: Xu, Liuyi, et al.
Published: (2026)
AgentProcessBench: Diagnosing Step-Level Process Quality in Tool-Using Agents
by: Fan, Shengda, et al.
Published: (2026)
by: Fan, Shengda, et al.
Published: (2026)
AQUAH: Automatic Quantification and Unified Agent in Hydrology
by: Yan, Songkun, et al.
Published: (2025)
by: Yan, Songkun, et al.
Published: (2025)
DynaTrust: Defending Multi-Agent Systems Against Sleeper Agents via Dynamic Trust Graphs
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
An Improved Multi-Agent Algorithm for Cooperative and Competitive Environments by Identifying and Encouraging Cooperation among Agents
by: Qi, Junjie, et al.
Published: (2025)
by: Qi, Junjie, et al.
Published: (2025)
Auditable Agents
by: Nian, Yi, et al.
Published: (2026)
by: Nian, Yi, et al.
Published: (2026)
Diagnosing Memorization in Chain-of-Thought Reasoning, One Token at a Time
by: Li, Huihan, et al.
Published: (2025)
by: Li, Huihan, et al.
Published: (2025)
What Do LLM Agents Know About Their World? Task2Quiz: A Paradigm for Studying Environment Understanding
by: Liu, Siyuan, et al.
Published: (2026)
by: Liu, Siyuan, et al.
Published: (2026)
EconAgent: Large Language Model-Empowered Agents for Simulating Macroeconomic Activities
by: Li, Nian, et al.
Published: (2023)
by: Li, Nian, et al.
Published: (2023)
SPGNN: Recognizing Salient Subgraph Patterns via Enhanced Graph Convolution and Pooling
by: Dong, Zehao, et al.
Published: (2024)
by: Dong, Zehao, et al.
Published: (2024)
CureAgent: A Training-Free Executor-Analyst Framework for Clinical Reasoning
by: Xie, Ting-Ting, et al.
Published: (2025)
by: Xie, Ting-Ting, et al.
Published: (2025)
A Large-Scale Empirical Study on Improving the Fairness of Image Classification Models
by: Yang, Junjie, et al.
Published: (2024)
by: Yang, Junjie, et al.
Published: (2024)
Similar Items
-
SliceGraph: Mapping Process Isomers in Multi-Run Chain-of-Thought Reasoning
by: Chen, Kang, et al.
Published: (2026) -
ARM: Role-Conditioned Neuron Transplantation for Training-Free Generalist LLM Agent Merging
by: Feng, Zhuoka, et al.
Published: (2026) -
NEX: Neuron Explore-Exploit Scoring for Label-Free Chain-of-Thought Selection and Model Ranking
by: Chen, Kang, et al.
Published: (2026) -
Thinking Traps in Long Chain-of-Thought: A Measurable Study and Trap-Aware Adaptive Restart
by: Chen, Kang, et al.
Published: (2026) -
AgentRx: Diagnosing AI Agent Failures from Execution Trajectories
by: Barke, Shraddha, et al.
Published: (2026)