TraceSIR: A Multi-Agent Framework for Structured Analysis and Reporting of Agentic Execution Traces
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Shu-Xun, Wang, Cunxiang, Zhang, Haoke, Yu, Wenbo, Wu, Lindong, Gui, Jiayi, Yang, Dayong, Cen, Yukuo, Feng, Zhuoer, Wen, Bosi, Wang, Yidong, Zhong, Lucen, Ren, Jiamin, Zhang, Linfeng, Tang, Jie |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RLAR: An Agentic Reward System for Multi-task Reinforcement Learning on Large Language Models
by: Feng, Andrew Zhuoer, et al.
Published: (2026)
by: Feng, Andrew Zhuoer, et al.
Published: (2026)
UDA: Unsupervised Debiasing Alignment for Pair-wise LLM-as-a-Judge
by: Zhang, Yang, et al.
Published: (2025)
by: Zhang, Yang, et al.
Published: (2025)
RAVEL: Reasoning Agents for Validating and Evaluating LLM Text Synthesis
by: Feng, Andrew Zhuoer, et al.
Published: (2026)
by: Feng, Andrew Zhuoer, et al.
Published: (2026)
PST-Bench: Tracing and Benchmarking the Source of Publications
by: Zhang, Fanjin, et al.
Published: (2024)
by: Zhang, Fanjin, et al.
Published: (2024)
StepMathAgent: A Step-Wise Agent for Evaluating Mathematical Processes through Tree-of-Error
by: Yang, Shu-Xun, et al.
Published: (2025)
by: Yang, Shu-Xun, et al.
Published: (2025)
Unlocking Recursive Thinking of LLMs: Alignment via Refinement
by: Zhang, Haoke, et al.
Published: (2025)
by: Zhang, Haoke, et al.
Published: (2025)
TraceLLM: Security Diagnosis Through Traces and Smart Contracts in Ethereum
by: Wang, Shuzheng, et al.
Published: (2025)
by: Wang, Shuzheng, et al.
Published: (2025)
Personalized Forgetting Mechanism with Concept-Driven Knowledge Tracing
by: Wang, Shanshan, et al.
Published: (2024)
by: Wang, Shanshan, et al.
Published: (2024)
Dual-State Personalized Knowledge Tracing with Emotional Incorporation
by: Wang, Shanshan, et al.
Published: (2024)
by: Wang, Shanshan, et al.
Published: (2024)
No Certificate, No Execution: Certified Traces as a Foundation for Trustworthy AI Agents
by: Yanglet, Xiao-Yang Liu, et al.
Published: (2026)
by: Yanglet, Xiao-Yang Liu, et al.
Published: (2026)
Transformed $\ell_1$ Regularizations for Robust Principal Component Analysis: Toward a Fine-Grained Understanding
by: Zhao, Kun, et al.
Published: (2025)
by: Zhao, Kun, et al.
Published: (2025)
Modello Cosmologico Diaframmico basato su E = Mp^2: Origine cosciente dell'universo
by: Lucen, et al.
Published: (2025)
by: Lucen, et al.
Published: (2025)
Demystifying Errors in LLM Reasoning Traces: An Empirical Study of Code Execution Simulation
by: Abdollahi, Mohammad, et al.
Published: (2025)
by: Abdollahi, Mohammad, et al.
Published: (2025)
DVD: A Robust Method for Detecting Variant Contamination in Large Language Model Evaluation
by: Liang, Renzhao, et al.
Published: (2026)
by: Liang, Renzhao, et al.
Published: (2026)
Buchi neri come diaframmi cosmici: permeabilità informazionale, materia oscura e coerenza tra universi
by: Lucen, et al.
Published: (2025)
by: Lucen, et al.
Published: (2025)
ReasoningShield: Safety Detection over Reasoning Traces of Large Reasoning Models
by: Li, Changyi, et al.
Published: (2025)
by: Li, Changyi, et al.
Published: (2025)
Enhancing the Corrosion Resistance of Cu–Fe Alloy by Trace Yttrium Addition
by: Jing Xu, et al.
Published: (2026)
by: Jing Xu, et al.
Published: (2026)
StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Trace is the Next AutoDiff: Generative Optimization with Rich Feedback, Execution Traces, and LLMs
by: Cheng, Ching-An, et al.
Published: (2024)
by: Cheng, Ching-An, et al.
Published: (2024)
Executable Boundary Contracts for Sound Event Traces
by: Alpay, Faruk, et al.
Published: (2026)
by: Alpay, Faruk, et al.
Published: (2026)
Simple Fault Localization using Execution Traces
by: Prenner, Julian Aron, et al.
Published: (2025)
by: Prenner, Julian Aron, et al.
Published: (2025)
TraceTrans: Translation and Spatial Tracing for Surgical Prediction
by: Luo, Xiyu, et al.
Published: (2025)
by: Luo, Xiyu, et al.
Published: (2025)
Pre-Training and Prompting for Few-Shot Node Classification on Text-Attributed Graphs
by: Zhao, Huanjing, et al.
Published: (2024)
by: Zhao, Huanjing, et al.
Published: (2024)
Artefact for Revisiting the Combination of Static Analysis Error Traces and Dynamic Symbolic Execution
by: Xu, Yihua, et al.
Published: (2026)
by: Xu, Yihua, et al.
Published: (2026)
TraceMem: Weaving Narrative Memory Schemata from User Conversational Traces
by: Shu, Yiming, et al.
Published: (2026)
by: Shu, Yiming, et al.
Published: (2026)
threaTrace: Detecting and Tracing Host-based Threats in Node Level Through Provenance Graph Learning
by: Wang, Su, et al.
Published: (2021)
by: Wang, Su, et al.
Published: (2021)
Can Brand Activism Help? The Effects of Fit and Consumer Trait in Building Brand Trust
by: Cen Wang, et al.
Published: (2025)
by: Cen Wang, et al.
Published: (2025)
IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation
by: Wen, Bosi, et al.
Published: (2026)
by: Wen, Bosi, et al.
Published: (2026)
TRAIL: Trace Reasoning and Agentic Issue Localization
by: Deshpande, Darshan, et al.
Published: (2025)
by: Deshpande, Darshan, et al.
Published: (2025)
PrefRAG: Preference-Driven Multi-Source Retrieval Augmented Generation
by: Zhao, Qingfei, et al.
Published: (2024)
by: Zhao, Qingfei, et al.
Published: (2024)
Runtime Execution Traces Guided Automated Program Repair with Multi-Agent Debate
by: Wu, Jiaqing, et al.
Published: (2026)
by: Wu, Jiaqing, et al.
Published: (2026)
TraceCoder: A Trace-Driven Multi-Agent Framework for Automated Debugging of LLM-Generated Code
by: Huang, Jiangping, et al.
Published: (2026)
by: Huang, Jiangping, et al.
Published: (2026)
GraphAlign: Pretraining One Graph Neural Network on Multiple Graphs via Feature Alignment
by: Hou, Zhenyu, et al.
Published: (2024)
by: Hou, Zhenyu, et al.
Published: (2024)
LLM-driven Effective Knowledge Tracing by Integrating Dual-channel Difficulty
by: Cen, Jiahui, et al.
Published: (2025)
by: Cen, Jiahui, et al.
Published: (2025)
How Likely Do LLMs with CoT Mimic Human Reasoning?
by: Bao, Guangsheng, et al.
Published: (2024)
by: Bao, Guangsheng, et al.
Published: (2024)
Fast Quantum Algorithms for Trace Distance Estimation
by: Wang, Qisheng, et al.
Published: (2023)
by: Wang, Qisheng, et al.
Published: (2023)
TraceFlow: Dynamic 3D Reconstruction of Specular Scenes Driven by Ray Tracing
by: Tao, Jiachen, et al.
Published: (2025)
by: Tao, Jiachen, et al.
Published: (2025)
Willful Disobedience: Automatically Detecting Failures in Agentic Traces
by: Sharma, Reshabh K, et al.
Published: (2026)
by: Sharma, Reshabh K, et al.
Published: (2026)
Teaching LLMs Program Semantics via Symbolic Execution Traces
by: Bayer, Jonas, et al.
Published: (2026)
by: Bayer, Jonas, et al.
Published: (2026)
GeoBrowse: A Geolocation Benchmark for Agentic Tool Use with Expert-Annotated Reasoning Traces
by: Geng, Xinyu, et al.
Published: (2026)
by: Geng, Xinyu, et al.
Published: (2026)
Similar Items
-
RLAR: An Agentic Reward System for Multi-task Reinforcement Learning on Large Language Models
by: Feng, Andrew Zhuoer, et al.
Published: (2026) -
UDA: Unsupervised Debiasing Alignment for Pair-wise LLM-as-a-Judge
by: Zhang, Yang, et al.
Published: (2025) -
RAVEL: Reasoning Agents for Validating and Evaluating LLM Text Synthesis
by: Feng, Andrew Zhuoer, et al.
Published: (2026) -
PST-Bench: Tracing and Benchmarking the Source of Publications
by: Zhang, Fanjin, et al.
Published: (2024) -
StepMathAgent: A Step-Wise Agent for Evaluating Mathematical Processes through Tree-of-Error
by: Yang, Shu-Xun, et al.
Published: (2025)