When Only the Final Text Survives: Implicit Execution Tracing for Multi-Agent Attribution
Fuente:
arXiv
Saved in:
| Main Authors: | Nian, Yi, Cao, Haosen, Zhu, Shenzhe, Zou, Henry Peng, Luan, Qingqing, Zhao, Yue |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HarmTransform: Transforming Explicit Harmful Queries into Stealthy via Multi-Agent Debate
by: Zhu, Shenzhe
Published: (2025)
by: Zhu, Shenzhe
Published: (2025)
The Automated but Risky Game: Modeling and Benchmarking Agent-to-Agent Negotiations and Transactions in Consumer Markets
by: Zhu, Shenzhe, et al.
Published: (2025)
by: Zhu, Shenzhe, et al.
Published: (2025)
TAGFN: A Text-Attributed Graph Dataset for Fake News Detection in the Age of LLMs
by: Liu, Kay, et al.
Published: (2025)
by: Liu, Kay, et al.
Published: (2025)
ImplicitAVE: An Open-Source Dataset and Multimodal LLMs Benchmark for Implicit Attribute Value Extraction
by: Zou, Henry Peng, et al.
Published: (2024)
by: Zou, Henry Peng, et al.
Published: (2024)
MADIAVE: Multi-Agent Debate for Implicit Attribute Value Extraction
by: Huang, Wei-Chieh, et al.
Published: (2025)
by: Huang, Wei-Chieh, et al.
Published: (2025)
TraceBack: Multi-Agent Decomposition for Fine-Grained Table Attribution
by: Anvekar, Tejas, et al.
Published: (2026)
by: Anvekar, Tejas, et al.
Published: (2026)
TraceSIR: A Multi-Agent Framework for Structured Analysis and Reporting of Agentic Execution Traces
by: Yang, Shu-Xun, et al.
Published: (2026)
by: Yang, Shu-Xun, et al.
Published: (2026)
Multi-User Large Language Model Agents
by: Yang, Shu, et al.
Published: (2026)
by: Yang, Shu, et al.
Published: (2026)
EIVEN: Efficient Implicit Attribute Value Extraction using Multimodal LLM
by: Zou, Henry Peng, et al.
Published: (2024)
by: Zou, Henry Peng, et al.
Published: (2024)
CITE: A Comprehensive Benchmark for Heterogeneous Text-Attributed Graphs on Catalytic Materials
by: Zhang, Chenghao, et al.
Published: (2025)
by: Zhang, Chenghao, et al.
Published: (2025)
Positive Text Reframing under Multi-strategy Optimization
by: Jia, Shutong, et al.
Published: (2024)
by: Jia, Shutong, et al.
Published: (2024)
When Users Change Their Mind: Evaluating Interruptible Agents in Long-Horizon Web Navigation
by: Zou, Henry Peng, et al.
Published: (2026)
by: Zou, Henry Peng, et al.
Published: (2026)
When Agents Trade: Live Multi-Market Trading Benchmark for LLM Agents
by: Qian, Lingfei, et al.
Published: (2025)
by: Qian, Lingfei, et al.
Published: (2025)
You Only Align Once: Propagating Cooperative Behaviors in Multi-Agent Systems through Seed Agents
by: Hsing, Nicole, et al.
Published: (2026)
by: Hsing, Nicole, et al.
Published: (2026)
Which Agent Causes Task Failures and When? On Automated Failure Attribution of LLM Multi-Agent Systems
by: Zhang, Shaokun, et al.
Published: (2025)
by: Zhang, Shaokun, et al.
Published: (2025)
MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems
by: Deng, Xinle, et al.
Published: (2026)
by: Deng, Xinle, et al.
Published: (2026)
Abduct, Act, Predict: Scaffolding Causal Inference for Automated Failure Attribution in Multi-Agent Systems
by: West, Alva, et al.
Published: (2025)
by: West, Alva, et al.
Published: (2025)
You Only Read Once (YORO): Learning to Internalize Database Knowledge for Text-to-SQL
by: Kobayashi, Hideo, et al.
Published: (2024)
by: Kobayashi, Hideo, et al.
Published: (2024)
VerifyMAS: Hypothesis Verification for Failure Attribution in LLM Multi-Agent Systems
by: Qiao, Hezhe, et al.
Published: (2026)
by: Qiao, Hezhe, et al.
Published: (2026)
Multi-Intent Attribute-Aware Text Matching in Searching
by: Li, Mingzhe, et al.
Published: (2024)
by: Li, Mingzhe, et al.
Published: (2024)
Ask Only When Needed: Proactive Retrieval from Memory and Skills for Experience-Driven Lifelong Agents
by: Cai, Yuxuan, et al.
Published: (2026)
by: Cai, Yuxuan, et al.
Published: (2026)
MALIBU Benchmark: Multi-Agent LLM Implicit Bias Uncovered
by: Mirza, Imran, et al.
Published: (2025)
by: Mirza, Imran, et al.
Published: (2025)
Evolving and Executing Research Plans via Double-Loop Multi-Agent Collaboration
by: Zhang, Zhi, et al.
Published: (2025)
by: Zhang, Zhi, et al.
Published: (2025)
GLEAN: Active Generalized Category Discovery with Diverse LLM Feedback
by: Zou, Henry Peng, et al.
Published: (2025)
by: Zou, Henry Peng, et al.
Published: (2025)
Q-Mirror: Unlocking the Multi-Modal Potential of Scientific Text-Only QA Pairs
by: Wang, Junying, et al.
Published: (2025)
by: Wang, Junying, et al.
Published: (2025)
FlexSQL: Flexible Exploration and Execution Make Better Text-to-SQL Agents
by: Pham, Quang Hieu, et al.
Published: (2026)
by: Pham, Quang Hieu, et al.
Published: (2026)
Can AI Truly Represent Your Voice in Deliberations? A Comprehensive Study of Large-Scale Opinion Aggregation with LLMs
by: Zhu, Shenzhe, et al.
Published: (2025)
by: Zhu, Shenzhe, et al.
Published: (2025)
Towards Data Contamination Detection for Modern Large Language Models: Limitations, Inconsistencies, and Oracle Challenges
by: Samuel, Vinay, et al.
Published: (2024)
by: Samuel, Vinay, et al.
Published: (2024)
ViHateT5: Enhancing Hate Speech Detection in Vietnamese With A Unified Text-to-Text Transformer Model
by: Nguyen, Luan Thanh
Published: (2024)
by: Nguyen, Luan Thanh
Published: (2024)
Beyond Negative Rollouts: Positive-Only Policy Optimization with Implicit Negative Gradients
by: Xu, Mingwei, et al.
Published: (2026)
by: Xu, Mingwei, et al.
Published: (2026)
Invoke Interfaces Only When Needed: Adaptive Invocation for Large Language Models in Question Answering
by: Zhao, Jihao, et al.
Published: (2025)
by: Zhao, Jihao, et al.
Published: (2025)
Shepherd: A Runtime Substrate Empowering Meta-Agents with a Formalized Execution Trace
by: Yu, Simon, et al.
Published: (2026)
by: Yu, Simon, et al.
Published: (2026)
GenerationPrograms: Fine-grained Attribution with Executable Programs
by: Wan, David, et al.
Published: (2025)
by: Wan, David, et al.
Published: (2025)
Towards Lightweight, Adaptive and Attribute-Aware Multi-Aspect Controllable Text Generation with Large Language Models
by: Zhu, Chenyu, et al.
Published: (2025)
by: Zhu, Chenyu, et al.
Published: (2025)
Think Only When You Need with Large Hybrid-Reasoning Models
by: Jiang, Lingjie, et al.
Published: (2025)
by: Jiang, Lingjie, et al.
Published: (2025)
Grounding Spatial Relations in Text-Only Language Models
by: Azkune, Gorka, et al.
Published: (2024)
by: Azkune, Gorka, et al.
Published: (2024)
Reason Only When Needed: Efficient Generative Reward Modeling via Model-Internal Uncertainty
by: Xue, Chao, et al.
Published: (2026)
by: Xue, Chao, et al.
Published: (2026)
Controllable Data Augmentation for Few-Shot Text Mining with Chain-of-Thought Attribute Manipulation
by: Peng, Letian, et al.
Published: (2023)
by: Peng, Letian, et al.
Published: (2023)
Nexus: Execution-Grounded Multi-Agent Test Oracle Synthesis
by: Huang, Dong, et al.
Published: (2025)
by: Huang, Dong, et al.
Published: (2025)
Atomic Skills are the Prerequisite: When Reinforcement Learning Synthesizes Compositional Reasoning, and When It Only Amplifies
by: Cheng, Sitao, et al.
Published: (2025)
by: Cheng, Sitao, et al.
Published: (2025)
Similar Items
-
HarmTransform: Transforming Explicit Harmful Queries into Stealthy via Multi-Agent Debate
by: Zhu, Shenzhe
Published: (2025) -
The Automated but Risky Game: Modeling and Benchmarking Agent-to-Agent Negotiations and Transactions in Consumer Markets
by: Zhu, Shenzhe, et al.
Published: (2025) -
TAGFN: A Text-Attributed Graph Dataset for Fake News Detection in the Age of LLMs
by: Liu, Kay, et al.
Published: (2025) -
ImplicitAVE: An Open-Source Dataset and Multimodal LLMs Benchmark for Implicit Attribute Value Extraction
by: Zou, Henry Peng, et al.
Published: (2024) -
MADIAVE: Multi-Agent Debate for Implicit Attribute Value Extraction
by: Huang, Wei-Chieh, et al.
Published: (2025)