What Do Agents Learn from Trajectory-SFT: Semantics or Interfaces?

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Gu, Weizheng, Li, Chengze, Yu, Zhuohao, Sun, Mengyuan, Yang, Zhibang, Wang, Wei, Jia, Hongrui, Zhang, Shikun, Ye, Wei
Format: Preprint
Publié: 2026
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866910008442093568
author Gu, Weizheng
Li, Chengze
Yu, Zhuohao
Sun, Mengyuan
Yang, Zhibang
Wang, Wei
Jia, Hongrui
Zhang, Shikun
Ye, Wei
author_facet Gu, Weizheng
Li, Chengze
Yu, Zhuohao
Sun, Mengyuan
Yang, Zhibang
Wang, Wei
Jia, Hongrui
Zhang, Shikun
Ye, Wei
contents Large language models are increasingly evaluated as interactive agents, yet standard agent benchmarks conflate two qualitatively distinct sources of success: semantic tool-use and interface-specific interaction pattern memorization. Because both mechanisms can yield identical task success on the original interface, benchmark scores alone are not identifiable evidence of environment-invariant capability. We propose PIPE, a protocol-level evaluation augmentation for diagnosing interface reliance by minimally rewriting environment interfaces while preserving task semantics and execution behavior. Across 16 environments from AgentBench and AgentGym and a range of open-source and API-based agents, PIPE reveals that trajectory-SFT substantially amplifies interface shortcutting: trained agents degrade sharply under minimal interface rewrites, while non-trajectory-trained models remain largely stable. We further introduce Interface Reliance (IR), a counterbalanced alias-based metric that quantifies preference for training-time interfaces, and show that interface shortcutting exhibits environment-dependent, non-monotonic training dynamics that remain invisible under standard evaluation. Our code is available at https://anonymous.4open.science/r/What-Do-Agents-Learn-from-Trajectory-SFT-Semantics-or-Interfaces--0831/.
format Preprint
id arxiv_https___arxiv_org_abs_2602_01611
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle What Do Agents Learn from Trajectory-SFT: Semantics or Interfaces?
Gu, Weizheng
Li, Chengze
Yu, Zhuohao
Sun, Mengyuan
Yang, Zhibang
Wang, Wei
Jia, Hongrui
Zhang, Shikun
Ye, Wei
Machine Learning
Large language models are increasingly evaluated as interactive agents, yet standard agent benchmarks conflate two qualitatively distinct sources of success: semantic tool-use and interface-specific interaction pattern memorization. Because both mechanisms can yield identical task success on the original interface, benchmark scores alone are not identifiable evidence of environment-invariant capability. We propose PIPE, a protocol-level evaluation augmentation for diagnosing interface reliance by minimally rewriting environment interfaces while preserving task semantics and execution behavior. Across 16 environments from AgentBench and AgentGym and a range of open-source and API-based agents, PIPE reveals that trajectory-SFT substantially amplifies interface shortcutting: trained agents degrade sharply under minimal interface rewrites, while non-trajectory-trained models remain largely stable. We further introduce Interface Reliance (IR), a counterbalanced alias-based metric that quantifies preference for training-time interfaces, and show that interface shortcutting exhibits environment-dependent, non-monotonic training dynamics that remain invisible under standard evaluation. Our code is available at https://anonymous.4open.science/r/What-Do-Agents-Learn-from-Trajectory-SFT-Semantics-or-Interfaces--0831/.
title What Do Agents Learn from Trajectory-SFT: Semantics or Interfaces?
topic Machine Learning
url https://arxiv.org/abs/2602.01611