Saved in:
| Main Authors: | Zhang, Guijia, Yang, Shu, Gong, Xilin, Wang, Di |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.10286 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hallucination as Exploit: Evidence-Carrying Multimodal Agents
by: Zhang, Guijia, et al.
Published: (2026)
by: Zhang, Guijia, et al.
Published: (2026)
Skill or Skip? Learning Selective Skill Invocation in Agentic Tasks via Dual-Granularity Preference Learning
by: Chen, Chishui, et al.
Published: (2026)
by: Chen, Chishui, et al.
Published: (2026)
Logic Agent: Enhancing Validity with Logic Rule Invocation
by: Liu, Hanmeng, et al.
Published: (2024)
by: Liu, Hanmeng, et al.
Published: (2024)
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills
by: Hou, Yingyong, et al.
Published: (2026)
by: Hou, Yingyong, et al.
Published: (2026)
When Agents Fail to Act: A Diagnostic Framework for Tool Invocation Reliability in Multi-Agent LLM Systems
by: Huang, Donghao, et al.
Published: (2026)
by: Huang, Donghao, et al.
Published: (2026)
MONICA: Real-Time Monitoring and Calibration of Chain-of-Thought Sycophancy in Large Reasoning Models
by: Hu, Jingyu, et al.
Published: (2025)
by: Hu, Jingyu, et al.
Published: (2025)
Counterfactual Trace Auditing of LLM Agent Skills
by: Zhou, Xiaolin, et al.
Published: (2026)
by: Zhou, Xiaolin, et al.
Published: (2026)
Red-Teaming Coding Agents from a Tool-Invocation Perspective: An Empirical Security Assessment
by: Xie, Yuchong, et al.
Published: (2025)
by: Xie, Yuchong, et al.
Published: (2025)
When Modalities Conflict: How Unimodal Reasoning Uncertainty Governs Preference Dynamics in MLLMs
by: Zhang, Zhuoran, et al.
Published: (2025)
by: Zhang, Zhuoran, et al.
Published: (2025)
DroidCall: A Dataset for LLM-powered Android Intent Invocation
by: Xie, Weikai, et al.
Published: (2024)
by: Xie, Weikai, et al.
Published: (2024)
Analyzing Message-Code Inconsistency in AI Coding Agent-Authored Pull Requests
by: Gong, Jingzhi, et al.
Published: (2026)
by: Gong, Jingzhi, et al.
Published: (2026)
SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces
by: Jin, Chang, et al.
Published: (2026)
by: Jin, Chang, et al.
Published: (2026)
Advancing and Benchmarking Personalized Tool Invocation for LLMs
by: Huang, Xu, et al.
Published: (2025)
by: Huang, Xu, et al.
Published: (2025)
Structured Security Auditing and Robustness Enhancement for Untrusted Agent Skills
by: Lv, Lijia, et al.
Published: (2026)
by: Lv, Lijia, et al.
Published: (2026)
Auditable Agents
by: Nian, Yi, et al.
Published: (2026)
by: Nian, Yi, et al.
Published: (2026)
SkillMaster: Toward Autonomous Skill Mastery in LLM Agents
by: Yang, Min, et al.
Published: (2026)
by: Yang, Min, et al.
Published: (2026)
Trigger without Trace: Towards Stealthy Backdoor Attack on Text-to-Image Diffusion Models
by: Zhang, Jie, et al.
Published: (2025)
by: Zhang, Jie, et al.
Published: (2025)
Request-Only Optimization for Recommendation Systems
by: Guo, Liang, et al.
Published: (2025)
by: Guo, Liang, et al.
Published: (2025)
SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision
by: Liu, Yuxuan, et al.
Published: (2026)
by: Liu, Yuxuan, et al.
Published: (2026)
GeoMind: An Agentic Workflow for Lithology Classification with Reasoned Tool Invocation
by: Zhou, Yitong, et al.
Published: (2026)
by: Zhou, Yitong, et al.
Published: (2026)
ScaleSim: Serving Large-Scale Multi-Agent Simulation with Invocation Distance-Based Memory Management
by: Pan, Zaifeng, et al.
Published: (2026)
by: Pan, Zaifeng, et al.
Published: (2026)
Agent Audit: A Security Analysis System for LLM Agent Applications
by: Zhang, Haiyue, et al.
Published: (2026)
by: Zhang, Haiyue, et al.
Published: (2026)
Semia: Auditing Agent Skills via Constraint-Guided Representation Synthesis
by: Wen, Hongbo, et al.
Published: (2026)
by: Wen, Hongbo, et al.
Published: (2026)
Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents
by: Zhang, Xi, et al.
Published: (2026)
by: Zhang, Xi, et al.
Published: (2026)
SkillOpt: Executive Strategy for Self-Evolving Agent Skills
by: Yang, Yifan, et al.
Published: (2026)
by: Yang, Yifan, et al.
Published: (2026)
Requesting Expert Reasoning: Augmenting LLM Agents with Learned Collaborative Intervention
by: Wang, Zhiming, et al.
Published: (2026)
by: Wang, Zhiming, et al.
Published: (2026)
From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills
by: Huang, Zisu, et al.
Published: (2026)
by: Huang, Zisu, et al.
Published: (2026)
SkillGenBench: Benchmarking Skill Generation Pipelines for LLM Agents
by: Zhou, Yifan, et al.
Published: (2026)
by: Zhou, Yifan, et al.
Published: (2026)
ClawKeeper: Comprehensive Safety Protection for OpenClaw Agents Through Skills, Plugins, and Watchers
by: Liu, Songyang, et al.
Published: (2026)
by: Liu, Songyang, et al.
Published: (2026)
EmbodiSkill: Skill-Aware Reflection for Self-Evolving Embodied Agents
by: Ju, Ruofei, et al.
Published: (2026)
by: Ju, Ruofei, et al.
Published: (2026)
MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structural Anomaly Detection
by: Tan, Zhewen, et al.
Published: (2026)
by: Tan, Zhewen, et al.
Published: (2026)
Memento-Skills: Let Agents Design Agents
by: Zhou, Huichi, et al.
Published: (2026)
by: Zhou, Huichi, et al.
Published: (2026)
Towards Security-Auditable LLM Agents: A Unified Graph Representation
by: Li, Chaofan, et al.
Published: (2026)
by: Li, Chaofan, et al.
Published: (2026)
Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents
by: Mi, Qirui, et al.
Published: (2026)
by: Mi, Qirui, et al.
Published: (2026)
SkillMOO: Multi-Objective Optimization of Agent Skills for Software Engineering
by: Gong, Jingzhi, et al.
Published: (2026)
by: Gong, Jingzhi, et al.
Published: (2026)
Automated Creation and Enrichment Framework for Improved Invocation of Enterprise APIs as Tools
by: Agarwal, Prerna, et al.
Published: (2025)
by: Agarwal, Prerna, et al.
Published: (2025)
From Skill Text to Skill Structure: The Scheduling-Structural-Logical Representation for Agent Skills
by: Liang, Qiliang, et al.
Published: (2026)
by: Liang, Qiliang, et al.
Published: (2026)
Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations
by: Liu, Yilong, et al.
Published: (2026)
by: Liu, Yilong, et al.
Published: (2026)
Single-Node Trigger Backdoor Attacks in Graph-Based Recommendation Systems
by: Li, Runze, et al.
Published: (2025)
by: Li, Runze, et al.
Published: (2025)
Similar Items
-
Hallucination as Exploit: Evidence-Carrying Multimodal Agents
by: Zhang, Guijia, et al.
Published: (2026) -
Skill or Skip? Learning Selective Skill Invocation in Agentic Tasks via Dual-Granularity Preference Learning
by: Chen, Chishui, et al.
Published: (2026) -
Logic Agent: Enhancing Validity with Logic Rule Invocation
by: Liu, Hanmeng, et al.
Published: (2024) -
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills
by: Hou, Yingyong, et al.
Published: (2026) -
When Agents Fail to Act: A Diagnostic Framework for Tool Invocation Reliability in Multi-Agent LLM Systems
by: Huang, Donghao, et al.
Published: (2026)