SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jia, Xiaojun, Liao, Jie, Qin, Simeng, Gu, Jindong, Ren, Wenqi, Cao, Xiaochun, Liu, Yang, Torr, Philip |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SkillAttack: Automated Red Teaming of Agent Skills through Attack Path Refinement
von: Duan, Zenghao, et al.
Veröffentlicht: (2026)
von: Duan, Zenghao, et al.
Veröffentlicht: (2026)
When Skills Lie: Hidden-Comment Injection in LLM Agents
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
Sealing the Audit-Runtime Gap for LLM Skills
von: Shen, Tingda, et al.
Veröffentlicht: (2026)
von: Shen, Tingda, et al.
Veröffentlicht: (2026)
BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
SkillScope: Toward Fine-Grained Least-Privilege Enforcement for Agent Skills
von: Wu, Jiangrong, et al.
Veröffentlicht: (2026)
von: Wu, Jiangrong, et al.
Veröffentlicht: (2026)
SkillTester: Benchmarking Utility and Security of Agent Skills
von: Wang, Leye, et al.
Veröffentlicht: (2026)
von: Wang, Leye, et al.
Veröffentlicht: (2026)
Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks
von: Schmotz, David, et al.
Veröffentlicht: (2026)
von: Schmotz, David, et al.
Veröffentlicht: (2026)
SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems
von: Feng, Yunhao, et al.
Veröffentlicht: (2026)
von: Feng, Yunhao, et al.
Veröffentlicht: (2026)
Prompt Injection Attacks on Agentic Coding Assistants: A Systematic Analysis of Vulnerabilities in Skills, Tools, and Protocol Ecosystems
von: Maloyan, Narek, et al.
Veröffentlicht: (2026)
von: Maloyan, Narek, et al.
Veröffentlicht: (2026)
Defenses & Enablers For Skill Injection Attacks on Terminal Based Agents
von: Fujinuma, Yoshinari, et al.
Veröffentlicht: (2026)
von: Fujinuma, Yoshinari, et al.
Veröffentlicht: (2026)
Does Few-shot Learning Suffer from Backdoor Attacks?
von: Liu, Xinwei, et al.
Veröffentlicht: (2023)
von: Liu, Xinwei, et al.
Veröffentlicht: (2023)
HarmfulSkillBench: How Do Harmful Skills Weaponize Your Agents?
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
SkillProbe: Security Auditing for Emerging Agent Skill Marketplaces via Multi-Agent Collaboration
von: Guo, Zihan, et al.
Veröffentlicht: (2026)
von: Guo, Zihan, et al.
Veröffentlicht: (2026)
SkillSieve: A Hierarchical Triage Framework for Detecting Malicious AI Agent Skills
von: Hou, Yinghan, et al.
Veröffentlicht: (2026)
von: Hou, Yinghan, et al.
Veröffentlicht: (2026)
LLM Jailbreak Detection for (Almost) Free!
von: Chen, Guorui, et al.
Veröffentlicht: (2025)
von: Chen, Guorui, et al.
Veröffentlicht: (2025)
Obscure but Effective: Classical Chinese Jailbreak Prompt Optimization via Bio-Inspired Search
von: Huang, Xun, et al.
Veröffentlicht: (2026)
von: Huang, Xun, et al.
Veröffentlicht: (2026)
Structured Security Auditing and Robustness Enhancement for Untrusted Agent Skills
von: Lv, Lijia, et al.
Veröffentlicht: (2026)
von: Lv, Lijia, et al.
Veröffentlicht: (2026)
Harmless Yet Harmful: Neutral Prompting Attacks for Stealthy Hallucination Steering in Agent Skills
von: Hsu, Chia-Yi, et al.
Veröffentlicht: (2026)
von: Hsu, Chia-Yi, et al.
Veröffentlicht: (2026)
When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems
von: Wang, Su, et al.
Veröffentlicht: (2026)
von: Wang, Su, et al.
Veröffentlicht: (2026)
SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents
von: Ouyang, Yipeng, et al.
Veröffentlicht: (2026)
von: Ouyang, Yipeng, et al.
Veröffentlicht: (2026)
Do Skill Descriptions Tell the Truth? Detecting Undisclosed Security Behaviors in Code-Backed LLM Skills
von: He, Wenhui, et al.
Veröffentlicht: (2026)
von: He, Wenhui, et al.
Veröffentlicht: (2026)
Trust Me, Import This: Dependency Steering Attacks via Malicious Agent Skills
von: Liu, Yiyong, et al.
Veröffentlicht: (2026)
von: Liu, Yiyong, et al.
Veröffentlicht: (2026)
OmniSafeBench-MM: A Unified Benchmark and Toolbox for Multimodal Jailbreak Attack-Defense Evaluation
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
von: Zhu, Kaijie, et al.
Veröffentlicht: (2025)
von: Zhu, Kaijie, et al.
Veröffentlicht: (2025)
ObliInjection: Order-Oblivious Prompt Injection Attack to LLM Agents with Multi-source Data
von: Wang, Reachal, et al.
Veröffentlicht: (2025)
von: Wang, Reachal, et al.
Veröffentlicht: (2025)
Black-Box Skill Stealing Attack from Proprietary LLM Agents: An Empirical Study
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem
von: Beurer-Kellner, Luca, et al.
Veröffentlicht: (2026)
von: Beurer-Kellner, Luca, et al.
Veröffentlicht: (2026)
No Attack Required: Semantic Fuzzing for Specification Violations in Agent Skills
von: Li, Ying, et al.
Veröffentlicht: (2026)
von: Li, Ying, et al.
Veröffentlicht: (2026)
AgentWatcher: A Rule-based Prompt Injection Monitor
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
Behavioral Integrity Verification for AI Agent Skills
von: Wu, Yuhao, et al.
Veröffentlicht: (2026)
von: Wu, Yuhao, et al.
Veröffentlicht: (2026)
What Skills Do Cyber Security Professionals Need?
von: Ullah, Faheem, et al.
Veröffentlicht: (2025)
von: Ullah, Faheem, et al.
Veröffentlicht: (2025)
AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills
von: Zhuang, Haomin, et al.
Veröffentlicht: (2026)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2026)
Exploiting LLM Agent Supply Chains via Payload-less Skills
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
Towards Secure Agent Skills: Architecture, Threat Taxonomy, and Security Analysis
von: Li, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Li, Zhiyuan, et al.
Veröffentlicht: (2026)
Proteus: A Self-Evolving Red Team for Agent Skill Ecosystems
von: Zhou, Zhaojiacheng
Veröffentlicht: (2026)
von: Zhou, Zhaojiacheng
Veröffentlicht: (2026)
Context Matters: Repository-Aware Security Analysis of the Agent Skill Ecosystem
von: Holzbauer, Florian, et al.
Veröffentlicht: (2026)
von: Holzbauer, Florian, et al.
Veröffentlicht: (2026)
RouteGuard: Internal-Signal Detection of Skill Poisoning in LLM Agents
von: Xiao, Wenjie, et al.
Veröffentlicht: (2026)
von: Xiao, Wenjie, et al.
Veröffentlicht: (2026)
Evolving Skill-Structured Attack Memory Enhances LLM Jailbreaking
von: Zhang, Junke, et al.
Veröffentlicht: (2026)
von: Zhang, Junke, et al.
Veröffentlicht: (2026)
Improved Techniques for Optimization-Based Jailbreaking on Large Language Models
von: Jia, Xiaojun, et al.
Veröffentlicht: (2024)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2024)
DeepSeek Robustness Against Semantic-Character Dual-Space Mutated Prompt Injection
von: Ren, Junyu, et al.
Veröffentlicht: (2026)
von: Ren, Junyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SkillAttack: Automated Red Teaming of Agent Skills through Attack Path Refinement
von: Duan, Zenghao, et al.
Veröffentlicht: (2026) -
When Skills Lie: Hidden-Comment Injection in LLM Agents
von: Wang, Qianli, et al.
Veröffentlicht: (2026) -
Sealing the Audit-Runtime Gap for LLM Skills
von: Shen, Tingda, et al.
Veröffentlicht: (2026) -
BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning
von: Tie, Guiyao, et al.
Veröffentlicht: (2026) -
SkillScope: Toward Fine-Grained Least-Privilege Enforcement for Agent Skills
von: Wu, Jiangrong, et al.
Veröffentlicht: (2026)