Imprompter: Tricking LLM Agents into Improper Tool Use
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fu, Xiaohan, Li, Shuheng, Wang, Zihan, Liu, Yihao, Gupta, Rajesh K., Berg-Kirkpatrick, Taylor, Fernandes, Earlence |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fun-tuning: Characterizing the Vulnerability of Proprietary LLMs to Optimization-based Prompt Injection Attacks via the Fine-Tuning Interface
von: Labunets, Andrey, et al.
Veröffentlicht: (2025)
von: Labunets, Andrey, et al.
Veröffentlicht: (2025)
ceLLMate: Sandboxing Browser AI Agents
von: Meng, Luoxi, et al.
Veröffentlicht: (2025)
von: Meng, Luoxi, et al.
Veröffentlicht: (2025)
RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments
von: Fu, Yuchuan, et al.
Veröffentlicht: (2025)
von: Fu, Yuchuan, et al.
Veröffentlicht: (2025)
AgentGuard: An Attribute-Based Access Control Framework for Tool-Use LLM-Based Agent
von: Luo, Jiaqi, et al.
Veröffentlicht: (2026)
von: Luo, Jiaqi, et al.
Veröffentlicht: (2026)
MalTool: Malicious Tool Attacks on LLM Agents
von: Hu, Yuepeng, et al.
Veröffentlicht: (2026)
von: Hu, Yuepeng, et al.
Veröffentlicht: (2026)
Agent Security is a Systems Problem
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2026)
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2026)
Tricking LLM-Based NPCs into Spilling Secrets
von: Shiomi, Kyohei, et al.
Veröffentlicht: (2025)
von: Shiomi, Kyohei, et al.
Veröffentlicht: (2025)
May I have your Attention? Breaking Fine-Tuning based Prompt Injection Defenses using Architecture-Aware Attacks
von: Pandya, Nishit V., et al.
Veröffentlicht: (2025)
von: Pandya, Nishit V., et al.
Veröffentlicht: (2025)
Black-Box Skill Stealing Attack from Proprietary LLM Agents: An Empirical Study
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
Systems Security Foundations for Agentic Computing
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2025)
von: Christodorescu, Mihai, et al.
Veröffentlicht: (2025)
Les Dissonances: Cross-Tool Harvesting and Polluting in Pool-of-Tools Empowered LLM Agents
von: Li, Zichuan, et al.
Veröffentlicht: (2025)
von: Li, Zichuan, et al.
Veröffentlicht: (2025)
Trust No Tool: Evaluating and Defending LLM Agents under Untrusted Tool Feedback
von: Yan, Lecheng, et al.
Veröffentlicht: (2026)
von: Yan, Lecheng, et al.
Veröffentlicht: (2026)
Exposing LLM User Privacy via Traffic Fingerprint Analysis: A Study of Privacy Risks in LLM Agent Interactions
von: Zhang, Yixiang, et al.
Veröffentlicht: (2025)
von: Zhang, Yixiang, et al.
Veröffentlicht: (2025)
Prompt Injection Attack to Tool Selection in LLM Agents
von: Shi, Jiawen, et al.
Veröffentlicht: (2025)
von: Shi, Jiawen, et al.
Veröffentlicht: (2025)
AttriGuard: Defeating Indirect Prompt Injection in LLM Agents via Causal Attribution of Tool Invocations
von: He, Yu, et al.
Veröffentlicht: (2026)
von: He, Yu, et al.
Veröffentlicht: (2026)
AgentDID: Trustless Identity Authentication for AI Agents
von: Xu, Minghui, et al.
Veröffentlicht: (2026)
von: Xu, Minghui, et al.
Veröffentlicht: (2026)
Governing Dynamic Capabilities: Cryptographic Binding and Reproducibility Verification for AI Agent Tool Use
von: Zhou, Ziling
Veröffentlicht: (2026)
von: Zhou, Ziling
Veröffentlicht: (2026)
ToolTweak: An Attack on Tool Selection in LLM-based Agents
von: Sneh, Jonathan, et al.
Veröffentlicht: (2025)
von: Sneh, Jonathan, et al.
Veröffentlicht: (2025)
Your LLM Agent Can Leak Your Data: Data Exfiltration via Backdoored Tool Use
von: Zhang, Wuyang, et al.
Veröffentlicht: (2026)
von: Zhang, Wuyang, et al.
Veröffentlicht: (2026)
OpenClaw PRISM: A Zero-Fork, Defense-in-Depth Runtime Security Layer for Tool-Augmented LLM Agents
von: Li, Frank
Veröffentlicht: (2026)
von: Li, Frank
Veröffentlicht: (2026)
Teaching an Old Dog New Tricks: Verifiable FHE Using Commodity Hardware
von: Drean, Jules, et al.
Veröffentlicht: (2024)
von: Drean, Jules, et al.
Veröffentlicht: (2024)
Reframing LLM Agent Security as an Agent-Human Interaction Problem
von: Wang, Peiran, et al.
Veröffentlicht: (2026)
von: Wang, Peiran, et al.
Veröffentlicht: (2026)
The Verifier Tax: Horizon Dependent Safety Success Tradeoffs in Tool Using LLM Agents
von: Sah, Tanmay, et al.
Veröffentlicht: (2026)
von: Sah, Tanmay, et al.
Veröffentlicht: (2026)
Memory-Induced Tool-Drift in LLM Agents
von: Dabas, Mahavir, et al.
Veröffentlicht: (2026)
von: Dabas, Mahavir, et al.
Veröffentlicht: (2026)
AgentTrust: Runtime Safety Evaluation and Interception for AI Agent Tool Use
von: Yang, Chenglin
Veröffentlicht: (2026)
von: Yang, Chenglin
Veröffentlicht: (2026)
PentestAgent: Incorporating LLM Agents to Automated Penetration Testing
von: Shen, Xiangmin, et al.
Veröffentlicht: (2024)
von: Shen, Xiangmin, et al.
Veröffentlicht: (2024)
Victim as a Service: Designing a System for Engaging with Interactive Scammers
von: Spokoyny, Daniel, et al.
Veröffentlicht: (2025)
von: Spokoyny, Daniel, et al.
Veröffentlicht: (2025)
Defensible Design for OpenClaw: Securing Autonomous Tool-Invoking Agents
von: Li, Zongwei, et al.
Veröffentlicht: (2026)
von: Li, Zongwei, et al.
Veröffentlicht: (2026)
Bag of Tricks: Benchmarking of Jailbreak Attacks on LLMs
von: Xu, Zhao, et al.
Veröffentlicht: (2024)
von: Xu, Zhao, et al.
Veröffentlicht: (2024)
AgentVisor: Defending LLM Agents Against Prompt Injection via Semantic Virtualization
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
Tricking Retrievers with Influential Tokens: An Efficient Black-Box Corpus Poisoning Attack
von: Wang, Cheng, et al.
Veröffentlicht: (2025)
von: Wang, Cheng, et al.
Veröffentlicht: (2025)
AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents
von: Rassul, Yassin H., et al.
Veröffentlicht: (2026)
von: Rassul, Yassin H., et al.
Veröffentlicht: (2026)
SoK: Agentic Skills -- Beyond Tool Use in LLM Agents
von: Jiang, Yanna, et al.
Veröffentlicht: (2026)
von: Jiang, Yanna, et al.
Veröffentlicht: (2026)
Bag of Tricks for Subverting Reasoning-based Safety Guardrails
von: Chen, Shuo, et al.
Veröffentlicht: (2025)
von: Chen, Shuo, et al.
Veröffentlicht: (2025)
KryptoPilot: An Open-World Knowledge-Augmented LLM Agent for Automated Cryptographic Exploitation
von: Liu, Xiaonan, et al.
Veröffentlicht: (2026)
von: Liu, Xiaonan, et al.
Veröffentlicht: (2026)
Evaluating Privilege Usage of Agents with Real-World Tools
von: Zhang, Quan, et al.
Veröffentlicht: (2026)
von: Zhang, Quan, et al.
Veröffentlicht: (2026)
HarnessAgent: Scaling Automatic Fuzzing Harness Construction with Tool-Augmented LLM Pipelines
von: Yang, Kang, et al.
Veröffentlicht: (2025)
von: Yang, Kang, et al.
Veröffentlicht: (2025)
Causality Laundering: Denial-Feedback Leakage in Tool-Calling LLM Agents
von: Chinaei, Mohammad Hossein
Veröffentlicht: (2026)
von: Chinaei, Mohammad Hossein
Veröffentlicht: (2026)
FlowSteer: Prompt-Only Workflow Steering Exposes Planning-Time Vulnerabilities in Multi-Agent LLM Systems
von: Li, Fanxiao, et al.
Veröffentlicht: (2026)
von: Li, Fanxiao, et al.
Veröffentlicht: (2026)
VIGIL: Defending LLM Agents Against Tool Stream Injection via Verify-Before-Commit
von: Lin, Junda, et al.
Veröffentlicht: (2026)
von: Lin, Junda, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Fun-tuning: Characterizing the Vulnerability of Proprietary LLMs to Optimization-based Prompt Injection Attacks via the Fine-Tuning Interface
von: Labunets, Andrey, et al.
Veröffentlicht: (2025) -
ceLLMate: Sandboxing Browser AI Agents
von: Meng, Luoxi, et al.
Veröffentlicht: (2025) -
RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments
von: Fu, Yuchuan, et al.
Veröffentlicht: (2025) -
AgentGuard: An Attribute-Based Access Control Framework for Tool-Use LLM-Based Agent
von: Luo, Jiaqi, et al.
Veröffentlicht: (2026) -
MalTool: Malicious Tool Attacks on LLM Agents
von: Hu, Yuepeng, et al.
Veröffentlicht: (2026)