CommandSans: Securing AI Agents with Surgical Precision Prompt Sanitization
Fuente:
arXiv
Saved in:
| Main Authors: | Das, Debeshee, Beurer-Kellner, Luca, Fischer, Marc, Baader, Maximilian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Trojan Hippo: Weaponizing Agent Memory for Data Exfiltration
by: Das, Debeshee, et al.
Published: (2026)
by: Das, Debeshee, et al.
Published: (2026)
AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents
by: Debenedetti, Edoardo, et al.
Published: (2024)
by: Debenedetti, Edoardo, et al.
Published: (2024)
Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks
by: Schmotz, David, et al.
Published: (2026)
by: Schmotz, David, et al.
Published: (2026)
Design Patterns for Securing LLM Agents against Prompt Injections
by: Beurer-Kellner, Luca, et al.
Published: (2025)
by: Beurer-Kellner, Luca, et al.
Published: (2025)
Evading Data Contamination Detection for Language Models is (too) Easy
by: Dekoninck, Jasper, et al.
Published: (2024)
by: Dekoninck, Jasper, et al.
Published: (2024)
Ward: Provable RAG Dataset Inference via LLM Watermarks
by: Jovanović, Nikola, et al.
Published: (2024)
by: Jovanović, Nikola, et al.
Published: (2024)
Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem
by: Beurer-Kellner, Luca, et al.
Published: (2026)
by: Beurer-Kellner, Luca, et al.
Published: (2026)
CHAI: Command Hijacking against embodied AI
by: Burbano, Luis, et al.
Published: (2025)
by: Burbano, Luis, et al.
Published: (2025)
AutoBaxBuilder: Bootstrapping Code Security Benchmarking
by: von Arx, Tobias, et al.
Published: (2025)
by: von Arx, Tobias, et al.
Published: (2025)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
by: Geng, Runpeng, et al.
Published: (2025)
by: Geng, Runpeng, et al.
Published: (2025)
Trust No AI: Prompt Injection Along The CIA Security Triad
by: Rehberger, Johann
Published: (2024)
by: Rehberger, Johann
Published: (2024)
Breaking Agent Backbones: Evaluating the Security of Backbone LLMs in AI Agents
by: Bazinska, Julia, et al.
Published: (2025)
by: Bazinska, Julia, et al.
Published: (2025)
ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?
by: Wang, Zhun, et al.
Published: (2026)
by: Wang, Zhun, et al.
Published: (2026)
BaxBench: Can LLMs Generate Correct and Secure Backends?
by: Vero, Mark, et al.
Published: (2025)
by: Vero, Mark, et al.
Published: (2025)
EVMbench: Evaluating AI Agents on Smart Contract Security
by: Wang, Justin, et al.
Published: (2026)
by: Wang, Justin, et al.
Published: (2026)
Reconstruction of Differentially Private Text Sanitization via Large Language Models
by: Pang, Shuchao, et al.
Published: (2024)
by: Pang, Shuchao, et al.
Published: (2024)
BrowseSafe: Understanding and Preventing Prompt Injection Within AI Browser Agents
by: Zhang, Kaiyuan, et al.
Published: (2025)
by: Zhang, Kaiyuan, et al.
Published: (2025)
CySecBench: Generative AI-based CyberSecurity-focused Prompt Dataset for Benchmarking Large Language Models
by: Wahréus, Johan, et al.
Published: (2025)
by: Wahréus, Johan, et al.
Published: (2025)
Security Considerations for Artificial Intelligence Agents
by: Li, Ninghui, et al.
Published: (2026)
by: Li, Ninghui, et al.
Published: (2026)
An Empirical Evaluation of LLM-Generated Code Security Across Prompting Methods
by: Kharma, Mohammed, et al.
Published: (2026)
by: Kharma, Mohammed, et al.
Published: (2026)
SAMEP: A Secure Protocol for Persistent Context Sharing Across AI Agents
by: Masoor, Hari
Published: (2025)
by: Masoor, Hari
Published: (2025)
Membership Inference Attacks Cannot Prove that a Model Was Trained On Your Data
by: Zhang, Jie, et al.
Published: (2024)
by: Zhang, Jie, et al.
Published: (2024)
SVAgent: AI Agent for Hardware Security Verification Assertion
by: Guo, Rui, et al.
Published: (2025)
by: Guo, Rui, et al.
Published: (2025)
Evaluating Prompt Injection Defenses for Educational LLM Tutors: Security-Usability-Latency Trade-offs
by: Maiorano, Alexandre Cristovão
Published: (2026)
by: Maiorano, Alexandre Cristovão
Published: (2026)
SoK: Security and Privacy Risks of Healthcare AI
by: Chang, Yuanhaur, et al.
Published: (2024)
by: Chang, Yuanhaur, et al.
Published: (2024)
Logic layer Prompt Control Injection (LPCI): A Novel Security Vulnerability Class in Agentic Systems
by: Atta, Hammad, et al.
Published: (2025)
by: Atta, Hammad, et al.
Published: (2025)
Blind Baselines Beat Membership Inference Attacks for Foundation Models
by: Das, Debeshee, et al.
Published: (2024)
by: Das, Debeshee, et al.
Published: (2024)
Smart Water Security with AI and Blockchain-Enhanced Digital Twins
by: Homaei, Mohammadhossein, et al.
Published: (2025)
by: Homaei, Mohammadhossein, et al.
Published: (2025)
Position: AI Security Policy Should Target Systems, Not Models
by: Riegler, Michael A., et al.
Published: (2026)
by: Riegler, Michael A., et al.
Published: (2026)
SAGA: A Security Architecture for Governing AI Agentic Systems
by: Syros, Georgios, et al.
Published: (2025)
by: Syros, Georgios, et al.
Published: (2025)
ADR: An Agentic Detection System for Enterprise Agentic AI Security
by: Li, Chenning, et al.
Published: (2026)
by: Li, Chenning, et al.
Published: (2026)
Privacy and Security Implications of Cloud-Based AI Services : A Survey
by: Luqman, Alka, et al.
Published: (2024)
by: Luqman, Alka, et al.
Published: (2024)
WebSP-Eval: Evaluating Web Agents on Website Security and Privacy Tasks
by: Ramesh, Guruprasad Viswanathan, et al.
Published: (2026)
by: Ramesh, Guruprasad Viswanathan, et al.
Published: (2026)
AgentGuardian: Learning Access Control Policies to Govern AI Agent Behavior
by: Abaev, Nadya, et al.
Published: (2026)
by: Abaev, Nadya, et al.
Published: (2026)
PoolFlip: A Multi-Agent Reinforcement Learning Security Environment for Cyber Defense
by: Cadet, Xavier, et al.
Published: (2025)
by: Cadet, Xavier, et al.
Published: (2025)
Command & Control (C2) Traffic Detection Via Algorithm Generated Domain (Dga) Classification Using Deep Learning And Natural Language Processing
by: Felix, Maria Milena Araujo
Published: (2025)
by: Felix, Maria Milena Araujo
Published: (2025)
Probing Latent Subspaces in LLM for AI Security: Identifying and Manipulating Adversarial States
by: Chia, Xin Wei, et al.
Published: (2025)
by: Chia, Xin Wei, et al.
Published: (2025)
Taming Data Challenges in ML-based Security Tasks Using Generative AI
by: Kanchi, Shravya, et al.
Published: (2025)
by: Kanchi, Shravya, et al.
Published: (2025)
SME-TEAM: Leveraging Trust and Ethics for Secure and Responsible Use of AI and LLMs in SMEs
by: Sarker, Iqbal H., et al.
Published: (2025)
by: Sarker, Iqbal H., et al.
Published: (2025)
Towards Trustworthy AI: Secure Deepfake Detection using CNNs and Zero-Knowledge Proofs
by: Islam, H M Mohaimanul, et al.
Published: (2025)
by: Islam, H M Mohaimanul, et al.
Published: (2025)
Similar Items
-
Trojan Hippo: Weaponizing Agent Memory for Data Exfiltration
by: Das, Debeshee, et al.
Published: (2026) -
AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents
by: Debenedetti, Edoardo, et al.
Published: (2024) -
Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks
by: Schmotz, David, et al.
Published: (2026) -
Design Patterns for Securing LLM Agents against Prompt Injections
by: Beurer-Kellner, Luca, et al.
Published: (2025) -
Evading Data Contamination Detection for Language Models is (too) Easy
by: Dekoninck, Jasper, et al.
Published: (2024)