A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Qianshan, Yang, Tengchao, Wang, Yaochen, Li, Xinfeng, Li, Lijun, Yin, Zhenfei, Zhan, Yi, Holz, Thorsten, Lin, Zhiqiang, Wang, XiaoFeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Vision for Access Control in LLM-based Agent Systems
by: Li, Xinfeng, et al.
Published: (2025)
by: Li, Xinfeng, et al.
Published: (2025)
Proactive Hardening of LLM Defenses with HASTE
by: Chen, Henry, et al.
Published: (2026)
by: Chen, Henry, et al.
Published: (2026)
MindGuard: Intrinsic Decision Inspection for Securing LLM Agents Against Metadata Poisoning
by: Wang, Zhiqiang, et al.
Published: (2025)
by: Wang, Zhiqiang, et al.
Published: (2025)
Rethinking Side-Channel Analysis: Automated Discovery and Analysis of Side-Channel Leakage with LLM-Assisted Agents
by: Xu, Zhen, et al.
Published: (2026)
by: Xu, Zhen, et al.
Published: (2026)
LocalAlign: Enabling Generalizable Prompt Injection Defense via Generation of Near-Target Adversarial Examples for Alignment Training
by: Gong, Yuyang, et al.
Published: (2026)
by: Gong, Yuyang, et al.
Published: (2026)
ProtoGuard-SL: Prototype Consistency Based Backdoor Defense for Vertical Split Learning
by: Shui, Yuhan, et al.
Published: (2026)
by: Shui, Yuhan, et al.
Published: (2026)
MemLineage: Lineage-Guided Enforcement for LLM Agent Memory
by: Ouyang, Ciyan, et al.
Published: (2026)
by: Ouyang, Ciyan, et al.
Published: (2026)
AttriGuard: Defeating Indirect Prompt Injection in LLM Agents via Causal Attribution of Tool Invocations
by: He, Yu, et al.
Published: (2026)
by: He, Yu, et al.
Published: (2026)
LeakGuard: Detecting Memory Leaks Accurately and Scalably
by: Liang, Hongliang, et al.
Published: (2025)
by: Liang, Hongliang, et al.
Published: (2025)
Memory Poisoning Attack and Defense on Memory Based LLM-Agents
by: Sunil, Balachandra Devarangadi, et al.
Published: (2026)
by: Sunil, Balachandra Devarangadi, et al.
Published: (2026)
TrapSuffix: Proactive Defense Against Adversarial Suffixes in Jailbreaking
by: Du, Mengyao, et al.
Published: (2026)
by: Du, Mengyao, et al.
Published: (2026)
MemPrivacy: Privacy-Preserving Personalized Memory Management for Edge-Cloud Agents
by: Chen, Yining, et al.
Published: (2026)
by: Chen, Yining, et al.
Published: (2026)
An Explorative Study of Pig Butchering Scams
by: Acharya, Bhupendra, et al.
Published: (2024)
by: Acharya, Bhupendra, et al.
Published: (2024)
WebAgentGuard: A Reasoning-Driven Guard Model for Detecting Prompt Injection Attacks in Web Agents
by: Chen, Yulin, et al.
Published: (2026)
by: Chen, Yulin, et al.
Published: (2026)
A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
by: Kong, Dezhang, et al.
Published: (2025)
by: Kong, Dezhang, et al.
Published: (2025)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
by: Zhang, Hanrong, et al.
Published: (2024)
by: Zhang, Hanrong, et al.
Published: (2024)
Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts
by: Hasan, Md. Mehedi, et al.
Published: (2025)
by: Hasan, Md. Mehedi, et al.
Published: (2025)
AegisAgent: An Autonomous Defense Agent Against Prompt Injection Attacks in LLM-HARs
by: Wang, Yihan, et al.
Published: (2025)
by: Wang, Yihan, et al.
Published: (2025)
RerouteGuard: Understanding and Mitigating Adversarial Risks for LLM Routing
by: Zhang, Wenhui, et al.
Published: (2026)
by: Zhang, Wenhui, et al.
Published: (2026)
Consiglieres in the Shadow: Understanding the Use of Uncensored Large Language Models in Cybercrimes
by: Lin, Zilong, et al.
Published: (2025)
by: Lin, Zilong, et al.
Published: (2025)
OpenClaw PRISM: A Zero-Fork, Defense-in-Depth Runtime Security Layer for Tool-Augmented LLM Agents
by: Li, Frank
Published: (2026)
by: Li, Frank
Published: (2026)
Disassembling Obfuscated Executables with LLM
by: Rong, Huanyao, et al.
Published: (2024)
by: Rong, Huanyao, et al.
Published: (2024)
CCX: Enabling Unmodified Intel SGX Applications on Arm CCA
by: Schulze, Matti, et al.
Published: (2026)
by: Schulze, Matti, et al.
Published: (2026)
DualGuard: Dual-stream Large Language Model Watermarking Defense against Paraphrase and Spoofing Attack
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
From Defender to Devil? Unintended Risk Interactions Induced by LLM Defenses
by: Meng, Xiangtao, et al.
Published: (2025)
by: Meng, Xiangtao, et al.
Published: (2025)
The Landscape of Prompt Injection Threats in LLM Agents: From Taxonomy to Analysis
by: Wang, Peiran, et al.
Published: (2026)
by: Wang, Peiran, et al.
Published: (2026)
AgentGuard: An Attribute-Based Access Control Framework for Tool-Use LLM-Based Agent
by: Luo, Jiaqi, et al.
Published: (2026)
by: Luo, Jiaqi, et al.
Published: (2026)
Taxonomy, Evaluation and Exploitation of IPI-Centric LLM Agent Defense Frameworks
by: Ji, Zimo, et al.
Published: (2025)
by: Ji, Zimo, et al.
Published: (2025)
ClawGuard: Out-of-Band Detection of LLM Agent Workflow Hijacking via EM Side Channel
by: Gan, Leo Linqian, et al.
Published: (2026)
by: Gan, Leo Linqian, et al.
Published: (2026)
DistillGuard: Evaluating Defenses Against LLM Knowledge Distillation
by: Jiang, Bo
Published: (2026)
by: Jiang, Bo
Published: (2026)
Autonomous LLM Agent Worms: Cross-Platform Propagation, Automated Discovery and Temporal Re-Entry Defense
by: Zha, Mingming, et al.
Published: (2026)
by: Zha, Mingming, et al.
Published: (2026)
LLM-Enhanced Software Patch Localization
by: Yu, Jinhong, et al.
Published: (2024)
by: Yu, Jinhong, et al.
Published: (2024)
Cybersecurity through Entropy Injection: A Paradigm Shift from Reactive Defense to Proactive Uncertainty
by: Janani, Kush
Published: (2025)
by: Janani, Kush
Published: (2025)
Collaborative Shadows: Distributed Backdoor Attacks in LLM-Based Multi-Agent Systems
by: Zhu, Pengyu, et al.
Published: (2025)
by: Zhu, Pengyu, et al.
Published: (2025)
Leaky Cauldron on the Dark Land: Understanding Memory Side-Channel Hazards in SGX
by: Wang, Wenhao, et al.
Published: (2017)
by: Wang, Wenhao, et al.
Published: (2017)
Clues in Tweets: Twitter-Guided Discovery and Analysis of SMS Spam
by: Tang, Siyuan, et al.
Published: (2022)
by: Tang, Siyuan, et al.
Published: (2022)
JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks
by: Zhang, Xiaoyu, et al.
Published: (2023)
by: Zhang, Xiaoyu, et al.
Published: (2023)
HarmRLVR: Weaponizing Verifiable Rewards for Harmful LLM Alignment
by: Liu, Yuexiao, et al.
Published: (2025)
by: Liu, Yuexiao, et al.
Published: (2025)
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
by: Zhan, Qiusi, et al.
Published: (2025)
by: Zhan, Qiusi, et al.
Published: (2025)
Blockchain Security Risk Assessment in Quantum Era, Migration Strategies and Proactive Defense
by: Baseri, Yaser, et al.
Published: (2025)
by: Baseri, Yaser, et al.
Published: (2025)
Similar Items
-
A Vision for Access Control in LLM-based Agent Systems
by: Li, Xinfeng, et al.
Published: (2025) -
Proactive Hardening of LLM Defenses with HASTE
by: Chen, Henry, et al.
Published: (2026) -
MindGuard: Intrinsic Decision Inspection for Securing LLM Agents Against Metadata Poisoning
by: Wang, Zhiqiang, et al.
Published: (2025) -
Rethinking Side-Channel Analysis: Automated Discovery and Analysis of Side-Channel Leakage with LLM-Assisted Agents
by: Xu, Zhen, et al.
Published: (2026) -
LocalAlign: Enabling Generalizable Prompt Injection Defense via Generation of Near-Target Adversarial Examples for Alignment Training
by: Gong, Yuyang, et al.
Published: (2026)