Mind Your HEARTBEAT! Claw Background Execution Inherently Enables Silent Memory Pollution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yechao, Zhao, Shiqian, Zhang, Jie, Deng, Gelei, Zhang, Jiawen, Liu, Xiaogeng, Xiao, Chaowei, Zhang, Tianwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AutoDAN-Reasoning: Enhancing Strategies Exploration based Jailbreak Attacks with Test-Time Scaling
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2025)
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2025)
Oedipus: LLM-enchanced Reasoning CAPTCHA Solver
von: Deng, Gelei, et al.
Veröffentlicht: (2024)
von: Deng, Gelei, et al.
Veröffentlicht: (2024)
ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathologically Long Reasoning in Large Reasoning Models
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2026)
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2026)
DECEIVE-AFC: Adversarial Claim Attacks against Search-Enabled LLM-based Fact-Checking Systems
von: Ou, Haoran, et al.
Veröffentlicht: (2026)
von: Ou, Haoran, et al.
Veröffentlicht: (2026)
IRCopilot: Automated Incident Response with Large Language Models
von: Lin, Xihuan, et al.
Veröffentlicht: (2025)
von: Lin, Xihuan, et al.
Veröffentlicht: (2025)
OET: Optimization-based prompt injection Evaluation Toolkit
von: Pan, Jinsheng, et al.
Veröffentlicht: (2025)
von: Pan, Jinsheng, et al.
Veröffentlicht: (2025)
RePD: Defending Jailbreak Attack through a Retrieval-based Prompt Decomposition Process
von: Wang, Peiran, et al.
Veröffentlicht: (2024)
von: Wang, Peiran, et al.
Veröffentlicht: (2024)
Don't Let the Claw Grip Your Hand: A Security Analysis and Defense Framework for OpenClaw
von: Shan, Zhengyang, et al.
Veröffentlicht: (2026)
von: Shan, Zhengyang, et al.
Veröffentlicht: (2026)
Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning
von: Deng, Gelei, et al.
Veröffentlicht: (2024)
von: Deng, Gelei, et al.
Veröffentlicht: (2024)
When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models
von: Ou, Haoran, et al.
Veröffentlicht: (2025)
von: Ou, Haoran, et al.
Veröffentlicht: (2025)
Don't Listen To Me: Understanding and Exploring Jailbreak Prompts of Large Language Models
von: Yu, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Yu, Zhiyuan, et al.
Veröffentlicht: (2024)
Enhancing Model Defense Against Jailbreaks with Proactive Safety Reasoning
von: Yang, Xianglin, et al.
Veröffentlicht: (2025)
von: Yang, Xianglin, et al.
Veröffentlicht: (2025)
What Makes a Good LLM Agent for Real-world Penetration Testing?
von: Deng, Gelei, et al.
Veröffentlicht: (2026)
von: Deng, Gelei, et al.
Veröffentlicht: (2026)
DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
AdaSteer: Your Aligned LLM is Inherently an Adaptive Jailbreak Defender
von: Zhao, Weixiang, et al.
Veröffentlicht: (2025)
von: Zhao, Weixiang, et al.
Veröffentlicht: (2025)
MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots
von: Deng, Gelei, et al.
Veröffentlicht: (2023)
von: Deng, Gelei, et al.
Veröffentlicht: (2023)
AdaShield: Safeguarding Multimodal Large Language Models from Structure-based Attack via Adaptive Shield Prompting
von: Wang, Yu, et al.
Veröffentlicht: (2024)
von: Wang, Yu, et al.
Veröffentlicht: (2024)
AgentDyn: Are Your Agent Security Defenses Deployable in Real-World Dynamic Environments?
von: Li, Hao, et al.
Veröffentlicht: (2026)
von: Li, Hao, et al.
Veröffentlicht: (2026)
SAME: Sample Reconstruction against Model Extraction Attacks
von: Xie, Yi, et al.
Veröffentlicht: (2023)
von: Xie, Yi, et al.
Veröffentlicht: (2023)
Turning Bias into Bugs: Bandit-Guided Style Manipulation Attacks on LLM Judges
von: Yang, Xianglin, et al.
Veröffentlicht: (2026)
von: Yang, Xianglin, et al.
Veröffentlicht: (2026)
Towards Effective Prompt Stealing Attack against Text-to-Image Diffusion Models
von: Zhao, Shiqian, et al.
Veröffentlicht: (2025)
von: Zhao, Shiqian, et al.
Veröffentlicht: (2025)
AgentSys: Secure and Dynamic LLM Agents Through Explicit Hierarchical Memory Management
von: Wen, Ruoyao, et al.
Veröffentlicht: (2026)
von: Wen, Ruoyao, et al.
Veröffentlicht: (2026)
PentestEval: Benchmarking LLM-based Penetration Testing with Modular and Stage-Level Design
von: Yang, Ruozhao, et al.
Veröffentlicht: (2025)
von: Yang, Ruozhao, et al.
Veröffentlicht: (2025)
AutoEG: Exploiting Known Third-Party Vulnerabilities in Black-Box Web Applications
von: Yang, Ruozhao, et al.
Veröffentlicht: (2026)
von: Yang, Ruozhao, et al.
Veröffentlicht: (2026)
Code Agent can be an End-to-end System Hacker: Benchmarking Real-world Threats of Computer-use Agent
von: Luo, Weidi, et al.
Veröffentlicht: (2025)
von: Luo, Weidi, et al.
Veröffentlicht: (2025)
Security, Privacy, and Ethical Risks in OpenClaw
von: Jin, Yutong, et al.
Veröffentlicht: (2026)
von: Jin, Yutong, et al.
Veröffentlicht: (2026)
JailBreakV: A Benchmark for Assessing the Robustness of MultiModal Large Language Models against Jailbreak Attacks
von: Luo, Weidi, et al.
Veröffentlicht: (2024)
von: Luo, Weidi, et al.
Veröffentlicht: (2024)
SafeClaw-R: Towards Safe and Secure Multi-Agent Personal Assistants
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
When Memory Becomes a Vulnerability: Towards Multi-turn Jailbreak Attacks against Text-to-Image Generation Systems
von: Zhao, Shiqian, et al.
Veröffentlicht: (2025)
von: Zhao, Shiqian, et al.
Veröffentlicht: (2025)
Robust-Wide: Robust Watermarking against Instruction-driven Image Editing
von: Hu, Runyi, et al.
Veröffentlicht: (2024)
von: Hu, Runyi, et al.
Veröffentlicht: (2024)
ClawKeeper: Comprehensive Safety Protection for OpenClaw Agents Through Skills, Plugins, and Watchers
von: Liu, Songyang, et al.
Veröffentlicht: (2026)
von: Liu, Songyang, et al.
Veröffentlicht: (2026)
Silent Guardians: Independent and Secure Decision Tree Evaluation Without Chatter
von: Li, Jinyuan, et al.
Veröffentlicht: (2026)
von: Li, Jinyuan, et al.
Veröffentlicht: (2026)
Low Rank Comes with Low Security: Gradient Assembly Poisoning Attacks against Distributed LoRA-based LLM Systems
von: Dong, Yueyan, et al.
Veröffentlicht: (2026)
von: Dong, Yueyan, et al.
Veröffentlicht: (2026)
Doxing via the Lens: Revealing Location-related Privacy Leakage on Multi-modal Large Reasoning Models
von: Luo, Weidi, et al.
Veröffentlicht: (2025)
von: Luo, Weidi, et al.
Veröffentlicht: (2025)
Why Does Little Robustness Help? A Further Step Towards Understanding Adversarial Transferability
von: Zhang, Yechao, et al.
Veröffentlicht: (2023)
von: Zhang, Yechao, et al.
Veröffentlicht: (2023)
PoseGuard: Pose-Guided Generation with Safety Guardrails
von: Wang, Kongxin, et al.
Veröffentlicht: (2025)
von: Wang, Kongxin, et al.
Veröffentlicht: (2025)
Silent Guardian: Protecting Text from Malicious Exploitation by Large Language Models
von: Zhao, Jiawei, et al.
Veröffentlicht: (2023)
von: Zhao, Jiawei, et al.
Veröffentlicht: (2023)
InferDPT: Privacy-Preserving Inference for Closed-box Large Language Model
von: Tong, Meng, et al.
Veröffentlicht: (2023)
von: Tong, Meng, et al.
Veröffentlicht: (2023)
"MCP Does Not Stand for Misuse Cryptography Protocol": Uncovering Cryptographic Misuse in Model Context Protocol at Scale
von: Yan, Biwei, et al.
Veröffentlicht: (2025)
von: Yan, Biwei, et al.
Veröffentlicht: (2025)
ART: Automatic Red-teaming for Text-to-Image Models to Protect Benign Users
von: Li, Guanlin, et al.
Veröffentlicht: (2024)
von: Li, Guanlin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AutoDAN-Reasoning: Enhancing Strategies Exploration based Jailbreak Attacks with Test-Time Scaling
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2025) -
Oedipus: LLM-enchanced Reasoning CAPTCHA Solver
von: Deng, Gelei, et al.
Veröffentlicht: (2024) -
ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathologically Long Reasoning in Large Reasoning Models
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2026) -
DECEIVE-AFC: Adversarial Claim Attacks against Search-Enabled LLM-based Fact-Checking Systems
von: Ou, Haoran, et al.
Veröffentlicht: (2026) -
IRCopilot: Automated Incident Response with Large Language Models
von: Lin, Xihuan, et al.
Veröffentlicht: (2025)