From Assistant to Double Agent: Formalizing and Benchmarking Attacks on OpenClaw for Personalized Local AI Agent
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yuhang, Xu, Feiming, Lin, Zheng, He, Guangyu, Huang, Yuzhe, Gao, Haichang, Niu, Zhenxing, Lian, Shiguo, Liu, Zhaoxiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Systematic Security Evaluation of OpenClaw and Its Variants
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
A Security Analysis of the OpenClaw AI Agent Framework
von: Suwansathit, Surada, et al.
Veröffentlicht: (2026)
von: Suwansathit, Surada, et al.
Veröffentlicht: (2026)
OpenClaw-RL: Train Any Agent Simply by Talking
von: Wang, Yinjie, et al.
Veröffentlicht: (2026)
von: Wang, Yinjie, et al.
Veröffentlicht: (2026)
ClawKeeper: Comprehensive Safety Protection for OpenClaw Agents Through Skills, Plugins, and Watchers
von: Liu, Songyang, et al.
Veröffentlicht: (2026)
von: Liu, Songyang, et al.
Veröffentlicht: (2026)
ICU-Bench:Benchmarking Continual Unlearning in Multimodal Large Language Models
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
Defensible Design for OpenClaw: Securing Autonomous Tool-Invoking Agents
von: Li, Zongwei, et al.
Veröffentlicht: (2026)
von: Li, Zongwei, et al.
Veröffentlicht: (2026)
Your Agent, Their Asset: A Real-World Safety Analysis of OpenClaw
von: Wang, Zijun, et al.
Veröffentlicht: (2026)
von: Wang, Zijun, et al.
Veröffentlicht: (2026)
QuantClaw: Precision Where It Matters for OpenClaw
von: Zhang, Manyi, et al.
Veröffentlicht: (2026)
von: Zhang, Manyi, et al.
Veröffentlicht: (2026)
OpenClaw AI Agents as Informal Learners at Moltbook: Characterizing an Emergent Learning Community at Scale
von: Chen, Eason, et al.
Veröffentlicht: (2026)
von: Chen, Eason, et al.
Veröffentlicht: (2026)
Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw
von: Yao, Hongwei, et al.
Veröffentlicht: (2026)
von: Yao, Hongwei, et al.
Veröffentlicht: (2026)
Taming OpenClaw: Security Analysis and Mitigation of Autonomous LLM Agent Threats
von: Deng, Xinhao, et al.
Veröffentlicht: (2026)
von: Deng, Xinhao, et al.
Veröffentlicht: (2026)
MATRA: Modeling the Attack Surface of Agentic AI Systems -- OpenClaw Case Study
von: Van hamme, Tim, et al.
Veröffentlicht: (2026)
von: Van hamme, Tim, et al.
Veröffentlicht: (2026)
Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
Re-Triggering Safeguards within LLMs for Jailbreak Detection
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
OpenClaw Agents on Moltbook: Risky Instruction Sharing and Norm Enforcement in an Agent-Only Social Network
von: Manik, Md Motaleb Hossen, et al.
Veröffentlicht: (2026)
von: Manik, Md Motaleb Hossen, et al.
Veröffentlicht: (2026)
Agent Privilege Separation in OpenClaw: A Structural Defense Against Prompt Injection
von: Cheng, Darren, et al.
Veröffentlicht: (2026)
von: Cheng, Darren, et al.
Veröffentlicht: (2026)
Clawdrain: Exploiting Tool-Calling Chains for Stealthy Token Exhaustion in OpenClaw Agents
von: Dong, Ben, et al.
Veröffentlicht: (2026)
von: Dong, Ben, et al.
Veröffentlicht: (2026)
Benchmarks for Trajectory Safety Evaluation and Diagnosis in OpenClaw and Codex: ATBench-Claw and ATBench-Codex
von: Yang, Zhonghao, et al.
Veröffentlicht: (2026)
von: Yang, Zhonghao, et al.
Veröffentlicht: (2026)
Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study
von: Xu, Luyao, et al.
Veröffentlicht: (2026)
von: Xu, Luyao, et al.
Veröffentlicht: (2026)
Security, Privacy, and Ethical Risks in OpenClaw
von: Jin, Yutong, et al.
Veröffentlicht: (2026)
von: Jin, Yutong, et al.
Veröffentlicht: (2026)
Null Space Constrained Contrastive Visual Forgetting for MLLM Unlearning
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
Robust MLLM Unlearning via Visual Knowledge Distillation
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
Foundations for Agentic AI Investigations from the Forensic Analysis of OpenClaw
von: Gruber, Jan, et al.
Veröffentlicht: (2026)
von: Gruber, Jan, et al.
Veröffentlicht: (2026)
Uncovering Security Threats and Architecting Defenses in Autonomous Agents: A Case Study of OpenClaw
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
von: Ying, Zonghao, et al.
Veröffentlicht: (2026)
When AI Agents Learn from Each Other: Insights from Emergent AI Agent Communities on OpenClaw for Human-AI Partnership in Education
von: Chen, Eason, et al.
Veröffentlicht: (2026)
von: Chen, Eason, et al.
Veröffentlicht: (2026)
GlitchMiner: Mining Glitch Tokens in Large Language Models via Gradient-based Discrete Optimization
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
Execution Is the New Attack Surface: Survivability-Aware Agentic Crypto Trading with OpenClaw-Style Local Executors
von: Borjigin, Ailiya, et al.
Veröffentlicht: (2026)
von: Borjigin, Ailiya, et al.
Veröffentlicht: (2026)
A Trajectory-Based Safety Audit of Clawdbot (OpenClaw)
von: Chen, Tianyu, et al.
Veröffentlicht: (2026)
von: Chen, Tianyu, et al.
Veröffentlicht: (2026)
Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
Autonomous Agent-Orchestrated Digital Twins (AADT): Leveraging the OpenClaw Framework for State Synchronization in Rare Genetic Disorders
von: Chen, Hongzhuo, et al.
Veröffentlicht: (2026)
von: Chen, Hongzhuo, et al.
Veröffentlicht: (2026)
OpenClaw PRISM: A Zero-Fork, Defense-in-Depth Runtime Security Layer for Tool-Augmented LLM Agents
von: Li, Frank
Veröffentlicht: (2026)
von: Li, Frank
Veröffentlicht: (2026)
Don't Let the Claw Grip Your Hand: A Security Analysis and Defense Framework for OpenClaw
von: Shan, Zhengyang, et al.
Veröffentlicht: (2026)
von: Shan, Zhengyang, et al.
Veröffentlicht: (2026)
From Agent-Only Social Networks to Autonomous Scientific Research: Lessons from OpenClaw and Moltbook, and the Architecture of ClawdLab and Beach.Science
von: Weidener, Lukas, et al.
Veröffentlicht: (2026)
von: Weidener, Lukas, et al.
Veröffentlicht: (2026)
SafeClaw-R: Towards Safe and Secure Multi-Agent Personal Assistants
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
OpenGo: An OpenClaw-Based Robotic Dog with Real-Time Skill Switching
von: Li, Hanbing, et al.
Veröffentlicht: (2026)
von: Li, Hanbing, et al.
Veröffentlicht: (2026)
ClawTrap: A MITM-Based Red-Teaming Framework for Real-World OpenClaw Security Evaluation
von: Zhao, Haochen, et al.
Veröffentlicht: (2026)
von: Zhao, Haochen, et al.
Veröffentlicht: (2026)
MediaClaw: Multimodal Intelligent-Agent Platform Technical Report
von: Zhao, Shaoan, et al.
Veröffentlicht: (2026)
von: Zhao, Shaoan, et al.
Veröffentlicht: (2026)
Trojan's Whisper: Stealthy Manipulation of OpenClaw through Injected Bootstrapped Guidance
von: Liu, Fazhong, et al.
Veröffentlicht: (2026)
von: Liu, Fazhong, et al.
Veröffentlicht: (2026)
Automating Computational Chemistry Workflows via OpenClaw and Domain-Specific Skills
von: Ding, Mingwei, et al.
Veröffentlicht: (2026)
von: Ding, Mingwei, et al.
Veröffentlicht: (2026)
ROSClaw: An OpenClaw ROS 2 Framework for Agentic Robot Control and Interaction
von: Cardenas, Irvin Steve, et al.
Veröffentlicht: (2026)
von: Cardenas, Irvin Steve, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Systematic Security Evaluation of OpenClaw and Its Variants
von: Wang, Yuhang, et al.
Veröffentlicht: (2026) -
A Security Analysis of the OpenClaw AI Agent Framework
von: Suwansathit, Surada, et al.
Veröffentlicht: (2026) -
OpenClaw-RL: Train Any Agent Simply by Talking
von: Wang, Yinjie, et al.
Veröffentlicht: (2026) -
ClawKeeper: Comprehensive Safety Protection for OpenClaw Agents Through Skills, Plugins, and Watchers
von: Liu, Songyang, et al.
Veröffentlicht: (2026) -
ICU-Bench:Benchmarking Continual Unlearning in Multimodal Large Language Models
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)