RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Yeke, Doguhuan, Zhou, Yanming, Lin, Leo Y., Cai, Hongyu, Bianchi, Antonio, Celik, Z. Berkay |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SafeAgentBench: A Benchmark for Safe Task Planning of Embodied LLM Agents
by: Yin, Sheng, et al.
Published: (2024)
by: Yin, Sheng, et al.
Published: (2024)
Revisiting Adversarial Perception Attacks and Defense Methods on Autonomous Driving Systems
by: Chen, Cheng, et al.
Published: (2025)
by: Chen, Cheng, et al.
Published: (2025)
AGENTSAFE: Benchmarking the Safety of Embodied Agents on Hazardous Instructions
by: Ying, Zonghao, et al.
Published: (2025)
by: Ying, Zonghao, et al.
Published: (2025)
Rethinking How to Evaluate Language Model Jailbreak
by: Cai, Hongyu, et al.
Published: (2024)
by: Cai, Hongyu, et al.
Published: (2024)
Exploring and Developing a Pre-Model Safeguard with Draft Models
by: Cai, Hongyu, et al.
Published: (2026)
by: Cai, Hongyu, et al.
Published: (2026)
Enhancing LLM-based Autonomous Driving Agents to Mitigate Perception Attacks
by: Song, Ruoyu, et al.
Published: (2024)
by: Song, Ruoyu, et al.
Published: (2024)
AdvGrasp: Adversarial Attacks on Robotic Grasping from a Physical Perspective
by: Wang, Xiaofei, et al.
Published: (2025)
by: Wang, Xiaofei, et al.
Published: (2025)
The Yes-Man Syndrome: Benchmarking Abstention in Embodied Robotic Agents
by: Yeke, Doguhan, et al.
Published: (2026)
by: Yeke, Doguhan, et al.
Published: (2026)
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
by: Li, Xiao, et al.
Published: (2026)
by: Li, Xiao, et al.
Published: (2026)
MT-JailBench: A Modular Benchmark for Understanding Multi-Turn Jailbreak Attacks
by: Zhang, Xinkai, et al.
Published: (2026)
by: Zhang, Xinkai, et al.
Published: (2026)
Diagnosis-guided Attack Recovery for Securing Robotic Vehicles from Sensor Deception Attacks
by: Dash, Pritam, et al.
Published: (2022)
by: Dash, Pritam, et al.
Published: (2022)
Investigating the Impact of Dark Patterns on LLM-Based Web Agents
by: Ersoy, Devin, et al.
Published: (2025)
by: Ersoy, Devin, et al.
Published: (2025)
Systematic Discovery of Semantic Attacks in Online Map Construction through Conditional Diffusion
by: Wang, Chenyi, et al.
Published: (2026)
by: Wang, Chenyi, et al.
Published: (2026)
Towards Robust and Secure Embodied AI: A Survey on Vulnerabilities and Attacks
by: Xing, Wenpeng, et al.
Published: (2025)
by: Xing, Wenpeng, et al.
Published: (2025)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
by: Zhang, Hanrong, et al.
Published: (2024)
by: Zhang, Hanrong, et al.
Published: (2024)
LM-Scout: Analyzing the Security of Language Model Integration in Android Apps
by: Ibrahim, Muhammad, et al.
Published: (2025)
by: Ibrahim, Muhammad, et al.
Published: (2025)
Propagating Unsafe Actions in LLM Controlled Multi-Robot Collaboration via Single Robot Compromise
by: Huang, Zhen, et al.
Published: (2026)
by: Huang, Zhen, et al.
Published: (2026)
The Shawshank Redemption of Embodied AI: Understanding and Benchmarking Indirect Environmental Jailbreaks
by: Li, Chunyang, et al.
Published: (2025)
by: Li, Chunyang, et al.
Published: (2025)
RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors
by: Bai, Fengshuo, et al.
Published: (2024)
by: Bai, Fengshuo, et al.
Published: (2024)
Detection of Adversarial Attacks in Robotic Perception
by: Sharawy, Ziad, et al.
Published: (2026)
by: Sharawy, Ziad, et al.
Published: (2026)
Beyond Model Jailbreak: Systematic Dissection of the "Ten DeadlySins" in Embodied Intelligence
by: Huang, Yuhang, et al.
Published: (2025)
by: Huang, Yuhang, et al.
Published: (2025)
BlackboxBench: A Comprehensive Benchmark of Black-box Adversarial Attacks
by: Zheng, Meixi, et al.
Published: (2023)
by: Zheng, Meixi, et al.
Published: (2023)
International Students and Scams: At Risk Abroad
by: Zhang, Katherine, et al.
Published: (2025)
by: Zhang, Katherine, et al.
Published: (2025)
MAUI: Reconstructing Private Client Data in Federated Transfer Learning
by: Dabholkar, Ahaan, et al.
Published: (2025)
by: Dabholkar, Ahaan, et al.
Published: (2025)
Not What You Asked For: Typographic Attacks in Household Robot Manipulation
by: Iranmanesh, Ali, et al.
Published: (2026)
by: Iranmanesh, Ali, et al.
Published: (2026)
Offensive Robot Cybersecurity
by: Mayoral-Vilches, Víctor
Published: (2025)
by: Mayoral-Vilches, Víctor
Published: (2025)
How Vulnerable Is My Learned Policy? Universal Adversarial Perturbation Attacks On Modern Behavior Cloning Policies
by: Kalra, Akansha, et al.
Published: (2025)
by: Kalra, Akansha, et al.
Published: (2025)
Integrating Robotic Navigation with Blockchain: A Novel PoS-Based Approach for Heterogeneous Robotic Teams
by: Paykari, Nasim, et al.
Published: (2025)
by: Paykari, Nasim, et al.
Published: (2025)
SpecGuard: Specification Aware Recovery for Robotic Autonomous Vehicles from Physical Attacks
by: Dash, Pritam, et al.
Published: (2024)
by: Dash, Pritam, et al.
Published: (2024)
On the Feasibility of Fingerprinting Collaborative Robot Network Traffic
by: Tang, Cheng, et al.
Published: (2023)
by: Tang, Cheng, et al.
Published: (2023)
Deep Learning Model Inversion Attacks and Defenses: A Comprehensive Survey
by: Yang, Wencheng, et al.
Published: (2025)
by: Yang, Wencheng, et al.
Published: (2025)
GPS-IDS: An Anomaly-based GPS Spoofing Attack Detection Framework for Autonomous Vehicles
by: Abrar, Murad Mehrab, et al.
Published: (2024)
by: Abrar, Murad Mehrab, et al.
Published: (2024)
Time-Constrained Intelligent Adversaries for Automation Vulnerability Testing: A Multi-Robot Patrol Case Study
by: Ward, James C., et al.
Published: (2025)
by: Ward, James C., et al.
Published: (2025)
EM-MIAs: Enhancing Membership Inference Attacks in Large Language Models through Ensemble Modeling
by: Song, Zichen, et al.
Published: (2024)
by: Song, Zichen, et al.
Published: (2024)
Channel State Information Analysis for Jamming Attack Detection in Static and Dynamic UAV Networks -- An Experimental Study
by: Mykytyn, Pavlo, et al.
Published: (2025)
by: Mykytyn, Pavlo, et al.
Published: (2025)
Cybersecurity and Embodiment Integrity for Modern Robots: A Conceptual Framework
by: Giaretta, Alberto, et al.
Published: (2024)
by: Giaretta, Alberto, et al.
Published: (2024)
JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks
by: Zhang, Xiaoyu, et al.
Published: (2023)
by: Zhang, Xiaoyu, et al.
Published: (2023)
Implementing a Robot Intrusion Prevention System (RIPS) for ROS 2
by: Soriano-Salvador, Enrique, et al.
Published: (2024)
by: Soriano-Salvador, Enrique, et al.
Published: (2024)
Drones that Think on their Feet: Sudden Landing Decisions with Embodied AI
by: Barbosa, Diego Ortiz, et al.
Published: (2025)
by: Barbosa, Diego Ortiz, et al.
Published: (2025)
Property-Guided Cyber-Physical Reduction and Surrogation for Safety Analysis in Robotic Vehicles
by: Sayom, Nazmus Shakib, et al.
Published: (2025)
by: Sayom, Nazmus Shakib, et al.
Published: (2025)
Similar Items
-
SafeAgentBench: A Benchmark for Safe Task Planning of Embodied LLM Agents
by: Yin, Sheng, et al.
Published: (2024) -
Revisiting Adversarial Perception Attacks and Defense Methods on Autonomous Driving Systems
by: Chen, Cheng, et al.
Published: (2025) -
AGENTSAFE: Benchmarking the Safety of Embodied Agents on Hazardous Instructions
by: Ying, Zonghao, et al.
Published: (2025) -
Rethinking How to Evaluate Language Model Jailbreak
by: Cai, Hongyu, et al.
Published: (2024) -
Exploring and Developing a Pre-Model Safeguard with Draft Models
by: Cai, Hongyu, et al.
Published: (2026)