SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cheng, Hao, Miao, Changtao, Song, Tianle, Wu, Yin, Liu, He, Xiao, Erjia, Chen, Junchi, Shi, Xiaoyu, Wang, Yichi, Yang, Jing, Wang, Taowen, Duan, Jinhao, Sun, Mengshu, Dong, Peiyan, Shen, Xuan, Cao, Yang, Xu, Renjing, Xu, Kaidi, Gu, Jindong, Zhang, Bo, Zhang, Jize, Lin, Chenhao, Torr, Philip, Shen, Chao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unveiling Typographic Deceptions: Insights of the Typographic Vulnerability in Large Vision-Language Model
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
Manipulation Facing Threats: Evaluating Physical Vulnerabilities in End-to-End Vision Language Action Models
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
Transfer Attack for Bad and Good: Explain and Boost Adversarial Transferability across Multimodal Large Language Models
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
Exploring Typographic Visual Prompts Injection Threats in Cross-Modality Generation Models
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
Not Just Text: Uncovering Vision Modality Typographic Threats in Image Generation Models
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
Jailbreak-AudioBench: In-Depth Evaluation and Analysis of Jailbreak Threats for Large Audio Language Models
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
Gaining the Sparse Rewards by Exploring Lottery Tickets in Spiking Neural Network
von: Cheng, Hao, et al.
Veröffentlicht: (2023)
von: Cheng, Hao, et al.
Veröffentlicht: (2023)
RRAM-Based Bio-Inspired Circuits for Mobile Epileptic Correlation Extraction and Seizure Prediction
von: Wang, Hao, et al.
Veröffentlicht: (2024)
von: Wang, Hao, et al.
Veröffentlicht: (2024)
Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
von: Duan, Jinhao, et al.
Veröffentlicht: (2023)
von: Duan, Jinhao, et al.
Veröffentlicht: (2023)
ACT-Diffusion: Efficient Adversarial Consistency Training for One-step Diffusion Models
von: Kong, Fei, et al.
Veröffentlicht: (2023)
von: Kong, Fei, et al.
Veröffentlicht: (2023)
Multi-Floor Zero-Shot Object Navigation Policy
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2024)
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2024)
On-the-Fly VLA Adaptation via Test-Time Reinforcement Learning
von: Liu, Changyu, et al.
Veröffentlicht: (2026)
von: Liu, Changyu, et al.
Veröffentlicht: (2026)
TriHelper: Zero-Shot Object Navigation with Dynamic Assistance
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2024)
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2024)
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
von: Fan, Haozhi, et al.
Veröffentlicht: (2026)
von: Fan, Haozhi, et al.
Veröffentlicht: (2026)
Sparse Neurons Carry Strong Signals of Question Ambiguity in LLMs
von: Zhang, Zhuoxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Zhuoxuan, et al.
Veröffentlicht: (2025)
SpecHFD
von: Shengyu, Wang, et al.
Veröffentlicht: (2026)
von: Shengyu, Wang, et al.
Veröffentlicht: (2026)
Conformal Lesion Segmentation for 3D Medical Images
von: Tan, Binyu, et al.
Veröffentlicht: (2025)
von: Tan, Binyu, et al.
Veröffentlicht: (2025)
Adversarial Example Soups: Improving Transferability and Stealthiness for Free
von: Yang, Bo, et al.
Veröffentlicht: (2024)
von: Yang, Bo, et al.
Veröffentlicht: (2024)
UniGround: Universal 3D Visual Grounding via Training-Free Scene Parsing
von: Zhang, Jiaxi, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaxi, et al.
Veröffentlicht: (2026)
Bi-Level Control of Weaving Sections in Mixed Traffic Environments with Connected and Automated Vehicles
von: Yan, Longhao, et al.
Veröffentlicht: (2024)
von: Yan, Longhao, et al.
Veröffentlicht: (2024)
Can Knowledge-Graph-based Retrieval Augmented Generation Really Retrieve What You Need?
von: Yu, Junchi, et al.
Veröffentlicht: (2025)
von: Yu, Junchi, et al.
Veröffentlicht: (2025)
Use as Many Surrogates as You Want: Selective Ensemble Attack to Unleash Transferability without Sacrificing Resource Efficiency
von: Yang, Bo, et al.
Veröffentlicht: (2025)
von: Yang, Bo, et al.
Veröffentlicht: (2025)
Platoon-Centric Green Light Optimal Speed Advisory Using Safe Reinforcement Learning
von: Yang, Ruining, et al.
Veröffentlicht: (2025)
von: Yang, Ruining, et al.
Veröffentlicht: (2025)
Collaboration of Fusion and Independence: Hypercomplex-driven Robust Multi-Modal Knowledge Graph Completion
von: Liu, Zhiqiang, et al.
Veröffentlicht: (2025)
von: Liu, Zhiqiang, et al.
Veröffentlicht: (2025)
Unleashing the Potential of SAM2 for Biomedical Images and Videos: A Survey
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
ClawSafety: "Safe" LLMs, Unsafe Agents
von: Wei, Bowen, et al.
Veröffentlicht: (2026)
von: Wei, Bowen, et al.
Veröffentlicht: (2026)
T2MBench: A Benchmark for Out-of-Distribution Text-to-Motion Generation
von: Yang, Bin, et al.
Veröffentlicht: (2026)
von: Yang, Bin, et al.
Veröffentlicht: (2026)
ConU: Conformal Uncertainty in Large Language Models with Correctness Coverage Guarantees
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2024)
DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation
von: Wang, Yunheng, et al.
Veröffentlicht: (2025)
von: Wang, Yunheng, et al.
Veröffentlicht: (2025)
TraceDet: Hallucination Detection from the Decoding Trace of Diffusion Large Language Models
von: Chang, Shenxu, et al.
Veröffentlicht: (2025)
von: Chang, Shenxu, et al.
Veröffentlicht: (2025)
Enhancing High-Speed Cruising Performance of Autonomous Vehicles through Integrated Deep Reinforcement Learning Framework
von: Liang, Jinhao, et al.
Veröffentlicht: (2024)
von: Liang, Jinhao, et al.
Veröffentlicht: (2024)
Self-Discovering Interpretable Diffusion Latent Directions for Responsible Text-to-Image Generation
von: Li, Hang, et al.
Veröffentlicht: (2023)
von: Li, Hang, et al.
Veröffentlicht: (2023)
Enforcing Cooperative Safety for Reinforcement Learning-based Mixed-Autonomy Platoon Control
von: Zhou, Jingyuan, et al.
Veröffentlicht: (2024)
von: Zhou, Jingyuan, et al.
Veröffentlicht: (2024)
PostAlign: Multimodal Grounding as a Corrective Lens for MLLMs
von: Wu, Yixuan, et al.
Veröffentlicht: (2025)
von: Wu, Yixuan, et al.
Veröffentlicht: (2025)
Learning Robust Generalizable Radiance Field with Visibility and Feature Augmented Point Representation
von: Wang, Jiaxu, et al.
Veröffentlicht: (2024)
von: Wang, Jiaxu, et al.
Veröffentlicht: (2024)
DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation
von: Hu, Wenhao, et al.
Veröffentlicht: (2025)
von: Hu, Wenhao, et al.
Veröffentlicht: (2025)
Privacy on the Fly: A Predictive Adversarial Transformation Network for Mobile Sensor Data
von: Song, Tianle, et al.
Veröffentlicht: (2025)
von: Song, Tianle, et al.
Veröffentlicht: (2025)
Don't Let the Claw Grip Your Hand: A Security Analysis and Defense Framework for OpenClaw
von: Shan, Zhengyang, et al.
Veröffentlicht: (2026)
von: Shan, Zhengyang, et al.
Veröffentlicht: (2026)
Toward Efficient Data-Free Unlearning
von: Zhang, Chenhao, et al.
Veröffentlicht: (2024)
von: Zhang, Chenhao, et al.
Veröffentlicht: (2024)
A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly
von: Yao, Yifan, et al.
Veröffentlicht: (2023)
von: Yao, Yifan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Unveiling Typographic Deceptions: Insights of the Typographic Vulnerability in Large Vision-Language Model
von: Cheng, Hao, et al.
Veröffentlicht: (2024) -
Manipulation Facing Threats: Evaluating Physical Vulnerabilities in End-to-End Vision Language Action Models
von: Cheng, Hao, et al.
Veröffentlicht: (2024) -
Transfer Attack for Bad and Good: Explain and Boost Adversarial Transferability across Multimodal Large Language Models
von: Cheng, Hao, et al.
Veröffentlicht: (2024) -
Exploring Typographic Visual Prompts Injection Threats in Cross-Modality Generation Models
von: Cheng, Hao, et al.
Veröffentlicht: (2025) -
Not Just Text: Uncovering Vision Modality Typographic Threats in Image Generation Models
von: Cheng, Hao, et al.
Veröffentlicht: (2024)