A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Ada, Wu, Yongjiang, Zhang, Junyuan, Xiao, Jingyu, Yang, Shu, Huang, Jen-tse, Wang, Kun, Wang, Wenxuan, Wang, Shuai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Jailbreak Distillation: Renewable Safety Benchmarking
por: Zhang, Jingyu, et al.
Publicado: (2025)
por: Zhang, Jingyu, et al.
Publicado: (2025)
Argus: Reorchestrating Static Analysis via a Multi-Agent Ensemble for Full-Chain Security Vulnerability Detection
por: Liang, Zi, et al.
Publicado: (2026)
por: Liang, Zi, et al.
Publicado: (2026)
Exploring the Security Threats of Retriever Backdoors in Retrieval-Augmented Code Generation
por: Li, Tian, et al.
Publicado: (2025)
por: Li, Tian, et al.
Publicado: (2025)
Exploring the Security Threats of Knowledge Base Poisoning in Retrieval-Augmented Code Generation
por: Lin, Bo, et al.
Publicado: (2025)
por: Lin, Bo, et al.
Publicado: (2025)
ProSec: Fortifying Code LLMs with Proactive Security Alignment
por: Xu, Xiangzhe, et al.
Publicado: (2024)
por: Xu, Xiangzhe, et al.
Publicado: (2024)
Agent Skills in the Wild: An Empirical Study of Security Vulnerabilities at Scale
por: Liu, Yi, et al.
Publicado: (2026)
por: Liu, Yi, et al.
Publicado: (2026)
Bugdar: AI-Augmented Secure Code Review for GitHub Pull Requests
por: Naulty, John, et al.
Publicado: (2025)
por: Naulty, John, et al.
Publicado: (2025)
Deep Learning Model Security: Threats and Defenses
por: Wang, Tianyang, et al.
Publicado: (2024)
por: Wang, Tianyang, et al.
Publicado: (2024)
Web Agents Should Adopt the Plan-Then-Execute Paradigm
por: Piet, Julien, et al.
Publicado: (2026)
por: Piet, Julien, et al.
Publicado: (2026)
An Extensive Comparison of Static Application Security Testing Tools
por: Esposito, Matteo, et al.
Publicado: (2024)
por: Esposito, Matteo, et al.
Publicado: (2024)
Risks and Compliance with the EU's Core Cyber Security Legislation
por: Ruohonen, Jukka, et al.
Publicado: (2025)
por: Ruohonen, Jukka, et al.
Publicado: (2025)
An Overview of Cyber Security Funding for Open Source Software
por: Ruohonen, Jukka, et al.
Publicado: (2024)
por: Ruohonen, Jukka, et al.
Publicado: (2024)
Security Is Relative: Training-Free Vulnerability Detection via Multi-Agent Behavioral Contract Synthesis
por: Wang, Yongchao, et al.
Publicado: (2026)
por: Wang, Yongchao, et al.
Publicado: (2026)
Towards Personalizing Secure Programming Education with LLM-Injected Vulnerabilities
por: Frazier, Matthew, et al.
Publicado: (2026)
por: Frazier, Matthew, et al.
Publicado: (2026)
SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios
por: Chen, Junkai, et al.
Publicado: (2025)
por: Chen, Junkai, et al.
Publicado: (2025)
Exploring Privacy and Security as Drivers for Environmental Sustainability in Cloud-Based Office Solutions
por: Kayembe, Jason, et al.
Publicado: (2025)
por: Kayembe, Jason, et al.
Publicado: (2025)
Training Language Model Agents to Find Vulnerabilities with CTF-Dojo
por: Zhuo, Terry Yue, et al.
Publicado: (2025)
por: Zhuo, Terry Yue, et al.
Publicado: (2025)
Managing Security Evidence in Safety-Critical Organizations
por: Mohamad, Mazen, et al.
Publicado: (2024)
por: Mohamad, Mazen, et al.
Publicado: (2024)
The Security Performance Analysis of Blockchain System Based on Post-Quantum Cryptography -- A Case Study of Cryptocurrency Exchanges
por: Chen, Abel C. H.
Publicado: (2024)
por: Chen, Abel C. H.
Publicado: (2024)
An Interview Study on Third-Party Cyber Threat Hunting Processes in the U.S. Department of Homeland Security
por: Maxam III, William P., et al.
Publicado: (2024)
por: Maxam III, William P., et al.
Publicado: (2024)
AC4: Algebraic Computation Checker for Circuit Constraints in ZKPs
por: Yang, Qizhe, et al.
Publicado: (2024)
por: Yang, Qizhe, et al.
Publicado: (2024)
Chain-of-Jailbreak Attack for Image Generation Models via Editing Step by Step
por: Wang, Wenxuan, et al.
Publicado: (2024)
por: Wang, Wenxuan, et al.
Publicado: (2024)
SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
por: Hu, Qi, et al.
Publicado: (2026)
por: Hu, Qi, et al.
Publicado: (2026)
A Survey of Web Application Security Tutorials
por: Chembakottu, Bhagya, et al.
Publicado: (2026)
por: Chembakottu, Bhagya, et al.
Publicado: (2026)
Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Agents
por: Li, Xu, et al.
Publicado: (2026)
por: Li, Xu, et al.
Publicado: (2026)
Understanding, Implementing, and Supporting Security Assurance Cases in Safety-Critical Domains
por: Mohamad, Mazen
Publicado: (2025)
por: Mohamad, Mazen
Publicado: (2025)
How to Secure Existing C and C++ Software without Memory Safety
por: Erlingsson, Úlfar
Publicado: (2025)
por: Erlingsson, Úlfar
Publicado: (2025)
Extending the OWASP Multi-Agentic System Threat Modeling Guide: Insights from Multi-Agent Security Research
por: Krawiecka, Klaudia, et al.
Publicado: (2025)
por: Krawiecka, Klaudia, et al.
Publicado: (2025)
LightSC: The Making of a Usable Security Classification Tool for DevSecOps
por: Shrestha, Manish, et al.
Publicado: (2024)
por: Shrestha, Manish, et al.
Publicado: (2024)
Civil Servants as Builders: Enabling Non-IT Staff to Develop Secure Python and R Tools
por: Sharma, Prashant
Publicado: (2025)
por: Sharma, Prashant
Publicado: (2025)
Reading Between the Code Lines: On the Use of Self-Admitted Technical Debt for Security Analysis
por: Ferreyra, Nicolás E. Díaz, et al.
Publicado: (2026)
por: Ferreyra, Nicolás E. Díaz, et al.
Publicado: (2026)
AutoTestForge: A Multidimensional Automated Testing Framework for Natural Language Processing Models
por: Xing, Hengrui, et al.
Publicado: (2025)
por: Xing, Hengrui, et al.
Publicado: (2025)
Security Concerns in Generative AI Coding Assistants: Insights from Online Discussions on GitHub Copilot
por: Ferreyra, Nicolás E. Díaz, et al.
Publicado: (2026)
por: Ferreyra, Nicolás E. Díaz, et al.
Publicado: (2026)
Give LLMs a Security Course: Securing Retrieval-Augmented Code Generation via Knowledge Injection
por: Lin, Bo, et al.
Publicado: (2025)
por: Lin, Bo, et al.
Publicado: (2025)
RealSec-bench: A Benchmark for Evaluating Secure Code Generation in Real-World Repositories
por: Wang, Yanlin, et al.
Publicado: (2026)
por: Wang, Yanlin, et al.
Publicado: (2026)
No Silver Bullet: Towards Demonstrating Secure Software Development for Danish Small and Medium Enterprises in a Business-to-Business Model
por: Asadi, Raha, et al.
Publicado: (2025)
por: Asadi, Raha, et al.
Publicado: (2025)
SkillProbe: Security Auditing for Emerging Agent Skill Marketplaces via Multi-Agent Collaboration
por: Guo, Zihan, et al.
Publicado: (2026)
por: Guo, Zihan, et al.
Publicado: (2026)
BraveGuard: From Open-World Threats to Safer Computer-Use Agents
por: Feng, Yunhao, et al.
Publicado: (2026)
por: Feng, Yunhao, et al.
Publicado: (2026)
Bridging Safety and Security in Complex Systems: A Model-Based Approach with SAFT-GT Toolchain
por: Pekaric, Irdin, et al.
Publicado: (2026)
por: Pekaric, Irdin, et al.
Publicado: (2026)
SafeToolBench: Pioneering a Prospective Benchmark to Evaluating Tool Utilization Safety in LLMs
por: Xia, Hongfei, et al.
Publicado: (2025)
por: Xia, Hongfei, et al.
Publicado: (2025)
Ejemplares similares
-
Jailbreak Distillation: Renewable Safety Benchmarking
por: Zhang, Jingyu, et al.
Publicado: (2025) -
Argus: Reorchestrating Static Analysis via a Multi-Agent Ensemble for Full-Chain Security Vulnerability Detection
por: Liang, Zi, et al.
Publicado: (2026) -
Exploring the Security Threats of Retriever Backdoors in Retrieval-Augmented Code Generation
por: Li, Tian, et al.
Publicado: (2025) -
Exploring the Security Threats of Knowledge Base Poisoning in Retrieval-Augmented Code Generation
por: Lin, Bo, et al.
Publicado: (2025) -
ProSec: Fortifying Code LLMs with Proactive Security Alignment
por: Xu, Xiangzhe, et al.
Publicado: (2024)