Multi-Agent Systems Execute Arbitrary Malicious Code
Fuente:
arXiv
Saved in:
| Main Authors: | Triedman, Harold, Jha, Rishi, Shmatikov, Vitaly |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Breaking and Fixing Defenses Against Control-Flow Hijacking in Multi-Agent Systems
by: Jha, Rishi, et al.
Published: (2025)
by: Jha, Rishi, et al.
Published: (2025)
Agent Meltdowns: The Road to Hell Is Paved with Helpful Agents
by: Jha, Rishi, et al.
Published: (2026)
by: Jha, Rishi, et al.
Published: (2026)
Deep-Research Agents Can Be Poisoned via User-Generated Content
by: Zhang, Tingwei, et al.
Published: (2026)
by: Zhang, Tingwei, et al.
Published: (2026)
Adversarial Illusions in Multi-Modal Embeddings
by: Zhang, Tingwei, et al.
Published: (2023)
by: Zhang, Tingwei, et al.
Published: (2023)
Rerouting LLM Routers
by: Shafran, Avital, et al.
Published: (2025)
by: Shafran, Avital, et al.
Published: (2025)
Adversarial Decoding: Generating Readable Documents for Adversarial Objectives
by: Zhang, Collin, et al.
Published: (2024)
by: Zhang, Collin, et al.
Published: (2024)
Machine Against the RAG: Jamming Retrieval-Augmented Generation with Blocker Documents
by: Shafran, Avital, et al.
Published: (2024)
by: Shafran, Avital, et al.
Published: (2024)
MillStone: How Open-Minded Are LLMs?
by: Triedman, Harold, et al.
Published: (2025)
by: Triedman, Harold, et al.
Published: (2025)
Adversarial Hubness in Multi-Modal Retrieval
by: Zhang, Tingwei, et al.
Published: (2024)
by: Zhang, Tingwei, et al.
Published: (2024)
Differential Degradation Vulnerabilities in Censorship Circumvention Systems
by: Sun, Zhen, et al.
Published: (2024)
by: Sun, Zhen, et al.
Published: (2024)
Self-interpreting Adversarial Images
by: Zhang, Tingwei, et al.
Published: (2024)
by: Zhang, Tingwei, et al.
Published: (2024)
Towards Quantum Machine Learning for Malicious Code Analysis
by: Lopez, Jesus, et al.
Published: (2025)
by: Lopez, Jesus, et al.
Published: (2025)
MIP against Agent: Malicious Image Patches Hijacking Multimodal OS Agents
by: Aichberger, Lukas, et al.
Published: (2025)
by: Aichberger, Lukas, et al.
Published: (2025)
FreeMOCA: Memory-Free Continual Learning for Malicious Code Analysis
by: Asadi, Zahra, et al.
Published: (2026)
by: Asadi, Zahra, et al.
Published: (2026)
MAIDS: Malicious Agent Identification-based Data Security Model for Cloud Environments
by: Gupta, Kishu, et al.
Published: (2024)
by: Gupta, Kishu, et al.
Published: (2024)
Localizing Malicious Outputs from CodeLLM
by: Borana, Mayukh, et al.
Published: (2025)
by: Borana, Mayukh, et al.
Published: (2025)
From Past to Present: A Survey of Malicious URL Detection Techniques, Datasets and Code Repositories
by: Tian, Ye, et al.
Published: (2025)
by: Tian, Ye, et al.
Published: (2025)
Continuous Multi-Task Pre-training for Malicious URL Detection and Webpage Classification
by: Li, Yujie, et al.
Published: (2024)
by: Li, Yujie, et al.
Published: (2024)
ML Study of MaliciousTransactions in Ethereum
by: Katz, Natan
Published: (2024)
by: Katz, Natan
Published: (2024)
How to Steal Reasoning Without Reasoning Traces
by: Zhang, Tingwei, et al.
Published: (2026)
by: Zhang, Tingwei, et al.
Published: (2026)
Detecting Malicious AI Agents Through Simulated Interactions
by: Pi, Yulu, et al.
Published: (2025)
by: Pi, Yulu, et al.
Published: (2025)
MOCHA: Are Code Language Models Robust Against Multi-Turn Malicious Coding Prompts?
by: Wahed, Muntasir, et al.
Published: (2025)
by: Wahed, Muntasir, et al.
Published: (2025)
Malicious and Unintentional Disclosure Risks in Large Language Models for Code Generation
by: Rabin, Rafiqul, et al.
Published: (2025)
by: Rabin, Rafiqul, et al.
Published: (2025)
Exploiting Leaderboards for Large-Scale Distribution of Malicious Models
by: Suri, Anshuman, et al.
Published: (2025)
by: Suri, Anshuman, et al.
Published: (2025)
Toward More Generalized Malicious URL Detection Models
by: Tsai, YunDa, et al.
Published: (2022)
by: Tsai, YunDa, et al.
Published: (2022)
A New Dataset and Methodology for Malicious URL Classification
by: Schvartzman, Ilan, et al.
Published: (2024)
by: Schvartzman, Ilan, et al.
Published: (2024)
GasTrace: Detecting Sandwich Attack Malicious Accounts in Ethereum
by: Liu, Zekai, et al.
Published: (2024)
by: Liu, Zekai, et al.
Published: (2024)
RobPI: Robust Private Inference against Malicious Client
by: Xue, Jiaqi, et al.
Published: (2026)
by: Xue, Jiaqi, et al.
Published: (2026)
Malicious Internet Entity Detection Using Local Graph Inference
by: Mandlik, Simon, et al.
Published: (2024)
by: Mandlik, Simon, et al.
Published: (2024)
CleanBase: Detecting Malicious Documents in RAG Knowledge Databases
by: Jin, Weifei, et al.
Published: (2026)
by: Jin, Weifei, et al.
Published: (2026)
Fake or Compromised? Making Sense of Malicious Clients in Federated Learning
by: Mozaffari, Hamid, et al.
Published: (2024)
by: Mozaffari, Hamid, et al.
Published: (2024)
WebGuard++:Interpretable Malicious URL Detection via Bidirectional Fusion of HTML Subgraphs and Multi-Scale Convolutional BERT
by: Tian, Ye, et al.
Published: (2025)
by: Tian, Ye, et al.
Published: (2025)
Byzantine Outside, Curious Inside: Reconstructing Data Through Malicious Updates
by: Yue, Kai, et al.
Published: (2025)
by: Yue, Kai, et al.
Published: (2025)
Efficient and Adaptable Detection of Malicious LLM Prompts via Bootstrap Aggregation
by: Hassan, Shayan Ali, et al.
Published: (2026)
by: Hassan, Shayan Ali, et al.
Published: (2026)
VLMGuard: Defending VLMs against Malicious Prompts via Unlabeled Data
by: Du, Xuefeng, et al.
Published: (2024)
by: Du, Xuefeng, et al.
Published: (2024)
PDFInspect: A Unified Feature Extraction Framework for Malicious Document Detection
by: P, Sharmila S
Published: (2026)
by: P, Sharmila S
Published: (2026)
Towards Novel Malicious Packet Recognition: A Few-Shot Learning Approach
by: Stein, Kyle, et al.
Published: (2024)
by: Stein, Kyle, et al.
Published: (2024)
A Consensus-Bayesian Framework for Detecting Malicious Activity in Enterprise Directory Access Graphs
by: Uppuluri, Pratyush, et al.
Published: (2026)
by: Uppuluri, Pratyush, et al.
Published: (2026)
Privacy-Constrained Policies via Mutual Information Regularized Policy Gradients
by: Cundy, Chris, et al.
Published: (2020)
by: Cundy, Chris, et al.
Published: (2020)
Examining the Rat in the Tunnel: Interpretable Multi-Label Classification of Tor-based Malware
by: Karunanayake, Ishan, et al.
Published: (2024)
by: Karunanayake, Ishan, et al.
Published: (2024)
Similar Items
-
Breaking and Fixing Defenses Against Control-Flow Hijacking in Multi-Agent Systems
by: Jha, Rishi, et al.
Published: (2025) -
Agent Meltdowns: The Road to Hell Is Paved with Helpful Agents
by: Jha, Rishi, et al.
Published: (2026) -
Deep-Research Agents Can Be Poisoned via User-Generated Content
by: Zhang, Tingwei, et al.
Published: (2026) -
Adversarial Illusions in Multi-Modal Embeddings
by: Zhang, Tingwei, et al.
Published: (2023) -
Rerouting LLM Routers
by: Shafran, Avital, et al.
Published: (2025)