PLAGUE: Plug-and-play framework for Lifelong Adaptive Generation of Multi-turn Exploits
Fuente:
arXiv
Saved in:
| Main Authors: | Bhuiya, Neeladri, Aggarwal, Madhav, Purwar, Diptanshu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents
by: Rahman, Salman, et al.
Published: (2025)
by: Rahman, Salman, et al.
Published: (2025)
TrinityGuard: A Unified Framework for Safeguarding Multi-Agent Systems
by: Wang, Kai, et al.
Published: (2026)
by: Wang, Kai, et al.
Published: (2026)
SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces
by: Jin, Chang, et al.
Published: (2026)
by: Jin, Chang, et al.
Published: (2026)
Wolfpack Adversarial Attack for Robust Multi-Agent Reinforcement Learning
by: Lee, Sunwoo, et al.
Published: (2025)
by: Lee, Sunwoo, et al.
Published: (2025)
CuDA2: An approach for Incorporating Traitor Agents into Cooperative Multi-Agent Systems
by: Chen, Zhen, et al.
Published: (2024)
by: Chen, Zhen, et al.
Published: (2024)
Hierarchical Adversarially-Resilient Multi-Agent Reinforcement Learning for Cyber-Physical Systems Security
by: Alqithami, Saad
Published: (2025)
by: Alqithami, Saad
Published: (2025)
Beyond Single-Agent Alignment: Preventing Context-Fragmented Violations in Multi-Agent Systems
by: Wu, Jie, et al.
Published: (2026)
by: Wu, Jie, et al.
Published: (2026)
Learning Communication Between Heterogeneous Agents in Multi-Agent Reinforcement Learning for Autonomous Cyber Defence
by: Popa, Alex, et al.
Published: (2026)
by: Popa, Alex, et al.
Published: (2026)
A Call to Action for a Secure-by-Design Generative AI Paradigm
by: Alharthi, Dalal, et al.
Published: (2025)
by: Alharthi, Dalal, et al.
Published: (2025)
LegalSim: Multi-Agent Simulation of Legal Systems for Discovering Procedural Exploits
by: Badhe, Sanket
Published: (2025)
by: Badhe, Sanket
Published: (2025)
Differentially Private Distributed Inference
by: Papachristou, Marios, et al.
Published: (2024)
by: Papachristou, Marios, et al.
Published: (2024)
Cloud Investigation Automation Framework (CIAF): An AI-Driven Approach to Cloud Forensics
by: Alharthi, Dalal, et al.
Published: (2025)
by: Alharthi, Dalal, et al.
Published: (2025)
Can LLMs get help from other LLMs without revealing private information?
by: Hartmann, Florian, et al.
Published: (2024)
by: Hartmann, Florian, et al.
Published: (2024)
Differentially Private Reinforcement Learning with Self-Play
by: Qiao, Dan, et al.
Published: (2024)
by: Qiao, Dan, et al.
Published: (2024)
CRAKEN: Cybersecurity LLM Agent with Knowledge-Based Execution
by: Shao, Minghao, et al.
Published: (2025)
by: Shao, Minghao, et al.
Published: (2025)
Robustness of Agentic AI Systems via Adversarially-Aligned Jacobian Regularization
by: Mumcu, Furkan, et al.
Published: (2026)
by: Mumcu, Furkan, et al.
Published: (2026)
Optimal Cost Constrained Adversarial Attacks For Multiple Agent Systems
by: Lu, Ziqing, et al.
Published: (2023)
by: Lu, Ziqing, et al.
Published: (2023)
AI-Driven Adaptive Adversaries and the Erosion of Cryptographic Trust in Public Key Systems
by: Radanliev, Petar
Published: (2026)
by: Radanliev, Petar
Published: (2026)
A Neuro-Symbolic Multi-Agent Approach to Legal-Cybersecurity Knowledge Integration
by: Bonfanti, Chiara, et al.
Published: (2025)
by: Bonfanti, Chiara, et al.
Published: (2025)
PsySafe: A Comprehensive Framework for Psychological-based Attack, Defense, and Evaluation of Multi-agent System Safety
by: Zhang, Zaibin, et al.
Published: (2024)
by: Zhang, Zaibin, et al.
Published: (2024)
Attack the Messages, Not the Agents: A Multi-round Adaptive Stealthy Tampering Framework for LLM-MAS
by: Yan, Bingyu, et al.
Published: (2025)
by: Yan, Bingyu, et al.
Published: (2025)
CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking
by: Rani, Nanda, et al.
Published: (2026)
by: Rani, Nanda, et al.
Published: (2026)
Securing Multi-Agent Systems Against Corruptions via Node Contribution Backpropagation
by: Wu, Chengcan, et al.
Published: (2025)
by: Wu, Chengcan, et al.
Published: (2025)
FLARE: Adaptive Multi-Dimensional Reputation for Robust Client Reliability in Federated Learning
by: Younesi, Abolfazl, et al.
Published: (2025)
by: Younesi, Abolfazl, et al.
Published: (2025)
Emerging Safety Attack and Defense in Federated Instruction Tuning of Large Language Models
by: Ye, Rui, et al.
Published: (2024)
by: Ye, Rui, et al.
Published: (2024)
Effective Red-Teaming of Policy-Adherent Agents
by: Nakash, Itay, et al.
Published: (2025)
by: Nakash, Itay, et al.
Published: (2025)
ScamAgents: How AI Agents Can Simulate Human-Level Scam Calls
by: Badhe, Sanket
Published: (2025)
by: Badhe, Sanket
Published: (2025)
Multi-Agent Actor-Critics in Autonomous Cyber Defense
by: Wang, Mingjun, et al.
Published: (2024)
by: Wang, Mingjun, et al.
Published: (2024)
Chronology of Multi-Agent Interactions for Provenance of Evolving Information
by: Chang, Ching-Chun, et al.
Published: (2025)
by: Chang, Ching-Chun, et al.
Published: (2025)
Collaborative penetration testing suite for emerging generative AI algorithms
by: Radanliev, Petar
Published: (2025)
by: Radanliev, Petar
Published: (2025)
Operationalising Artificial Intelligence Bills of Materials (AIBOMs) for Verifiable AI Provenance and Lifecycle Assurance
by: Radanliev, Petar, et al.
Published: (2026)
by: Radanliev, Petar, et al.
Published: (2026)
ClawWorm: Self-Propagating Attacks Across LLM Agent Ecosystems
by: Zhang, Yihao, et al.
Published: (2026)
by: Zhang, Yihao, et al.
Published: (2026)
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
by: Kong, Dezhang, et al.
Published: (2025)
by: Kong, Dezhang, et al.
Published: (2025)
AMLNet: A Knowledge-Based Multi-Agent Framework to Generate and Detect Realistic Money Laundering Transactions
by: Huda, Sabin, et al.
Published: (2025)
by: Huda, Sabin, et al.
Published: (2025)
Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems
by: Lee, Donghyun, et al.
Published: (2024)
by: Lee, Donghyun, et al.
Published: (2024)
From Cloud-Native to Trust-Native: A Protocol for Verifiable Multi-Agent Systems
by: Li, Muyang
Published: (2025)
by: Li, Muyang
Published: (2025)
Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents
by: de Witt, Christian Schroeder, et al.
Published: (2025)
by: de Witt, Christian Schroeder, et al.
Published: (2025)
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
by: Patil, KrishnaSaiReddy
Published: (2026)
by: Patil, KrishnaSaiReddy
Published: (2026)
GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems
by: Mateo-Torrejón, Pablo, et al.
Published: (2026)
by: Mateo-Torrejón, Pablo, et al.
Published: (2026)
Explainable and Fine-Grained Safeguarding of LLM Multi-Agent Systems via Bi-Level Graph Anomaly Detection
by: Pan, Junjun, et al.
Published: (2025)
by: Pan, Junjun, et al.
Published: (2025)
Similar Items
-
X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents
by: Rahman, Salman, et al.
Published: (2025) -
TrinityGuard: A Unified Framework for Safeguarding Multi-Agent Systems
by: Wang, Kai, et al.
Published: (2026) -
SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces
by: Jin, Chang, et al.
Published: (2026) -
Wolfpack Adversarial Attack for Robust Multi-Agent Reinforcement Learning
by: Lee, Sunwoo, et al.
Published: (2025) -
CuDA2: An approach for Incorporating Traitor Agents into Cooperative Multi-Agent Systems
by: Chen, Zhen, et al.
Published: (2024)