Red-Teaming LLM Multi-Agent Systems via Communication Attacks
Fuente:
arXiv
Salvato in:
| Autori principali: | He, Pengfei, Lin, Yupin, Dong, Shen, Xu, Han, Xing, Yue, Liu, Hui |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems
di: He, Pengfei, et al.
Pubblicazione: (2025)
di: He, Pengfei, et al.
Pubblicazione: (2025)
SkillAttack: Automated Red Teaming of Agent Skills through Attack Path Refinement
di: Duan, Zenghao, et al.
Pubblicazione: (2026)
di: Duan, Zenghao, et al.
Pubblicazione: (2026)
Comprehensive Vulnerability Analysis is Necessary for Trustworthy LLM-MAS
di: He, Pengfei, et al.
Pubblicazione: (2025)
di: He, Pengfei, et al.
Pubblicazione: (2025)
RedTWIZ: Diverse LLM Red Teaming via Adaptive Attack Planning
di: Horal, Artur, et al.
Pubblicazione: (2025)
di: Horal, Artur, et al.
Pubblicazione: (2025)
Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents
di: He, Pengfei, et al.
Pubblicazione: (2026)
di: He, Pengfei, et al.
Pubblicazione: (2026)
Multi-Faceted Studies on Data Poisoning can Advance LLM Development
di: He, Pengfei, et al.
Pubblicazione: (2025)
di: He, Pengfei, et al.
Pubblicazione: (2025)
SafeSearch: Automated Red-Teaming of LLM-Based Search Agents
di: Dong, Jianshuo, et al.
Pubblicazione: (2025)
di: Dong, Jianshuo, et al.
Pubblicazione: (2025)
Stealthy Backdoor Attack via Confidence-driven Sampling
di: He, Pengfei, et al.
Pubblicazione: (2023)
di: He, Pengfei, et al.
Pubblicazione: (2023)
Data Poisoning for In-context Learning
di: He, Pengfei, et al.
Pubblicazione: (2024)
di: He, Pengfei, et al.
Pubblicazione: (2024)
PRM-Free Security Alignment of Large Models via Red Teaming and Adversarial Training
di: Du, Pengfei
Pubblicazione: (2025)
di: Du, Pengfei
Pubblicazione: (2025)
CyberSleuth: Autonomous Blue-Team LLM Agent for Web Attack Forensics
di: Fumero, Stefano, et al.
Pubblicazione: (2025)
di: Fumero, Stefano, et al.
Pubblicazione: (2025)
Sharpness-Aware Data Poisoning Attack
di: He, Pengfei, et al.
Pubblicazione: (2023)
di: He, Pengfei, et al.
Pubblicazione: (2023)
Autonomous Adversary: Red-Teaming in the age of LLM
di: Mamun, Mohammad, et al.
Pubblicazione: (2026)
di: Mamun, Mohammad, et al.
Pubblicazione: (2026)
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
di: Kong, Dezhang, et al.
Pubblicazione: (2025)
di: Kong, Dezhang, et al.
Pubblicazione: (2025)
Incalmo: An Autonomous LLM-assisted System for Red Teaming Multi-Host Networks
di: Singer, Brian, et al.
Pubblicazione: (2025)
di: Singer, Brian, et al.
Pubblicazione: (2025)
Unveiling Privacy Risks in LLM Agent Memory
di: Wang, Bo, et al.
Pubblicazione: (2025)
di: Wang, Bo, et al.
Pubblicazione: (2025)
CheatAgent: Attacking LLM-Empowered Recommender Systems via LLM Agent
di: Ning, Liang-bo, et al.
Pubblicazione: (2025)
di: Ning, Liang-bo, et al.
Pubblicazione: (2025)
Red Teaming Methodology for Design Obfuscation
di: Liu, Yuntao, et al.
Pubblicazione: (2025)
di: Liu, Yuntao, et al.
Pubblicazione: (2025)
Collaborative Shadows: Distributed Backdoor Attacks in LLM-Based Multi-Agent Systems
di: Zhu, Pengyu, et al.
Pubblicazione: (2025)
di: Zhu, Pengyu, et al.
Pubblicazione: (2025)
AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration
di: Zhou, Andy, et al.
Pubblicazione: (2025)
di: Zhou, Andy, et al.
Pubblicazione: (2025)
Mitigating the Privacy Issues in Retrieval-Augmented Generation (RAG) via Pure Synthetic Data
di: Zeng, Shenglai, et al.
Pubblicazione: (2024)
di: Zeng, Shenglai, et al.
Pubblicazione: (2024)
Automatic Red Teaming LLM-based Agents with Model Context Protocol Tools
di: He, Ping, et al.
Pubblicazione: (2025)
di: He, Ping, et al.
Pubblicazione: (2025)
MUZZLE: Adaptive Agentic Red-Teaming of Web Agents Against Indirect Prompt Injection Attacks
di: Syros, Georgios, et al.
Pubblicazione: (2026)
di: Syros, Georgios, et al.
Pubblicazione: (2026)
Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw
di: Yao, Hongwei, et al.
Pubblicazione: (2026)
di: Yao, Hongwei, et al.
Pubblicazione: (2026)
The Trust Paradox in LLM-Based Multi-Agent Systems: When Collaboration Becomes a Security Vulnerability
di: Xu, Zijie, et al.
Pubblicazione: (2025)
di: Xu, Zijie, et al.
Pubblicazione: (2025)
Reasoning-Style Poisoning of LLM Agents via Stealthy Style Transfer: Process-Level Attacks and Runtime Monitoring in RSV Space
di: Zhou, Xingfu, et al.
Pubblicazione: (2025)
di: Zhou, Xingfu, et al.
Pubblicazione: (2025)
SIRAJ: Diverse and Efficient Red-Teaming for LLM Agents via Distilled Structured Reasoning
di: Zhou, Kaiwen, et al.
Pubblicazione: (2025)
di: Zhou, Kaiwen, et al.
Pubblicazione: (2025)
When Scanners Lie: Evaluator Instability in LLM Red-Teaming
di: Erez, Lidor, et al.
Pubblicazione: (2026)
di: Erez, Lidor, et al.
Pubblicazione: (2026)
Adversarial Contrastive Learning for LLM Quantization Attacks
di: Song, Dinghong, et al.
Pubblicazione: (2026)
di: Song, Dinghong, et al.
Pubblicazione: (2026)
Automated Progressive Red Teaming
di: Jiang, Bojian, et al.
Pubblicazione: (2024)
di: Jiang, Bojian, et al.
Pubblicazione: (2024)
Red Teaming GPT-4V: Are GPT-4V Safe Against Uni/Multi-Modal Jailbreak Attacks?
di: Chen, Shuo, et al.
Pubblicazione: (2024)
di: Chen, Shuo, et al.
Pubblicazione: (2024)
InputSnatch: Stealing Input in LLM Services via Timing Side-Channel Attacks
di: Zheng, Xinyao, et al.
Pubblicazione: (2024)
di: Zheng, Xinyao, et al.
Pubblicazione: (2024)
Prompt Optimization and Evaluation for LLM Automated Red Teaming
di: Freenor, Michael, et al.
Pubblicazione: (2025)
di: Freenor, Michael, et al.
Pubblicazione: (2025)
T-MAP: Red-Teaming LLM Agents with Trajectory-aware Evolutionary Search
di: Lee, Hyomin, et al.
Pubblicazione: (2026)
di: Lee, Hyomin, et al.
Pubblicazione: (2026)
LAAF: Logic-layer Automated Attack Framework A Systematic Red-Teaming Methodology for LPCI Vulnerabilities in Agentic Large Language Model Systems
di: Atta, Hammad, et al.
Pubblicazione: (2026)
di: Atta, Hammad, et al.
Pubblicazione: (2026)
AegisAgent: An Autonomous Defense Agent Against Prompt Injection Attacks in LLM-HARs
di: Wang, Yihan, et al.
Pubblicazione: (2025)
di: Wang, Yihan, et al.
Pubblicazione: (2025)
Black-Box Skill Stealing Attack from Proprietary LLM Agents: An Empirical Study
di: Wang, Zihan, et al.
Pubblicazione: (2026)
di: Wang, Zihan, et al.
Pubblicazione: (2026)
ProvAgent: Threat Detection Based on Identity-Behavior Binding and Multi-Agent Collaborative Attack Investigation
di: Yan, Wenhao, et al.
Pubblicazione: (2026)
di: Yan, Wenhao, et al.
Pubblicazione: (2026)
Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection
di: Debi, Tanusree, et al.
Pubblicazione: (2026)
di: Debi, Tanusree, et al.
Pubblicazione: (2026)
Visual Exclusivity Attacks: Automatic Multimodal Red Teaming via Agentic Planning
di: Zhang, Yunbei, et al.
Pubblicazione: (2026)
di: Zhang, Yunbei, et al.
Pubblicazione: (2026)
Documenti analoghi
-
To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems
di: He, Pengfei, et al.
Pubblicazione: (2025) -
SkillAttack: Automated Red Teaming of Agent Skills through Attack Path Refinement
di: Duan, Zenghao, et al.
Pubblicazione: (2026) -
Comprehensive Vulnerability Analysis is Necessary for Trustworthy LLM-MAS
di: He, Pengfei, et al.
Pubblicazione: (2025) -
RedTWIZ: Diverse LLM Red Teaming via Adaptive Attack Planning
di: Horal, Artur, et al.
Pubblicazione: (2025) -
Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents
di: He, Pengfei, et al.
Pubblicazione: (2026)