Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents
Fuente:
arXiv
Guardado en:
| Autores principales: | de Witt, Christian Schroeder, Krawiecka, Klaudia, Krawczuk, Igor, Hagag, Ben, Anderson, William L., Belcak, Peter, Bucknall, Ben, Cai, Xiaohong, Chopra, Ayush, Cohen, Doron, Del Rosario, Ron F., Draguns, Andis, Gray, Annie, Katz, Keren, Mavroudis, Vasilios, Mink, Jaron, Motwani, Sumeet Ramesh, Petit, Jonathan, Rembeck, Leif-Sebastian, Smith, Chandler, Sotiropoulos, John, Young, Steven, Scheffler, Sarah, Llewellyn, Mary |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Architecture Matters for Multi-Agent Security
por: Hagag, Ben, et al.
Publicado: (2026)
por: Hagag, Ben, et al.
Publicado: (2026)
Extending the OWASP Multi-Agentic System Threat Modeling Guide: Insights from Multi-Agent Security Research
por: Krawiecka, Klaudia, et al.
Publicado: (2025)
por: Krawiecka, Klaudia, et al.
Publicado: (2025)
Architecting Resilient LLM Agents: A Guide to Secure Plan-then-Execute Implementations
por: Del Rosario, Ron F., et al.
Publicado: (2025)
por: Del Rosario, Ron F., et al.
Publicado: (2025)
Referential Security as a New Paradigm for AI Evaluations
por: Ristea, Dan, et al.
Publicado: (2026)
por: Ristea, Dan, et al.
Publicado: (2026)
Limitations of Agents Simulated by Predictive Models
por: Douglas, Raymond, et al.
Publicado: (2024)
por: Douglas, Raymond, et al.
Publicado: (2024)
Unelicitable Backdoors in Language Models via Cryptographic Transformer Circuits
por: Draguns, Andis, et al.
Publicado: (2024)
por: Draguns, Andis, et al.
Publicado: (2024)
Mitigating the Influence of Distractor Tasks in LMs with Prior-Aware Decoding
por: Douglas, Raymond, et al.
Publicado: (2024)
por: Douglas, Raymond, et al.
Publicado: (2024)
Who Governs the Machine? A Machine Identity Governance Taxonomy (MIGT) for AI Systems Operating Across Enterprise and Geopolitical Boundaries
por: Kurtz, Andrew, et al.
Publicado: (2026)
por: Kurtz, Andrew, et al.
Publicado: (2026)
Zero-Trust Network Access (ZTNA)
por: Mavroudis, Vasilios
Publicado: (2024)
por: Mavroudis, Vasilios
Publicado: (2024)
Inherently Interpretable and Uncertainty-Aware Models for Online Learning in Cyber-Security Problems
por: Kolicic, Benjamin, et al.
Publicado: (2024)
por: Kolicic, Benjamin, et al.
Publicado: (2024)
CybORG++: An Enhanced Gym for the Development of Autonomous Cyber Agents
por: Emerson, Harry, et al.
Publicado: (2024)
por: Emerson, Harry, et al.
Publicado: (2024)
Thought Virus: Viral Misalignment via Subliminal Prompting in Multi-Agent Systems
por: Weckbecker, Moritz, et al.
Publicado: (2026)
por: Weckbecker, Moritz, et al.
Publicado: (2026)
HonestCyberEval: An AI Cyber Risk Benchmark for Automated Software Exploitation
por: Ristea, Dan, et al.
Publicado: (2024)
por: Ristea, Dan, et al.
Publicado: (2024)
Analysis of Publicly Accessible Operational Technology and Associated Risks
por: Rodda, Matthew, et al.
Publicado: (2025)
por: Rodda, Matthew, et al.
Publicado: (2025)
Quantifying Mix Network Privacy Erosion with Generative Models
por: Mavroudis, Vasilios, et al.
Publicado: (2025)
por: Mavroudis, Vasilios, et al.
Publicado: (2025)
An Attentive Graph Agent for Topology-Adaptive Cyber Defence
por: Sandoval, Ilya Orson, et al.
Publicado: (2025)
por: Sandoval, Ilya Orson, et al.
Publicado: (2025)
REAL: Benchmarking Autonomous Agents on Deterministic Simulations of Real Websites
por: Garg, Divyansh, et al.
Publicado: (2025)
por: Garg, Divyansh, et al.
Publicado: (2025)
Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
por: Putta, Pranav, et al.
Publicado: (2024)
por: Putta, Pranav, et al.
Publicado: (2024)
SoK (or SoLK?): On the Quantitative Study of Sociodemographic Factors and Computer Security Behaviors
por: Wei, Miranda, et al.
Publicado: (2024)
por: Wei, Miranda, et al.
Publicado: (2024)
On Efficient Bayesian Exploration in Model-Based Reinforcement Learning
por: Caron, Alberto, et al.
Publicado: (2025)
por: Caron, Alberto, et al.
Publicado: (2025)
Towards Causal Model-Based Policy Optimization
por: Caron, Alberto, et al.
Publicado: (2025)
por: Caron, Alberto, et al.
Publicado: (2025)
A View on Out-of-Distribution Identification from a Statistical Testing Theory Perspective
por: Caron, Alberto, et al.
Publicado: (2024)
por: Caron, Alberto, et al.
Publicado: (2024)
Nearest Neighbour with Bandit Feedback
por: Pasteris, Stephen, et al.
Publicado: (2023)
por: Pasteris, Stephen, et al.
Publicado: (2023)
Mitigating Deep Reinforcement Learning Backdoors in the Neural Activation Space
por: Vyas, Sanyam, et al.
Publicado: (2024)
por: Vyas, Sanyam, et al.
Publicado: (2024)
Beyond Rewards in Reinforcement Learning for Cyber Defence
por: Bates, Elizabeth, et al.
Publicado: (2026)
por: Bates, Elizabeth, et al.
Publicado: (2026)
Fairness with Exponential Weights
por: Pasteris, Stephen, et al.
Publicado: (2024)
por: Pasteris, Stephen, et al.
Publicado: (2024)
Less is more? Rewards in RL for Cyber Defence
por: Bates, Elizabeth, et al.
Publicado: (2025)
por: Bates, Elizabeth, et al.
Publicado: (2025)
Extraction Propagation
por: Pasteris, Stephen, et al.
Publicado: (2024)
por: Pasteris, Stephen, et al.
Publicado: (2024)
MALT: Improving Reasoning with Multi-Agent LLM Training
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
Position Paper: Model Access should be a Key Concern in AI Governance
por: Kembery, Edward, et al.
Publicado: (2024)
por: Kembery, Edward, et al.
Publicado: (2024)
Secret Collusion among AI Agents: Multi-Agent Deception via Steganography
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
"Unlimited Realm of Exploration and Experimentation": Methods and Motivations of AI-Generated Sexual Content Creators
por: Mink, Jaron, et al.
Publicado: (2026)
por: Mink, Jaron, et al.
Publicado: (2026)
An effect of abrupt current disruption
por: Dembovskis, Andis
Publicado: (2011)
por: Dembovskis, Andis
Publicado: (2011)
Who Owns This Agent? Tracing AI Agents Back to Their Owners
por: Chocron, Ruben, et al.
Publicado: (2026)
por: Chocron, Ruben, et al.
Publicado: (2026)
Emerging Practices in Frontier AI Safety Frameworks
por: Buhl, Marie Davidsen, et al.
Publicado: (2025)
por: Buhl, Marie Davidsen, et al.
Publicado: (2025)
Environment Complexity and Nash Equilibria in a Sequential Social Dilemma
por: Yasir, Mustafa, et al.
Publicado: (2024)
por: Yasir, Mustafa, et al.
Publicado: (2024)
Autonomous Network Defence using Reinforcement Learning
por: Foley, Myles, et al.
Publicado: (2024)
por: Foley, Myles, et al.
Publicado: (2024)
Online Convex Optimisation: The Optimal Switching Regret for all Segmentations Simultaneously
por: Pasteris, Stephen, et al.
Publicado: (2024)
por: Pasteris, Stephen, et al.
Publicado: (2024)
Security of AI Agents
por: He, Yifeng, et al.
Publicado: (2024)
por: He, Yifeng, et al.
Publicado: (2024)
The Myth of Slavery: Modern Study Shows that Slaves was Treated Much better than what was thought before
por: Hagag, Waleed
Publicado: (2026)
por: Hagag, Waleed
Publicado: (2026)
Ejemplares similares
-
Architecture Matters for Multi-Agent Security
por: Hagag, Ben, et al.
Publicado: (2026) -
Extending the OWASP Multi-Agentic System Threat Modeling Guide: Insights from Multi-Agent Security Research
por: Krawiecka, Klaudia, et al.
Publicado: (2025) -
Architecting Resilient LLM Agents: A Guide to Secure Plan-then-Execute Implementations
por: Del Rosario, Ron F., et al.
Publicado: (2025) -
Referential Security as a New Paradigm for AI Evaluations
por: Ristea, Dan, et al.
Publicado: (2026) -
Limitations of Agents Simulated by Predictive Models
por: Douglas, Raymond, et al.
Publicado: (2024)