Securing AI Systems: A Guide to Known Attacks and Impacts
Fuente:
arXiv
Guardado en:
| Autores principales: | Kiribuchi, Naoto, Zenitani, Kengo, Semitsu, Takayuki |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Automated Analysis of Global AI Safety Initiatives: A Taxonomy-Driven LLM Approach
por: Semitsu, Takayuki, et al.
Publicado: (2026)
por: Semitsu, Takayuki, et al.
Publicado: (2026)
Securing AI Agents Against Prompt Injection Attacks
por: Ramakrishnan, Badrinath, et al.
Publicado: (2025)
por: Ramakrishnan, Badrinath, et al.
Publicado: (2025)
AI Security Map: Holistic Organization of AI Security Technologies and Impacts on Stakeholders
por: Kato, Hiroya, et al.
Publicado: (2025)
por: Kato, Hiroya, et al.
Publicado: (2025)
Architecting Secure AI Agents: Perspectives on System-Level Defenses Against Indirect Prompt Injection Attacks
por: Xiang, Chong, et al.
Publicado: (2026)
por: Xiang, Chong, et al.
Publicado: (2026)
Breaking the Code: Security Assessment of AI Code Agents Through Systematic Jailbreaking Attacks
por: Saha, Shoumik, et al.
Publicado: (2025)
por: Saha, Shoumik, et al.
Publicado: (2025)
SPEAR: Security Posture Evaluation using AI Planner-Reasoning on Attack-Connectivity Hypergraphs
por: Podder, Rakesh, et al.
Publicado: (2025)
por: Podder, Rakesh, et al.
Publicado: (2025)
Towards Robust and Secure Embodied AI: A Survey on Vulnerabilities and Attacks
por: Xing, Wenpeng, et al.
Publicado: (2025)
por: Xing, Wenpeng, et al.
Publicado: (2025)
Security of Internet of Agents: Attacks and Countermeasures
por: Wang, Yuntao, et al.
Publicado: (2025)
por: Wang, Yuntao, et al.
Publicado: (2025)
Offensive Security for AI Systems: Concepts, Practices, and Applications
por: Harguess, Josh, et al.
Publicado: (2025)
por: Harguess, Josh, et al.
Publicado: (2025)
Security-First AI: Foundations for Robust and Trustworthy Systems
por: Tallam, Krti
Publicado: (2025)
por: Tallam, Krti
Publicado: (2025)
CAM-LDS: Cyber Attack Manifestations for Automatic Interpretation of System Logs and Security Alerts
por: Landauer, Max, et al.
Publicado: (2026)
por: Landauer, Max, et al.
Publicado: (2026)
Serverless AI Security: Attack Surface Analysis and Runtime Protection Mechanisms for FaaS-Based Machine Learning
por: Pathade, Chetan, et al.
Publicado: (2026)
por: Pathade, Chetan, et al.
Publicado: (2026)
Evolving Security in LLMs: A Study of Jailbreak Attacks and Defenses
por: Shang, Zhengchun, et al.
Publicado: (2025)
por: Shang, Zhengchun, et al.
Publicado: (2025)
CyberSentinel: An Emergent Threat Detection System for AI Security
por: Tallam, Krti
Publicado: (2025)
por: Tallam, Krti
Publicado: (2025)
The System Prompt Is the Attack Surface: How LLM Agent Configuration Shapes Security and Creates Exploitable Vulnerabilities
por: Litvak, Ron
Publicado: (2026)
por: Litvak, Ron
Publicado: (2026)
Securing Agentic AI: Threat Modeling and Risk Analysis for Network Monitoring Agentic AI System
por: Zambare, Pallavi, et al.
Publicado: (2025)
por: Zambare, Pallavi, et al.
Publicado: (2025)
Security of AI Agents
por: He, Yifeng, et al.
Publicado: (2024)
por: He, Yifeng, et al.
Publicado: (2024)
MCP Guardian: A Security-First Layer for Safeguarding MCP-Based AI System
por: Kumar, Sonu, et al.
Publicado: (2025)
por: Kumar, Sonu, et al.
Publicado: (2025)
CivicShield: A Cross-Domain Defense-in-Depth Framework for Securing Government-Facing AI Chatbots Against Multi-Turn Adversarial Attacks
por: Patil, KrishnaSaiReddy
Publicado: (2026)
por: Patil, KrishnaSaiReddy
Publicado: (2026)
RAG Security and Privacy: Formalizing the Threat Model and Attack Surface
por: Arzanipour, Atousa, et al.
Publicado: (2025)
por: Arzanipour, Atousa, et al.
Publicado: (2025)
Enhancing TinyML Security: Study of Adversarial Attack Transferability
por: Shah, Parin, et al.
Publicado: (2024)
por: Shah, Parin, et al.
Publicado: (2024)
Democratizing ML for Enterprise Security: A Self-Sustained Attack Detection Framework
por: Momeni, Sadegh, et al.
Publicado: (2025)
por: Momeni, Sadegh, et al.
Publicado: (2025)
Securing Retrieval-Augmented Generation: A Taxonomy of Attacks, Defenses, and Future Directions
por: Xu, Yuming, et al.
Publicado: (2026)
por: Xu, Yuming, et al.
Publicado: (2026)
SECUREVENT: Hybrid AI/ML Security Monitoring for Distributed Event-Based Systems
por: Liang, Eric
Publicado: (2026)
por: Liang, Eric
Publicado: (2026)
ATLAS: AI-Assisted Threat-to-Assertion Learning for System-on-Chip Security Verification
por: Tashdid, Ishraq, et al.
Publicado: (2026)
por: Tashdid, Ishraq, et al.
Publicado: (2026)
Agentic AI for Cyber Resilience: A New Security Paradigm and Its System-Theoretic Foundations
por: Li, Tao, et al.
Publicado: (2025)
por: Li, Tao, et al.
Publicado: (2025)
Security of and by Generative AI platforms
por: Hayagreevan, Hari, et al.
Publicado: (2024)
por: Hayagreevan, Hari, et al.
Publicado: (2024)
The AI Security Pyramid of Pain
por: Ward, Chris M., et al.
Publicado: (2024)
por: Ward, Chris M., et al.
Publicado: (2024)
Secure Multiparty Generative AI
por: Shrestha, Manil, et al.
Publicado: (2024)
por: Shrestha, Manil, et al.
Publicado: (2024)
MATRA: Modeling the Attack Surface of Agentic AI Systems -- OpenClaw Case Study
por: Van hamme, Tim, et al.
Publicado: (2026)
por: Van hamme, Tim, et al.
Publicado: (2026)
Distillability of LLM Security Logic: Predicting Attack Success Rate of Outline Filling Attack via Ranking Regression
por: Zhang, Tianyu, et al.
Publicado: (2025)
por: Zhang, Tianyu, et al.
Publicado: (2025)
Secure Tug-of-War (SecTOW): Iterative Defense-Attack Training with Reinforcement Learning for Multimodal Model Security
por: Dai, Muzhi, et al.
Publicado: (2025)
por: Dai, Muzhi, et al.
Publicado: (2025)
Concept-Guided Backdoor Attack on Vision Language Models
por: Shen, Haoyu, et al.
Publicado: (2025)
por: Shen, Haoyu, et al.
Publicado: (2025)
PBI-Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization
por: Cheng, Ruoxi, et al.
Publicado: (2024)
por: Cheng, Ruoxi, et al.
Publicado: (2024)
Meta SecAlign: A Secure Foundation LLM Against Prompt Injection Attacks
por: Chen, Sizhe, et al.
Publicado: (2025)
por: Chen, Sizhe, et al.
Publicado: (2025)
Autonomous Action Runtime Management(AARM):A System Specification for Securing AI-Driven Actions at Runtime
por: Errico, Herman
Publicado: (2026)
por: Errico, Herman
Publicado: (2026)
Towards Secure MLOps: Surveying Attacks, Mitigation Strategies, and Research Challenges
por: Patel, Raj, et al.
Publicado: (2025)
por: Patel, Raj, et al.
Publicado: (2025)
WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks
por: Evtimov, Ivan, et al.
Publicado: (2025)
por: Evtimov, Ivan, et al.
Publicado: (2025)
FedSecurity: Benchmarking Attacks and Defenses in Federated Learning and Federated LLMs
por: Han, Shanshan, et al.
Publicado: (2023)
por: Han, Shanshan, et al.
Publicado: (2023)
Cascade: Composing Software-Hardware Attack Gadgets for Adversarial Threat Amplification in Compound AI Systems
por: Banerjee, Sarbartha, et al.
Publicado: (2026)
por: Banerjee, Sarbartha, et al.
Publicado: (2026)
Ejemplares similares
-
Automated Analysis of Global AI Safety Initiatives: A Taxonomy-Driven LLM Approach
por: Semitsu, Takayuki, et al.
Publicado: (2026) -
Securing AI Agents Against Prompt Injection Attacks
por: Ramakrishnan, Badrinath, et al.
Publicado: (2025) -
AI Security Map: Holistic Organization of AI Security Technologies and Impacts on Stakeholders
por: Kato, Hiroya, et al.
Publicado: (2025) -
Architecting Secure AI Agents: Perspectives on System-Level Defenses Against Indirect Prompt Injection Attacks
por: Xiang, Chong, et al.
Publicado: (2026) -
Breaking the Code: Security Assessment of AI Code Agents Through Systematic Jailbreaking Attacks
por: Saha, Shoumik, et al.
Publicado: (2025)