CivicShield: A Cross-Domain Defense-in-Depth Framework for Securing Government-Facing AI Chatbots Against Multi-Turn Adversarial Attacks
Fuente:
arXiv
Guardado en:
| Autor principal: | Patil, KrishnaSaiReddy |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
por: Patil, KrishnaSaiReddy
Publicado: (2026)
por: Patil, KrishnaSaiReddy
Publicado: (2026)
RAGShield: Detecting Numerical Claim Manipulation in Government RAG Systems
por: Patil, KrishnaSaiReddy
Publicado: (2026)
por: Patil, KrishnaSaiReddy
Publicado: (2026)
Bidirectional Intention Inference Enhances LLMs' Defense Against Multi-Turn Jailbreak Attacks
por: Tong, Haibo, et al.
Publicado: (2025)
por: Tong, Haibo, et al.
Publicado: (2025)
Mirage: Defense against CrossPath Attacks in Software Defined Networks
por: Murtuza, Shariq, et al.
Publicado: (2024)
por: Murtuza, Shariq, et al.
Publicado: (2024)
Cross-Domain AI for Early Attack Detection and Defense Against Malicious Flows in O-RAN
por: Xavier, Bruno Missi, et al.
Publicado: (2024)
por: Xavier, Bruno Missi, et al.
Publicado: (2024)
Scam Shield: Multi-Model Voting and Fine-Tuned LLMs Against Adversarial Attacks
por: Chang, Chen-Wei, et al.
Publicado: (2025)
por: Chang, Chen-Wei, et al.
Publicado: (2025)
Securing Genomic Data Against Inference Attacks in Federated Learning Environments
por: Pathade, Chetan, et al.
Publicado: (2025)
por: Pathade, Chetan, et al.
Publicado: (2025)
Shielding Latent Face Representations From Privacy Attacks
por: Kaushik, Arjun Ramesh, et al.
Publicado: (2025)
por: Kaushik, Arjun Ramesh, et al.
Publicado: (2025)
A No-Defense Defense Against Gradient-Based Adversarial Attacks on ML-NIDS: Is Less More?
por: elShehaby, Mohamed, et al.
Publicado: (2026)
por: elShehaby, Mohamed, et al.
Publicado: (2026)
AntiFLipper: A Secure and Efficient Defense Against Label-Flipping Attacks in Federated Learning
por: Rahman, Aashnan, et al.
Publicado: (2025)
por: Rahman, Aashnan, et al.
Publicado: (2025)
The Aegis Protocol: A Foundational Security Framework for Autonomous AI Agents
por: Adapala, Sai Teja Reddy, et al.
Publicado: (2025)
por: Adapala, Sai Teja Reddy, et al.
Publicado: (2025)
Defense Against Prompt Injection Attack by Leveraging Attack Techniques
por: Chen, Yulin, et al.
Publicado: (2024)
por: Chen, Yulin, et al.
Publicado: (2024)
A Multi-task Adversarial Attack Against Face Authentication
por: Wang, Hanrui, et al.
Publicado: (2024)
por: Wang, Hanrui, et al.
Publicado: (2024)
A Defensive Framework Against Adversarial Attacks on Machine Learning-Based Network Intrusion Detection Systems
por: Tafreshian, Benyamin, et al.
Publicado: (2025)
por: Tafreshian, Benyamin, et al.
Publicado: (2025)
Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense
por: Hao, Shuyang, et al.
Publicado: (2025)
por: Hao, Shuyang, et al.
Publicado: (2025)
DYNAMITE: Dynamic Defense Selection for Enhancing Machine Learning-based Intrusion Detection Against Adversarial Attacks
por: Chen, Jing, et al.
Publicado: (2025)
por: Chen, Jing, et al.
Publicado: (2025)
MCP-Guard: A Multi-Stage Defense-in-Depth Framework for Securing Model Context Protocol in Agentic AI
por: Xing, Wenpeng, et al.
Publicado: (2025)
por: Xing, Wenpeng, et al.
Publicado: (2025)
Architecting Secure AI Agents: Perspectives on System-Level Defenses Against Indirect Prompt Injection Attacks
por: Xiang, Chong, et al.
Publicado: (2026)
por: Xiang, Chong, et al.
Publicado: (2026)
Surviving the Unseen: Predictive Defense for Novel Multi-Turn Multimodal Attacks
por: You, Doohee
Publicado: (2026)
por: You, Doohee
Publicado: (2026)
SecureLearn -- An Attack-agnostic Defense for Multiclass Machine Learning Against Data Poisoning Attacks
por: Paracha, Anum, et al.
Publicado: (2025)
por: Paracha, Anum, et al.
Publicado: (2025)
System Password Security: Attack and Defense Mechanisms
por: Shi, Chaofang, et al.
Publicado: (2025)
por: Shi, Chaofang, et al.
Publicado: (2025)
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
por: Brown, Hannah, et al.
Publicado: (2024)
por: Brown, Hannah, et al.
Publicado: (2024)
Modeling the Attack: Detecting AI-Generated Text by Quantifying Adversarial Perturbations
por: Teja, Lekkala Sai, et al.
Publicado: (2025)
por: Teja, Lekkala Sai, et al.
Publicado: (2025)
A Zero Trust Framework for Realization and Defense Against Generative AI Attacks in Power Grid
por: Munir, Md. Shirajum, et al.
Publicado: (2024)
por: Munir, Md. Shirajum, et al.
Publicado: (2024)
Towards Imperceptible Adversarial Defense: A Gradient-Driven Shield against Facial Manipulations
por: Li, Yue, et al.
Publicado: (2025)
por: Li, Yue, et al.
Publicado: (2025)
TrapSuffix: Proactive Defense Against Adversarial Suffixes in Jailbreaking
por: Du, Mengyao, et al.
Publicado: (2026)
por: Du, Mengyao, et al.
Publicado: (2026)
Beyond Surface-Level Patterns: An Essence-Driven Defense Framework Against Jailbreak Attacks in LLMs
por: Xiang, Shiyu, et al.
Publicado: (2025)
por: Xiang, Shiyu, et al.
Publicado: (2025)
Turn Your Face Into An Attack Surface: Screen Attack Using Facial Reflections in Video Conferencing
por: Huang, Yong, et al.
Publicado: (2026)
por: Huang, Yong, et al.
Publicado: (2026)
Behavior-Aware and Generalizable Defense Against Black-Box Adversarial Attacks for ML-Based IDS
por: Ennaji, Sabrine, et al.
Publicado: (2025)
por: Ennaji, Sabrine, et al.
Publicado: (2025)
MAS-Shield: A Defense Framework for Secure and Efficient LLM MAS
por: Wang, Kaixiang, et al.
Publicado: (2025)
por: Wang, Kaixiang, et al.
Publicado: (2025)
Securing AI Agents Against Prompt Injection Attacks
por: Ramakrishnan, Badrinath, et al.
Publicado: (2025)
por: Ramakrishnan, Badrinath, et al.
Publicado: (2025)
AutoAdv: Automated Adversarial Prompting for Multi-Turn Jailbreaking of Large Language Models
por: Reddy, Aashray, et al.
Publicado: (2025)
por: Reddy, Aashray, et al.
Publicado: (2025)
AI-Driven Cybersecurity Threats: A Survey of Emerging Risks and Defensive Strategies
por: Erukude, Sai Teja, et al.
Publicado: (2026)
por: Erukude, Sai Teja, et al.
Publicado: (2026)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
por: Zhu, Kaijie, et al.
Publicado: (2025)
por: Zhu, Kaijie, et al.
Publicado: (2025)
Cyber-Physical Security Vulnerabilities Identification and Classification in Smart Manufacturing -- A Defense-in-Depth Driven Framework and Taxonomy
por: Rahman, Md Habibor, et al.
Publicado: (2024)
por: Rahman, Md Habibor, et al.
Publicado: (2024)
Diffusion-Guided Adversarial Perturbation Injection for Generalizable Defense Against Facial Manipulations
por: Li, Yue, et al.
Publicado: (2026)
por: Li, Yue, et al.
Publicado: (2026)
Backdoor Attacks and Defenses in Computer Vision Domain: A Survey
por: Abbasi, Bilal Hussain, et al.
Publicado: (2025)
por: Abbasi, Bilal Hussain, et al.
Publicado: (2025)
SATversary: Adversarial Attacks and Defenses for Satellite Fingerprinting
por: Smailes, Joshua, et al.
Publicado: (2025)
por: Smailes, Joshua, et al.
Publicado: (2025)
Detection and Defense Against Prominent Attacks on Preconditioned LLM-Integrated Virtual Assistants
por: Chan, Chun Fai, et al.
Publicado: (2024)
por: Chan, Chun Fai, et al.
Publicado: (2024)
A Scheduling-Aware Defense Against Prefetching-Based Side-Channel Attacks
por: Schlüter, Till, et al.
Publicado: (2024)
por: Schlüter, Till, et al.
Publicado: (2024)
Ejemplares similares
-
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
por: Patil, KrishnaSaiReddy
Publicado: (2026) -
RAGShield: Detecting Numerical Claim Manipulation in Government RAG Systems
por: Patil, KrishnaSaiReddy
Publicado: (2026) -
Bidirectional Intention Inference Enhances LLMs' Defense Against Multi-Turn Jailbreak Attacks
por: Tong, Haibo, et al.
Publicado: (2025) -
Mirage: Defense against CrossPath Attacks in Software Defined Networks
por: Murtuza, Shariq, et al.
Publicado: (2024) -
Cross-Domain AI for Early Attack Detection and Defense Against Malicious Flows in O-RAN
por: Xavier, Bruno Missi, et al.
Publicado: (2024)