Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts
Fuente:
arXiv
Guardado en:
| Autores principales: | Hasan, Md. Mehedi, Mehedi, Sk Tanzir, Rahman, Ziaur, Mostafiz, Rafid, Hossain, Md. Abir |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CLIN-LLM: A Safety-Constrained Hybrid Framework for Clinical Diagnosis and Treatment Generation
por: Hasan, Md. Mehedi, et al.
Publicado: (2025)
por: Hasan, Md. Mehedi, et al.
Publicado: (2025)
ReDAG-RT: Global Rate-Priority Scheduling for Real-Time Multi-DAG Execution in ROS 2
por: Hasan, Md. Mehedi, et al.
Publicado: (2026)
por: Hasan, Md. Mehedi, et al.
Publicado: (2026)
FAARM: Firmware Attestation and Authentication Framework for Mali GPUs
por: Hasan, Md. Mehedi
Publicado: (2025)
por: Hasan, Md. Mehedi
Publicado: (2025)
DSTAN-Med: Dual-Channel Spatiotemporal Attention with Physiological Plausibility Filtering for False Data Injection Attack Detection in IoT-Based Medical Devices
por: Hasan, Md Mehedi, et al.
Publicado: (2026)
por: Hasan, Md Mehedi, et al.
Publicado: (2026)
Explainable but Vulnerable: Adversarial Attacks on XAI Explanation in Cybersecurity Applications
por: Mia, Maraz, et al.
Publicado: (2025)
por: Mia, Maraz, et al.
Publicado: (2025)
AbuseGPT: Abuse of Generative AI ChatBots to Create Smishing Campaigns
por: Shibli, Ashfak Md, et al.
Publicado: (2024)
por: Shibli, Ashfak Md, et al.
Publicado: (2024)
QUT-DV25: A Dataset for Dynamic Analysis of Next-Gen Software Supply Chain Attacks
por: Mehedi, Sk Tanzir, et al.
Publicado: (2025)
por: Mehedi, Sk Tanzir, et al.
Publicado: (2025)
DySec: A Machine Learning-based Dynamic Analysis for Detecting Malicious Packages in PyPI Ecosystem
por: Mehedi, Sk Tanzir, et al.
Publicado: (2025)
por: Mehedi, Sk Tanzir, et al.
Publicado: (2025)
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
por: Cao, Tri, et al.
Publicado: (2026)
por: Cao, Tri, et al.
Publicado: (2026)
DistillGuard: Evaluating Defenses Against LLM Knowledge Distillation
por: Jiang, Bo
Publicado: (2026)
por: Jiang, Bo
Publicado: (2026)
Bypassing Prompt Injection Detectors through Evasive Injections
por: Rahman, Md Jahedur, et al.
Publicado: (2026)
por: Rahman, Md Jahedur, et al.
Publicado: (2026)
Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
por: Lin, Lixing, et al.
Publicado: (2026)
por: Lin, Lixing, et al.
Publicado: (2026)
The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?
por: Bhatt, Manish, et al.
Publicado: (2026)
por: Bhatt, Manish, et al.
Publicado: (2026)
X-Guard: Multilingual Guard Agent for Content Moderation
por: Upadhayay, Bibek, et al.
Publicado: (2025)
por: Upadhayay, Bibek, et al.
Publicado: (2025)
ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection
por: Zhao, Wei, et al.
Publicado: (2026)
por: Zhao, Wei, et al.
Publicado: (2026)
Defenses Against Prompt Attacks Learn Surface Heuristics
por: Li, Shawn, et al.
Publicado: (2026)
por: Li, Shawn, et al.
Publicado: (2026)
Hacking Back the AI-Hacker: Prompt Injection as a Defense Against LLM-driven Cyberattacks
por: Pasquini, Dario, et al.
Publicado: (2024)
por: Pasquini, Dario, et al.
Publicado: (2024)
ExplainableGuard: Interpretable Adversarial Defense for Large Language Models Using Chain-of-Thought Reasoning
por: Guan, Shaowei, et al.
Publicado: (2025)
por: Guan, Shaowei, et al.
Publicado: (2025)
Beyond the Benchmark: Innovative Defenses Against Prompt Injection Attacks
por: Shaheer, Safwan, et al.
Publicado: (2025)
por: Shaheer, Safwan, et al.
Publicado: (2025)
GUARD-SLM: Token Activation-Based Defense Against Jailbreak Attacks for Small Language Models
por: Mia, Md Jueal, et al.
Publicado: (2026)
por: Mia, Md Jueal, et al.
Publicado: (2026)
A Survey on Agentic Security: Applications, Threats and Defenses
por: Shahriar, Asif, et al.
Publicado: (2025)
por: Shahriar, Asif, et al.
Publicado: (2025)
Fingerprinting Deep Learning Models via Network Traffic Patterns in Federated Learning
por: Shuvo, Md Nahid Hasan, et al.
Publicado: (2025)
por: Shuvo, Md Nahid Hasan, et al.
Publicado: (2025)
Detecting Prompt Injection Attacks Against Application Using Classifiers
por: Shaheer, Safwan, et al.
Publicado: (2025)
por: Shaheer, Safwan, et al.
Publicado: (2025)
A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
por: Wei, Qianshan, et al.
Publicado: (2025)
por: Wei, Qianshan, et al.
Publicado: (2025)
Privacy Guard & Token Parsimony by Prompt and Context Handling and LLM Routing
por: Langiu, Alessio
Publicado: (2026)
por: Langiu, Alessio
Publicado: (2026)
eDySec: A Deep Learning-based Explainable Dynamic Analysis Framework for Detecting Malicious Packages in PyPI Ecosystem
por: Mehedi, Sk Tanzir, et al.
Publicado: (2026)
por: Mehedi, Sk Tanzir, et al.
Publicado: (2026)
CANGuard: A Spatio-Temporal CNN-GRU-Attention Hybrid Architecture for Intrusion Detection in In-Vehicle CAN Networks
por: Sajib, Rakib Hossain, et al.
Publicado: (2026)
por: Sajib, Rakib Hossain, et al.
Publicado: (2026)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
por: Zhu, Kaijie, et al.
Publicado: (2025)
por: Zhu, Kaijie, et al.
Publicado: (2025)
A Real-Time, Self-Tuning Moderator Framework for Adversarial Prompt Detection
por: Zhang, Ivan
Publicado: (2025)
por: Zhang, Ivan
Publicado: (2025)
SDNGuardStack: An Explainable Ensemble Learning Framework for High-Accuracy Intrusion Detection in Software-Defined Networks
por: Ashikuzzaman, et al.
Publicado: (2026)
por: Ashikuzzaman, et al.
Publicado: (2026)
Agent Privilege Separation in OpenClaw: A Structural Defense Against Prompt Injection
por: Cheng, Darren, et al.
Publicado: (2026)
por: Cheng, Darren, et al.
Publicado: (2026)
An Optimized Decision Tree-Based Framework for Explainable IoT Anomaly Detection
por: Ashikuzzaman, et al.
Publicado: (2026)
por: Ashikuzzaman, et al.
Publicado: (2026)
Secure Energy Transactions Using Blockchain Leveraging AI for Fraud Detection and Energy Market Stability
por: Khan, Md Asif Ul Hoq, et al.
Publicado: (2025)
por: Khan, Md Asif Ul Hoq, et al.
Publicado: (2025)
Behavior-Aware and Generalizable Defense Against Black-Box Adversarial Attacks for ML-Based IDS
por: Ennaji, Sabrine, et al.
Publicado: (2025)
por: Ennaji, Sabrine, et al.
Publicado: (2025)
To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt
por: Wang, Zhilong, et al.
Publicado: (2025)
por: Wang, Zhilong, et al.
Publicado: (2025)
Attacks and Defenses Against LLM Fingerprinting
por: Kurian, Kevin, et al.
Publicado: (2025)
por: Kurian, Kevin, et al.
Publicado: (2025)
Toward Autonomous SOC Operations: End-to-End LLM Framework for Threat Detection, Query Generation, and Resolution in Security Operations
por: Saju, Md Hasan, et al.
Publicado: (2026)
por: Saju, Md Hasan, et al.
Publicado: (2026)
Real-time ML-based Defense Against Malicious Payload in Reconfigurable Embedded Systems
por: Stahle-Smith, Rye, et al.
Publicado: (2025)
por: Stahle-Smith, Rye, et al.
Publicado: (2025)
An advanced data fabric architecture leveraging homomorphic encryption and federated learning
por: Rieyan, Sakib Anwar, et al.
Publicado: (2024)
por: Rieyan, Sakib Anwar, et al.
Publicado: (2024)
Safety-Oriented Routing Analysis of Mixtral MoE Under Benign and Harmful Prompts
por: Siddiky, Md Nurul Absar
Publicado: (2026)
por: Siddiky, Md Nurul Absar
Publicado: (2026)
Ejemplares similares
-
CLIN-LLM: A Safety-Constrained Hybrid Framework for Clinical Diagnosis and Treatment Generation
por: Hasan, Md. Mehedi, et al.
Publicado: (2025) -
ReDAG-RT: Global Rate-Priority Scheduling for Real-Time Multi-DAG Execution in ROS 2
por: Hasan, Md. Mehedi, et al.
Publicado: (2026) -
FAARM: Firmware Attestation and Authentication Framework for Mali GPUs
por: Hasan, Md. Mehedi
Publicado: (2025) -
DSTAN-Med: Dual-Channel Spatiotemporal Attention with Physiological Plausibility Filtering for False Data Injection Attack Detection in IoT-Based Medical Devices
por: Hasan, Md Mehedi, et al.
Publicado: (2026) -
Explainable but Vulnerable: Adversarial Attacks on XAI Explanation in Cybersecurity Applications
por: Mia, Maraz, et al.
Publicado: (2025)