Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hasan, Md. Mehedi, Mehedi, Sk Tanzir, Rahman, Ziaur, Mostafiz, Rafid, Hossain, Md. Abir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CLIN-LLM: A Safety-Constrained Hybrid Framework for Clinical Diagnosis and Treatment Generation
von: Hasan, Md. Mehedi, et al.
Veröffentlicht: (2025)
von: Hasan, Md. Mehedi, et al.
Veröffentlicht: (2025)
ReDAG-RT: Global Rate-Priority Scheduling for Real-Time Multi-DAG Execution in ROS 2
von: Hasan, Md. Mehedi, et al.
Veröffentlicht: (2026)
von: Hasan, Md. Mehedi, et al.
Veröffentlicht: (2026)
FAARM: Firmware Attestation and Authentication Framework for Mali GPUs
von: Hasan, Md. Mehedi
Veröffentlicht: (2025)
von: Hasan, Md. Mehedi
Veröffentlicht: (2025)
DSTAN-Med: Dual-Channel Spatiotemporal Attention with Physiological Plausibility Filtering for False Data Injection Attack Detection in IoT-Based Medical Devices
von: Hasan, Md Mehedi, et al.
Veröffentlicht: (2026)
von: Hasan, Md Mehedi, et al.
Veröffentlicht: (2026)
Explainable but Vulnerable: Adversarial Attacks on XAI Explanation in Cybersecurity Applications
von: Mia, Maraz, et al.
Veröffentlicht: (2025)
von: Mia, Maraz, et al.
Veröffentlicht: (2025)
AbuseGPT: Abuse of Generative AI ChatBots to Create Smishing Campaigns
von: Shibli, Ashfak Md, et al.
Veröffentlicht: (2024)
von: Shibli, Ashfak Md, et al.
Veröffentlicht: (2024)
QUT-DV25: A Dataset for Dynamic Analysis of Next-Gen Software Supply Chain Attacks
von: Mehedi, Sk Tanzir, et al.
Veröffentlicht: (2025)
von: Mehedi, Sk Tanzir, et al.
Veröffentlicht: (2025)
DySec: A Machine Learning-based Dynamic Analysis for Detecting Malicious Packages in PyPI Ecosystem
von: Mehedi, Sk Tanzir, et al.
Veröffentlicht: (2025)
von: Mehedi, Sk Tanzir, et al.
Veröffentlicht: (2025)
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
von: Cao, Tri, et al.
Veröffentlicht: (2026)
von: Cao, Tri, et al.
Veröffentlicht: (2026)
DistillGuard: Evaluating Defenses Against LLM Knowledge Distillation
von: Jiang, Bo
Veröffentlicht: (2026)
von: Jiang, Bo
Veröffentlicht: (2026)
Bypassing Prompt Injection Detectors through Evasive Injections
von: Rahman, Md Jahedur, et al.
Veröffentlicht: (2026)
von: Rahman, Md Jahedur, et al.
Veröffentlicht: (2026)
Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
von: Lin, Lixing, et al.
Veröffentlicht: (2026)
von: Lin, Lixing, et al.
Veröffentlicht: (2026)
The Defense Trilemma: Why Prompt Injection Defense Wrappers Fail?
von: Bhatt, Manish, et al.
Veröffentlicht: (2026)
von: Bhatt, Manish, et al.
Veröffentlicht: (2026)
X-Guard: Multilingual Guard Agent for Content Moderation
von: Upadhayay, Bibek, et al.
Veröffentlicht: (2025)
von: Upadhayay, Bibek, et al.
Veröffentlicht: (2025)
ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection
von: Zhao, Wei, et al.
Veröffentlicht: (2026)
von: Zhao, Wei, et al.
Veröffentlicht: (2026)
Defenses Against Prompt Attacks Learn Surface Heuristics
von: Li, Shawn, et al.
Veröffentlicht: (2026)
von: Li, Shawn, et al.
Veröffentlicht: (2026)
Hacking Back the AI-Hacker: Prompt Injection as a Defense Against LLM-driven Cyberattacks
von: Pasquini, Dario, et al.
Veröffentlicht: (2024)
von: Pasquini, Dario, et al.
Veröffentlicht: (2024)
ExplainableGuard: Interpretable Adversarial Defense for Large Language Models Using Chain-of-Thought Reasoning
von: Guan, Shaowei, et al.
Veröffentlicht: (2025)
von: Guan, Shaowei, et al.
Veröffentlicht: (2025)
Beyond the Benchmark: Innovative Defenses Against Prompt Injection Attacks
von: Shaheer, Safwan, et al.
Veröffentlicht: (2025)
von: Shaheer, Safwan, et al.
Veröffentlicht: (2025)
GUARD-SLM: Token Activation-Based Defense Against Jailbreak Attacks for Small Language Models
von: Mia, Md Jueal, et al.
Veröffentlicht: (2026)
von: Mia, Md Jueal, et al.
Veröffentlicht: (2026)
A Survey on Agentic Security: Applications, Threats and Defenses
von: Shahriar, Asif, et al.
Veröffentlicht: (2025)
von: Shahriar, Asif, et al.
Veröffentlicht: (2025)
Fingerprinting Deep Learning Models via Network Traffic Patterns in Federated Learning
von: Shuvo, Md Nahid Hasan, et al.
Veröffentlicht: (2025)
von: Shuvo, Md Nahid Hasan, et al.
Veröffentlicht: (2025)
Detecting Prompt Injection Attacks Against Application Using Classifiers
von: Shaheer, Safwan, et al.
Veröffentlicht: (2025)
von: Shaheer, Safwan, et al.
Veröffentlicht: (2025)
A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
von: Wei, Qianshan, et al.
Veröffentlicht: (2025)
von: Wei, Qianshan, et al.
Veröffentlicht: (2025)
Privacy Guard & Token Parsimony by Prompt and Context Handling and LLM Routing
von: Langiu, Alessio
Veröffentlicht: (2026)
von: Langiu, Alessio
Veröffentlicht: (2026)
eDySec: A Deep Learning-based Explainable Dynamic Analysis Framework for Detecting Malicious Packages in PyPI Ecosystem
von: Mehedi, Sk Tanzir, et al.
Veröffentlicht: (2026)
von: Mehedi, Sk Tanzir, et al.
Veröffentlicht: (2026)
CANGuard: A Spatio-Temporal CNN-GRU-Attention Hybrid Architecture for Intrusion Detection in In-Vehicle CAN Networks
von: Sajib, Rakib Hossain, et al.
Veröffentlicht: (2026)
von: Sajib, Rakib Hossain, et al.
Veröffentlicht: (2026)
MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents
von: Zhu, Kaijie, et al.
Veröffentlicht: (2025)
von: Zhu, Kaijie, et al.
Veröffentlicht: (2025)
A Real-Time, Self-Tuning Moderator Framework for Adversarial Prompt Detection
von: Zhang, Ivan
Veröffentlicht: (2025)
von: Zhang, Ivan
Veröffentlicht: (2025)
SDNGuardStack: An Explainable Ensemble Learning Framework for High-Accuracy Intrusion Detection in Software-Defined Networks
von: Ashikuzzaman, et al.
Veröffentlicht: (2026)
von: Ashikuzzaman, et al.
Veröffentlicht: (2026)
Agent Privilege Separation in OpenClaw: A Structural Defense Against Prompt Injection
von: Cheng, Darren, et al.
Veröffentlicht: (2026)
von: Cheng, Darren, et al.
Veröffentlicht: (2026)
An Optimized Decision Tree-Based Framework for Explainable IoT Anomaly Detection
von: Ashikuzzaman, et al.
Veröffentlicht: (2026)
von: Ashikuzzaman, et al.
Veröffentlicht: (2026)
Secure Energy Transactions Using Blockchain Leveraging AI for Fraud Detection and Energy Market Stability
von: Khan, Md Asif Ul Hoq, et al.
Veröffentlicht: (2025)
von: Khan, Md Asif Ul Hoq, et al.
Veröffentlicht: (2025)
Behavior-Aware and Generalizable Defense Against Black-Box Adversarial Attacks for ML-Based IDS
von: Ennaji, Sabrine, et al.
Veröffentlicht: (2025)
von: Ennaji, Sabrine, et al.
Veröffentlicht: (2025)
To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt
von: Wang, Zhilong, et al.
Veröffentlicht: (2025)
von: Wang, Zhilong, et al.
Veröffentlicht: (2025)
Attacks and Defenses Against LLM Fingerprinting
von: Kurian, Kevin, et al.
Veröffentlicht: (2025)
von: Kurian, Kevin, et al.
Veröffentlicht: (2025)
Toward Autonomous SOC Operations: End-to-End LLM Framework for Threat Detection, Query Generation, and Resolution in Security Operations
von: Saju, Md Hasan, et al.
Veröffentlicht: (2026)
von: Saju, Md Hasan, et al.
Veröffentlicht: (2026)
Real-time ML-based Defense Against Malicious Payload in Reconfigurable Embedded Systems
von: Stahle-Smith, Rye, et al.
Veröffentlicht: (2025)
von: Stahle-Smith, Rye, et al.
Veröffentlicht: (2025)
An advanced data fabric architecture leveraging homomorphic encryption and federated learning
von: Rieyan, Sakib Anwar, et al.
Veröffentlicht: (2024)
von: Rieyan, Sakib Anwar, et al.
Veröffentlicht: (2024)
Safety-Oriented Routing Analysis of Mixtral MoE Under Benign and Harmful Prompts
von: Siddiky, Md Nurul Absar
Veröffentlicht: (2026)
von: Siddiky, Md Nurul Absar
Veröffentlicht: (2026)
Ähnliche Einträge
-
CLIN-LLM: A Safety-Constrained Hybrid Framework for Clinical Diagnosis and Treatment Generation
von: Hasan, Md. Mehedi, et al.
Veröffentlicht: (2025) -
ReDAG-RT: Global Rate-Priority Scheduling for Real-Time Multi-DAG Execution in ROS 2
von: Hasan, Md. Mehedi, et al.
Veröffentlicht: (2026) -
FAARM: Firmware Attestation and Authentication Framework for Mali GPUs
von: Hasan, Md. Mehedi
Veröffentlicht: (2025) -
DSTAN-Med: Dual-Channel Spatiotemporal Attention with Physiological Plausibility Filtering for False Data Injection Attack Detection in IoT-Based Medical Devices
von: Hasan, Md Mehedi, et al.
Veröffentlicht: (2026) -
Explainable but Vulnerable: Adversarial Attacks on XAI Explanation in Cybersecurity Applications
von: Mia, Maraz, et al.
Veröffentlicht: (2025)