Saved in:
| Main Author: | Zhang, Ivan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2508.07139 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts
by: Hasan, Md. Mehedi, et al.
Published: (2025)
by: Hasan, Md. Mehedi, et al.
Published: (2025)
Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
by: Lin, Lixing, et al.
Published: (2026)
by: Lin, Lixing, et al.
Published: (2026)
Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models
by: Ma, Jiachen, et al.
Published: (2024)
by: Ma, Jiachen, et al.
Published: (2024)
SEASONED: Semantic-Enhanced Self-Counterfactual Explainable Detection of Adversarial Exploiter Contracts
by: Ai, Xng, et al.
Published: (2025)
by: Ai, Xng, et al.
Published: (2025)
Adversarial-Resilient RF Fingerprinting: A CNN-GAN Framework for Rogue Transmitter Detection
by: Dhakal, Raju, et al.
Published: (2025)
by: Dhakal, Raju, et al.
Published: (2025)
Democratizing ML for Enterprise Security: A Self-Sustained Attack Detection Framework
by: Momeni, Sadegh, et al.
Published: (2025)
by: Momeni, Sadegh, et al.
Published: (2025)
RAG-targeted Adversarial Attack on LLM-based Threat Detection and Mitigation Framework
by: Ikbarieh, Seif, et al.
Published: (2025)
by: Ikbarieh, Seif, et al.
Published: (2025)
In-Browser LLM-Guided Fuzzing for Real-Time Prompt Injection Testing in Agentic AI Browsers
by: Cohen, Avihay
Published: (2025)
by: Cohen, Avihay
Published: (2025)
Fight Back Against Jailbreaking via Prompt Adversarial Tuning
by: Mo, Yichuan, et al.
Published: (2024)
by: Mo, Yichuan, et al.
Published: (2024)
Jailbreaking GPT-4V via Self-Adversarial Attacks with System Prompts
by: Wu, Yuanwei, et al.
Published: (2023)
by: Wu, Yuanwei, et al.
Published: (2023)
MALCDF: A Distributed Multi-Agent LLM Framework for Real-Time Cyber
by: Bhardwaj, Arth, et al.
Published: (2025)
by: Bhardwaj, Arth, et al.
Published: (2025)
Masked Language Model Based Textual Adversarial Example Detection
by: Zhang, Xiaomei, et al.
Published: (2023)
by: Zhang, Xiaomei, et al.
Published: (2023)
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
by: Cao, Tri, et al.
Published: (2026)
by: Cao, Tri, et al.
Published: (2026)
Scam Shield: Multi-Model Voting and Fine-Tuned LLMs Against Adversarial Attacks
by: Chang, Chen-Wei, et al.
Published: (2025)
by: Chang, Chen-Wei, et al.
Published: (2025)
Involuntary Jailbreak: On Self-Prompting Attacks
by: Guo, Yangyang, et al.
Published: (2025)
by: Guo, Yangyang, et al.
Published: (2025)
Advanced Real-Time Fraud Detection Using RAG-Based LLMs
by: Singh, Gurjot, et al.
Published: (2025)
by: Singh, Gurjot, et al.
Published: (2025)
Self-interpreting Adversarial Images
by: Zhang, Tingwei, et al.
Published: (2024)
by: Zhang, Tingwei, et al.
Published: (2024)
Stealthy Dual-Trigger Backdoors: Attacking Prompt Tuning in LM-Empowered Graph Foundation Models
by: Xue, Xiaoyu, et al.
Published: (2025)
by: Xue, Xiaoyu, et al.
Published: (2025)
FLAME: Flexible LLM-Assisted Moderation Engine
by: Bakulin, Ivan, et al.
Published: (2025)
by: Bakulin, Ivan, et al.
Published: (2025)
LumiMAS: A Comprehensive Framework for Real-Time Monitoring and Enhanced Observability in Multi-Agent Systems
by: Solomon, Ron, et al.
Published: (2025)
by: Solomon, Ron, et al.
Published: (2025)
Security Document Classification with a Fine-Tuned Local Large Language Model: Benchmark Data and an Open-Source System
by: Dobrovolskyi, Ivan
Published: (2026)
by: Dobrovolskyi, Ivan
Published: (2026)
I Don't Know You, But I Can Catch You: Real-Time Defense against Diverse Adversarial Patches for Object Detectors
by: Lin, Zijin, et al.
Published: (2024)
by: Lin, Zijin, et al.
Published: (2024)
Detecting Adversarial Fine-tuning with Auditing Agents
by: Egler, Sarah, et al.
Published: (2025)
by: Egler, Sarah, et al.
Published: (2025)
LogSHIELD: A Graph-based Real-time Anomaly Detection Framework using Frequency Analysis
by: Roy, Krishna Chandra, et al.
Published: (2024)
by: Roy, Krishna Chandra, et al.
Published: (2024)
A Novel and Practical Universal Adversarial Perturbations against Deep Reinforcement Learning based Intrusion Detection Systems
by: Zhang, H., et al.
Published: (2025)
by: Zhang, H., et al.
Published: (2025)
AI-Powered Anomaly Detection with Blockchain for Real-Time Security and Reliability in Autonomous Vehicles
by: Shit, Rathin Chandra, et al.
Published: (2025)
by: Shit, Rathin Chandra, et al.
Published: (2025)
ICON: Indirect Prompt Injection Defense for Agents based on Inference-Time Correction
by: Wang, Che, et al.
Published: (2026)
by: Wang, Che, et al.
Published: (2026)
Code Vulnerability Repair with Large Language Model using Context-Aware Prompt Tuning
by: Khan, Arshiya, et al.
Published: (2024)
by: Khan, Arshiya, et al.
Published: (2024)
Unified Threat Detection and Mitigation Framework (UTDMF): Combating Prompt Injection, Deception, and Bias in Enterprise-Scale Transformers
by: KumarRavindran, Santhosh
Published: (2025)
by: KumarRavindran, Santhosh
Published: (2025)
Adversarial Attacks against Windows PE Malware Detection: A Survey of the State-of-the-Art
by: Ling, Xiang, et al.
Published: (2021)
by: Ling, Xiang, et al.
Published: (2021)
Integrated Simulation Framework for Adversarial Attacks on Autonomous Vehicles
by: Anagnostopoulos, Christos, et al.
Published: (2025)
by: Anagnostopoulos, Christos, et al.
Published: (2025)
FAST-IDS: A Fast Two-Stage Intrusion Detection System with Hybrid Compression for Real-Time Threat Detection in Connected and Autonomous Vehicles
by: S, Devika, et al.
Published: (2025)
by: S, Devika, et al.
Published: (2025)
Doppelganger Method: Breaking Role Consistency in LLM Agent via Prompt-based Transferable Adversarial Attack
by: Kang, Daewon, et al.
Published: (2025)
by: Kang, Daewon, et al.
Published: (2025)
MULTI-LF: A Continuous Learning Framework for Real-Time Malicious Traffic Detection in Multi-Environment Networks
by: Rustam, Furqan, et al.
Published: (2025)
by: Rustam, Furqan, et al.
Published: (2025)
DataSentinel: A Game-Theoretic Detection of Prompt Injection Attacks
by: Liu, Yupei, et al.
Published: (2025)
by: Liu, Yupei, et al.
Published: (2025)
Adversarial Defense in Cybersecurity: A Systematic Review of GANs for Threat Detection and Mitigation
by: Ndayipfukamiye, Tharcisse, et al.
Published: (2025)
by: Ndayipfukamiye, Tharcisse, et al.
Published: (2025)
Adversarial Evasion in Non-Stationary Malware Detection: Minimizing Drift Signals through Similarity-Constrained Perturbations
by: Acharya, Pawan, et al.
Published: (2026)
by: Acharya, Pawan, et al.
Published: (2026)
How stealthy is stealthy? Studying the Efficacy of Black-Box Adversarial Attacks in the Real World
by: Panebianco, Francesco, et al.
Published: (2025)
by: Panebianco, Francesco, et al.
Published: (2025)
FTSmartAudit: A Knowledge Distillation-Enhanced Framework for Automated Smart Contract Auditing Using Fine-Tuned LLMs
by: Wei, Zhiyuan, et al.
Published: (2024)
by: Wei, Zhiyuan, et al.
Published: (2024)
ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection
by: Zhao, Wei, et al.
Published: (2026)
by: Zhao, Wei, et al.
Published: (2026)
Similar Items
-
Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts
by: Hasan, Md. Mehedi, et al.
Published: (2025) -
Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
by: Lin, Lixing, et al.
Published: (2026) -
Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models
by: Ma, Jiachen, et al.
Published: (2024) -
SEASONED: Semantic-Enhanced Self-Counterfactual Explainable Detection of Adversarial Exploiter Contracts
by: Ai, Xng, et al.
Published: (2025) -
Adversarial-Resilient RF Fingerprinting: A CNN-GAN Framework for Rogue Transmitter Detection
by: Dhakal, Raju, et al.
Published: (2025)