Proactive Adversarial Defense: Harnessing Prompt Tuning in Vision-Language Models to Detect Unseen Backdoored Images
Fuente:
arXiv
Saved in:
| Main Authors: | Stein, Kyle, Mahyari, Andrew Arash, Francia, Guillermo, El-Sheikh, Eman |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Packet Inspection Transformer: A Self-Supervised Journey to Unseen Malware Detection with Few Samples
by: Stein, Kyle, et al.
Published: (2024)
by: Stein, Kyle, et al.
Published: (2024)
A Transformer-Based Framework for Payload Malware Detection and Classification
by: Stein, Kyle, et al.
Published: (2024)
by: Stein, Kyle, et al.
Published: (2024)
Towards Novel Malicious Packet Recognition: A Few-Shot Learning Approach
by: Stein, Kyle, et al.
Published: (2024)
by: Stein, Kyle, et al.
Published: (2024)
Proactive Disentangled Modeling of Trigger-Object Pairings for Backdoor Defense
by: Stein, Kyle, et al.
Published: (2025)
by: Stein, Kyle, et al.
Published: (2025)
Visual Adaptive Prompting for Compositional Zero-Shot Learning
by: Stein, Kyle, et al.
Published: (2025)
by: Stein, Kyle, et al.
Published: (2025)
Adaptive Additive Parameter Updates of Vision Transformers for Few-Shot Continual Learning
by: Stein, Kyle, et al.
Published: (2025)
by: Stein, Kyle, et al.
Published: (2025)
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
by: Wei, Shaokui, et al.
Published: (2024)
by: Wei, Shaokui, et al.
Published: (2024)
Harnessing the Power of LLMs in Source Code Vulnerability Detection
by: Mahyari, Andrew A
Published: (2024)
by: Mahyari, Andrew A
Published: (2024)
Backdoor Mitigation in Object Detection via Adversarial Fine-Tuning
by: Dunnett, Kealan, et al.
Published: (2026)
by: Dunnett, Kealan, et al.
Published: (2026)
Let the Noise Speak: Harnessing Noise for a Unified Defense Against Adversarial and Backdoor Attacks
by: Shahriar, Md Hasan, et al.
Published: (2024)
by: Shahriar, Md Hasan, et al.
Published: (2024)
Transductive One-Shot Learning Meet Subspace Decomposition
by: Stein, Kyle, et al.
Published: (2025)
by: Stein, Kyle, et al.
Published: (2025)
TrapSuffix: Proactive Defense Against Adversarial Suffixes in Jailbreaking
by: Du, Mengyao, et al.
Published: (2026)
by: Du, Mengyao, et al.
Published: (2026)
BackdoorIndicator: Leveraging OOD Data for Proactive Backdoor Detection in Federated Learning
by: Li, Songze, et al.
Published: (2024)
by: Li, Songze, et al.
Published: (2024)
Transforming Cyber Defense: Harnessing Agentic and Frontier AI for Proactive, Ethical Threat Intelligence
by: Tallam, Krti
Published: (2025)
by: Tallam, Krti
Published: (2025)
Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt
by: Ying, Zonghao, et al.
Published: (2024)
by: Ying, Zonghao, et al.
Published: (2024)
Backdoor-Powered Prompt Injection Attacks Nullify Defense Methods
by: Chen, Yulin, et al.
Published: (2025)
by: Chen, Yulin, et al.
Published: (2025)
Harnessing the Power of Large Vision Language Models for Synthetic Image Detection
by: Keita, Mamadou, et al.
Published: (2024)
by: Keita, Mamadou, et al.
Published: (2024)
Adversarially Guided Stateful Defense Against Backdoor Attacks in Federated Deep Learning
by: Ali, Hassan, et al.
Published: (2024)
by: Ali, Hassan, et al.
Published: (2024)
Learning to Detect Unseen Jailbreak Attacks in Large Vision-Language Models
by: Liang, Shuang, et al.
Published: (2025)
by: Liang, Shuang, et al.
Published: (2025)
AttackVLA: Benchmarking Adversarial and Backdoor Attacks on Vision-Language-Action Models
by: Li, Jiayu, et al.
Published: (2025)
by: Li, Jiayu, et al.
Published: (2025)
R-TPT: Improving Adversarial Robustness of Vision-Language Models through Test-Time Prompt Tuning
by: Sheng, Lijun, et al.
Published: (2025)
by: Sheng, Lijun, et al.
Published: (2025)
Backdoor Attacks and Defenses in Computer Vision Domain: A Survey
by: Abbasi, Bilal Hussain, et al.
Published: (2025)
by: Abbasi, Bilal Hussain, et al.
Published: (2025)
Double Backdoored: Converting Code Large Language Model Backdoors to Traditional Malware via Adversarial Instruction Tuning Attacks
by: Hossen, Md Imran, et al.
Published: (2024)
by: Hossen, Md Imran, et al.
Published: (2024)
PAD-FT: A Lightweight Defense for Backdoor Attacks via Data Purification and Fine-Tuning
by: Xu, Yukai, et al.
Published: (2024)
by: Xu, Yukai, et al.
Published: (2024)
Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation
by: Zhong, Zhiyuan, et al.
Published: (2025)
by: Zhong, Zhiyuan, et al.
Published: (2025)
BeDKD: Backdoor Defense Based on Directional Mapping Module and Adversarial Knowledge Distillation
by: Wu, Zhengxian, et al.
Published: (2025)
by: Wu, Zhengxian, et al.
Published: (2025)
TrojanDam: Detection-Free Backdoor Defense in Federated Learning through Proactive Model Robustification utilizing OOD Data
by: Dai, Yanbo, et al.
Published: (2025)
by: Dai, Yanbo, et al.
Published: (2025)
Backdoor Directions in Vision Transformers
by: Karayalcin, Sengim, et al.
Published: (2026)
by: Karayalcin, Sengim, et al.
Published: (2026)
ODDR: Outlier Detection & Dimension Reduction Based Defense Against Adversarial Patches
by: Chattopadhyay, Nandish, et al.
Published: (2023)
by: Chattopadhyay, Nandish, et al.
Published: (2023)
Robust Image Classification: Defensive Strategies against FGSM and PGD Adversarial Attacks
by: Waghela, Hetvi, et al.
Published: (2024)
by: Waghela, Hetvi, et al.
Published: (2024)
Proactive Hardening of LLM Defenses with HASTE
by: Chen, Henry, et al.
Published: (2026)
by: Chen, Henry, et al.
Published: (2026)
Struggle with Adversarial Defense? Try Diffusion
by: Li, Yujie, et al.
Published: (2024)
by: Li, Yujie, et al.
Published: (2024)
Test-Time Attention Purification for Backdoored Large Vision Language Models
by: Zhang, Zhifang, et al.
Published: (2026)
by: Zhang, Zhifang, et al.
Published: (2026)
Robust Anti-Backdoor Instruction Tuning in LVLMs
by: Xun, Yuan, et al.
Published: (2025)
by: Xun, Yuan, et al.
Published: (2025)
Nearest is Not Dearest: Towards Practical Defense against Quantization-conditioned Backdoor Attacks
by: Li, Boheng, et al.
Published: (2024)
by: Li, Boheng, et al.
Published: (2024)
Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
Architectural Backdoors in Deep Learning: A Survey of Vulnerabilities, Detection, and Defense
by: Childress, Victoria, et al.
Published: (2025)
by: Childress, Victoria, et al.
Published: (2025)
Attack as Defense: Run-time Backdoor Implantation for Image Content Protection
by: Zhang, Haichuan, et al.
Published: (2024)
by: Zhang, Haichuan, et al.
Published: (2024)
Hidden Tail: Adversarial Image Causing Stealthy Resource Consumption in Vision-Language Models
by: Zhang, Rui, et al.
Published: (2025)
by: Zhang, Rui, et al.
Published: (2025)
Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization
by: Maljkovic, Igor, et al.
Published: (2026)
by: Maljkovic, Igor, et al.
Published: (2026)
Similar Items
-
Packet Inspection Transformer: A Self-Supervised Journey to Unseen Malware Detection with Few Samples
by: Stein, Kyle, et al.
Published: (2024) -
A Transformer-Based Framework for Payload Malware Detection and Classification
by: Stein, Kyle, et al.
Published: (2024) -
Towards Novel Malicious Packet Recognition: A Few-Shot Learning Approach
by: Stein, Kyle, et al.
Published: (2024) -
Proactive Disentangled Modeling of Trigger-Object Pairings for Backdoor Defense
by: Stein, Kyle, et al.
Published: (2025) -
Visual Adaptive Prompting for Compositional Zero-Shot Learning
by: Stein, Kyle, et al.
Published: (2025)