Interpretability-Guided Test-Time Adversarial Defense
Fuente:
arXiv
Saved in:
| Main Authors: | Kulkarni, Akshay, Weng, Tsui-Wei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Continual Adversarial Defense
by: Wang, Qian, et al.
Published: (2023)
by: Wang, Qian, et al.
Published: (2023)
Adversarially Guided Stateful Defense Against Backdoor Attacks in Federated Deep Learning
by: Ali, Hassan, et al.
Published: (2024)
by: Ali, Hassan, et al.
Published: (2024)
Adversarial Defenses via Vector Quantization
by: Dong, Zhiyi, et al.
Published: (2023)
by: Dong, Zhiyi, et al.
Published: (2023)
Ensemble Adversarial Defense via Integration of Multiple Dispersed Low Curvature Models
by: Zhao, Kaikang, et al.
Published: (2024)
by: Zhao, Kaikang, et al.
Published: (2024)
Adversarial Sparse Teacher: Defense Against Distillation-Based Model Stealing Attacks Using Adversarial Examples
by: Yilmaz, Eda, et al.
Published: (2024)
by: Yilmaz, Eda, et al.
Published: (2024)
Meta Invariance Defense Towards Generalizable Robustness to Unknown Adversarial Attacks
by: Zhang, Lei, et al.
Published: (2024)
by: Zhang, Lei, et al.
Published: (2024)
Data-free Defense of Black Box Models Against Adversarial Attacks
by: Nayak, Gaurav Kumar, et al.
Published: (2022)
by: Nayak, Gaurav Kumar, et al.
Published: (2022)
Let the Noise Speak: Harnessing Noise for a Unified Defense Against Adversarial and Backdoor Attacks
by: Shahriar, Md Hasan, et al.
Published: (2024)
by: Shahriar, Md Hasan, et al.
Published: (2024)
R-TPT: Improving Adversarial Robustness of Vision-Language Models through Test-Time Prompt Tuning
by: Sheng, Lijun, et al.
Published: (2025)
by: Sheng, Lijun, et al.
Published: (2025)
Soften to Defend: Towards Adversarial Robustness via Self-Guided Label Refinement
by: Yu, Daiwei, et al.
Published: (2024)
by: Yu, Daiwei, et al.
Published: (2024)
Adversary-Guided Motion Retargeting for Skeleton Anonymization
by: Carr, Thomas, et al.
Published: (2024)
by: Carr, Thomas, et al.
Published: (2024)
Evaluating Adversarial Robustness: A Comparison Of FGSM, Carlini-Wagner Attacks, And The Role of Distillation as Defense Mechanism
by: Sarkar, Trilokesh Ranjan, et al.
Published: (2024)
by: Sarkar, Trilokesh Ranjan, et al.
Published: (2024)
DIFFender: Diffusion-Based Adversarial Defense against Patch Attacks
by: Kang, Caixin, et al.
Published: (2023)
by: Kang, Caixin, et al.
Published: (2023)
Real-world Adversarial Defense against Patch Attacks based on Diffusion Model
by: Wei, Xingxing, et al.
Published: (2024)
by: Wei, Xingxing, et al.
Published: (2024)
MedBN: Robust Test-Time Adaptation against Malicious Test Samples
by: Park, Hyejin, et al.
Published: (2024)
by: Park, Hyejin, et al.
Published: (2024)
Adversarial Training against Location-Optimized Adversarial Patches
by: Rao, Sukrut, et al.
Published: (2020)
by: Rao, Sukrut, et al.
Published: (2020)
AICAttack: Adversarial Image Captioning Attack with Attention-Based Optimization
by: Li, Jiyao, et al.
Published: (2024)
by: Li, Jiyao, et al.
Published: (2024)
UnMarker: A Universal Attack on Defensive Image Watermarking
by: Kassis, Andre, et al.
Published: (2024)
by: Kassis, Andre, et al.
Published: (2024)
Jailbreaking the Non-Transferable Barrier via Test-Time Data Disguising
by: Xiang, Yongli, et al.
Published: (2025)
by: Xiang, Yongli, et al.
Published: (2025)
Adversarial Pruning: A Survey and Benchmark of Pruning Methods for Adversarial Robustness
by: Piras, Giorgio, et al.
Published: (2024)
by: Piras, Giorgio, et al.
Published: (2024)
Transcending Adversarial Perturbations: Manifold-Aided Adversarial Examples with Legitimate Semantics
by: Li, Shuai, et al.
Published: (2024)
by: Li, Shuai, et al.
Published: (2024)
Beyond Top Activations: Efficient and Reliable Crowdsourced Evaluation of Automated Interpretability
by: Oikarinen, Tuomas, et al.
Published: (2025)
by: Oikarinen, Tuomas, et al.
Published: (2025)
Interpretation of Neural Networks is Susceptible to Universal Adversarial Perturbations
by: Oskouie, Haniyeh Ehsani, et al.
Published: (2022)
by: Oskouie, Haniyeh Ehsani, et al.
Published: (2022)
AR-GAN: Generative Adversarial Network-Based Defense Method Against Adversarial Attacks on the Traffic Sign Classification System of Autonomous Vehicles
by: Salek, M Sabbir, et al.
Published: (2023)
by: Salek, M Sabbir, et al.
Published: (2023)
Exploring Adversarial Attacks against Latent Diffusion Model from the Perspective of Adversarial Transferability
by: Chen, Junxi, et al.
Published: (2024)
by: Chen, Junxi, et al.
Published: (2024)
AED-PADA:Improving Generalizability of Adversarial Example Detection via Principal Adversarial Domain Adaptation
by: Peng, Heqi, et al.
Published: (2024)
by: Peng, Heqi, et al.
Published: (2024)
Random Erasing vs. Model Inversion: A Promising Defense or a False Hope?
by: Tran, Viet-Hung, et al.
Published: (2024)
by: Tran, Viet-Hung, et al.
Published: (2024)
Revealing Vulnerabilities of Neural Networks in Parameter Learning and Defense Against Explanation-Aware Backdoors
by: Kadir, Md Abdul, et al.
Published: (2024)
by: Kadir, Md Abdul, et al.
Published: (2024)
Noise-Aware and Dynamically Adaptive Federated Defense Framework for SAR Image Target Recognition
by: Hou, Yuchao, et al.
Published: (2025)
by: Hou, Yuchao, et al.
Published: (2025)
AGNES: Abstraction-guided Framework for Deep Neural Networks Security
by: Dhonthi, Akshay, et al.
Published: (2023)
by: Dhonthi, Akshay, et al.
Published: (2023)
A Multi-Scale Isolation Forest Approach for Real-Time Detection and Filtering of FGSM Adversarial Attacks in Video Streams of Autonomous Vehicles
by: Abhulimhen, Richard, et al.
Published: (2025)
by: Abhulimhen, Richard, et al.
Published: (2025)
Evaluating the Evaluators: Trust in Adversarial Robustness Tests
by: Cinà, Antonio Emanuele, et al.
Published: (2025)
by: Cinà, Antonio Emanuele, et al.
Published: (2025)
PuriDefense: Randomized Local Implicit Adversarial Purification for Defending Black-box Query-based Attacks
by: Guo, Ping, et al.
Published: (2024)
by: Guo, Ping, et al.
Published: (2024)
Proactive Adversarial Defense: Harnessing Prompt Tuning in Vision-Language Models to Detect Unseen Backdoored Images
by: Stein, Kyle, et al.
Published: (2024)
by: Stein, Kyle, et al.
Published: (2024)
Transferable Adversarial Examples with Bayes Approach
by: Fan, Mingyuan, et al.
Published: (2022)
by: Fan, Mingyuan, et al.
Published: (2022)
Benchmarking Adversarial Patch Selection and Location
by: Kimhi, Shai, et al.
Published: (2025)
by: Kimhi, Shai, et al.
Published: (2025)
Why Does Little Robustness Help? A Further Step Towards Understanding Adversarial Transferability
by: Zhang, Yechao, et al.
Published: (2023)
by: Zhang, Yechao, et al.
Published: (2023)
Explainable Adversarial Attacks on Coarse-to-Fine Classifiers
by: Heidarizadeh, Akram, et al.
Published: (2025)
by: Heidarizadeh, Akram, et al.
Published: (2025)
Differentially Private Synthetic Data via Foundation Model APIs 1: Images
by: Lin, Zinan, et al.
Published: (2023)
by: Lin, Zinan, et al.
Published: (2023)
Transferable Adversarial Attacks on SAM and Its Downstream Models
by: Xia, Song, et al.
Published: (2024)
by: Xia, Song, et al.
Published: (2024)
Similar Items
-
Continual Adversarial Defense
by: Wang, Qian, et al.
Published: (2023) -
Adversarially Guided Stateful Defense Against Backdoor Attacks in Federated Deep Learning
by: Ali, Hassan, et al.
Published: (2024) -
Adversarial Defenses via Vector Quantization
by: Dong, Zhiyi, et al.
Published: (2023) -
Ensemble Adversarial Defense via Integration of Multiple Dispersed Low Curvature Models
by: Zhao, Kaikang, et al.
Published: (2024) -
Adversarial Sparse Teacher: Defense Against Distillation-Based Model Stealing Attacks Using Adversarial Examples
by: Yilmaz, Eda, et al.
Published: (2024)