Blending adversarial training and representation-conditional purification via aggregation improves adversarial robustness
Fuente:
arXiv
Saved in:
| Main Authors: | Ballarin, Emanuele, Ansuini, Alessio, Bortolussi, Luca |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deepfake detectors are DUMB: A benchmark to assess adversarial training robustness under transferability constraints
by: Serrano, Adrian, et al.
Published: (2026)
by: Serrano, Adrian, et al.
Published: (2026)
D2R: dual regularization loss with collaborative adversarial generation for model robustness
by: Liu, Zhenyu, et al.
Published: (2025)
by: Liu, Zhenyu, et al.
Published: (2025)
Trainwreck: A damaging adversarial attack on image classifiers
by: Zahálka, Jan
Published: (2023)
by: Zahálka, Jan
Published: (2023)
Robust width: A lightweight and certifiable adversarial defense
by: Peck, Jonathan, et al.
Published: (2024)
by: Peck, Jonathan, et al.
Published: (2024)
Verification of Neural Networks against Convolutional Perturbations via Parameterised Kernels
by: Brückner, Benedikt, et al.
Published: (2024)
by: Brückner, Benedikt, et al.
Published: (2024)
Evaluating the Evaluators: Trust in Adversarial Robustness Tests
by: Cinà, Antonio Emanuele, et al.
Published: (2025)
by: Cinà, Antonio Emanuele, et al.
Published: (2025)
ADBM: Adversarial diffusion bridge model for reliable adversarial purification
by: Li, Xiao, et al.
Published: (2024)
by: Li, Xiao, et al.
Published: (2024)
Pixel Seal: Adversarial-only training for invisible image and video watermarking
by: Souček, Tomáš, et al.
Published: (2025)
by: Souček, Tomáš, et al.
Published: (2025)
IoUCert: Robustness Verification for Anchor-based Object Detectors
by: Brückner, Benedikt, et al.
Published: (2026)
by: Brückner, Benedikt, et al.
Published: (2026)
Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack
by: Chen, Yukun, et al.
Published: (2025)
by: Chen, Yukun, et al.
Published: (2025)
Better Safe than Sorry: Pre-training CLIP against Targeted Data Poisoning and Backdoor Attacks
by: Yang, Wenhan, et al.
Published: (2023)
by: Yang, Wenhan, et al.
Published: (2023)
Are aligned neural networks adversarially aligned?
by: Carlini, Nicholas, et al.
Published: (2023)
by: Carlini, Nicholas, et al.
Published: (2023)
Evaluating the robustness of adversarial defenses in malware detection systems
by: Jafari, Mostafa, et al.
Published: (2025)
by: Jafari, Mostafa, et al.
Published: (2025)
AutoAdvExBench: Benchmarking autonomous exploitation of adversarial example defenses
by: Carlini, Nicholas, et al.
Published: (2025)
by: Carlini, Nicholas, et al.
Published: (2025)
Personalized Federated Learning via Backbone Self-Distillation
by: Wang, Pengju, et al.
Published: (2024)
by: Wang, Pengju, et al.
Published: (2024)
Harnessing Hyperbolic Geometry for Harmful Prompt Detection and Sanitization
by: Maljkovic, Igor, et al.
Published: (2026)
by: Maljkovic, Igor, et al.
Published: (2026)
REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
by: Chen, Yukun, et al.
Published: (2025)
by: Chen, Yukun, et al.
Published: (2025)
Exploring the Adversarial Frontier: Quantifying Robustness via Adversarial Hypervolume
by: Guo, Ping, et al.
Published: (2024)
by: Guo, Ping, et al.
Published: (2024)
Purify Unlearnable Examples via Rate-Constrained Variational Autoencoders
by: Yu, Yi, et al.
Published: (2024)
by: Yu, Yi, et al.
Published: (2024)
Towards Personalized Federated Learning via Comprehensive Knowledge Distillation
by: Wang, Pengju, et al.
Published: (2024)
by: Wang, Pengju, et al.
Published: (2024)
Explainable AI and susceptibility to adversarial attacks: a case study in classification of breast ultrasound images
by: Rasaee, Hamza, et al.
Published: (2021)
by: Rasaee, Hamza, et al.
Published: (2021)
Improving Robustness to Model Inversion Attacks via Sparse Coding Architectures
by: Dibbo, Sayanton V., et al.
Published: (2024)
by: Dibbo, Sayanton V., et al.
Published: (2024)
Hypnopaedia-Aware Machine Unlearning via Psychometrics of Artificial Mental Imagery
by: Chang, Ching-Chun, et al.
Published: (2024)
by: Chang, Ching-Chun, et al.
Published: (2024)
MOVE: Effective and Harmless Ownership Verification via Embedded External Features
by: Li, Yiming, et al.
Published: (2022)
by: Li, Yiming, et al.
Published: (2022)
Towards Sample-specific Backdoor Attack with Clean Labels via Attribute Trigger
by: Zhu, Mingyan, et al.
Published: (2023)
by: Zhu, Mingyan, et al.
Published: (2023)
CertDW: Towards Certified Dataset Ownership Verification via Conformal Prediction
by: Qiao, Ting, et al.
Published: (2025)
by: Qiao, Ting, et al.
Published: (2025)
FedFG: Privacy-Preserving and Robust Federated Learning via Flow-Matching Generation
by: Wang, Ruiyang, et al.
Published: (2026)
by: Wang, Ruiyang, et al.
Published: (2026)
FC-Attack: Jailbreaking Multimodal Large Language Models via Auto-Generated Flowcharts
by: Zhang, Ziyi, et al.
Published: (2025)
by: Zhang, Ziyi, et al.
Published: (2025)
Rotation-Invariant Spherical Watermarking via Third-Order SO(3) Representation Coupling
by: Chen, Pengzhen, et al.
Published: (2026)
by: Chen, Pengzhen, et al.
Published: (2026)
FreezeAsGuard: Mitigating Illegal Adaptation of Diffusion Models via Selective Tensor Freezing
by: Huang, Kai, et al.
Published: (2024)
by: Huang, Kai, et al.
Published: (2024)
Transferable Black-Box One-Shot Forging of Watermarks via Image Preference Models
by: Souček, Tomáš, et al.
Published: (2025)
by: Souček, Tomáš, et al.
Published: (2025)
Efficient Semi-Supervised Adversarial Training via Latent Clustering-Based Data Reduction
by: Ghosh, Somrita, et al.
Published: (2025)
by: Ghosh, Somrita, et al.
Published: (2025)
Efficient Black-box Adversarial Attacks via Bayesian Optimization Guided by a Function Prior
by: Cheng, Shuyu, et al.
Published: (2024)
by: Cheng, Shuyu, et al.
Published: (2024)
Value-Aligned Prompt Moderation via Zero-Shot Agentic Rewriting for Safe Image Generation
by: Zhao, Xin, et al.
Published: (2025)
by: Zhao, Xin, et al.
Published: (2025)
CLIP-Inspector: Model-Level Backdoor Detection for Prompt-Tuned CLIP via OOD Trigger Inversion
by: Jindal, Akshit, et al.
Published: (2026)
by: Jindal, Akshit, et al.
Published: (2026)
PointNCBW: Towards Dataset Ownership Verification for Point Clouds via Negative Clean-label Backdoor Watermark
by: Wei, Cheng, et al.
Published: (2024)
by: Wei, Cheng, et al.
Published: (2024)
Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection
by: Shao, Zedian, et al.
Published: (2026)
by: Shao, Zedian, et al.
Published: (2026)
CPR: Retrieval Augmented Generation for Copyright Protection
by: Golatkar, Aditya, et al.
Published: (2024)
by: Golatkar, Aditya, et al.
Published: (2024)
Accuracy of TextFooler black box adversarial attacks on 01 loss sign activation neural network ensemble
by: Xue, Yunzhe, et al.
Published: (2024)
by: Xue, Yunzhe, et al.
Published: (2024)
RAB$^2$-DEF: Dynamic and explainable defense against adversarial attacks in Federated Learning to fair poor clients
by: Rodríguez-Barroso, Nuria, et al.
Published: (2024)
by: Rodríguez-Barroso, Nuria, et al.
Published: (2024)
Similar Items
-
Deepfake detectors are DUMB: A benchmark to assess adversarial training robustness under transferability constraints
by: Serrano, Adrian, et al.
Published: (2026) -
D2R: dual regularization loss with collaborative adversarial generation for model robustness
by: Liu, Zhenyu, et al.
Published: (2025) -
Trainwreck: A damaging adversarial attack on image classifiers
by: Zahálka, Jan
Published: (2023) -
Robust width: A lightweight and certifiable adversarial defense
by: Peck, Jonathan, et al.
Published: (2024) -
Verification of Neural Networks against Convolutional Perturbations via Parameterised Kernels
by: Brückner, Benedikt, et al.
Published: (2024)