REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yukun, Shao, Shuo, Huang, Enhao, Li, Yiming, Chen, Pin-Yu, Qin, Zhan, Ren, Kui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
Backdoor Directions in Vision Transformers
von: Karayalcin, Sengim, et al.
Veröffentlicht: (2026)
von: Karayalcin, Sengim, et al.
Veröffentlicht: (2026)
PointNCBW: Towards Dataset Ownership Verification for Point Clouds via Negative Clean-label Backdoor Watermark
von: Wei, Cheng, et al.
Veröffentlicht: (2024)
von: Wei, Cheng, et al.
Veröffentlicht: (2024)
MIBench: A Comprehensive Framework for Benchmarking Model Inversion Attack and Defense
von: Qiu, Yixiang, et al.
Veröffentlicht: (2024)
von: Qiu, Yixiang, et al.
Veröffentlicht: (2024)
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
von: Wei, Shaokui, et al.
Veröffentlicht: (2024)
von: Wei, Shaokui, et al.
Veröffentlicht: (2024)
Nearest is Not Dearest: Towards Practical Defense against Quantization-conditioned Backdoor Attacks
von: Li, Boheng, et al.
Veröffentlicht: (2024)
von: Li, Boheng, et al.
Veröffentlicht: (2024)
VillanDiffusion: A Unified Backdoor Attack Framework for Diffusion Models
von: Chou, Sheng-Yen, et al.
Veröffentlicht: (2023)
von: Chou, Sheng-Yen, et al.
Veröffentlicht: (2023)
InverTune: Removing Backdoors from Multimodal Contrastive Learning Models via Trigger Inversion and Activation Tuning
von: Sun, Mengyuan, et al.
Veröffentlicht: (2025)
von: Sun, Mengyuan, et al.
Veröffentlicht: (2025)
Backdoor Defense in Diffusion Models via Spatial Attention Unlearning
von: Jha, Abha, et al.
Veröffentlicht: (2025)
von: Jha, Abha, et al.
Veröffentlicht: (2025)
Rethinking Data Protection in the (Generative) Artificial Intelligence Era
von: Li, Yiming, et al.
Veröffentlicht: (2025)
von: Li, Yiming, et al.
Veröffentlicht: (2025)
SurrogatePrompt: Bypassing the Safety Filter of Text-to-Image Models via Substitution
von: Ba, Zhongjie, et al.
Veröffentlicht: (2023)
von: Ba, Zhongjie, et al.
Veröffentlicht: (2023)
Guarding the Gate: ConceptGuard Battles Concept-Level Backdoors in Concept Bottleneck Models
von: Lai, Songning, et al.
Veröffentlicht: (2024)
von: Lai, Songning, et al.
Veröffentlicht: (2024)
Test-Time Attention Purification for Backdoored Large Vision Language Models
von: Zhang, Zhifang, et al.
Veröffentlicht: (2026)
von: Zhang, Zhifang, et al.
Veröffentlicht: (2026)
DiffMI: Breaking Face Recognition Privacy via Diffusion-Driven Training-Free Model Inversion
von: Wang, Hanrui, et al.
Veröffentlicht: (2025)
von: Wang, Hanrui, et al.
Veröffentlicht: (2025)
Towards Sample-specific Backdoor Attack with Clean Labels via Attribute Trigger
von: Zhu, Mingyan, et al.
Veröffentlicht: (2023)
von: Zhu, Mingyan, et al.
Veröffentlicht: (2023)
Model X-ray:Detecting Backdoored Models via Decision Boundary
von: Su, Yanghao, et al.
Veröffentlicht: (2024)
von: Su, Yanghao, et al.
Veröffentlicht: (2024)
Data Free Backdoor Attacks
von: Cao, Bochuan, et al.
Veröffentlicht: (2024)
von: Cao, Bochuan, et al.
Veröffentlicht: (2024)
Cert-SSBD: Certified Backdoor Defense with Sample-Specific Smoothing Noises
von: Qiao, Ting, et al.
Veröffentlicht: (2025)
von: Qiao, Ting, et al.
Veröffentlicht: (2025)
WMCopier: Forging Invisible Image Watermarks on Arbitrary Images
von: Dong, Ziping, et al.
Veröffentlicht: (2025)
von: Dong, Ziping, et al.
Veröffentlicht: (2025)
Backdoor Cleaning without External Guidance in MLLM Fine-tuning
von: Rong, Xuankun, et al.
Veröffentlicht: (2025)
von: Rong, Xuankun, et al.
Veröffentlicht: (2025)
IU: Imperceptible Universal Backdoor Attack
von: Lin, Hsin, et al.
Veröffentlicht: (2026)
von: Lin, Hsin, et al.
Veröffentlicht: (2026)
Trap-MID: Trapdoor-based Defense against Model Inversion Attacks
von: Liu, Zhen-Ting, et al.
Veröffentlicht: (2024)
von: Liu, Zhen-Ting, et al.
Veröffentlicht: (2024)
OmniSafeBench-MM: A Unified Benchmark and Toolbox for Multimodal Jailbreak Attack-Defense Evaluation
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
Inevitable Encounters: Backdoor Attacks Involving Lossy Compression
von: Li, Qian, et al.
Veröffentlicht: (2026)
von: Li, Qian, et al.
Veröffentlicht: (2026)
Clean-image Backdoor Attacks
von: Rong, Dazhong, et al.
Veröffentlicht: (2024)
von: Rong, Dazhong, et al.
Veröffentlicht: (2024)
CatchBackdoor: Backdoor Detection via Critical Trojan Neural Path Fuzzing
von: Jin, Haibo, et al.
Veröffentlicht: (2021)
von: Jin, Haibo, et al.
Veröffentlicht: (2021)
Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
PAD-FT: A Lightweight Defense for Backdoor Attacks via Data Purification and Fine-Tuning
von: Xu, Yukai, et al.
Veröffentlicht: (2024)
von: Xu, Yukai, et al.
Veröffentlicht: (2024)
Gradient Inversion of Federated Diffusion Models
von: Huang, Jiyue, et al.
Veröffentlicht: (2024)
von: Huang, Jiyue, et al.
Veröffentlicht: (2024)
An Inversion-based Measure of Memorization for Diffusion Models
von: Ma, Zhe, et al.
Veröffentlicht: (2024)
von: Ma, Zhe, et al.
Veröffentlicht: (2024)
Backdoor Attacks against No-Reference Image Quality Assessment Models via a Scalable Trigger
von: Yu, Yi, et al.
Veröffentlicht: (2024)
von: Yu, Yi, et al.
Veröffentlicht: (2024)
A Survey of Trojan Attacks and Defenses to Deep Neural Networks
von: Jin, Lingxin, et al.
Veröffentlicht: (2024)
von: Jin, Lingxin, et al.
Veröffentlicht: (2024)
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations
von: Ye, Mang, et al.
Veröffentlicht: (2025)
von: Ye, Mang, et al.
Veröffentlicht: (2025)
Ensemble Adversarial Defense via Integration of Multiple Dispersed Low Curvature Models
von: Zhao, Kaikang, et al.
Veröffentlicht: (2024)
von: Zhao, Kaikang, et al.
Veröffentlicht: (2024)
Random Erasing vs. Model Inversion: A Promising Defense or a False Hope?
von: Tran, Viet-Hung, et al.
Veröffentlicht: (2024)
von: Tran, Viet-Hung, et al.
Veröffentlicht: (2024)
Backdoor Attacks on Prompt-Driven Video Segmentation Foundation Models
von: Zhang, Zongmin, et al.
Veröffentlicht: (2025)
von: Zhang, Zongmin, et al.
Veröffentlicht: (2025)
Gungnir: Exploiting Stylistic Features in Images for Backdoor Attacks on Diffusion Models
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
Revocable Backdoor for Deep Model Trading
von: Xu, Yiran, et al.
Veröffentlicht: (2024)
von: Xu, Yiran, et al.
Veröffentlicht: (2024)
Adversarially Guided Stateful Defense Against Backdoor Attacks in Federated Deep Learning
von: Ali, Hassan, et al.
Veröffentlicht: (2024)
von: Ali, Hassan, et al.
Veröffentlicht: (2024)
A Proxy Attack-Free Strategy for Practically Improving the Poisoning Efficiency in Backdoor Attacks
von: Li, Ziqiang, et al.
Veröffentlicht: (2023)
von: Li, Ziqiang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack
von: Chen, Yukun, et al.
Veröffentlicht: (2025) -
Backdoor Directions in Vision Transformers
von: Karayalcin, Sengim, et al.
Veröffentlicht: (2026) -
PointNCBW: Towards Dataset Ownership Verification for Point Clouds via Negative Clean-label Backdoor Watermark
von: Wei, Cheng, et al.
Veröffentlicht: (2024) -
MIBench: A Comprehensive Framework for Benchmarking Model Inversion Attack and Defense
von: Qiu, Yixiang, et al.
Veröffentlicht: (2024) -
Mitigating Backdoor Attack by Injecting Proactive Defensive Backdoor
von: Wei, Shaokui, et al.
Veröffentlicht: (2024)