Improving Robustness to Model Inversion Attacks via Sparse Coding Architectures
Fuente:
arXiv
Guardado en:
| Autores principales: | Dibbo, Sayanton V., Breuer, Adam, Moore, Juston, Teti, Michael |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Are Neuro-Inspired Multi-Modal Vision-Language Models Resilient to Membership Inference Privacy Leakage?
por: Amebley, David, et al.
Publicado: (2025)
por: Amebley, David, et al.
Publicado: (2025)
LCANets++: Robust Audio Classification using Multi-layer Neural Networks with Lateral Competition
por: Dibbo, Sayanton V., et al.
Publicado: (2023)
por: Dibbo, Sayanton V., et al.
Publicado: (2023)
Beyond Attack Success Rate: A Multi-Metric Evaluation of Adversarial Transferability in Medical Imaging Models
por: Curl, Emily, et al.
Publicado: (2026)
por: Curl, Emily, et al.
Publicado: (2026)
Gradient Inversion of Federated Diffusion Models
por: Huang, Jiyue, et al.
Publicado: (2024)
por: Huang, Jiyue, et al.
Publicado: (2024)
Trap-MID: Trapdoor-based Defense against Model Inversion Attacks
por: Liu, Zhen-Ting, et al.
Publicado: (2024)
por: Liu, Zhen-Ting, et al.
Publicado: (2024)
Watertox: The Art of Simplicity in Universal Attacks A Cross-Model Framework for Robust Adversarial Generation
por: Gao, Zhenghao, et al.
Publicado: (2024)
por: Gao, Zhenghao, et al.
Publicado: (2024)
CtrlAttack: A Unified Attack on World-Model Control in Diffusion Models
por: Xu, Shuhan, et al.
Publicado: (2026)
por: Xu, Shuhan, et al.
Publicado: (2026)
Robustness Analysis against Adversarial Patch Attacks in Fully Unmanned Stores
por: Na, Hyunsik, et al.
Publicado: (2025)
por: Na, Hyunsik, et al.
Publicado: (2025)
Reason2Attack: Jailbreaking Text-to-Image Models via LLM Reasoning
por: Zhang, Chenyu, et al.
Publicado: (2025)
por: Zhang, Chenyu, et al.
Publicado: (2025)
AttackVLA: Benchmarking Adversarial and Backdoor Attacks on Vision-Language-Action Models
por: Li, Jiayu, et al.
Publicado: (2025)
por: Li, Jiayu, et al.
Publicado: (2025)
Silent Branding Attack: Trigger-free Data Poisoning Attack on Text-to-Image Diffusion Models
por: Jang, Sangwon, et al.
Publicado: (2025)
por: Jang, Sangwon, et al.
Publicado: (2025)
REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
por: Chen, Yukun, et al.
Publicado: (2025)
por: Chen, Yukun, et al.
Publicado: (2025)
Region-Guided Attack on the Segment Anything Model (SAM)
por: Liu, Xiaoliang, et al.
Publicado: (2024)
por: Liu, Xiaoliang, et al.
Publicado: (2024)
Metaphor-based Jailbreak Attacks on Text-to-Image Models
por: Zhang, Chenyu, et al.
Publicado: (2025)
por: Zhang, Chenyu, et al.
Publicado: (2025)
Fight Perturbations with Perturbations: Defending Adversarial Attacks via Neuron Influence
por: Chen, Ruoxi, et al.
Publicado: (2021)
por: Chen, Ruoxi, et al.
Publicado: (2021)
Fooling the Watchers: Breaking AIGC Detectors via Semantic Prompt Attacks
por: Hao, Run, et al.
Publicado: (2025)
por: Hao, Run, et al.
Publicado: (2025)
Black-Box Forgery Attacks on Semantic Watermarks for Diffusion Models
por: Müller, Andreas, et al.
Publicado: (2024)
por: Müller, Andreas, et al.
Publicado: (2024)
INK: Inheritable Natural Backdoor Attack Against Model Distillation
por: Liu, Xiaolei, et al.
Publicado: (2023)
por: Liu, Xiaolei, et al.
Publicado: (2023)
Everywhere Attack: Attacking Locally and Globally to Boost Targeted Transferability
por: Zeng, Hui, et al.
Publicado: (2025)
por: Zeng, Hui, et al.
Publicado: (2025)
CAAP: Capture-Aware Adversarial Patch Attacks on Palmprint Recognition Models
por: Liu, Renyang, et al.
Publicado: (2026)
por: Liu, Renyang, et al.
Publicado: (2026)
PLA: Prompt Learning Attack against Text-to-Image Generative Models
por: Lyu, Xinqi, et al.
Publicado: (2025)
por: Lyu, Xinqi, et al.
Publicado: (2025)
Learning to Detect Unseen Jailbreak Attacks in Large Vision-Language Models
por: Liang, Shuang, et al.
Publicado: (2025)
por: Liang, Shuang, et al.
Publicado: (2025)
A Method to Facilitate Membership Inference Attacks in Deep Learning Models
por: Chen, Zitao, et al.
Publicado: (2024)
por: Chen, Zitao, et al.
Publicado: (2024)
Adversarial Attacks and Defenses on Text-to-Image Diffusion Models: A Survey
por: Zhang, Chenyu, et al.
Publicado: (2024)
por: Zhang, Chenyu, et al.
Publicado: (2024)
Backdoor Attack Against Vision Transformers via Attention Gradient-Based Image Erosion
por: Guo, Ji, et al.
Publicado: (2024)
por: Guo, Ji, et al.
Publicado: (2024)
PuFace: Defending against Facial Cloaking Attacks for Facial Recognition Models
por: Wen, Jing
Publicado: (2024)
por: Wen, Jing
Publicado: (2024)
PAD-FT: A Lightweight Defense for Backdoor Attacks via Data Purification and Fine-Tuning
por: Xu, Yukai, et al.
Publicado: (2024)
por: Xu, Yukai, et al.
Publicado: (2024)
DiffZOO: A Purely Query-Based Black-Box Attack for Red-teaming Text-to-Image Generative Model via Zeroth Order Optimization
por: Dang, Pucheng, et al.
Publicado: (2024)
por: Dang, Pucheng, et al.
Publicado: (2024)
Superpixel Attack: Enhancing Black-box Adversarial Attack with Image-driven Division Areas
por: Oe, Issa, et al.
Publicado: (2025)
por: Oe, Issa, et al.
Publicado: (2025)
Sparse Autoencoder as a Zero-Shot Classifier for Concept Erasing in Text-to-Image Diffusion Models
por: Tian, Zhihua, et al.
Publicado: (2025)
por: Tian, Zhihua, et al.
Publicado: (2025)
Do We Really Need Quantum Machine Learning?: A Multidimensional Empirical Study
por: Vhaduri, Sudip, et al.
Publicado: (2026)
por: Vhaduri, Sudip, et al.
Publicado: (2026)
Towards Dataset Copyright Evasion Attack against Personalized Text-to-Image Diffusion Models
por: Gao, Kuofeng, et al.
Publicado: (2025)
por: Gao, Kuofeng, et al.
Publicado: (2025)
Data Free Backdoor Attacks
por: Cao, Bochuan, et al.
Publicado: (2024)
por: Cao, Bochuan, et al.
Publicado: (2024)
Adversarial Robustness of Vision in Open Foundation Models
por: Fox, Jonathon, et al.
Publicado: (2025)
por: Fox, Jonathon, et al.
Publicado: (2025)
Improving the Transferability of Adversarial Attacks by an Input Transpose
por: Wan, Qing, et al.
Publicado: (2025)
por: Wan, Qing, et al.
Publicado: (2025)
T2I-RiskyPrompt: A Benchmark for Safety Evaluation, Attack, and Defense on Text-to-Image Model
por: Zhang, Chenyu, et al.
Publicado: (2025)
por: Zhang, Chenyu, et al.
Publicado: (2025)
ROBIN: Robust and Invisible Watermarks for Diffusion Models with Adversarial Optimization
por: Huang, Huayang, et al.
Publicado: (2024)
por: Huang, Huayang, et al.
Publicado: (2024)
Backdoor Attacks against Image-to-Image Networks
por: Jiang, Wenbo, et al.
Publicado: (2024)
por: Jiang, Wenbo, et al.
Publicado: (2024)
Antelope: Potent and Concealed Jailbreak Attack Strategy
por: Zhao, Xin, et al.
Publicado: (2024)
por: Zhao, Xin, et al.
Publicado: (2024)
A Robust Attack: Displacement Backdoor Attack
por: Li, Yong, et al.
Publicado: (2025)
por: Li, Yong, et al.
Publicado: (2025)
Ejemplares similares
-
Are Neuro-Inspired Multi-Modal Vision-Language Models Resilient to Membership Inference Privacy Leakage?
por: Amebley, David, et al.
Publicado: (2025) -
LCANets++: Robust Audio Classification using Multi-layer Neural Networks with Lateral Competition
por: Dibbo, Sayanton V., et al.
Publicado: (2023) -
Beyond Attack Success Rate: A Multi-Metric Evaluation of Adversarial Transferability in Medical Imaging Models
por: Curl, Emily, et al.
Publicado: (2026) -
Gradient Inversion of Federated Diffusion Models
por: Huang, Jiyue, et al.
Publicado: (2024) -
Trap-MID: Trapdoor-based Defense against Model Inversion Attacks
por: Liu, Zhen-Ting, et al.
Publicado: (2024)