CENSOR: Defense Against Gradient Inversion via Orthogonal Subspace Bayesian Sampling
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Kaiyuan, Cheng, Siyuan, Shen, Guangyu, Ribeiro, Bruno, An, Shengwei, Chen, Pin-Yu, Zhang, Xiangyu, Li, Ninghui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MIST: Defending Against Membership Inference Attacks Through Membership-Invariant Subspace Training
by: Li, Jiacheng, et al.
Published: (2023)
by: Li, Jiacheng, et al.
Published: (2023)
UNIT: Backdoor Mitigation via Automated Neural Distribution Tightening
by: Cheng, Siyuan, et al.
Published: (2024)
by: Cheng, Siyuan, et al.
Published: (2024)
Rapid Optimization for Jailbreaking LLMs via Subconscious Exploitation and Echopraxia
by: Shen, Guangyu, et al.
Published: (2024)
by: Shen, Guangyu, et al.
Published: (2024)
ASPIRER: Bypassing System Prompts With Permutation-based Backdoors in LLMs
by: Yan, Lu, et al.
Published: (2024)
by: Yan, Lu, et al.
Published: (2024)
LLM Agents Should Employ Security Principles
by: Zhang, Kaiyuan, et al.
Published: (2025)
by: Zhang, Kaiyuan, et al.
Published: (2025)
Elijah: Eliminating Backdoors Injected in Diffusion Models via Distribution Shift
by: An, Shengwei, et al.
Published: (2023)
by: An, Shengwei, et al.
Published: (2023)
SOFT: Selective Data Obfuscation for Protecting LLM Fine-tuning against Membership Inference Attacks
by: Zhang, Kaiyuan, et al.
Published: (2025)
by: Zhang, Kaiyuan, et al.
Published: (2025)
LOTUS: Evasive and Resilient Backdoor Attacks through Sub-Partitioning
by: Cheng, Siyuan, et al.
Published: (2024)
by: Cheng, Siyuan, et al.
Published: (2024)
Window-based Membership Inference Attacks Against Fine-tuned Large Language Models
by: Chen, Yuetian, et al.
Published: (2026)
by: Chen, Yuetian, et al.
Published: (2026)
Opening A Pandora's Box: Things You Should Know in the Era of Custom GPTs
by: Tao, Guanhong, et al.
Published: (2023)
by: Tao, Guanhong, et al.
Published: (2023)
Defensive Prompt Patch: A Robust and Interpretable Defense of LLMs against Jailbreak Attacks
by: Xiong, Chen, et al.
Published: (2024)
by: Xiong, Chen, et al.
Published: (2024)
Cascading and Proxy Membership Inference Attacks
by: Du, Yuntao, et al.
Published: (2025)
by: Du, Yuntao, et al.
Published: (2025)
From Poisoned to Aware: Fostering Backdoor Self-Awareness in LLMs
by: Shen, Guangyu, et al.
Published: (2025)
by: Shen, Guangyu, et al.
Published: (2025)
MGC: A Compiler Framework Exploiting Compositional Blindness in Aligned LLMs for Malware Generation
by: Yan, Lu, et al.
Published: (2025)
by: Yan, Lu, et al.
Published: (2025)
Security Considerations for Artificial Intelligence Agents
by: Li, Ninghui, et al.
Published: (2026)
by: Li, Ninghui, et al.
Published: (2026)
SafeSteer: Adaptive Subspace Steering for Efficient Jailbreak Defense in Vision-Language Models
by: Zeng, Xiyu, et al.
Published: (2025)
by: Zeng, Xiyu, et al.
Published: (2025)
Activation Gradient based Poisoned Sample Detection Against Backdoor Attacks
by: Yuan, Danni, et al.
Published: (2023)
by: Yuan, Danni, et al.
Published: (2023)
SVDefense: Effective Defense against Gradient Inversion Attacks via Singular Value Decomposition
by: Luo, Chenxiang, et al.
Published: (2025)
by: Luo, Chenxiang, et al.
Published: (2025)
A New Federated Learning Framework Against Gradient Inversion Attacks
by: Guo, Pengxin, et al.
Published: (2024)
by: Guo, Pengxin, et al.
Published: (2024)
Evaluating Selective Encryption Against Gradient Inversion Attacks
by: Gu, Jiajun, et al.
Published: (2025)
by: Gu, Jiajun, et al.
Published: (2025)
A No-Defense Defense Against Gradient-Based Adversarial Attacks on ML-NIDS: Is Less More?
by: elShehaby, Mohamed, et al.
Published: (2026)
by: elShehaby, Mohamed, et al.
Published: (2026)
Imitative Membership Inference Attack
by: Du, Yuntao, et al.
Published: (2025)
by: Du, Yuntao, et al.
Published: (2025)
Deep Learning Model Inversion Attacks and Defenses: A Comprehensive Survey
by: Yang, Wencheng, et al.
Published: (2025)
by: Yang, Wencheng, et al.
Published: (2025)
Optimal Defenses Against Gradient Reconstruction Attacks
by: Chen, Yuxiao, et al.
Published: (2024)
by: Chen, Yuxiao, et al.
Published: (2024)
TED-LaST: Towards Robust Backdoor Defense Against Adaptive Attacks
by: Mo, Xiaoxing, et al.
Published: (2025)
by: Mo, Xiaoxing, et al.
Published: (2025)
Combining Stochastic Defenses to Resist Gradient Inversion: An Ablation Study
by: Scheliga, Daniel, et al.
Published: (2022)
by: Scheliga, Daniel, et al.
Published: (2022)
Understanding Deep Gradient Leakage via Inversion Influence Functions
by: Zhang, Haobo, et al.
Published: (2023)
by: Zhang, Haobo, et al.
Published: (2023)
REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
by: Chen, Yukun, et al.
Published: (2025)
by: Chen, Yukun, et al.
Published: (2025)
Enhancing Privacy of Spatiotemporal Federated Learning against Gradient Inversion Attacks
by: Zheng, Lele, et al.
Published: (2024)
by: Zheng, Lele, et al.
Published: (2024)
Defending Against Neural Network Model Inversion Attacks via Data Poisoning
by: Zhou, Shuai, et al.
Published: (2024)
by: Zhou, Shuai, et al.
Published: (2024)
Subspace Defense: Discarding Adversarial Perturbations by Learning a Subspace for Clean Signals
by: Zheng, Rui, et al.
Published: (2024)
by: Zheng, Rui, et al.
Published: (2024)
BrowseSafe: Understanding and Preventing Prompt Injection Within AI Browser Agents
by: Zhang, Kaiyuan, et al.
Published: (2025)
by: Zhang, Kaiyuan, et al.
Published: (2025)
MirrorShield: Towards Universal Defense Against Jailbreaks via Entropy-Guided Mirror Crafting
by: Pu, Rui, et al.
Published: (2025)
by: Pu, Rui, et al.
Published: (2025)
Beyond Surface-Level Patterns: An Essence-Driven Defense Framework Against Jailbreak Attacks in LLMs
by: Xiang, Shiyu, et al.
Published: (2025)
by: Xiang, Shiyu, et al.
Published: (2025)
AutoVerifier: An Agentic Automated Verification Framework Using Large Language Models
by: Du, Yuntao, et al.
Published: (2026)
by: Du, Yuntao, et al.
Published: (2026)
E-SAGE: Explainability-based Defense Against Backdoor Attacks on Graph Neural Networks
by: Yuan, Dingqiang, et al.
Published: (2024)
by: Yuan, Dingqiang, et al.
Published: (2024)
Federated Learning under Attack: Improving Gradient Inversion for Batch of Images
by: Leite, Luiz, et al.
Published: (2024)
by: Leite, Luiz, et al.
Published: (2024)
PECAN: A Deterministic Certified Defense Against Backdoor Attacks
by: Zhang, Yuhao, et al.
Published: (2023)
by: Zhang, Yuhao, et al.
Published: (2023)
Defense Against Model Stealing Based on Account-Aware Distribution Discrepancy
by: Mei, Jian-Ping, et al.
Published: (2025)
by: Mei, Jian-Ping, et al.
Published: (2025)
Defending Against Prompt Injection With a Few DefensiveTokens
by: Chen, Sizhe, et al.
Published: (2025)
by: Chen, Sizhe, et al.
Published: (2025)
Similar Items
-
MIST: Defending Against Membership Inference Attacks Through Membership-Invariant Subspace Training
by: Li, Jiacheng, et al.
Published: (2023) -
UNIT: Backdoor Mitigation via Automated Neural Distribution Tightening
by: Cheng, Siyuan, et al.
Published: (2024) -
Rapid Optimization for Jailbreaking LLMs via Subconscious Exploitation and Echopraxia
by: Shen, Guangyu, et al.
Published: (2024) -
ASPIRER: Bypassing System Prompts With Permutation-based Backdoors in LLMs
by: Yan, Lu, et al.
Published: (2024) -
LLM Agents Should Employ Security Principles
by: Zhang, Kaiyuan, et al.
Published: (2025)