CBD: A Certified Backdoor Detector Based on Local Dominant Probability
Fuente:
arXiv
Saved in:
| Main Authors: | Xiang, Zhen, Xiong, Zidi, Li, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models
by: Xiang, Zhen, et al.
Published: (2024)
by: Xiang, Zhen, et al.
Published: (2024)
Authority Backdoor: A Certifiable Backdoor Mechanism for Authoring DNNs
by: Yang, Han, et al.
Published: (2025)
by: Yang, Han, et al.
Published: (2025)
Under-confidence Backdoors Are Resilient and Stealthy Backdoors
by: Peng, Minlong, et al.
Published: (2022)
by: Peng, Minlong, et al.
Published: (2022)
PECAN: A Deterministic Certified Defense Against Backdoor Attacks
by: Zhang, Yuhao, et al.
Published: (2023)
by: Zhang, Yuhao, et al.
Published: (2023)
Fusing Pruned and Backdoored Models: Optimal Transport-based Data-free Backdoor Mitigation
by: Lin, Weilin, et al.
Published: (2024)
by: Lin, Weilin, et al.
Published: (2024)
Persistent Backdoor Attacks in Continual Learning
by: Guo, Zhen, et al.
Published: (2024)
by: Guo, Zhen, et al.
Published: (2024)
BackdoorBench: A Comprehensive Benchmark and Analysis of Backdoor Learning
by: Wu, Baoyuan, et al.
Published: (2024)
by: Wu, Baoyuan, et al.
Published: (2024)
DarkMind: Latent Chain-of-Thought Backdoor in Customized LLMs
by: Guo, Zhen, et al.
Published: (2025)
by: Guo, Zhen, et al.
Published: (2025)
CAMP in the Odyssey: Provably Robust Reinforcement Learning with Certified Radius Maximization
by: Wang, Derui, et al.
Published: (2025)
by: Wang, Derui, et al.
Published: (2025)
Certified Defense on the Fairness of Graph Neural Networks
by: Dong, Yushun, et al.
Published: (2023)
by: Dong, Yushun, et al.
Published: (2023)
Unlearning Backdoor Attacks through Gradient-Based Model Pruning
by: Dunnett, Kealan, et al.
Published: (2024)
by: Dunnett, Kealan, et al.
Published: (2024)
Backdoor Learning Curves: Explaining Backdoor Poisoning Beyond Influence Functions
by: Cinà, Antonio Emanuele, et al.
Published: (2021)
by: Cinà, Antonio Emanuele, et al.
Published: (2021)
Certified Unlearning for Neural Networks
by: Koloskova, Anastasia, et al.
Published: (2025)
by: Koloskova, Anastasia, et al.
Published: (2025)
Hardware-Triggered Backdoors
by: Möller, Jonas, et al.
Published: (2026)
by: Möller, Jonas, et al.
Published: (2026)
BlockDoor: Blocking Backdoor Based Watermarks in Deep Neural Networks
by: Puah, Yi Hao, et al.
Published: (2024)
by: Puah, Yi Hao, et al.
Published: (2024)
Is the Trigger Essential? A Feature-Based Triggerless Backdoor Attack in Vertical Federated Learning
by: Liu, Yige, et al.
Published: (2026)
by: Liu, Yige, et al.
Published: (2026)
MalRAG: A Retrieval-Augmented LLM Framework for Open-set Malicious Traffic Identification
by: Luo, Xiang, et al.
Published: (2025)
by: Luo, Xiang, et al.
Published: (2025)
Certified Machine Unlearning via Noisy Stochastic Gradient Descent
by: Chien, Eli, et al.
Published: (2024)
by: Chien, Eli, et al.
Published: (2024)
Universal Graph Backdoor Defense: A Feature-based Homophily Perspective
by: Pan, Mengting, et al.
Published: (2026)
by: Pan, Mengting, et al.
Published: (2026)
Game of Trojans: Adaptive Adversaries Against Output-based Trojaned-Model Detectors
by: Sahabandu, Dinuka, et al.
Published: (2024)
by: Sahabandu, Dinuka, et al.
Published: (2024)
RPP: A Certified Poisoned-Sample Detection Framework for Backdoor Attacks under Dataset Imbalance
by: Lin, Miao, et al.
Published: (2026)
by: Lin, Miao, et al.
Published: (2026)
Turn-Based Structural Triggers: Prompt-Free Backdoors in Multi-Turn LLMs
by: Lu, Yiyang, et al.
Published: (2026)
by: Lu, Yiyang, et al.
Published: (2026)
Instruction Backdoor Attacks Against Customized LLMs
by: Zhang, Rui, et al.
Published: (2024)
by: Zhang, Rui, et al.
Published: (2024)
Cross-Paradigm Graph Backdoor Attacks with Promptable Subgraph Triggers
by: Liu, Dongyi, et al.
Published: (2025)
by: Liu, Dongyi, et al.
Published: (2025)
Adaptive Backdoor Attacks with Reasonable Constraints on Graph Neural Networks
by: Dong, Xuewen, et al.
Published: (2025)
by: Dong, Xuewen, et al.
Published: (2025)
BadMerging: Backdoor Attacks Against Model Merging
by: Zhang, Jinghuai, et al.
Published: (2024)
by: Zhang, Jinghuai, et al.
Published: (2024)
A Certified Unlearning Approach without Access to Source Data
by: Basaran, Umit Yigit, et al.
Published: (2025)
by: Basaran, Umit Yigit, et al.
Published: (2025)
ArcGen: Generalizing Neural Backdoor Detection Across Diverse Architectures
by: Yang, Zhonghao, et al.
Published: (2025)
by: Yang, Zhonghao, et al.
Published: (2025)
On Using Certified Training towards Empirical Robustness
by: De Palma, Alessandro, et al.
Published: (2024)
by: De Palma, Alessandro, et al.
Published: (2024)
Cross-Input Certified Training for Universal Perturbations
by: Xu, Changming, et al.
Published: (2024)
by: Xu, Changming, et al.
Published: (2024)
Securing Federated Learning against Backdoor Threats with Foundation Model Integration
by: Bi, Xiaohuan, et al.
Published: (2024)
by: Bi, Xiaohuan, et al.
Published: (2024)
RigorLLM: Resilient Guardrails for Large Language Models against Undesired Content
by: Yuan, Zhuowen, et al.
Published: (2024)
by: Yuan, Zhuowen, et al.
Published: (2024)
AgentPoison: Red-teaming LLM Agents via Poisoning Memory or Knowledge Bases
by: Chen, Zhaorun, et al.
Published: (2024)
by: Chen, Zhaorun, et al.
Published: (2024)
EmInspector: Combating Backdoor Attacks in Federated Self-Supervised Learning Through Embedding Inspection
by: Qian, Yuwen, et al.
Published: (2024)
by: Qian, Yuwen, et al.
Published: (2024)
Real-World Adversarial Attacks on RF-Based Drone Detectors
by: Gazit, Omer, et al.
Published: (2025)
by: Gazit, Omer, et al.
Published: (2025)
Backdooring Masked Diffusion Language Models
by: Cao, Daniel Yiming, et al.
Published: (2026)
by: Cao, Daniel Yiming, et al.
Published: (2026)
Seal Your Backdoor with Variational Defense
by: Sabolić, Ivan, et al.
Published: (2025)
by: Sabolić, Ivan, et al.
Published: (2025)
Backdoor Attacks on Decentralised Post-Training
by: Ersoy, Oğuzhan, et al.
Published: (2026)
by: Ersoy, Oğuzhan, et al.
Published: (2026)
Robustness Inspired Graph Backdoor Defense
by: Zhang, Zhiwei, et al.
Published: (2024)
by: Zhang, Zhiwei, et al.
Published: (2024)
Self-Purification Mitigates Backdoors in Multimodal Diffusion Language Models
by: Wan, Guangnian, et al.
Published: (2026)
by: Wan, Guangnian, et al.
Published: (2026)
Similar Items
-
BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models
by: Xiang, Zhen, et al.
Published: (2024) -
Authority Backdoor: A Certifiable Backdoor Mechanism for Authoring DNNs
by: Yang, Han, et al.
Published: (2025) -
Under-confidence Backdoors Are Resilient and Stealthy Backdoors
by: Peng, Minlong, et al.
Published: (2022) -
PECAN: A Deterministic Certified Defense Against Backdoor Attacks
by: Zhang, Yuhao, et al.
Published: (2023) -
Fusing Pruned and Backdoored Models: Optimal Transport-based Data-free Backdoor Mitigation
by: Lin, Weilin, et al.
Published: (2024)