Backdooring Masked Diffusion Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cao, Daniel Yiming, Wang, Chengzhong, Chou, Sheng-Yen, Huang, Chengyu, Chen, Pin-Yu, An, Shengwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VillanDiffusion: A Unified Backdoor Attack Framework for Diffusion Models
von: Chou, Sheng-Yen, et al.
Veröffentlicht: (2023)
von: Chou, Sheng-Yen, et al.
Veröffentlicht: (2023)
Elijah: Eliminating Backdoors Injected in Diffusion Models via Distribution Shift
von: An, Shengwei, et al.
Veröffentlicht: (2023)
von: An, Shengwei, et al.
Veröffentlicht: (2023)
Rethinking Backdoor Attacks on Dataset Distillation: A Kernel Method Perspective
von: Chung, Ming-Yu, et al.
Veröffentlicht: (2023)
von: Chung, Ming-Yu, et al.
Veröffentlicht: (2023)
Self-Purification Mitigates Backdoors in Multimodal Diffusion Language Models
von: Wan, Guangnian, et al.
Veröffentlicht: (2026)
von: Wan, Guangnian, et al.
Veröffentlicht: (2026)
Backdoor Attacks on Discrete Graph Diffusion Models
von: Wang, Jiawen, et al.
Veröffentlicht: (2025)
von: Wang, Jiawen, et al.
Veröffentlicht: (2025)
Defending against Backdoor Attack on Deep Neural Networks
von: Cheng, Hao, et al.
Veröffentlicht: (2020)
von: Cheng, Hao, et al.
Veröffentlicht: (2020)
PSBD: Prediction Shift Uncertainty Unlocks Backdoor Detection
von: Li, Wei, et al.
Veröffentlicht: (2024)
von: Li, Wei, et al.
Veröffentlicht: (2024)
Diff-Cleanse: Identifying and Mitigating Backdoor Attacks in Diffusion Models
von: Hao, Jiang, et al.
Veröffentlicht: (2024)
von: Hao, Jiang, et al.
Veröffentlicht: (2024)
Imperio: Language-Guided Backdoor Attacks for Arbitrary Model Control
von: Chow, Ka-Ho, et al.
Veröffentlicht: (2024)
von: Chow, Ka-Ho, et al.
Veröffentlicht: (2024)
REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
UIBDiffusion: Universal Imperceptible Backdoor Attack for Diffusion Models
von: Han, Yuning, et al.
Veröffentlicht: (2024)
von: Han, Yuning, et al.
Veröffentlicht: (2024)
ArcGen: Generalizing Neural Backdoor Detection Across Diverse Architectures
von: Yang, Zhonghao, et al.
Veröffentlicht: (2025)
von: Yang, Zhonghao, et al.
Veröffentlicht: (2025)
BadRSSD: Backdoor Attacks on Regularized Self-Supervised Diffusion Models
von: Wang, Jiayao, et al.
Veröffentlicht: (2026)
von: Wang, Jiayao, et al.
Veröffentlicht: (2026)
CENSOR: Defense Against Gradient Inversion via Orthogonal Subspace Bayesian Sampling
von: Zhang, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Kaiyuan, et al.
Veröffentlicht: (2025)
Awakening the Hydra: Stabilizing Multi-Concept Backdoor Injection in Text-to-Image Diffusion Models
von: Wang, Kai, et al.
Veröffentlicht: (2026)
von: Wang, Kai, et al.
Veröffentlicht: (2026)
TERD: A Unified Framework for Safeguarding Diffusion Models Against Backdoors
von: Mo, Yichuan, et al.
Veröffentlicht: (2024)
von: Mo, Yichuan, et al.
Veröffentlicht: (2024)
IBD-PSC: Input-level Backdoor Detection via Parameter-oriented Scaling Consistency
von: Hou, Linshan, et al.
Veröffentlicht: (2024)
von: Hou, Linshan, et al.
Veröffentlicht: (2024)
DiffusionHijack: Supply-Chain PRNG Backdoor Attack on Diffusion Models and Quantum Random Number Defense
von: You, Ziyang, et al.
Veröffentlicht: (2026)
von: You, Ziyang, et al.
Veröffentlicht: (2026)
Unelicitable Backdoors in Language Models via Cryptographic Transformer Circuits
von: Draguns, Andis, et al.
Veröffentlicht: (2024)
von: Draguns, Andis, et al.
Veröffentlicht: (2024)
Injecting Undetectable Backdoors in Obfuscated Neural Networks and Language Models
von: Kalavasis, Alkis, et al.
Veröffentlicht: (2024)
von: Kalavasis, Alkis, et al.
Veröffentlicht: (2024)
BackdoorBench: A Comprehensive Benchmark and Analysis of Backdoor Learning
von: Wu, Baoyuan, et al.
Veröffentlicht: (2024)
von: Wu, Baoyuan, et al.
Veröffentlicht: (2024)
FilterFL: Knowledge Filtering-based Data-Free Backdoor Defense for Federated Learning
von: Yang, Yanxin, et al.
Veröffentlicht: (2023)
von: Yang, Yanxin, et al.
Veröffentlicht: (2023)
BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models
von: Xiang, Zhen, et al.
Veröffentlicht: (2024)
von: Xiang, Zhen, et al.
Veröffentlicht: (2024)
The 'Sure' Trap: Multi-Scale Poisoning Analysis of Stealthy Compliance-Only Backdoors in Fine-Tuned Large Language Models
von: Tan, Yuting, et al.
Veröffentlicht: (2025)
von: Tan, Yuting, et al.
Veröffentlicht: (2025)
BADTV: Unveiling Backdoor Threats in Third-Party Task Vectors
von: Hsu, Chia-Yi, et al.
Veröffentlicht: (2025)
von: Hsu, Chia-Yi, et al.
Veröffentlicht: (2025)
Composite Backdoor Attacks Against Large Language Models
von: Huang, Hai, et al.
Veröffentlicht: (2023)
von: Huang, Hai, et al.
Veröffentlicht: (2023)
Winter Soldier: Backdooring Language Models at Pre-Training with Indirect Data Poisoning
von: Bouaziz, Wassim, et al.
Veröffentlicht: (2025)
von: Bouaziz, Wassim, et al.
Veröffentlicht: (2025)
Is the Trigger Essential? A Feature-Based Triggerless Backdoor Attack in Vertical Federated Learning
von: Liu, Yige, et al.
Veröffentlicht: (2026)
von: Liu, Yige, et al.
Veröffentlicht: (2026)
CodePurify: Defend Backdoor Attacks on Neural Code Models via Entropy-based Purification
von: Mu, Fangwen, et al.
Veröffentlicht: (2024)
von: Mu, Fangwen, et al.
Veröffentlicht: (2024)
Fusing Pruned and Backdoored Models: Optimal Transport-based Data-free Backdoor Mitigation
von: Lin, Weilin, et al.
Veröffentlicht: (2024)
von: Lin, Weilin, et al.
Veröffentlicht: (2024)
SteganoBackdoor: Stealthy and Data-Efficient Backdoor Attacks on Language Models
von: Xue, Eric, et al.
Veröffentlicht: (2025)
von: Xue, Eric, et al.
Veröffentlicht: (2025)
Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models
von: Clop, Cody, et al.
Veröffentlicht: (2024)
von: Clop, Cody, et al.
Veröffentlicht: (2024)
Shortcuts Everywhere and Nowhere: Exploring Multi-Trigger Backdoor Attacks
von: Li, Yige, et al.
Veröffentlicht: (2024)
von: Li, Yige, et al.
Veröffentlicht: (2024)
MADE: Graph Backdoor Defense with Masked Unlearning
von: Lin, Xiao, et al.
Veröffentlicht: (2024)
von: Lin, Xiao, et al.
Veröffentlicht: (2024)
Backdooring Bias ($B^2$) into Stable Diffusion Models
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
State Backdoor: Towards Stealthy Real-world Poisoning Attack on Vision-Language-Action Model in State Space
von: Guo, Ji, et al.
Veröffentlicht: (2026)
von: Guo, Ji, et al.
Veröffentlicht: (2026)
Universal Graph Backdoor Defense: A Feature-based Homophily Perspective
von: Pan, Mengting, et al.
Veröffentlicht: (2026)
von: Pan, Mengting, et al.
Veröffentlicht: (2026)
Stealthy and Adjustable Text-Guided Backdoor Attacks on Multimodal Pretrained Models
von: Zhang, Yiyang, et al.
Veröffentlicht: (2026)
von: Zhang, Yiyang, et al.
Veröffentlicht: (2026)
Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
von: Xu, Jiashu, et al.
Veröffentlicht: (2023)
von: Xu, Jiashu, et al.
Veröffentlicht: (2023)
Planting Undetectable Backdoors in Machine Learning Models
von: Goldwasser, Shafi, et al.
Veröffentlicht: (2022)
von: Goldwasser, Shafi, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
VillanDiffusion: A Unified Backdoor Attack Framework for Diffusion Models
von: Chou, Sheng-Yen, et al.
Veröffentlicht: (2023) -
Elijah: Eliminating Backdoors Injected in Diffusion Models via Distribution Shift
von: An, Shengwei, et al.
Veröffentlicht: (2023) -
Rethinking Backdoor Attacks on Dataset Distillation: A Kernel Method Perspective
von: Chung, Ming-Yu, et al.
Veröffentlicht: (2023) -
Self-Purification Mitigates Backdoors in Multimodal Diffusion Language Models
von: Wan, Guangnian, et al.
Veröffentlicht: (2026) -
Backdoor Attacks on Discrete Graph Diffusion Models
von: Wang, Jiawen, et al.
Veröffentlicht: (2025)