Coward: Collision-based OOD Watermarking for Practical Proactive Federated Backdoor Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Wenjie, Gu, Siying, Li, Yiming, Li, Shuxin, Chen, Zhili, Zhang, Tianwei, Xia, Shu-Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BackdoorIndicator: Leveraging OOD Data for Proactive Backdoor Detection in Federated Learning
von: Li, Songze, et al.
Veröffentlicht: (2024)
von: Li, Songze, et al.
Veröffentlicht: (2024)
CBW: Towards Dataset Ownership Verification for Speaker Verification via Clustering-based Backdoor Watermarking
von: Li, Yiming, et al.
Veröffentlicht: (2025)
von: Li, Yiming, et al.
Veröffentlicht: (2025)
TrojanDam: Detection-Free Backdoor Defense in Federated Learning through Proactive Model Robustification utilizing OOD Data
von: Dai, Yanbo, et al.
Veröffentlicht: (2025)
von: Dai, Yanbo, et al.
Veröffentlicht: (2025)
Proactive Detection of Voice Cloning with Localized Watermarking
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
On the Weaknesses of Backdoor-based Model Watermarking: An Information-theoretic Perspective
von: Hu, Aoting, et al.
Veröffentlicht: (2024)
von: Hu, Aoting, et al.
Veröffentlicht: (2024)
SSCL-BW: Sample-Specific Clean-Label Backdoor Watermarking for Dataset Ownership Verification
von: Wang, Yingjia, et al.
Veröffentlicht: (2025)
von: Wang, Yingjia, et al.
Veröffentlicht: (2025)
BadEdit: Backdooring large language models by model editing
von: Li, Yanzhou, et al.
Veröffentlicht: (2024)
von: Li, Yanzhou, et al.
Veröffentlicht: (2024)
You Can Backdoor Personalized Federated Learning
von: Ye, Tiandi, et al.
Veröffentlicht: (2023)
von: Ye, Tiandi, et al.
Veröffentlicht: (2023)
WGLE:Backdoor-free and Multi-bit Black-box Watermarking for Graph Neural Networks
von: Li, Tingzhi, et al.
Veröffentlicht: (2025)
von: Li, Tingzhi, et al.
Veröffentlicht: (2025)
Backdoor Sentinel: Detecting and Detoxifying Backdoors in Diffusion Models via Temporal Noise Consistency
von: Wang, Bingzheng, et al.
Veröffentlicht: (2026)
von: Wang, Bingzheng, et al.
Veröffentlicht: (2026)
When Backdoors Speak: Understanding LLM Backdoor Attacks Through Model-Generated Explanations
von: Ge, Huaizhi, et al.
Veröffentlicht: (2024)
von: Ge, Huaizhi, et al.
Veröffentlicht: (2024)
Buffer is All You Need: Defending Federated Learning against Backdoor Attacks under Non-iids via Buffering
von: Lyu, Xingyu, et al.
Veröffentlicht: (2025)
von: Lyu, Xingyu, et al.
Veröffentlicht: (2025)
Harmless Backdoor-based Client-side Watermarking in Federated Learning
von: Luo, Kaijing, et al.
Veröffentlicht: (2024)
von: Luo, Kaijing, et al.
Veröffentlicht: (2024)
Secure On-Device Video OOD Detection Without Backpropagation
von: Li, Shawn, et al.
Veröffentlicht: (2025)
von: Li, Shawn, et al.
Veröffentlicht: (2025)
Learning to Watermark: A Selective Watermarking Framework for Large Language Models via Multi-Objective Optimization
von: Wang, Chenrui, et al.
Veröffentlicht: (2025)
von: Wang, Chenrui, et al.
Veröffentlicht: (2025)
Towards Sample-specific Backdoor Attack with Clean Labels via Attribute Trigger
von: Zhu, Mingyan, et al.
Veröffentlicht: (2023)
von: Zhu, Mingyan, et al.
Veröffentlicht: (2023)
PCDiff: Proactive Control for Ownership Protection in Diffusion Models with Watermark Compatibility
von: Gai, Keke, et al.
Veröffentlicht: (2025)
von: Gai, Keke, et al.
Veröffentlicht: (2025)
Modification and Generated-Text Detection: Achieving Dual Detection Capabilities for the Outputs of LLM by Watermark
von: Cai, Yuhang, et al.
Veröffentlicht: (2025)
von: Cai, Yuhang, et al.
Veröffentlicht: (2025)
Enhancing Model Defense Against Jailbreaks with Proactive Safety Reasoning
von: Yang, Xianglin, et al.
Veröffentlicht: (2025)
von: Yang, Xianglin, et al.
Veröffentlicht: (2025)
BitHydra: Towards Bit-flip Inference Cost Attack against Large Language Models
von: Yan, Xiaobei, et al.
Veröffentlicht: (2025)
von: Yan, Xiaobei, et al.
Veröffentlicht: (2025)
Chain-of-Scrutiny: Detecting Backdoor Attacks for Large Language Models
von: Li, Xi, et al.
Veröffentlicht: (2024)
von: Li, Xi, et al.
Veröffentlicht: (2024)
SWAP: Towards Copyright Auditing of Soft Prompts via Sequential Watermarking
von: Yang, Wenyuan, et al.
Veröffentlicht: (2025)
von: Yang, Wenyuan, et al.
Veröffentlicht: (2025)
FLClear: Visually Verifiable Multi-Client Watermarking for Federated Learning
von: Gu, Chen, et al.
Veröffentlicht: (2025)
von: Gu, Chen, et al.
Veröffentlicht: (2025)
SFIBA: Spatial-based Full-target Invisible Backdoor Attacks
von: Yin, Yangxu, et al.
Veröffentlicht: (2025)
von: Yin, Yangxu, et al.
Veröffentlicht: (2025)
AutoBackdoor: Automating Backdoor Attacks via LLM Agents
von: Li, Yige, et al.
Veröffentlicht: (2025)
von: Li, Yige, et al.
Veröffentlicht: (2025)
TimeGuard: Channel-wise Pool Training for Backdoor Defense in Time Series Forecasting
von: Nguyen, Quang Duc, et al.
Veröffentlicht: (2026)
von: Nguyen, Quang Duc, et al.
Veröffentlicht: (2026)
Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
von: Chen, Yukun, et al.
Veröffentlicht: (2025)
CEFW: A Comprehensive Evaluation Framework for Watermark in Large Language Models
von: Zhang, Shuhao, et al.
Veröffentlicht: (2025)
von: Zhang, Shuhao, et al.
Veröffentlicht: (2025)
Robust Client-Server Watermarking for Split Federated Learning
von: Tang, Jiaxiong, et al.
Veröffentlicht: (2025)
von: Tang, Jiaxiong, et al.
Veröffentlicht: (2025)
Backdoor4Good: Benchmarking Beneficial Uses of Backdoors in LLMs
von: Li, Yige, et al.
Veröffentlicht: (2026)
von: Li, Yige, et al.
Veröffentlicht: (2026)
ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs
von: Yang, Yuchen, et al.
Veröffentlicht: (2024)
von: Yang, Yuchen, et al.
Veröffentlicht: (2024)
Smark: A Watermark for Text-to-Speech Diffusion Models via Discrete Wavelet Transform
von: Zhang, Yichuan, et al.
Veröffentlicht: (2025)
von: Zhang, Yichuan, et al.
Veröffentlicht: (2025)
BURN: Backdoor Unlearning via Adversarial Boundary Analysis
von: Su, Yanghao, et al.
Veröffentlicht: (2025)
von: Su, Yanghao, et al.
Veröffentlicht: (2025)
FFCBA: Feature-based Full-target Clean-label Backdoor Attacks
von: Yin, Yangxu, et al.
Veröffentlicht: (2025)
von: Yin, Yangxu, et al.
Veröffentlicht: (2025)
Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
PPBFL: A Privacy Protected Blockchain-based Federated Learning Model
von: Li, Yang, et al.
Veröffentlicht: (2024)
von: Li, Yang, et al.
Veröffentlicht: (2024)
Backdoors in RLVR: Jailbreak Backdoors in LLMs From Verifiable Reward
von: Guo, Weiyang, et al.
Veröffentlicht: (2026)
von: Guo, Weiyang, et al.
Veröffentlicht: (2026)
PSBD: Prediction Shift Uncertainty Unlocks Backdoor Detection
von: Li, Wei, et al.
Veröffentlicht: (2024)
von: Li, Wei, et al.
Veröffentlicht: (2024)
Fluent: Round-efficient Secure Aggregation for Private Federated Learning
von: Li, Xincheng, et al.
Veröffentlicht: (2024)
von: Li, Xincheng, et al.
Veröffentlicht: (2024)
Deciphering the Interplay between Attack and Protection Complexity in Privacy-Preserving Federated Learning
von: Zhang, Xiaojin, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaojin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
BackdoorIndicator: Leveraging OOD Data for Proactive Backdoor Detection in Federated Learning
von: Li, Songze, et al.
Veröffentlicht: (2024) -
CBW: Towards Dataset Ownership Verification for Speaker Verification via Clustering-based Backdoor Watermarking
von: Li, Yiming, et al.
Veröffentlicht: (2025) -
TrojanDam: Detection-Free Backdoor Defense in Federated Learning through Proactive Model Robustification utilizing OOD Data
von: Dai, Yanbo, et al.
Veröffentlicht: (2025) -
Proactive Detection of Voice Cloning with Localized Watermarking
von: Roman, Robin San, et al.
Veröffentlicht: (2024) -
On the Weaknesses of Backdoor-based Model Watermarking: An Information-theoretic Perspective
von: Hu, Aoting, et al.
Veröffentlicht: (2024)