Class-feature Watermark: A Resilient Black-box Watermark Against Model Extraction Attacks
Fuente:
arXiv
Saved in:
| Main Authors: | Xiao, Yaxin, Ye, Qingqing, Liang, Zi, Li, Haoyang, Li, RongHua, Zheng, Huadi, Hu, Haibo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Sample-Level Evaluation and Generative Framework for Model Inversion Attacks
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
Why Are My Prompts Leaked? Unraveling Prompt Extraction Threats in Customized Large Language Models
by: Liang, Zi, et al.
Published: (2024)
by: Liang, Zi, et al.
Published: (2024)
MER-Inspector: Assessing model extraction risks from an attack-agnostic perspective
by: Zhang, Xinwei, et al.
Published: (2025)
by: Zhang, Xinwei, et al.
Published: (2025)
Performance-lossless Black-box Model Watermarking
by: Zhao, Na, et al.
Published: (2023)
by: Zhao, Na, et al.
Published: (2023)
Does Low Rank Adaptation Lead to Lower Robustness against Training-Time Attacks?
by: Liang, Zi, et al.
Published: (2025)
by: Liang, Zi, et al.
Published: (2025)
Enhancing LLM Watermark Resilience Against Both Scrubbing and Spoofing Attacks
by: Shen, Huanming, et al.
Published: (2025)
by: Shen, Huanming, et al.
Published: (2025)
Traceable Black-box Watermarks for Federated Learning
by: Xu, Jiahao, et al.
Published: (2025)
by: Xu, Jiahao, et al.
Published: (2025)
"Yes, My LoRD." Guiding Language Model Extraction with Locality Reinforced Distillation
by: Liang, Zi, et al.
Published: (2024)
by: Liang, Zi, et al.
Published: (2024)
EmbTracker: Traceable Black-box Watermarking for Federated Language Models
by: Zhao, Haodong, et al.
Published: (2026)
by: Zhao, Haodong, et al.
Published: (2026)
Bounding-box Watermarking: Defense against Model Extraction Attacks on Object Detectors
by: Koda, Satoru, et al.
Published: (2024)
by: Koda, Satoru, et al.
Published: (2024)
The Efficacy of Transfer-based No-box Attacks on Image Watermarking: A Pragmatic Analysis
by: Wu, Qilong, et al.
Published: (2024)
by: Wu, Qilong, et al.
Published: (2024)
Membership Inference Attacks and Defenses in Federated Learning: A Survey
by: Bai, Li, et al.
Published: (2024)
by: Bai, Li, et al.
Published: (2024)
RLSpoofer: A Lightweight Evaluator for LLM Watermark Spoofing Resilience
by: Huang, Hanbo, et al.
Published: (2026)
by: Huang, Hanbo, et al.
Published: (2026)
LoRAGuard: An Effective Black-box Watermarking Approach for LoRAs
by: Lv, Peizhuo, et al.
Published: (2025)
by: Lv, Peizhuo, et al.
Published: (2025)
MEA-Defender: A Robust Watermark against Model Extraction Attack
by: Lv, Peizhuo, et al.
Published: (2024)
by: Lv, Peizhuo, et al.
Published: (2024)
Audio Pirates: Black-box Audio Watermark Removal via Diffusion Priors
by: Yao, Lingfeng, et al.
Published: (2026)
by: Yao, Lingfeng, et al.
Published: (2026)
SEW: Strengthening Robustness of Black-box DNN Watermarking via Specificity Enhancement
by: Qiu, Huming, et al.
Published: (2026)
by: Qiu, Huming, et al.
Published: (2026)
ModelShield: Adaptive and Robust Watermark against Model Extraction Attack
by: Pang, Kaiyi, et al.
Published: (2024)
by: Pang, Kaiyi, et al.
Published: (2024)
Defending LLM Watermarking Against Spoofing Attacks with Contrastive Representation Learning
by: An, Li, et al.
Published: (2025)
by: An, Li, et al.
Published: (2025)
A Game Between the Defender and the Attacker for Trigger-based Black-box Model Watermarking
by: Huang, Chaoyue, et al.
Published: (2025)
by: Huang, Chaoyue, et al.
Published: (2025)
WGLE:Backdoor-free and Multi-bit Black-box Watermarking for Graph Neural Networks
by: Li, Tingzhi, et al.
Published: (2025)
by: Li, Tingzhi, et al.
Published: (2025)
Neural Dehydration: Effective Erasure of Black-box Watermarks from DNNs with Limited Data
by: Lu, Yifan, et al.
Published: (2023)
by: Lu, Yifan, et al.
Published: (2023)
RLCracker: Evaluating the Worst-Case Vulnerability of LLM Watermarks with Adaptive RL Attacks
by: Huang, Hanbo, et al.
Published: (2025)
by: Huang, Hanbo, et al.
Published: (2025)
AGATE: Stealthy Black-box Watermarking for Multimodal Model Copyright Protection
by: Gao, Jianbo, et al.
Published: (2025)
by: Gao, Jianbo, et al.
Published: (2025)
Toward Efficient Inference Attacks: Shadow Model Sharing via Mixture-of-Experts
by: Bai, Li, et al.
Published: (2025)
by: Bai, Li, et al.
Published: (2025)
Provable Watermarking for Data Poisoning Attacks
by: Zhu, Yifan, et al.
Published: (2025)
by: Zhu, Yifan, et al.
Published: (2025)
MarkSweep: A No-box Removal Attack on AI-Generated Image Watermarking via Noise Intensification and Frequency-aware Denoising
by: Cao, Jie, et al.
Published: (2026)
by: Cao, Jie, et al.
Published: (2026)
Cluster-Aware Attacks on Graph Watermarks
by: Nemecek, Alexander, et al.
Published: (2025)
by: Nemecek, Alexander, et al.
Published: (2025)
Neural Honeytrace: Plug&Play Watermarking Framework against Model Extraction Attacks
by: Xu, Yixiao, et al.
Published: (2025)
by: Xu, Yixiao, et al.
Published: (2025)
Dual Utilization of Perturbation for Stream Data Publication under Local Differential Privacy
by: Du, Rong, et al.
Published: (2025)
by: Du, Rong, et al.
Published: (2025)
Robust and Imperceptible Black-box DNN Watermarking Based on Fourier Perturbation Analysis and Frequency Sensitivity Clustering
by: Liu, Yong, et al.
Published: (2022)
by: Liu, Yong, et al.
Published: (2022)
When There Is No Decoder: Removing Watermarks from Stable Diffusion Models in a No-box Setting
by: Wu, Xiaodong, et al.
Published: (2025)
by: Wu, Xiaodong, et al.
Published: (2025)
Watermarks Attack Watermarks: Re-Watermarking as a Generic Removal Strategy
by: Bulychev, Maria, et al.
Published: (2026)
by: Bulychev, Maria, et al.
Published: (2026)
Interactive Trimming against Evasive Online Data Manipulation Attacks: A Game-Theoretic Approach
by: Fu, Yue, et al.
Published: (2024)
by: Fu, Yue, et al.
Published: (2024)
NSmark: Null Space Based Black-box Watermarking Defense Framework for Language Models
by: Zhao, Haodong, et al.
Published: (2024)
by: Zhao, Haodong, et al.
Published: (2024)
Reliable Model Watermarking: Defending Against Theft without Compromising on Evasion
by: Zhu, Hongyu, et al.
Published: (2024)
by: Zhu, Hongyu, et al.
Published: (2024)
Can a Single Message Paralyze the AI Infrastructure? The Rise of AbO-DDoS Attacks through Targeted Mobius Injection
by: Liang, Zi, et al.
Published: (2026)
by: Liang, Zi, et al.
Published: (2026)
Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking
by: You, Ziyang, et al.
Published: (2026)
by: You, Ziyang, et al.
Published: (2026)
Yours or Mine? Overwriting Attacks Against Neural Audio Watermarking
by: Yao, Lingfeng, et al.
Published: (2025)
by: Yao, Lingfeng, et al.
Published: (2025)
Autoregressive Images Watermarking through Lexical Biasing: An Approach Resistant to Regeneration Attack
by: Hui, Siqi, et al.
Published: (2025)
by: Hui, Siqi, et al.
Published: (2025)
Similar Items
-
A Sample-Level Evaluation and Generative Framework for Model Inversion Attacks
by: Li, Haoyang, et al.
Published: (2025) -
Why Are My Prompts Leaked? Unraveling Prompt Extraction Threats in Customized Large Language Models
by: Liang, Zi, et al.
Published: (2024) -
MER-Inspector: Assessing model extraction risks from an attack-agnostic perspective
by: Zhang, Xinwei, et al.
Published: (2025) -
Performance-lossless Black-box Model Watermarking
by: Zhao, Na, et al.
Published: (2023) -
Does Low Rank Adaptation Lead to Lower Robustness against Training-Time Attacks?
by: Liang, Zi, et al.
Published: (2025)