ModelShield: Adaptive and Robust Watermark against Model Extraction Attack
Fuente:
arXiv
Saved in:
| Main Authors: | Pang, Kaiyi, Qi, Tao, Wu, Chuhan, Bai, Minhao, Jiang, Minghu, Huang, Yongfeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learnable Linguistic Watermarks for Tracing Model Extraction Attacks on Large Language Models
by: Bai, Minhao, et al.
Published: (2024)
by: Bai, Minhao, et al.
Published: (2024)
Semantic Steganography: A Framework for Robust and High-Capacity Information Hiding using Large Language Models
by: Bai, Minhao, et al.
Published: (2024)
by: Bai, Minhao, et al.
Published: (2024)
Shifting-Merging: Secure, High-Capacity and Efficient Steganography via Large Language Models
by: Bai, Minhao, et al.
Published: (2025)
by: Bai, Minhao, et al.
Published: (2025)
Let Watermarks Speak: A Robust and Unforgeable Watermark for Language Models
by: Bai, Minhao
Published: (2024)
by: Bai, Minhao
Published: (2024)
Provable Secure Steganography Based on Adaptive Dynamic Sampling
by: Pang, Kaiyi, et al.
Published: (2025)
by: Pang, Kaiyi, et al.
Published: (2025)
Provably Secure Steganography Based on List Decoding
by: Pang, Kaiyi, et al.
Published: (2026)
by: Pang, Kaiyi, et al.
Published: (2026)
Provably Robust and Secure Steganography in Asymmetric Resource Scenario
by: Bai, Minhao, et al.
Published: (2024)
by: Bai, Minhao, et al.
Published: (2024)
Towards Next-Generation Steganalysis: LLMs Unleash the Power of Detecting Steganography
by: Yang, Minhao Bai. Jinshuai, et al.
Published: (2024)
by: Yang, Minhao Bai. Jinshuai, et al.
Published: (2024)
MEA-Defender: A Robust Watermark against Model Extraction Attack
by: Lv, Peizhuo, et al.
Published: (2024)
by: Lv, Peizhuo, et al.
Published: (2024)
Optimizing Adaptive Attacks against Watermarks for Language Models
by: Diaa, Abdulrahman, et al.
Published: (2024)
by: Diaa, Abdulrahman, et al.
Published: (2024)
Neural Honeytrace: Plug&Play Watermarking Framework against Model Extraction Attacks
by: Xu, Yixiao, et al.
Published: (2025)
by: Xu, Yixiao, et al.
Published: (2025)
External Data Extraction Attacks against Retrieval-Augmented Large Language Models
by: He, Yu, et al.
Published: (2025)
by: He, Yu, et al.
Published: (2025)
A Watermark-Conditioned Diffusion Model for IP Protection
by: Min, Rui, et al.
Published: (2024)
by: Min, Rui, et al.
Published: (2024)
A Plug-and-Play Method for Improving Imperceptibility and Capacity in Practical Generative Text Steganography
by: Pang, Kaiyi
Published: (2024)
by: Pang, Kaiyi
Published: (2024)
Bounding-box Watermarking: Defense against Model Extraction Attacks on Object Detectors
by: Koda, Satoru, et al.
Published: (2024)
by: Koda, Satoru, et al.
Published: (2024)
Model Extraction Attacks Revisited
by: Liang, Jiacheng, et al.
Published: (2023)
by: Liang, Jiacheng, et al.
Published: (2023)
A Robust Semantics-based Watermark for Large Language Model against Paraphrasing
by: Ren, Jie, et al.
Published: (2023)
by: Ren, Jie, et al.
Published: (2023)
SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking
by: Yao, Lingfeng, et al.
Published: (2025)
by: Yao, Lingfeng, et al.
Published: (2025)
Black-box Membership Inference Attacks against Fine-tuned Diffusion Models
by: Pang, Yan, et al.
Published: (2023)
by: Pang, Yan, et al.
Published: (2023)
Robustness of Watermarking on Text-to-Image Diffusion Models
by: Wu, Xiaodong, et al.
Published: (2024)
by: Wu, Xiaodong, et al.
Published: (2024)
White-box Membership Inference Attacks against Diffusion Models
by: Pang, Yan, et al.
Published: (2023)
by: Pang, Yan, et al.
Published: (2023)
MelShield: Robust Mel-Domain Audio Watermarking for Provenance Attribution of AI Generated Synthesized Speech
by: Jin, Yutong, et al.
Published: (2026)
by: Jin, Yutong, et al.
Published: (2026)
Enhancing Watermarking Quality for LLMs via Contextual Generation States Awareness
by: Yang, Peiru, et al.
Published: (2025)
by: Yang, Peiru, et al.
Published: (2025)
Robust-Wide: Robust Watermarking against Instruction-driven Image Editing
by: Hu, Runyi, et al.
Published: (2024)
by: Hu, Runyi, et al.
Published: (2024)
DivQAT: Enhancing Robustness of Quantized Convolutional Neural Networks against Model Extraction Attacks
by: Khaled, Kacem, et al.
Published: (2025)
by: Khaled, Kacem, et al.
Published: (2025)
MorphMark: Flexible Adaptive Watermarking for Large Language Models
by: Wang, Zongqi, et al.
Published: (2025)
by: Wang, Zongqi, et al.
Published: (2025)
Denial-of-Service Poisoning Attacks against Large Language Models
by: Gao, Kuofeng, et al.
Published: (2024)
by: Gao, Kuofeng, et al.
Published: (2024)
Provably Robust Explainable Graph Neural Networks against Graph Perturbation Attacks
by: Li, Jiate, et al.
Published: (2025)
by: Li, Jiate, et al.
Published: (2025)
Adaptive and Robust Watermark for Generative Tabular Data
by: Ngo, Dung Daniel, et al.
Published: (2024)
by: Ngo, Dung Daniel, et al.
Published: (2024)
RLCracker: Evaluating the Worst-Case Vulnerability of LLM Watermarks with Adaptive RL Attacks
by: Huang, Hanbo, et al.
Published: (2025)
by: Huang, Hanbo, et al.
Published: (2025)
EnCAgg: Enhanced Clustering Aggregation for Robust Federated Learning against Dynamic Model Poisoning
by: Zhang, Tianyun, et al.
Published: (2026)
by: Zhang, Tianyun, et al.
Published: (2026)
VideoMark: A Distortion-Free Robust Watermarking Framework for Video Diffusion Models
by: Hu, Xuming, et al.
Published: (2025)
by: Hu, Xuming, et al.
Published: (2025)
An Ensemble Framework for Unbiased Language Model Watermarking
by: Wu, Yihan, et al.
Published: (2025)
by: Wu, Yihan, et al.
Published: (2025)
Analyzing and Evaluating Unbiased Language Model Watermark
by: Wu, Yihan, et al.
Published: (2025)
by: Wu, Yihan, et al.
Published: (2025)
ShieldMMU: Detecting and Defending against Controlled-Channel Attacks in Shielding Memory System
by: Liu, Gang, et al.
Published: (2025)
by: Liu, Gang, et al.
Published: (2025)
DiffusionShield: A Watermark for Copyright Protection against Generative Diffusion Models
by: Cui, Yingqian, et al.
Published: (2023)
by: Cui, Yingqian, et al.
Published: (2023)
Vanishing Watermarks: Diffusion-Based Image Editing Undermines Robust Invisible Watermarking
by: Guo, Fan, et al.
Published: (2026)
by: Guo, Fan, et al.
Published: (2026)
AdaShield: Safeguarding Multimodal Large Language Models from Structure-based Attack via Adaptive Shield Prompting
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
Leveraging Optimization for Adaptive Attacks on Image Watermarks
by: Lukas, Nils, et al.
Published: (2023)
by: Lukas, Nils, et al.
Published: (2023)
System Prompt Extraction Attacks and Defenses in Large Language Models
by: Das, Badhan Chandra, et al.
Published: (2025)
by: Das, Badhan Chandra, et al.
Published: (2025)
Similar Items
-
Learnable Linguistic Watermarks for Tracing Model Extraction Attacks on Large Language Models
by: Bai, Minhao, et al.
Published: (2024) -
Semantic Steganography: A Framework for Robust and High-Capacity Information Hiding using Large Language Models
by: Bai, Minhao, et al.
Published: (2024) -
Shifting-Merging: Secure, High-Capacity and Efficient Steganography via Large Language Models
by: Bai, Minhao, et al.
Published: (2025) -
Let Watermarks Speak: A Robust and Unforgeable Watermark for Language Models
by: Bai, Minhao
Published: (2024) -
Provable Secure Steganography Based on Adaptive Dynamic Sampling
by: Pang, Kaiyi, et al.
Published: (2025)