Neural Dehydration: Effective Erasure of Black-box Watermarks from DNNs with Limited Data
Fuente:
arXiv
Salvato in:
| Autori principali: | Lu, Yifan, Li, Wenxuan, Zhang, Mi, Pan, Xudong, Yang, Min |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
WGLE:Backdoor-free and Multi-bit Black-box Watermarking for Graph Neural Networks
di: Li, Tingzhi, et al.
Pubblicazione: (2025)
di: Li, Tingzhi, et al.
Pubblicazione: (2025)
AGATE: Stealthy Black-box Watermarking for Multimodal Model Copyright Protection
di: Gao, Jianbo, et al.
Pubblicazione: (2025)
di: Gao, Jianbo, et al.
Pubblicazione: (2025)
BELT: Old-School Backdoor Attacks can Evade the State-of-the-Art Defense with Backdoor Exclusivity Lifting
di: Qiu, Huming, et al.
Pubblicazione: (2023)
di: Qiu, Huming, et al.
Pubblicazione: (2023)
StruPhantom: Evolutionary Injection Attacks on Black-Box Tabular Agents Powered by Large Language Models
di: Feng, Yang, et al.
Pubblicazione: (2025)
di: Feng, Yang, et al.
Pubblicazione: (2025)
NSmark: Null Space Based Black-box Watermarking Defense Framework for Language Models
di: Zhao, Haodong, et al.
Pubblicazione: (2024)
di: Zhao, Haodong, et al.
Pubblicazione: (2024)
MEraser: An Effective Fingerprint Erasure Approach for Large Language Models
di: Zhang, Jingxuan, et al.
Pubblicazione: (2025)
di: Zhang, Jingxuan, et al.
Pubblicazione: (2025)
AutoControl Arena: Synthesizing Executable Test Environments for Frontier AI Risk Evaluation
di: Li, Changyi, et al.
Pubblicazione: (2026)
di: Li, Changyi, et al.
Pubblicazione: (2026)
SEAL: Entangled White-box Watermarks on Low-Rank Adaptation
di: Oh, Giyeong, et al.
Pubblicazione: (2025)
di: Oh, Giyeong, et al.
Pubblicazione: (2025)
SEW: Strengthening Robustness of Black-box DNN Watermarking via Specificity Enhancement
di: Qiu, Huming, et al.
Pubblicazione: (2026)
di: Qiu, Huming, et al.
Pubblicazione: (2026)
Invisible Threats from Model Context Protocol: Generating Stealthy Injection Payload via Tree-based Adaptive Search
di: Shen, Yulin, et al.
Pubblicazione: (2026)
di: Shen, Yulin, et al.
Pubblicazione: (2026)
Understanding Byzantine Robustness in Federated Learning with A Black-box Server
di: Zhao, Fangyuan, et al.
Pubblicazione: (2024)
di: Zhao, Fangyuan, et al.
Pubblicazione: (2024)
SentinelNet: Safeguarding Multi-Agent Collaboration Through Credit-Based Dynamic Threat Detection
di: Feng, Yang, et al.
Pubblicazione: (2025)
di: Feng, Yang, et al.
Pubblicazione: (2025)
The Coding Limits of Robust Watermarking for Generative Models
di: Francati, Danilo, et al.
Pubblicazione: (2025)
di: Francati, Danilo, et al.
Pubblicazione: (2025)
Revisiting Backdoor Attacks on Time Series Classification in the Frequency Domain
di: Huang, Yuanmin, et al.
Pubblicazione: (2025)
di: Huang, Yuanmin, et al.
Pubblicazione: (2025)
Learning to Watermark: A Selective Watermarking Framework for Large Language Models via Multi-Objective Optimization
di: Wang, Chenrui, et al.
Pubblicazione: (2025)
di: Wang, Chenrui, et al.
Pubblicazione: (2025)
SSL-WM: A Black-Box Watermarking Approach for Encoders Pre-trained by Self-supervised Learning
di: Lv, Peizhuo, et al.
Pubblicazione: (2022)
di: Lv, Peizhuo, et al.
Pubblicazione: (2022)
Probabilistically Robust Watermarking of Neural Networks
di: Pautov, Mikhail, et al.
Pubblicazione: (2024)
di: Pautov, Mikhail, et al.
Pubblicazione: (2024)
LLMPirate: LLMs for Black-box Hardware IP Piracy
di: Gohil, Vasudev, et al.
Pubblicazione: (2024)
di: Gohil, Vasudev, et al.
Pubblicazione: (2024)
HogVul: Black-box Adversarial Code Generation Framework Against LM-based Vulnerability Detectors
di: Yang, Jingxiao, et al.
Pubblicazione: (2026)
di: Yang, Jingxiao, et al.
Pubblicazione: (2026)
Character-Level Perturbations Disrupt LLM Watermarks
di: Zhang, Zhaoxi, et al.
Pubblicazione: (2025)
di: Zhang, Zhaoxi, et al.
Pubblicazione: (2025)
Multi-task Adversarial Attacks against Black-box Model with Few-shot Queries
di: Wang, Wenqiang, et al.
Pubblicazione: (2025)
di: Wang, Wenqiang, et al.
Pubblicazione: (2025)
Large Language Model Watermark Stealing With Mixed Integer Programming
di: Zhang, Zhaoxi, et al.
Pubblicazione: (2024)
di: Zhang, Zhaoxi, et al.
Pubblicazione: (2024)
Revisiting the Information Capacity of Neural Network Watermarks: Upper Bound Estimation and Beyond
di: Li, Fangqi, et al.
Pubblicazione: (2024)
di: Li, Fangqi, et al.
Pubblicazione: (2024)
AgentTypo: Adaptive Typographic Prompt Injection Attacks against Black-box Multimodal Agents
di: Li, Yanjie, et al.
Pubblicazione: (2025)
di: Li, Yanjie, et al.
Pubblicazione: (2025)
KG-DF: A Black-box Defense Framework against Jailbreak Attacks Based on Knowledge Graphs
di: Liu, Shuyuan, et al.
Pubblicazione: (2025)
di: Liu, Shuyuan, et al.
Pubblicazione: (2025)
CODE ACROSTIC: Robust Watermarking for Code Generation
di: Lin, Li, et al.
Pubblicazione: (2025)
di: Lin, Li, et al.
Pubblicazione: (2025)
CyberEvolver: Structured Self-Evolution for Cybersecurity Agents On the Fly
di: Fan, Yihe, et al.
Pubblicazione: (2026)
di: Fan, Yihe, et al.
Pubblicazione: (2026)
Neural Honeytrace: Plug&Play Watermarking Framework against Model Extraction Attacks
di: Xu, Yixiao, et al.
Pubblicazione: (2025)
di: Xu, Yixiao, et al.
Pubblicazione: (2025)
A General Black-box Adversarial Attack on Graph-based Fake News Detectors
di: Zhu, Peican, et al.
Pubblicazione: (2024)
di: Zhu, Peican, et al.
Pubblicazione: (2024)
Watermarking Graph Neural Networks via Explanations for Ownership Protection
di: Downer, Jane, et al.
Pubblicazione: (2025)
di: Downer, Jane, et al.
Pubblicazione: (2025)
Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking
di: You, Ziyang, et al.
Pubblicazione: (2026)
di: You, Ziyang, et al.
Pubblicazione: (2026)
Hot-Swap MarkBoard: An Efficient Black-box Watermarking Approach for Large-scale Model Distribution
di: Zhang, Zhicheng, et al.
Pubblicazione: (2025)
di: Zhang, Zhicheng, et al.
Pubblicazione: (2025)
JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs
di: Li, Hongyi, et al.
Pubblicazione: (2024)
di: Li, Hongyi, et al.
Pubblicazione: (2024)
RAG-WM: An Efficient Black-Box Watermarking Approach for Retrieval-Augmented Generation of Large Language Models
di: Lv, Peizhuo, et al.
Pubblicazione: (2025)
di: Lv, Peizhuo, et al.
Pubblicazione: (2025)
FBA$^2$D: Frequency-based Black-box Attack for AI-generated Image Detection
di: Chen, Xiaojing, et al.
Pubblicazione: (2025)
di: Chen, Xiaojing, et al.
Pubblicazione: (2025)
AgentMark: Utility-Preserving Behavioral Watermarking for Agents
di: Huang, Kaibo, et al.
Pubblicazione: (2026)
di: Huang, Kaibo, et al.
Pubblicazione: (2026)
WebTrap Park: An Automated Platform for Systematic Security Evaluation of Web Agents
di: Wu, Xinyi, et al.
Pubblicazione: (2026)
di: Wu, Xinyi, et al.
Pubblicazione: (2026)
ArmSSL: Adversarial Robust Black-Box Watermarking for Self-Supervised Learning Pre-trained Encoders
di: Jiang, Yongqi, et al.
Pubblicazione: (2026)
di: Jiang, Yongqi, et al.
Pubblicazione: (2026)
Protecting Deep Neural Network Intellectual Property with Chaos-Based White-Box Watermarking
di: B, Sangeeth, et al.
Pubblicazione: (2025)
di: B, Sangeeth, et al.
Pubblicazione: (2025)
Injection, Attack and Erasure: Revocable Backdoor Attacks via Machine Unlearning
di: Song, Baogang, et al.
Pubblicazione: (2025)
di: Song, Baogang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
WGLE:Backdoor-free and Multi-bit Black-box Watermarking for Graph Neural Networks
di: Li, Tingzhi, et al.
Pubblicazione: (2025) -
AGATE: Stealthy Black-box Watermarking for Multimodal Model Copyright Protection
di: Gao, Jianbo, et al.
Pubblicazione: (2025) -
BELT: Old-School Backdoor Attacks can Evade the State-of-the-Art Defense with Backdoor Exclusivity Lifting
di: Qiu, Huming, et al.
Pubblicazione: (2023) -
StruPhantom: Evolutionary Injection Attacks on Black-Box Tabular Agents Powered by Large Language Models
di: Feng, Yang, et al.
Pubblicazione: (2025) -
NSmark: Null Space Based Black-box Watermarking Defense Framework for Language Models
di: Zhao, Haodong, et al.
Pubblicazione: (2024)