On the Difficulty of Constructing a Robust and Publicly-Detectable Watermark
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fairoze, Jaiden, Ortiz-Jimenez, Guillermo, Vecerik, Mel, Jha, Somesh, Gowal, Sven |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Publicly-Detectable Watermarking for Language Models
von: Fairoze, Jaiden, et al.
Veröffentlicht: (2023)
von: Fairoze, Jaiden, et al.
Veröffentlicht: (2023)
Bypassing Prompt Guards in Production with Controlled-Release Prompting
von: Fairoze, Jaiden, et al.
Veröffentlicht: (2025)
von: Fairoze, Jaiden, et al.
Veröffentlicht: (2025)
SoK: Watermarking for AI-Generated Content
von: Zhao, Xuandong, et al.
Veröffentlicht: (2024)
von: Zhao, Xuandong, et al.
Veröffentlicht: (2024)
SLVR: Securely Leveraging Client Validation for Robust Federated Learning
von: Choi, Jihye, et al.
Veröffentlicht: (2025)
von: Choi, Jihye, et al.
Veröffentlicht: (2025)
How Not to Detect Prompt Injections with an LLM
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
Is The Watermarking Of LLM-Generated Code Robust?
von: Suresh, Tarun, et al.
Veröffentlicht: (2024)
von: Suresh, Tarun, et al.
Veröffentlicht: (2024)
Refined Detection for Gumbel Watermarking
von: Lattimore, Tor
Veröffentlicht: (2026)
von: Lattimore, Tor
Veröffentlicht: (2026)
Robust Spectral Watermark for Synthetic Tabular Data
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
DRGW: Learning Disentangled Representations for Robust Graph Watermarking
von: Li, Jiasen, et al.
Veröffentlicht: (2026)
von: Li, Jiasen, et al.
Veröffentlicht: (2026)
Pr$εε$mpt: Sanitizing Sensitive Prompts for LLMs
von: Chowdhury, Amrita Roy, et al.
Veröffentlicht: (2025)
von: Chowdhury, Amrita Roy, et al.
Veröffentlicht: (2025)
Black-Box Detection of Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
Robust GNN Watermarking via Implicit Perception of Topological Invariants
von: Li, Jipeng, et al.
Veröffentlicht: (2025)
von: Li, Jipeng, et al.
Veröffentlicht: (2025)
Functional Homotopy: Smoothing Discrete Optimization via Continuous Parameters for LLM Jailbreak Attacks
von: Wang, Zi, et al.
Veröffentlicht: (2024)
von: Wang, Zi, et al.
Veröffentlicht: (2024)
What Really is a Member? Discrediting Membership Inference via Poisoning
von: Mangaokar, Neal, et al.
Veröffentlicht: (2025)
von: Mangaokar, Neal, et al.
Veröffentlicht: (2025)
A Crack in the Bark: Leveraging Public Knowledge to Remove Tree-Ring Watermarks
von: Lin, Junhua, et al.
Veröffentlicht: (2025)
von: Lin, Junhua, et al.
Veröffentlicht: (2025)
MEA-Defender: A Robust Watermark against Model Extraction Attack
von: Lv, Peizhuo, et al.
Veröffentlicht: (2024)
von: Lv, Peizhuo, et al.
Veröffentlicht: (2024)
Watermark under Fire: A Robustness Evaluation of LLM Watermarking
von: Liang, Jiacheng, et al.
Veröffentlicht: (2024)
von: Liang, Jiacheng, et al.
Veröffentlicht: (2024)
Revisiting the Robustness of Watermarking to Paraphrasing Attacks
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024)
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024)
Distortion-free Watermarks are not Truly Distortion-free under Watermark Key Collisions
von: Wu, Yihan, et al.
Veröffentlicht: (2024)
von: Wu, Yihan, et al.
Veröffentlicht: (2024)
Hashed Watermark as a Filter: Defeating Forging and Overwriting Attacks in Weight-based Neural Network Watermarking
von: Yao, Yuan, et al.
Veröffentlicht: (2025)
von: Yao, Yuan, et al.
Veröffentlicht: (2025)
Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2026)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2026)
PVMark: Enabling Public Verifiability for LLM Watermarking Schemes
von: Duan, Haohua, et al.
Veröffentlicht: (2025)
von: Duan, Haohua, et al.
Veröffentlicht: (2025)
Robust Distortion-free Watermarks for Language Models
von: Kuditipudi, Rohith, et al.
Veröffentlicht: (2023)
von: Kuditipudi, Rohith, et al.
Veröffentlicht: (2023)
Watermarking Generative Categorical Data
von: Gu, Bochao, et al.
Veröffentlicht: (2024)
von: Gu, Bochao, et al.
Veröffentlicht: (2024)
Towards Watermarking of Open-Source LLMs
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Provable Watermarking for Data Poisoning Attacks
von: Zhu, Yifan, et al.
Veröffentlicht: (2025)
von: Zhu, Yifan, et al.
Veröffentlicht: (2025)
Signal Watermark on Large Language Models
von: Xu, Zhenyu, et al.
Veröffentlicht: (2024)
von: Xu, Zhenyu, et al.
Veröffentlicht: (2024)
Unforgeable Watermarks for Language Models via Robust Signatures
von: Lin, Huijia, et al.
Veröffentlicht: (2026)
von: Lin, Huijia, et al.
Veröffentlicht: (2026)
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing
von: Feng, Ryan, et al.
Veröffentlicht: (2022)
von: Feng, Ryan, et al.
Veröffentlicht: (2022)
Ideal Attribution and Faithful Watermarks for Language Models
von: Song, Min Jae, et al.
Veröffentlicht: (2025)
von: Song, Min Jae, et al.
Veröffentlicht: (2025)
Traceable Black-box Watermarks for Federated Learning
von: Xu, Jiahao, et al.
Veröffentlicht: (2025)
von: Xu, Jiahao, et al.
Veröffentlicht: (2025)
LLM Fingerprinting via Semantically Conditioned Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Leveraging Optimization for Adaptive Attacks on Image Watermarks
von: Lukas, Nils, et al.
Veröffentlicht: (2023)
von: Lukas, Nils, et al.
Veröffentlicht: (2023)
Robust Data Watermarking in Language Models by Injecting Fictitious Knowledge
von: Cui, Xinyue, et al.
Veröffentlicht: (2025)
von: Cui, Xinyue, et al.
Veröffentlicht: (2025)
Edit Distance Robust Watermarks via Indexing Pseudorandom Codes
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
Watermarking Counterfactual Explanations
von: Guo, Hangzhi, et al.
Veröffentlicht: (2024)
von: Guo, Hangzhi, et al.
Veröffentlicht: (2024)
Breaking Distortion-free Watermarks in Large Language Models
von: Reynolds, Shayleen, et al.
Veröffentlicht: (2025)
von: Reynolds, Shayleen, et al.
Veröffentlicht: (2025)
LLM Watermarking Using Mixtures and Statistical-to-Computational Gaps
von: Abdalla, Pedro, et al.
Veröffentlicht: (2025)
von: Abdalla, Pedro, et al.
Veröffentlicht: (2025)
GENIE: Watermarking Graph Neural Networks for Link Prediction
von: Bachina, Venkata Sai Pranav, et al.
Veröffentlicht: (2024)
von: Bachina, Venkata Sai Pranav, et al.
Veröffentlicht: (2024)
PRO: Enabling Precise and Robust Text Watermark for Open-Source LLMs
von: Xue, Jiaqi, et al.
Veröffentlicht: (2025)
von: Xue, Jiaqi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Publicly-Detectable Watermarking for Language Models
von: Fairoze, Jaiden, et al.
Veröffentlicht: (2023) -
Bypassing Prompt Guards in Production with Controlled-Release Prompting
von: Fairoze, Jaiden, et al.
Veröffentlicht: (2025) -
SoK: Watermarking for AI-Generated Content
von: Zhao, Xuandong, et al.
Veröffentlicht: (2024) -
SLVR: Securely Leveraging Client Validation for Robust Federated Learning
von: Choi, Jihye, et al.
Veröffentlicht: (2025) -
How Not to Detect Prompt Injections with an LLM
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)