Transferable Black-Box One-Shot Forging of Watermarks via Image Preference Models
Fuente:
arXiv
Saved in:
| Main Authors: | Souček, Tomáš, Rebuffi, Sylvestre-Alvise, Fernandez, Pierre, Jovanović, Nikola, Elsahar, Hady, Lacatusu, Valeriu, Tran, Tuan, Mourachko, Alexandre |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Watermark in the Latent Space of Generative Models
by: Rebuffi, Sylvestre-Alvise, et al.
Published: (2026)
by: Rebuffi, Sylvestre-Alvise, et al.
Published: (2026)
Geometric Image Synchronization with Deep Watermarking
by: Fernandez, Pierre, et al.
Published: (2025)
by: Fernandez, Pierre, et al.
Published: (2025)
How Good is Post-Hoc Watermarking With Language Model Rephrasing?
by: Fernandez, Pierre, et al.
Published: (2025)
by: Fernandez, Pierre, et al.
Published: (2025)
Pixel Seal: Adversarial-only training for invisible image and video watermarking
by: Souček, Tomáš, et al.
Published: (2025)
by: Souček, Tomáš, et al.
Published: (2025)
TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection
by: Sander, Tom, et al.
Published: (2026)
by: Sander, Tom, et al.
Published: (2026)
We Can Hide More Bits: The Unused Watermarking Capacity in Theory and in Practice
by: Petrov, Aleksandar, et al.
Published: (2025)
by: Petrov, Aleksandar, et al.
Published: (2025)
Proactive Detection of Voice Cloning with Localized Watermarking
by: Roman, Robin San, et al.
Published: (2024)
by: Roman, Robin San, et al.
Published: (2024)
Watermarking Autoregressive Image Generation
by: Jovanović, Nikola, et al.
Published: (2025)
by: Jovanović, Nikola, et al.
Published: (2025)
Video Seal: Open and Efficient Video Watermarking
by: Fernandez, Pierre, et al.
Published: (2024)
by: Fernandez, Pierre, et al.
Published: (2024)
Black-Box Detection of Language Model Watermarks
by: Gloaguen, Thibaud, et al.
Published: (2024)
by: Gloaguen, Thibaud, et al.
Published: (2024)
Watermark Stealing in Large Language Models
by: Jovanović, Nikola, et al.
Published: (2024)
by: Jovanović, Nikola, et al.
Published: (2024)
Towards Watermarking of Open-Source LLMs
by: Gloaguen, Thibaud, et al.
Published: (2025)
by: Gloaguen, Thibaud, et al.
Published: (2025)
LLM Fingerprinting via Semantically Conditioned Watermarks
by: Gloaguen, Thibaud, et al.
Published: (2025)
by: Gloaguen, Thibaud, et al.
Published: (2025)
Watermarking Diffusion Language Models
by: Gloaguen, Thibaud, et al.
Published: (2025)
by: Gloaguen, Thibaud, et al.
Published: (2025)
A Unified Framework for LLM Watermarks
by: Gloaguen, Thibaud, et al.
Published: (2026)
by: Gloaguen, Thibaud, et al.
Published: (2026)
Discovering Spoofing Attempts on Language Model Watermarks
by: Gloaguen, Thibaud, et al.
Published: (2024)
by: Gloaguen, Thibaud, et al.
Published: (2024)
Ward: Provable RAG Dataset Inference via LLM Watermarks
by: Jovanović, Nikola, et al.
Published: (2024)
by: Jovanović, Nikola, et al.
Published: (2024)
A Watermark for Black-Box Language Models
by: Bahri, Dara, et al.
Published: (2024)
by: Bahri, Dara, et al.
Published: (2024)
ComMark: Covert and Robust Black-Box Model Watermarking with Compressed Samples
by: Yang, Yunfei, et al.
Published: (2025)
by: Yang, Yunfei, et al.
Published: (2025)
Cross-Lingual Summarization as a Black-Box Watermark Removal Attack
by: Ganesan, Gokul
Published: (2025)
by: Ganesan, Gokul
Published: (2025)
Towards Traitor Tracing in Black-and-White-Box DNN Watermarking with Tardos-based Codes
by: Rodriguez-Lois, Elena, et al.
Published: (2023)
by: Rodriguez-Lois, Elena, et al.
Published: (2023)
Forging the Unforgeable: On the Feasibility of Counterfeit Watermarks in Backdoor-Based Dataset Ownership Verification
by: Li, Zhiying, et al.
Published: (2024)
by: Li, Zhiying, et al.
Published: (2024)
WMCopier: Forging Invisible Image Watermarks on Arbitrary Images
by: Dong, Ziping, et al.
Published: (2025)
by: Dong, Ziping, et al.
Published: (2025)
Hashed Watermark as a Filter: Defeating Forging and Overwriting Attacks in Weight-based Neural Network Watermarking
by: Yao, Yuan, et al.
Published: (2025)
by: Yao, Yuan, et al.
Published: (2025)
SLIM: Stealthy Low-Coverage Black-Box Watermarking via Latent-Space Confusion Zones
by: Wu, Hengyu, et al.
Published: (2026)
by: Wu, Hengyu, et al.
Published: (2026)
Black-Box Forgery Attacks on Semantic Watermarks for Diffusion Models
by: Müller, Andreas, et al.
Published: (2024)
by: Müller, Andreas, et al.
Published: (2024)
Quantifying the Risk of Transferred Black Box Attacks
by: Cox, Disesdi Susanna, et al.
Published: (2025)
by: Cox, Disesdi Susanna, et al.
Published: (2025)
RAG-WM: An Efficient Black-Box Watermarking Approach for Retrieval-Augmented Generation of Large Language Models
by: Lv, Peizhuo, et al.
Published: (2025)
by: Lv, Peizhuo, et al.
Published: (2025)
ArmSSL: Adversarial Robust Black-Box Watermarking for Self-Supervised Learning Pre-trained Encoders
by: Jiang, Yongqi, et al.
Published: (2026)
by: Jiang, Yongqi, et al.
Published: (2026)
Rethinking LLM Watermark Detection in Black-Box Settings: A Non-Intrusive Third-Party Framework
by: Wang, Zhuoshang, et al.
Published: (2026)
by: Wang, Zhuoshang, et al.
Published: (2026)
SSL-WM: A Black-Box Watermarking Approach for Encoders Pre-trained by Self-supervised Learning
by: Lv, Peizhuo, et al.
Published: (2022)
by: Lv, Peizhuo, et al.
Published: (2022)
A4O: All Trigger for One sample
by: Vu, Duc Anh, et al.
Published: (2025)
by: Vu, Duc Anh, et al.
Published: (2025)
SemBind: Binding Diffusion Watermarks to Semantics Against Black-Box Forgery Attacks
by: Zhang, Xin, et al.
Published: (2026)
by: Zhang, Xin, et al.
Published: (2026)
DeMark: A Query-Free Black-Box Attack on Deepfake Watermarking Defenses
by: Song, Wei, et al.
Published: (2026)
by: Song, Wei, et al.
Published: (2026)
Why Neural Structural Obfuscation Can't Kill White-Box Watermarks for Good!
by: Jiang, Yanna, et al.
Published: (2026)
by: Jiang, Yanna, et al.
Published: (2026)
Fading the Digital Ink: A Universal Black-Box Attack Framework for 3DGS Watermarking Systems
by: Zeng, Qingyuan, et al.
Published: (2025)
by: Zeng, Qingyuan, et al.
Published: (2025)
Performance-lossless Black-box Model Watermarking
by: Zhao, Na, et al.
Published: (2023)
by: Zhao, Na, et al.
Published: (2023)
One Prompt to Verify Your Models: Black-Box Text-to-Image Models Verification via Non-Transferable Adversarial Attacks
by: Guo, Ji, et al.
Published: (2024)
by: Guo, Ji, et al.
Published: (2024)
DeepEclipse: How to Break White-Box DNN-Watermarking Schemes
by: Pegoraro, Alessandro, et al.
Published: (2024)
by: Pegoraro, Alessandro, et al.
Published: (2024)
Removing Box-Free Watermarks for Image-to-Image Models via Query-Based Reverse Engineering
by: An, Haonan, et al.
Published: (2025)
by: An, Haonan, et al.
Published: (2025)
Similar Items
-
Learning to Watermark in the Latent Space of Generative Models
by: Rebuffi, Sylvestre-Alvise, et al.
Published: (2026) -
Geometric Image Synchronization with Deep Watermarking
by: Fernandez, Pierre, et al.
Published: (2025) -
How Good is Post-Hoc Watermarking With Language Model Rephrasing?
by: Fernandez, Pierre, et al.
Published: (2025) -
Pixel Seal: Adversarial-only training for invisible image and video watermarking
by: Souček, Tomáš, et al.
Published: (2025) -
TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection
by: Sander, Tom, et al.
Published: (2026)