Detecting Benchmark Contamination Through Watermarking
Fuente:
arXiv
Saved in:
| Main Authors: | Sander, Tom, Fernandez, Pierre, Mahloujifar, Saeed, Durmus, Alain, Guo, Chuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Watermarking Makes Language Models Radioactive
by: Sander, Tom, et al.
Published: (2024)
by: Sander, Tom, et al.
Published: (2024)
Watermark Anything with Localized Messages
by: Sander, Tom, et al.
Published: (2024)
by: Sander, Tom, et al.
Published: (2024)
Proactive Detection of Voice Cloning with Localized Watermarking
by: Roman, Robin San, et al.
Published: (2024)
by: Roman, Robin San, et al.
Published: (2024)
Watermarking across Modalities for Content Tracing and Generative AI
by: Fernandez, Pierre
Published: (2025)
by: Fernandez, Pierre
Published: (2025)
A New Linear Scaling Rule for Private Adaptive Hyperparameter Optimization
by: Panda, Ashwinee, et al.
Published: (2022)
by: Panda, Ashwinee, et al.
Published: (2022)
On The Fragility of Benchmark Contamination Detection in Reasoning Models
by: Wang, Han, et al.
Published: (2025)
by: Wang, Han, et al.
Published: (2025)
Learning to Watermark in the Latent Space of Generative Models
by: Rebuffi, Sylvestre-Alvise, et al.
Published: (2026)
by: Rebuffi, Sylvestre-Alvise, et al.
Published: (2026)
Functional Invariants to Watermark Large Transformers
by: Fernandez, Pierre, et al.
Published: (2023)
by: Fernandez, Pierre, et al.
Published: (2023)
Modification and Generated-Text Detection: Achieving Dual Detection Capabilities for the Outputs of LLM by Watermark
by: Cai, Yuhang, et al.
Published: (2025)
by: Cai, Yuhang, et al.
Published: (2025)
WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks
by: Evtimov, Ivan, et al.
Published: (2025)
by: Evtimov, Ivan, et al.
Published: (2025)
Invariant-based Robust Weights Watermark for Large Language Models
by: Guo, Qingxiao, et al.
Published: (2025)
by: Guo, Qingxiao, et al.
Published: (2025)
MetaSeal: Defending Against Image Attribution Forgery Through Content-Dependent Cryptographic Watermarks
by: Zhou, Tong, et al.
Published: (2025)
by: Zhou, Tong, et al.
Published: (2025)
Publicly-Detectable Watermarking for Language Models
by: Fairoze, Jaiden, et al.
Published: (2023)
by: Fairoze, Jaiden, et al.
Published: (2023)
Coward: Collision-based OOD Watermarking for Practical Proactive Federated Backdoor Detection
by: Li, Wenjie, et al.
Published: (2025)
by: Li, Wenjie, et al.
Published: (2025)
Marking Code Without Breaking It: Code Watermarking for Detecting LLM-Generated Code
by: Kim, Jungin, et al.
Published: (2025)
by: Kim, Jungin, et al.
Published: (2025)
LLM Benchmark Datasets Should Be Contamination-Resistant
by: Al-Lawati, Ali, et al.
Published: (2026)
by: Al-Lawati, Ali, et al.
Published: (2026)
SoK: How Robust is Audio Watermarking in Generative AI models?
by: Wen, Yizhu, et al.
Published: (2025)
by: Wen, Yizhu, et al.
Published: (2025)
Machine Learning with Privacy for Protected Attributes
by: Mahloujifar, Saeed, et al.
Published: (2025)
by: Mahloujifar, Saeed, et al.
Published: (2025)
Your Semantic-Independent Watermark is Fragile: A Semantic Perturbation Attack against EaaS Watermark
by: Fei, Zekun, et al.
Published: (2024)
by: Fei, Zekun, et al.
Published: (2024)
Learning to Watermark: A Selective Watermarking Framework for Large Language Models via Multi-Objective Optimization
by: Wang, Chenrui, et al.
Published: (2025)
by: Wang, Chenrui, et al.
Published: (2025)
Sequential Behavioral Watermarking for LLM Agents
by: An, Hyeseon, et al.
Published: (2026)
by: An, Hyeseon, et al.
Published: (2026)
Probabilistically Robust Watermarking of Neural Networks
by: Pautov, Mikhail, et al.
Published: (2024)
by: Pautov, Mikhail, et al.
Published: (2024)
A Survey of Fragile Model Watermarking
by: Gao, Zhenzhe, et al.
Published: (2024)
by: Gao, Zhenzhe, et al.
Published: (2024)
The Coding Limits of Robust Watermarking for Generative Models
by: Francati, Danilo, et al.
Published: (2025)
by: Francati, Danilo, et al.
Published: (2025)
Character-Level Perturbations Disrupt LLM Watermarks
by: Zhang, Zhaoxi, et al.
Published: (2025)
by: Zhang, Zhaoxi, et al.
Published: (2025)
CODE ACROSTIC: Robust Watermarking for Code Generation
by: Lin, Li, et al.
Published: (2025)
by: Lin, Li, et al.
Published: (2025)
Watermark Overwriting Attack on StegaStamp algorithm
by: Serzhenko, I. F., et al.
Published: (2025)
by: Serzhenko, I. F., et al.
Published: (2025)
LLM Watermark Evasion via Bias Inversion
by: Hwang, Jeongyeon, et al.
Published: (2025)
by: Hwang, Jeongyeon, et al.
Published: (2025)
Asking Back: Interaction-Layer Antidistillation Watermarks
by: Yang, Guang, et al.
Published: (2026)
by: Yang, Guang, et al.
Published: (2026)
Multi-Designated Detector Watermarking for Language Models
by: Huang, Zhengan, et al.
Published: (2024)
by: Huang, Zhengan, et al.
Published: (2024)
Functional Subspace Watermarking for Large Language Models
by: Ding, Zikang, et al.
Published: (2026)
by: Ding, Zikang, et al.
Published: (2026)
Privacy Amplification for the Gaussian Mechanism via Bounded Support
by: Hu, Shengyuan, et al.
Published: (2024)
by: Hu, Shengyuan, et al.
Published: (2024)
Sandcastles in the Storm: Revisiting the (Im)possibility of Strong Watermarking
by: Harel-Canada, Fabrice Y, et al.
Published: (2025)
by: Harel-Canada, Fabrice Y, et al.
Published: (2025)
The Feasibility of Topic-Based Watermarking on Academic Peer Reviews
by: Nemecek, Alexander, et al.
Published: (2025)
by: Nemecek, Alexander, et al.
Published: (2025)
ChainMarks: Securing DNN Watermark with Cryptographic Chain
by: Choi, Brian, et al.
Published: (2025)
by: Choi, Brian, et al.
Published: (2025)
Robust Client-Server Watermarking for Split Federated Learning
by: Tang, Jiaxiong, et al.
Published: (2025)
by: Tang, Jiaxiong, et al.
Published: (2025)
Optimizing Adaptive Attacks against Watermarks for Language Models
by: Diaa, Abdulrahman, et al.
Published: (2024)
by: Diaa, Abdulrahman, et al.
Published: (2024)
On Protecting Agentic Systems' Intellectual Property via Watermarking
by: Wang, Liwen, et al.
Published: (2026)
by: Wang, Liwen, et al.
Published: (2026)
AgentMark: Utility-Preserving Behavioral Watermarking for Agents
by: Huang, Kaibo, et al.
Published: (2026)
by: Huang, Kaibo, et al.
Published: (2026)
Investigating Deep Watermark Security: An Adversarial Transferability Perspective
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Similar Items
-
Watermarking Makes Language Models Radioactive
by: Sander, Tom, et al.
Published: (2024) -
Watermark Anything with Localized Messages
by: Sander, Tom, et al.
Published: (2024) -
Proactive Detection of Voice Cloning with Localized Watermarking
by: Roman, Robin San, et al.
Published: (2024) -
Watermarking across Modalities for Content Tracing and Generative AI
by: Fernandez, Pierre
Published: (2025) -
A New Linear Scaling Rule for Private Adaptive Hyperparameter Optimization
by: Panda, Ashwinee, et al.
Published: (2022)