Who Gets Flagged? The Pluralistic Evaluation Gap in AI Content Watermarking
Fuente:
arXiv
Saved in:
| Main Authors: | Nemecek, Alexander, Zafar, Osama, Xu, Yuqiao, Li, Wenbiao, Ayday, Erman |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Topic-Based Watermarks for Large Language Models
by: Nemecek, Alexander, et al.
Published: (2024)
by: Nemecek, Alexander, et al.
Published: (2024)
Authenticated Contradictions from Desynchronized Provenance and Watermarking
by: Nemecek, Alexander, et al.
Published: (2026)
by: Nemecek, Alexander, et al.
Published: (2026)
The End of Trust: How Agentic AI Breaks Security Assumptions
by: Zafar, Osama, et al.
Published: (2026)
by: Zafar, Osama, et al.
Published: (2026)
Watermarking Without Standards Is Not AI Governance
by: Nemecek, Alexander, et al.
Published: (2025)
by: Nemecek, Alexander, et al.
Published: (2025)
Digital Agriculture Sandbox for Collaborative Research
by: Zafar, Osama, et al.
Published: (2025)
by: Zafar, Osama, et al.
Published: (2025)
Privacy-Preserving Data Linkage Across Private and Public Datasets for Collaborative Agriculture Research
by: Zafar, Osama, et al.
Published: (2024)
by: Zafar, Osama, et al.
Published: (2024)
Cluster-Aware Attacks on Graph Watermarks
by: Nemecek, Alexander, et al.
Published: (2025)
by: Nemecek, Alexander, et al.
Published: (2025)
The Feasibility of Topic-Based Watermarking on Academic Peer Reviews
by: Nemecek, Alexander, et al.
Published: (2025)
by: Nemecek, Alexander, et al.
Published: (2025)
Watermark-based Attribution of AI-Generated Content
by: Jiang, Zhengyuan, et al.
Published: (2024)
by: Jiang, Zhengyuan, et al.
Published: (2024)
Get my drift? Catching LLM Task Drift with Activation Deltas
by: Abdelnabi, Sahar, et al.
Published: (2024)
by: Abdelnabi, Sahar, et al.
Published: (2024)
Bridging the Copyright Gap: Do Large Vision-Language Models Recognize and Respect Copyrighted Content?
by: Xu, Naen, et al.
Published: (2025)
by: Xu, Naen, et al.
Published: (2025)
A User-Centric, Privacy-Preserving, and Verifiable Ecosystem for Personal Data Management and Utilization
by: Zafar, Osama, et al.
Published: (2025)
by: Zafar, Osama, et al.
Published: (2025)
Quantifying Memorization and Privacy Risks in Genomic Language Models
by: Nemecek, Alexander, et al.
Published: (2026)
by: Nemecek, Alexander, et al.
Published: (2026)
SWAN: Semantic Watermarking with Abstract Meaning Representation
by: Ye, Ziping, et al.
Published: (2026)
by: Ye, Ziping, et al.
Published: (2026)
Iteratively Prompting Multimodal LLMs to Reproduce Natural and AI-Generated Images
by: Naseh, Ali, et al.
Published: (2024)
by: Naseh, Ali, et al.
Published: (2024)
Privacy Policy Enforcement Guardrails for Data-Sensitive Retrieval-Augmented Generation
by: Zafar, Osama, et al.
Published: (2026)
by: Zafar, Osama, et al.
Published: (2026)
Watermarks Attack Watermarks: Re-Watermarking as a Generic Removal Strategy
by: Bulychev, Maria, et al.
Published: (2026)
by: Bulychev, Maria, et al.
Published: (2026)
Privacy-Preserving AI-Enabled Decentralized Learning and Employment Records System
by: Xu, Yuqiao, et al.
Published: (2026)
by: Xu, Yuqiao, et al.
Published: (2026)
AI Agents May Always Fall for Prompt Injections
by: Abdelnabi, Sahar, et al.
Published: (2026)
by: Abdelnabi, Sahar, et al.
Published: (2026)
SafeGen: Mitigating Sexually Explicit Content Generation in Text-to-Image Models
by: Li, Xinfeng, et al.
Published: (2024)
by: Li, Xinfeng, et al.
Published: (2024)
Secure and Robust Watermarking for AI-generated Images: A Comprehensive Survey
by: Cao, Jie, et al.
Published: (2025)
by: Cao, Jie, et al.
Published: (2025)
k-SemStamp: A Clustering-Based Semantic Watermark for Detection of Machine-Generated Text
by: Hou, Abe Bohan, et al.
Published: (2024)
by: Hou, Abe Bohan, et al.
Published: (2024)
Ethical Challenges in Computer Vision: Ensuring Privacy and Mitigating Bias in Publicly Available Datasets
by: Tahir, Ghalib Ahmed
Published: (2024)
by: Tahir, Ghalib Ahmed
Published: (2024)
SimMark: A Robust Sentence-Level Similarity-Based Watermarking Algorithm for Large Language Models
by: Dabiriaghdam, Amirhossein, et al.
Published: (2025)
by: Dabiriaghdam, Amirhossein, et al.
Published: (2025)
IAG: Input-aware Backdoor Attack on VLM-based Visual Grounding
by: Li, Junxian, et al.
Published: (2025)
by: Li, Junxian, et al.
Published: (2025)
Watermarking Large Language Models and the Generated Content: Opportunities and Challenges
by: Zhang, Ruisi, et al.
Published: (2024)
by: Zhang, Ruisi, et al.
Published: (2024)
ZKPROV: A Zero-Knowledge Approach to Dataset Provenance for Large Language Models
by: Namazi, Mina, et al.
Published: (2025)
by: Namazi, Mina, et al.
Published: (2025)
DLOVE: A new Security Evaluation Tool for Deep Learning Based Watermarking Techniques
by: Padhi, Sudev Kumar, et al.
Published: (2024)
by: Padhi, Sudev Kumar, et al.
Published: (2024)
Robust Watermarks Leak: Channel-Aware Feature Extraction Enables Adversarial Watermark Manipulation
by: Ba, Zhongjie, et al.
Published: (2025)
by: Ba, Zhongjie, et al.
Published: (2025)
Invisible Watermarks, Visible Gains: Steering Machine Unlearning with Bi-Level Watermarking Design
by: Sun, Yuhao, et al.
Published: (2025)
by: Sun, Yuhao, et al.
Published: (2025)
The Orthogonal Vulnerabilities of Generative AI Watermarks: A Comparative Empirical Benchmark of Spatial and Latent Provenance
by: Yu, Jesse, et al.
Published: (2026)
by: Yu, Jesse, et al.
Published: (2026)
Guidance Watermarking for Diffusion Models
by: Gesny, Enoal, et al.
Published: (2025)
by: Gesny, Enoal, et al.
Published: (2025)
Watermark Anything with Localized Messages
by: Sander, Tom, et al.
Published: (2024)
by: Sander, Tom, et al.
Published: (2024)
Detecting Malicious Concepts without Image Generation in AI-Generated Content (AIGC)
by: Xu, Kun, et al.
Published: (2025)
by: Xu, Kun, et al.
Published: (2025)
Neighborhood Blending: A Lightweight Inference-Time Defense Against Membership Inference Attacks
by: Zafar, Osama, et al.
Published: (2026)
by: Zafar, Osama, et al.
Published: (2026)
Watermarking Conditional Text Generation for AI Detection: Unveiling Challenges and a Semantic-Aware Watermark Remedy
by: Fu, Yu, et al.
Published: (2023)
by: Fu, Yu, et al.
Published: (2023)
One-shot Face Sketch Synthesis in the Wild via Generative Diffusion Prior and Instruction Tuning
by: Wu, Han, et al.
Published: (2025)
by: Wu, Han, et al.
Published: (2025)
Are GUI Agents Focused Enough? Automated Distraction via Semantic-level UI Element Injection
by: Yang, Wenkui, et al.
Published: (2026)
by: Yang, Wenkui, et al.
Published: (2026)
Image-Based Geolocation Using Large Vision-Language Models
by: Liu, Yi, et al.
Published: (2024)
by: Liu, Yi, et al.
Published: (2024)
MLLM-Protector: Ensuring MLLM's Safety without Hurting Performance
by: Pi, Renjie, et al.
Published: (2024)
by: Pi, Renjie, et al.
Published: (2024)
Similar Items
-
Topic-Based Watermarks for Large Language Models
by: Nemecek, Alexander, et al.
Published: (2024) -
Authenticated Contradictions from Desynchronized Provenance and Watermarking
by: Nemecek, Alexander, et al.
Published: (2026) -
The End of Trust: How Agentic AI Breaks Security Assumptions
by: Zafar, Osama, et al.
Published: (2026) -
Watermarking Without Standards Is Not AI Governance
by: Nemecek, Alexander, et al.
Published: (2025) -
Digital Agriculture Sandbox for Collaborative Research
by: Zafar, Osama, et al.
Published: (2025)