Watermarking LLM-Generated Datasets in Downstream Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Yugeng, Cong, Tianshuo, Backes, Michael, Li, Zheng, Zhang, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robustness Over Time: Understanding Adversarial Examples' Effectiveness on Longitudinal Versions of Large Language Models
by: Liu, Yugeng, et al.
Published: (2023)
by: Liu, Yugeng, et al.
Published: (2023)
Amplifying Machine Learning Attacks Through Strategic Compositions
by: Liu, Yugeng, et al.
Published: (2025)
by: Liu, Yugeng, et al.
Published: (2025)
$\texttt{ModSCAN}$: Measuring Stereotypical Bias in Large Vision-Language Models from Vision and Language Modalities
by: Jiang, Yukun, et al.
Published: (2024)
by: Jiang, Yukun, et al.
Published: (2024)
Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications
by: Wu, Yixin, et al.
Published: (2025)
by: Wu, Yixin, et al.
Published: (2025)
On Evaluating The Performance of Watermarked Machine-Generated Texts Under Adversarial Attacks
by: Liu, Zesen, et al.
Published: (2024)
by: Liu, Zesen, et al.
Published: (2024)
Adjacent Words, Divergent Intents: Jailbreaking Large Language Models via Task Concurrency
by: Jiang, Yukun, et al.
Published: (2025)
by: Jiang, Yukun, et al.
Published: (2025)
DPDSyn: Improving Differentially Private Dataset Synthesis for Model Training by Downstream Task Guidance
by: Jia, Mingxuan, et al.
Published: (2026)
by: Jia, Mingxuan, et al.
Published: (2026)
From Defender to Devil? Unintended Risk Interactions Induced by LLM Defenses
by: Meng, Xiangtao, et al.
Published: (2025)
by: Meng, Xiangtao, et al.
Published: (2025)
JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
by: Chu, Junjie, et al.
Published: (2024)
by: Chu, Junjie, et al.
Published: (2024)
Efficient and Universal Watermarking for LLM-Generated Code Detection
by: Li, Boquan, et al.
Published: (2024)
by: Li, Boquan, et al.
Published: (2024)
A Reinforcement Learning Framework for Robust and Secure LLM Watermarking
by: An, Li, et al.
Published: (2025)
by: An, Li, et al.
Published: (2025)
Watermarking LLM Agent Trajectories
by: Meng, Wenlong, et al.
Published: (2026)
by: Meng, Wenlong, et al.
Published: (2026)
RLCracker: Evaluating the Worst-Case Vulnerability of LLM Watermarks with Adaptive RL Attacks
by: Huang, Hanbo, et al.
Published: (2025)
by: Huang, Hanbo, et al.
Published: (2025)
Generative Model Watermarking Suppressing High-Frequency Artifacts
by: Zhang, Li, et al.
Published: (2023)
by: Zhang, Li, et al.
Published: (2023)
RLSpoofer: A Lightweight Evaluator for LLM Watermark Spoofing Resilience
by: Huang, Hanbo, et al.
Published: (2026)
by: Huang, Hanbo, et al.
Published: (2026)
Watermarking Text Data on Large Language Models for Dataset Copyright
by: Liu, Yixin, et al.
Published: (2023)
by: Liu, Yixin, et al.
Published: (2023)
Understanding LLM Behavior When Encountering User-Supplied Harmful Content in Harmless Tasks
by: Chu, Junjie, et al.
Published: (2026)
by: Chu, Junjie, et al.
Published: (2026)
Matrix Kloosterman Sums, Random Matrix Statistics, and Cryptography
by: Yang, Tianshuo
Published: (2026)
by: Yang, Tianshuo
Published: (2026)
Efficient Data-Free Model Stealing with Label Diversity
by: Liu, Yiyong, et al.
Published: (2024)
by: Liu, Yiyong, et al.
Published: (2024)
CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers
by: Zheng, Jingyi, et al.
Published: (2024)
by: Zheng, Jingyi, et al.
Published: (2024)
Membership Inference Attacks Against In-Context Learning
by: Wen, Rui, et al.
Published: (2024)
by: Wen, Rui, et al.
Published: (2024)
Towards Generalized and Stealthy Watermarking for Generative Code Models
by: Li, Haoxuan, et al.
Published: (2025)
by: Li, Haoxuan, et al.
Published: (2025)
MCGMark: An Encodable and Robust Online Watermark for Tracing LLM-Generated Malicious Code
by: Ning, Kaiwen, et al.
Published: (2024)
by: Ning, Kaiwen, et al.
Published: (2024)
SEAL: Subspace-Anchored Watermarks for LLM Ownership
by: Dai, Yanbo, et al.
Published: (2025)
by: Dai, Yanbo, et al.
Published: (2025)
On the Proactive Generation of Unsafe Images From Text-To-Image Models Using Benign Prompts
by: Wu, Yixin, et al.
Published: (2023)
by: Wu, Yixin, et al.
Published: (2023)
TabularMark: Watermarking Tabular Datasets for Machine Learning
by: Zheng, Yihao, et al.
Published: (2024)
by: Zheng, Yihao, et al.
Published: (2024)
Defending LLM Watermarking Against Spoofing Attacks with Contrastive Representation Learning
by: An, Li, et al.
Published: (2025)
by: An, Li, et al.
Published: (2025)
DWBench: Holistic Evaluation of Watermark for Dataset Copyright Auditing
by: Ren, Xiao, et al.
Published: (2026)
by: Ren, Xiao, et al.
Published: (2026)
Forging the Unforgeable: On the Feasibility of Counterfeit Watermarks in Backdoor-Based Dataset Ownership Verification
by: Li, Zhiying, et al.
Published: (2024)
by: Li, Zhiying, et al.
Published: (2024)
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns
by: Shen, Xinyue, et al.
Published: (2025)
by: Shen, Xinyue, et al.
Published: (2025)
Is The Watermarking Of LLM-Generated Code Robust?
by: Suresh, Tarun, et al.
Published: (2024)
by: Suresh, Tarun, et al.
Published: (2024)
SiGRRW: A Single-Watermark Robust Reversible Watermarking Framework with Guiding Strategy
by: Xu, Zikai, et al.
Published: (2026)
by: Xu, Zikai, et al.
Published: (2026)
Cryptanalysis of Pseudorandom Error-Correcting Codes
by: Wang, Tianrui, et al.
Published: (2025)
by: Wang, Tianrui, et al.
Published: (2025)
Character-Level Perturbations Disrupt LLM Watermarks
by: Zhang, Zhaoxi, et al.
Published: (2025)
by: Zhang, Zhaoxi, et al.
Published: (2025)
Prompt Stealing Attacks Against Text-to-Image Generation Models
by: Shen, Xinyue, et al.
Published: (2023)
by: Shen, Xinyue, et al.
Published: (2023)
Attacking Attention of Foundation Models Disrupts Downstream Tasks
by: Silva, Hondamunige Prasanna, et al.
Published: (2025)
by: Silva, Hondamunige Prasanna, et al.
Published: (2025)
Local Privacy-preserving Mechanisms and Applications in Machine Learning
by: Qin, Likun, et al.
Published: (2024)
by: Qin, Likun, et al.
Published: (2024)
AMUSE: Adaptive Multi-Segment Encoding for Dataset Watermarking
by: Alvar, Saeed Ranjbar, et al.
Published: (2024)
by: Alvar, Saeed Ranjbar, et al.
Published: (2024)
DuCodeMark: Dual-Purpose Code Dataset Watermarking via Style-Aware Watermark-Poison Design
by: Chen, Yuchen, et al.
Published: (2026)
by: Chen, Yuchen, et al.
Published: (2026)
Removing the Watermark Is Not Enough: Forensic Stealth in Generative-AI Watermark Removal
by: Goonatilake, Yevin Nikhel, et al.
Published: (2026)
by: Goonatilake, Yevin Nikhel, et al.
Published: (2026)
Similar Items
-
Robustness Over Time: Understanding Adversarial Examples' Effectiveness on Longitudinal Versions of Large Language Models
by: Liu, Yugeng, et al.
Published: (2023) -
Amplifying Machine Learning Attacks Through Strategic Compositions
by: Liu, Yugeng, et al.
Published: (2025) -
$\texttt{ModSCAN}$: Measuring Stereotypical Bias in Large Vision-Language Models from Vision and Language Modalities
by: Jiang, Yukun, et al.
Published: (2024) -
Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications
by: Wu, Yixin, et al.
Published: (2025) -
On Evaluating The Performance of Watermarked Machine-Generated Texts Under Adversarial Attacks
by: Liu, Zesen, et al.
Published: (2024)