From Compression to Accountability: Harmless Copyright Protection for Dataset Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Liang, Yan, Yang, Ziyuan, Sun, Mengyu, Zhou, Joey Tianyi, Zhang, Yi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dark Distillation: Backdooring Distilled Datasets without Accessing Raw Data
by: Yang, Ziyuan, et al.
Published: (2025)
by: Yang, Ziyuan, et al.
Published: (2025)
Beyond Dataset Watermarking: Model-Level Copyright Protection for Code Summarization Models
by: Zhang, Jiale, et al.
Published: (2024)
by: Zhang, Jiale, et al.
Published: (2024)
Watermarking Text Data on Large Language Models for Dataset Copyright
by: Liu, Yixin, et al.
Published: (2023)
by: Liu, Yixin, et al.
Published: (2023)
DWBench: Holistic Evaluation of Watermark for Dataset Copyright Auditing
by: Ren, Xiao, et al.
Published: (2026)
by: Ren, Xiao, et al.
Published: (2026)
Adversarial Example Based Fingerprinting for Robust Copyright Protection in Split Learning
by: Lin, Zhangting, et al.
Published: (2025)
by: Lin, Zhangting, et al.
Published: (2025)
Plaintext-Free Deep Learning for Privacy-Preserving Medical Image Analysis via Frequency Information Embedding
by: Sun, Mengyu, et al.
Published: (2024)
by: Sun, Mengyu, et al.
Published: (2024)
FairCMS: Cloud Media Sharing with Fair Copyright Protection
by: Xiao, Xiangli, et al.
Published: (2021)
by: Xiao, Xiangli, et al.
Published: (2021)
DataSafe: Copyright Protection with PUF Watermarking and Blockchain Tracking
by: Xue, Xiaolong, et al.
Published: (2024)
by: Xue, Xiaolong, et al.
Published: (2024)
Copyright Protection for Large Language Models: A Survey of Methods, Challenges, and Trends
by: Xu, Zhenhua, et al.
Published: (2025)
by: Xu, Zhenhua, et al.
Published: (2025)
SoK: Dataset Copyright Auditing in Machine Learning Systems
by: Du, Linkang, et al.
Published: (2024)
by: Du, Linkang, et al.
Published: (2024)
Data Taggants: Dataset Ownership Verification via Harmless Targeted Data Poisoning
by: Bouaziz, Wassim, et al.
Published: (2024)
by: Bouaziz, Wassim, et al.
Published: (2024)
Geometry-Aware Localized Watermarking for Copyright Protection in Embedding-as-a-Service
by: Chen, Zhimin, et al.
Published: (2026)
by: Chen, Zhimin, et al.
Published: (2026)
Harmless Backdoor-based Client-side Watermarking in Federated Learning
by: Luo, Kaijing, et al.
Published: (2024)
by: Luo, Kaijing, et al.
Published: (2024)
Strong Copyright Protection for Language Models via Adaptive Model Fusion
by: Abad, Javier, et al.
Published: (2024)
by: Abad, Javier, et al.
Published: (2024)
CopyrightMeter: Revisiting Copyright Protection in Text-to-image Models
by: Xu, Naen, et al.
Published: (2024)
by: Xu, Naen, et al.
Published: (2024)
CoTGuard: Using Chain-of-Thought Triggering for Copyright Protection in Multi-Agent LLM Systems
by: Wen, Yan, et al.
Published: (2025)
by: Wen, Yan, et al.
Published: (2025)
From Essence to Defense: Adaptive Semantic-aware Watermarking for Embedding-as-a-Service Copyright Protection
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
WATER-GS: Toward Copyright Protection for 3D Gaussian Splatting via Universal Watermarking
by: Tan, Yuqi, et al.
Published: (2024)
by: Tan, Yuqi, et al.
Published: (2024)
Protecting Deep Learning Model Copyrights with Adversarial Example-Free Reuse Detection
by: Luan, Xiaokun, et al.
Published: (2024)
by: Luan, Xiaokun, et al.
Published: (2024)
FFT: Towards Harmlessness Evaluation and Analysis for LLMs with Factuality, Fairness, Toxicity
by: Cui, Shiyao, et al.
Published: (2023)
by: Cui, Shiyao, et al.
Published: (2023)
Are You Copying My Prompt? Protecting the Copyright of Vision Prompt for VPaaS via Watermark
by: Ren, Huali, et al.
Published: (2024)
by: Ren, Huali, et al.
Published: (2024)
Harmless Yet Harmful: Neutral Prompting Attacks for Stealthy Hallucination Steering in Agent Skills
by: Hsu, Chia-Yi, et al.
Published: (2026)
by: Hsu, Chia-Yi, et al.
Published: (2026)
Understanding LLM Behavior When Encountering User-Supplied Harmful Content in Harmless Tasks
by: Chu, Junjie, et al.
Published: (2026)
by: Chu, Junjie, et al.
Published: (2026)
VICTOR: Dataset Copyright Auditing in Video Recognition Systems
by: Yuan, Quan, et al.
Published: (2025)
by: Yuan, Quan, et al.
Published: (2025)
Accountable Liveness
by: Lewis-Pye, Andrew, et al.
Published: (2025)
by: Lewis-Pye, Andrew, et al.
Published: (2025)
Protecting Copyrighted Material with Unique Identifiers in Large Language Model Training
by: Zhao, Shuai, et al.
Published: (2024)
by: Zhao, Shuai, et al.
Published: (2024)
AGATE: Stealthy Black-box Watermarking for Multimodal Model Copyright Protection
by: Gao, Jianbo, et al.
Published: (2025)
by: Gao, Jianbo, et al.
Published: (2025)
Deep Learning-based Dual Watermarking for Image Copyright Protection and Authentication
by: Padhi, Sudev Kumar, et al.
Published: (2025)
by: Padhi, Sudev Kumar, et al.
Published: (2025)
Copyright-Protected Language Generation via Adaptive Model Fusion
by: Abad, Javier, et al.
Published: (2024)
by: Abad, Javier, et al.
Published: (2024)
PREE: Towards Harmless and Adaptive Fingerprint Editing in Large Language Models via Knowledge Prefix Enhancement
by: Yue, Xubin, et al.
Published: (2025)
by: Yue, Xubin, et al.
Published: (2025)
The Trojan Knowledge: Bypassing Commercial LLM Guardrails via Harmless Prompt Weaving and Adaptive Tree Search
by: Wei, Rongzhe, et al.
Published: (2025)
by: Wei, Rongzhe, et al.
Published: (2025)
DP-GENG : Differentially Private Dataset Distillation Guided by DP-Generated Data
by: Shi, Shuo, et al.
Published: (2025)
by: Shi, Shuo, et al.
Published: (2025)
ME: Trigger Element Combination Backdoor Attack on Copyright Infringement
by: Yang, Feiyu, et al.
Published: (2025)
by: Yang, Feiyu, et al.
Published: (2025)
RegionMarker: A Region-Triggered Semantic Watermarking Framework for Embedding-as-a-Service Copyright Protection
by: Yang, Shufan, et al.
Published: (2025)
by: Yang, Shufan, et al.
Published: (2025)
AquaLoRA: Toward White-box Protection for Customized Stable Diffusion Models via Watermark LoRA
by: Feng, Weitao, et al.
Published: (2024)
by: Feng, Weitao, et al.
Published: (2024)
DataCook: Crafting Anti-Adversarial Examples for Healthcare Data Copyright Protection
by: Shang, Sihan, et al.
Published: (2024)
by: Shang, Sihan, et al.
Published: (2024)
Lossless Copyright Protection via Intrinsic Model Fingerprinting
by: Chen, Lingxiao, et al.
Published: (2026)
by: Chen, Lingxiao, et al.
Published: (2026)
Copyright Protection in Generative AI: A Technical Perspective
by: Ren, Jie, et al.
Published: (2024)
by: Ren, Jie, et al.
Published: (2024)
RTLMarker: Protecting LLM-Generated RTL Copyright via a Hardware Watermarking Framework
by: Wang, Kun, et al.
Published: (2025)
by: Wang, Kun, et al.
Published: (2025)
Sandi: A System for Accountability
by: Durak, F. Betül, et al.
Published: (2024)
by: Durak, F. Betül, et al.
Published: (2024)
Similar Items
-
Dark Distillation: Backdooring Distilled Datasets without Accessing Raw Data
by: Yang, Ziyuan, et al.
Published: (2025) -
Beyond Dataset Watermarking: Model-Level Copyright Protection for Code Summarization Models
by: Zhang, Jiale, et al.
Published: (2024) -
Watermarking Text Data on Large Language Models for Dataset Copyright
by: Liu, Yixin, et al.
Published: (2023) -
DWBench: Holistic Evaluation of Watermark for Dataset Copyright Auditing
by: Ren, Xiao, et al.
Published: (2026) -
Adversarial Example Based Fingerprinting for Robust Copyright Protection in Split Learning
by: Lin, Zhangting, et al.
Published: (2025)