Dataset Protection via Watermarked Canaries in Retrieval-Augmented LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yepeng, Zhao, Xuandong, Song, Dawn, Bu, Yuheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption
von: Liu, Yepeng, et al.
Veröffentlicht: (2025)
von: Liu, Yepeng, et al.
Veröffentlicht: (2025)
Defending LLM Watermarking Against Spoofing Attacks with Contrastive Representation Learning
von: An, Li, et al.
Veröffentlicht: (2025)
von: An, Li, et al.
Veröffentlicht: (2025)
The Landscape of Memorization in LLMs: Mechanisms, Measurement, and Mitigation
von: Xiong, Alexander, et al.
Veröffentlicht: (2025)
von: Xiong, Alexander, et al.
Veröffentlicht: (2025)
In-Context Watermarks for Large Language Models
von: Liu, Yepeng, et al.
Veröffentlicht: (2025)
von: Liu, Yepeng, et al.
Veröffentlicht: (2025)
Are You Getting What You Pay For? Auditing Model Substitution in LLM APIs
von: Cai, Will, et al.
Veröffentlicht: (2025)
von: Cai, Will, et al.
Veröffentlicht: (2025)
A Reinforcement Learning Framework for Robust and Secure LLM Watermarking
von: An, Li, et al.
Veröffentlicht: (2025)
von: An, Li, et al.
Veröffentlicht: (2025)
An Undetectable Watermark for Generative Image Models
von: Gunn, Sam, et al.
Veröffentlicht: (2024)
von: Gunn, Sam, et al.
Veröffentlicht: (2024)
Improving LLM Safety Alignment with Dual-Objective Optimization
von: Zhao, Xuandong, et al.
Veröffentlicht: (2025)
von: Zhao, Xuandong, et al.
Veröffentlicht: (2025)
Permute-and-Flip: An optimally stable and watermarkable decoder for LLMs
von: Zhao, Xuandong, et al.
Veröffentlicht: (2024)
von: Zhao, Xuandong, et al.
Veröffentlicht: (2024)
Geometry-Aware Localized Watermarking for Copyright Protection in Embedding-as-a-Service
von: Chen, Zhimin, et al.
Veröffentlicht: (2026)
von: Chen, Zhimin, et al.
Veröffentlicht: (2026)
SafeKey: Amplifying Aha-Moment Insights for Safety Reasoning
von: Zhou, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zhou, Kaiwen, et al.
Veröffentlicht: (2025)
Behavioral Canaries: Auditing Private Retrieved Context Usage in RL Fine-Tuning
von: Chen, Chaoran, et al.
Veröffentlicht: (2026)
von: Chen, Chaoran, et al.
Veröffentlicht: (2026)
Robust LLM Watermarking with Minimal Semantic Distortion for IP Protection
von: Dang, Kieu, et al.
Veröffentlicht: (2026)
von: Dang, Kieu, et al.
Veröffentlicht: (2026)
Distributional Information Embedding: A Framework for Multi-bit Watermarking
von: He, Haiyun, et al.
Veröffentlicht: (2025)
von: He, Haiyun, et al.
Veröffentlicht: (2025)
SoK: Are Watermarks in LLMs Ready for Deployment?
von: Dang, Kieu, et al.
Veröffentlicht: (2025)
von: Dang, Kieu, et al.
Veröffentlicht: (2025)
Image Watermarks are Removable Using Controllable Regeneration from Clean Noise
von: Liu, Yepeng, et al.
Veröffentlicht: (2024)
von: Liu, Yepeng, et al.
Veröffentlicht: (2024)
Theoretically Grounded Framework for LLM Watermarking: A Distribution-Adaptive Approach
von: He, Haiyun, et al.
Veröffentlicht: (2024)
von: He, Haiyun, et al.
Veröffentlicht: (2024)
Bileve: Securing Text Provenance in Large Language Models Against Spoofing with Bi-level Signature
von: Zhou, Tong, et al.
Veröffentlicht: (2024)
von: Zhou, Tong, et al.
Veröffentlicht: (2024)
Implicit Identity Technologies for LLMs: Fingerprinting and Watermarking across Datasets, Models, and Generated Content
von: Liu, Bing, et al.
Veröffentlicht: (2026)
von: Liu, Bing, et al.
Veröffentlicht: (2026)
Disabling Self-Correction in Retrieval-Augmented Generation via Stealthy Retriever Poisoning
von: Dai, Yanbo, et al.
Veröffentlicht: (2025)
von: Dai, Yanbo, et al.
Veröffentlicht: (2025)
EmMark: Robust Watermarks for IP Protection of Embedded Quantized Large Language Models
von: Zhang, Ruisi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruisi, et al.
Veröffentlicht: (2024)
From Essence to Defense: Adaptive Semantic-aware Watermarking for Embedding-as-a-Service Copyright Protection
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
Data Extraction Attacks in Retrieval-Augmented Generation via Backdoors
von: Peng, Yuefeng, et al.
Veröffentlicht: (2024)
von: Peng, Yuefeng, et al.
Veröffentlicht: (2024)
RAG Safety: Exploring Knowledge Poisoning Attacks to Retrieval-Augmented Generation
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2025)
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2025)
RegionMarker: A Region-Triggered Semantic Watermarking Framework for Embedding-as-a-Service Copyright Protection
von: Yang, Shufan, et al.
Veröffentlicht: (2025)
von: Yang, Shufan, et al.
Veröffentlicht: (2025)
Watermarks for Language Models via Probabilistic Automata
von: Wang, Yangkun, et al.
Veröffentlicht: (2025)
von: Wang, Yangkun, et al.
Veröffentlicht: (2025)
STAMP Your Content: Proving Dataset Membership via Watermarked Rephrasings
von: Rastogi, Saksham, et al.
Veröffentlicht: (2025)
von: Rastogi, Saksham, et al.
Veröffentlicht: (2025)
Protecting Your LLMs with Information Bottleneck
von: Liu, Zichuan, et al.
Veröffentlicht: (2024)
von: Liu, Zichuan, et al.
Veröffentlicht: (2024)
Watermarking LLM Agent Trajectories
von: Meng, Wenlong, et al.
Veröffentlicht: (2026)
von: Meng, Wenlong, et al.
Veröffentlicht: (2026)
Toward Copyright Integrity and Verifiability via Multi-Bit Watermarking for Intelligent Transportation Systems
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
Less is More: Sparse Watermarking in LLMs with Enhanced Text Quality
von: Hoang, Duy C., et al.
Veröffentlicht: (2024)
von: Hoang, Duy C., et al.
Veröffentlicht: (2024)
Beyond Text: Unveiling Privacy Vulnerabilities in Multi-modal Retrieval-Augmented Generation
von: Zhang, Jiankun, et al.
Veröffentlicht: (2025)
von: Zhang, Jiankun, et al.
Veröffentlicht: (2025)
Adaptive Text Watermark for Large Language Models
von: Liu, Yepeng, et al.
Veröffentlicht: (2024)
von: Liu, Yepeng, et al.
Veröffentlicht: (2024)
AutoBnB-RAG: Enhancing Multi-Agent Incident Response with Retrieval-Augmented Generation
von: Liu, Zefang, et al.
Veröffentlicht: (2025)
von: Liu, Zefang, et al.
Veröffentlicht: (2025)
TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection
von: Sander, Tom, et al.
Veröffentlicht: (2026)
von: Sander, Tom, et al.
Veröffentlicht: (2026)
Self-Sovereign Agent
von: Qu, Wenjie, et al.
Veröffentlicht: (2026)
von: Qu, Wenjie, et al.
Veröffentlicht: (2026)
Watermarking Conditional Text Generation for AI Detection: Unveiling Challenges and a Semantic-Aware Watermark Remedy
von: Fu, Yu, et al.
Veröffentlicht: (2023)
von: Fu, Yu, et al.
Veröffentlicht: (2023)
More Haste, Less Speed: Weaker Single-Layer Watermark Improves Distortion-Free Watermark Ensembles
von: Chen, Ruibo, et al.
Veröffentlicht: (2026)
von: Chen, Ruibo, et al.
Veröffentlicht: (2026)
TrojanRAG: Retrieval-Augmented Generation Can Be Backdoor Driver in Large Language Models
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2024)
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2024)
WARDEN: Multi-Directional Backdoor Watermarks for Embedding-as-a-Service Copyright Protection
von: Shetty, Anudeex, et al.
Veröffentlicht: (2024)
von: Shetty, Anudeex, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption
von: Liu, Yepeng, et al.
Veröffentlicht: (2025) -
Defending LLM Watermarking Against Spoofing Attacks with Contrastive Representation Learning
von: An, Li, et al.
Veröffentlicht: (2025) -
The Landscape of Memorization in LLMs: Mechanisms, Measurement, and Mitigation
von: Xiong, Alexander, et al.
Veröffentlicht: (2025) -
In-Context Watermarks for Large Language Models
von: Liu, Yepeng, et al.
Veröffentlicht: (2025) -
Are You Getting What You Pay For? Auditing Model Substitution in LLM APIs
von: Cai, Will, et al.
Veröffentlicht: (2025)