Ward: Provable RAG Dataset Inference via LLM Watermarks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jovanović, Nikola, Staab, Robin, Baader, Maximilian, Vechev, Martin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Watermark Stealing in Large Language Models
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024)
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024)
A Unified Framework for LLM Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
Watermarking Diffusion Language Models
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Discovering Spoofing Attempts on Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
LLM Fingerprinting via Semantically Conditioned Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Towards Watermarking of Open-Source LLMs
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Black-Box Detection of Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
Watch your steps: Dormant Adversarial Behaviors that Activate upon LLM Finetuning
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Exploiting LLM Quantization
von: Egashira, Kazuki, et al.
Veröffentlicht: (2024)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2024)
Fewer Weights, More Problems: A Practical Attack on LLM Pruning
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
Watermarking Autoregressive Image Generation
von: Jovanović, Nikola, et al.
Veröffentlicht: (2025)
von: Jovanović, Nikola, et al.
Veröffentlicht: (2025)
Every Bit, Everywhere, All at Once: A Binomial Multibit LLM Watermark
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
BaxBench: Can LLMs Generate Correct and Secure Backends?
von: Vero, Mark, et al.
Veröffentlicht: (2025)
von: Vero, Mark, et al.
Veröffentlicht: (2025)
Pay Attention to the Triggers: Constructing Backdoors That Survive Distillation
von: De Muri, Giovanni, et al.
Veröffentlicht: (2025)
von: De Muri, Giovanni, et al.
Veröffentlicht: (2025)
Mind the Gap: A Practical Attack on GGUF Quantization
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
Evading Data Contamination Detection for Language Models is (too) Easy
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
AutoBaxBuilder: Bootstrapping Code Security Benchmarking
von: von Arx, Tobias, et al.
Veröffentlicht: (2025)
von: von Arx, Tobias, et al.
Veröffentlicht: (2025)
SoK: Data Minimization in Machine Learning
von: Staab, Robin, et al.
Veröffentlicht: (2025)
von: Staab, Robin, et al.
Veröffentlicht: (2025)
CommandSans: Securing AI Agents with Surgical Precision Prompt Sanitization
von: Das, Debeshee, et al.
Veröffentlicht: (2025)
von: Das, Debeshee, et al.
Veröffentlicht: (2025)
Learning to Watermark LLM-generated Text via Reinforcement Learning
von: Xu, Xiaojun, et al.
Veröffentlicht: (2024)
von: Xu, Xiaojun, et al.
Veröffentlicht: (2024)
Transferable Black-Box One-Shot Forging of Watermarks via Image Preference Models
von: Souček, Tomáš, et al.
Veröffentlicht: (2025)
von: Souček, Tomáš, et al.
Veröffentlicht: (2025)
CLASP: Training-Free LLM-Assisted Source Code Watermarking via Semantic-Preserving Transformations
von: Xu, Rui, et al.
Veröffentlicht: (2025)
von: Xu, Rui, et al.
Veröffentlicht: (2025)
Honeyval: A Comprehensive Evaluation Framework for LLM-powered HTTP Honeypots
von: Vero, Mark, et al.
Veröffentlicht: (2026)
von: Vero, Mark, et al.
Veröffentlicht: (2026)
Double-I Watermark: Protecting Model Copyright for LLM Fine-tuning
von: Li, Shen, et al.
Veröffentlicht: (2024)
von: Li, Shen, et al.
Veröffentlicht: (2024)
Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
Unforgeable Watermarks for Language Models via Robust Signatures
von: Lin, Huijia, et al.
Veröffentlicht: (2026)
von: Lin, Huijia, et al.
Veröffentlicht: (2026)
Back to the Drawing Board for Fair Representation Learning
von: Pouget, Angéline, et al.
Veröffentlicht: (2024)
von: Pouget, Angéline, et al.
Veröffentlicht: (2024)
Hiding in Plain Sight: Disguising Data Stealing Attacks in Federated Learning
von: Garov, Kostadin, et al.
Veröffentlicht: (2023)
von: Garov, Kostadin, et al.
Veröffentlicht: (2023)
MarkTune: Improving the Quality-Detectability Trade-off in Open-Weight LLM Watermarking
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
von: Zhao, Yizhou, et al.
Veröffentlicht: (2025)
Edit Distance Robust Watermarks via Indexing Pseudorandom Codes
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
SWAP: Towards Copyright Auditing of Soft Prompts via Sequential Watermarking
von: Yang, Wenyuan, et al.
Veröffentlicht: (2025)
von: Yang, Wenyuan, et al.
Veröffentlicht: (2025)
Mitigating Watermark Forgery in Generative Models via Randomized Key Selection
von: Aremu, Toluwani, et al.
Veröffentlicht: (2025)
von: Aremu, Toluwani, et al.
Veröffentlicht: (2025)
Large Language Models are Advanced Anonymizers
von: Staab, Robin, et al.
Veröffentlicht: (2024)
von: Staab, Robin, et al.
Veröffentlicht: (2024)
MPC-Minimized Secure LLM Inference
von: Rathee, Deevashwer, et al.
Veröffentlicht: (2024)
von: Rathee, Deevashwer, et al.
Veröffentlicht: (2024)
Instruction Tuning for Secure Code Generation
von: He, Jingxuan, et al.
Veröffentlicht: (2024)
von: He, Jingxuan, et al.
Veröffentlicht: (2024)
Provable Privacy with Non-Private Pre-Processing
von: Hu, Yaxi, et al.
Veröffentlicht: (2024)
von: Hu, Yaxi, et al.
Veröffentlicht: (2024)
Sigil: Server-Enforced Watermarking in U-Shaped Split Federated Learning via Gradient Injection
von: Dai, Zhengchunmin, et al.
Veröffentlicht: (2025)
von: Dai, Zhengchunmin, et al.
Veröffentlicht: (2025)
LLM Benchmark Datasets Should Be Contamination-Resistant
von: Al-Lawati, Ali, et al.
Veröffentlicht: (2026)
von: Al-Lawati, Ali, et al.
Veröffentlicht: (2026)
FHE-Agent: Automating CKKS Configuration for Practical Encrypted Inference via an LLM-Guided Agentic Framework
von: Xu, Nuo, et al.
Veröffentlicht: (2025)
von: Xu, Nuo, et al.
Veröffentlicht: (2025)
SoK: Watermarking for AI-Generated Content
von: Zhao, Xuandong, et al.
Veröffentlicht: (2024)
von: Zhao, Xuandong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Watermark Stealing in Large Language Models
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024) -
A Unified Framework for LLM Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026) -
Watermarking Diffusion Language Models
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025) -
Discovering Spoofing Attempts on Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024) -
LLM Fingerprinting via Semantically Conditioned Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)