Watermark Stealing in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jovanović, Nikola, Staab, Robin, Vechev, Martin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Watermarking Diffusion Language Models
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Discovering Spoofing Attempts on Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
A Unified Framework for LLM Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
Ward: Provable RAG Dataset Inference via LLM Watermarks
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024)
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024)
Black-Box Detection of Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
Towards Watermarking of Open-Source LLMs
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
LLM Fingerprinting via Semantically Conditioned Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Watermarking Autoregressive Image Generation
von: Jovanović, Nikola, et al.
Veröffentlicht: (2025)
von: Jovanović, Nikola, et al.
Veröffentlicht: (2025)
Watch your steps: Dormant Adversarial Behaviors that Activate upon LLM Finetuning
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Pay Attention to the Triggers: Constructing Backdoors That Survive Distillation
von: De Muri, Giovanni, et al.
Veröffentlicht: (2025)
von: De Muri, Giovanni, et al.
Veröffentlicht: (2025)
Exploiting LLM Quantization
von: Egashira, Kazuki, et al.
Veröffentlicht: (2024)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2024)
Fewer Weights, More Problems: A Practical Attack on LLM Pruning
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
Mind the Gap: A Practical Attack on GGUF Quantization
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
von: Egashira, Kazuki, et al.
Veröffentlicht: (2025)
Every Bit, Everywhere, All at Once: A Binomial Multibit LLM Watermark
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
Hiding in Plain Sight: Disguising Data Stealing Attacks in Federated Learning
von: Garov, Kostadin, et al.
Veröffentlicht: (2023)
von: Garov, Kostadin, et al.
Veröffentlicht: (2023)
SoK: Data Minimization in Machine Learning
von: Staab, Robin, et al.
Veröffentlicht: (2025)
von: Staab, Robin, et al.
Veröffentlicht: (2025)
Large Language Models are Advanced Anonymizers
von: Staab, Robin, et al.
Veröffentlicht: (2024)
von: Staab, Robin, et al.
Veröffentlicht: (2024)
Large Language Model Watermark Stealing With Mixed Integer Programming
von: Zhang, Zhaoxi, et al.
Veröffentlicht: (2024)
von: Zhang, Zhaoxi, et al.
Veröffentlicht: (2024)
BaxBench: Can LLMs Generate Correct and Secure Backends?
von: Vero, Mark, et al.
Veröffentlicht: (2025)
von: Vero, Mark, et al.
Veröffentlicht: (2025)
DMark: Order-Agnostic Watermarking for Diffusion Large Language Models
von: Wu, Linyu, et al.
Veröffentlicht: (2025)
von: Wu, Linyu, et al.
Veröffentlicht: (2025)
Beyond A Fixed Seal: Adaptive Stealing Watermark in Large Language Models
von: Zhang, Shuhao, et al.
Veröffentlicht: (2026)
von: Zhang, Shuhao, et al.
Veröffentlicht: (2026)
Security and Detectability Analysis of Unicode Text Watermarking Methods Against Large Language Models
von: Hellmeier, Malte
Veröffentlicht: (2025)
von: Hellmeier, Malte
Veröffentlicht: (2025)
Architectural Backdoors for Within-Batch Data Stealing and Model Inference Manipulation
von: Küchler, Nicolas, et al.
Veröffentlicht: (2025)
von: Küchler, Nicolas, et al.
Veröffentlicht: (2025)
Duwak: Dual Watermarks in Large Language Models
von: Zhu, Chaoyi, et al.
Veröffentlicht: (2024)
von: Zhu, Chaoyi, et al.
Veröffentlicht: (2024)
Unforgeable Watermarks for Language Models via Robust Signatures
von: Lin, Huijia, et al.
Veröffentlicht: (2026)
von: Lin, Huijia, et al.
Veröffentlicht: (2026)
Large Language Models for Link Stealing Attacks Against Graph Neural Networks
von: Guan, Faqian, et al.
Veröffentlicht: (2024)
von: Guan, Faqian, et al.
Veröffentlicht: (2024)
Transferable Black-Box One-Shot Forging of Watermarks via Image Preference Models
von: Souček, Tomáš, et al.
Veröffentlicht: (2025)
von: Souček, Tomáš, et al.
Veröffentlicht: (2025)
Teach LLMs to Phish: Stealing Private Information from Language Models
von: Panda, Ashwinee, et al.
Veröffentlicht: (2024)
von: Panda, Ashwinee, et al.
Veröffentlicht: (2024)
Evading Data Contamination Detection for Language Models is (too) Easy
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
Exploring Query Efficient Data Generation towards Data-free Model Stealing in Hard Label Setting
von: Pei, Gaozheng, et al.
Veröffentlicht: (2024)
von: Pei, Gaozheng, et al.
Veröffentlicht: (2024)
Protecting Copyright of Medical Pre-trained Language Models: Training-Free Backdoor Model Watermarking
von: Kong, Cong, et al.
Veröffentlicht: (2024)
von: Kong, Cong, et al.
Veröffentlicht: (2024)
PostMark: A Robust Blackbox Watermark for Large Language Models
von: Chang, Yapei, et al.
Veröffentlicht: (2024)
von: Chang, Yapei, et al.
Veröffentlicht: (2024)
Back to the Drawing Board for Fair Representation Learning
von: Pouget, Angéline, et al.
Veröffentlicht: (2024)
von: Pouget, Angéline, et al.
Veröffentlicht: (2024)
Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
von: Shao, Shuo, et al.
Veröffentlicht: (2024)
Watermarking Makes Language Models Radioactive
von: Sander, Tom, et al.
Veröffentlicht: (2024)
von: Sander, Tom, et al.
Veröffentlicht: (2024)
Refine, Discriminate and Align: Stealing Encoders via Sample-Wise Prototypes and Multi-Relational Extraction
von: Wu, Shuchi, et al.
Veröffentlicht: (2023)
von: Wu, Shuchi, et al.
Veröffentlicht: (2023)
LinkThief: Combining Generalized Structure Knowledge with Node Similarity for Link Stealing Attack against GNN
von: Zhang, Yuxing, et al.
Veröffentlicht: (2024)
von: Zhang, Yuxing, et al.
Veröffentlicht: (2024)
GraphIP-Bench: How Hard Is It to Steal a Graph Neural Network, and Can We Stop It?
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2026)
von: Zhao, Kaixiang, et al.
Veröffentlicht: (2026)
Improved Large Language Model Jailbreak Detection via Pretrained Embeddings
von: Galinkin, Erick, et al.
Veröffentlicht: (2024)
von: Galinkin, Erick, et al.
Veröffentlicht: (2024)
Generative Models are Self-Watermarked: Declaring Model Authentication through Re-Generation
von: Desu, Aditya, et al.
Veröffentlicht: (2024)
von: Desu, Aditya, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Watermarking Diffusion Language Models
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025) -
Discovering Spoofing Attempts on Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024) -
A Unified Framework for LLM Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026) -
Ward: Provable RAG Dataset Inference via LLM Watermarks
von: Jovanović, Nikola, et al.
Veröffentlicht: (2024) -
Black-Box Detection of Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)