Vaporizer: Breaking Watermarking Schemes for Large Language Model Outputs
Fuente:
arXiv
Saved in:
| Main Authors: | Ng, Jonathan Hong Jin, Ngo, Anh Tu, Chattopadhyay, Anupam |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BlockDoor: Blocking Backdoor Based Watermarks in Deep Neural Networks
by: Puah, Yi Hao, et al.
Published: (2024)
by: Puah, Yi Hao, et al.
Published: (2024)
Persistence of Backdoor-based Watermarks for Neural Networks: A Comprehensive Evaluation
by: Ngo, Anh Tu, et al.
Published: (2025)
by: Ngo, Anh Tu, et al.
Published: (2025)
Invariant-based Robust Weights Watermark for Large Language Models
by: Guo, Qingxiao, et al.
Published: (2025)
by: Guo, Qingxiao, et al.
Published: (2025)
Cryptographic Backdoor for Neural Networks: Boon and Bane
by: Ngo, Anh Tu, et al.
Published: (2025)
by: Ngo, Anh Tu, et al.
Published: (2025)
Functional Subspace Watermarking for Large Language Models
by: Ding, Zikang, et al.
Published: (2026)
by: Ding, Zikang, et al.
Published: (2026)
Efficient and Encrypted Inference using Binarized Neural Networks within In-Memory Computing Architectures
by: Rajendran, Gokulnath, et al.
Published: (2025)
by: Rajendran, Gokulnath, et al.
Published: (2025)
Watermarking Techniques for Large Language Models: A Survey
by: Liang, Yuqing, et al.
Published: (2024)
by: Liang, Yuqing, et al.
Published: (2024)
Learning Nonlinearity of Boolean Functions: An Experimentation with Neural Networks
by: Ranga, Sriram, et al.
Published: (2025)
by: Ranga, Sriram, et al.
Published: (2025)
Learning to Watermark: A Selective Watermarking Framework for Large Language Models via Multi-Objective Optimization
by: Wang, Chenrui, et al.
Published: (2025)
by: Wang, Chenrui, et al.
Published: (2025)
Large Language Model Watermark Stealing With Mixed Integer Programming
by: Zhang, Zhaoxi, et al.
Published: (2024)
by: Zhang, Zhaoxi, et al.
Published: (2024)
DiffuseTrace: A Transparent and Flexible Watermarking Scheme for Latent Diffusion Model
by: Lei, Liangqi, et al.
Published: (2024)
by: Lei, Liangqi, et al.
Published: (2024)
CEFW: A Comprehensive Evaluation Framework for Watermark in Large Language Models
by: Zhang, Shuhao, et al.
Published: (2025)
by: Zhang, Shuhao, et al.
Published: (2025)
Watermark Stealing in Large Language Models
by: Jovanović, Nikola, et al.
Published: (2024)
by: Jovanović, Nikola, et al.
Published: (2024)
Beyond A Fixed Seal: Adaptive Stealing Watermark in Large Language Models
by: Zhang, Shuhao, et al.
Published: (2026)
by: Zhang, Shuhao, et al.
Published: (2026)
StealthInk: A Multi-bit and Stealthy Watermark for Large Language Models
by: Jiang, Ya, et al.
Published: (2025)
by: Jiang, Ya, et al.
Published: (2025)
DERMARK: A Dynamic, Efficient and Robust Multi-bit Watermark for Large Language Models
by: Lin, Qihao, et al.
Published: (2025)
by: Lin, Qihao, et al.
Published: (2025)
Modification and Generated-Text Detection: Achieving Dual Detection Capabilities for the Outputs of LLM by Watermark
by: Cai, Yuhang, et al.
Published: (2025)
by: Cai, Yuhang, et al.
Published: (2025)
Multi-Designated Detector Watermarking for Language Models
by: Huang, Zhengan, et al.
Published: (2024)
by: Huang, Zhengan, et al.
Published: (2024)
Breaking Minds, Breaking Systems: Jailbreaking Large Language Models via Human-like Psychological Manipulation
by: Liu, Zehao, et al.
Published: (2025)
by: Liu, Zehao, et al.
Published: (2025)
Optimized Couplings for Watermarking Large Language Models
by: Tsur, Dor, et al.
Published: (2025)
by: Tsur, Dor, et al.
Published: (2025)
Disappearing Ink: Obfuscation Breaks N-gram Code Watermarks in Theory and Practice
by: Zhang, Gehao, et al.
Published: (2025)
by: Zhang, Gehao, et al.
Published: (2025)
Marking Code Without Breaking It: Code Watermarking for Detecting LLM-Generated Code
by: Kim, Jungin, et al.
Published: (2025)
by: Kim, Jungin, et al.
Published: (2025)
NeuroBreak: Unveil Internal Jailbreak Mechanisms in Large Language Models
by: Zhang, Chuhan, et al.
Published: (2025)
by: Zhang, Chuhan, et al.
Published: (2025)
Optimizing Adaptive Attacks against Watermarks for Language Models
by: Diaa, Abdulrahman, et al.
Published: (2024)
by: Diaa, Abdulrahman, et al.
Published: (2024)
Privacy and Security Implications of Cloud-Based AI Services : A Survey
by: Luqman, Alka, et al.
Published: (2024)
by: Luqman, Alka, et al.
Published: (2024)
RAG-WM: An Efficient Black-Box Watermarking Approach for Retrieval-Augmented Generation of Large Language Models
by: Lv, Peizhuo, et al.
Published: (2025)
by: Lv, Peizhuo, et al.
Published: (2025)
Defending Large Language Models Against Jailbreak Exploits with Responsible AI Considerations
by: Wong, Ryan, et al.
Published: (2025)
by: Wong, Ryan, et al.
Published: (2025)
DMark: Order-Agnostic Watermarking for Diffusion Large Language Models
by: Wu, Linyu, et al.
Published: (2025)
by: Wu, Linyu, et al.
Published: (2025)
PrivComp-KG : Leveraging Knowledge Graph and Large Language Models for Privacy Policy Compliance Verification
by: Garza, Leon, et al.
Published: (2024)
by: Garza, Leon, et al.
Published: (2024)
Watermarking Diffusion Language Models
by: Gloaguen, Thibaud, et al.
Published: (2025)
by: Gloaguen, Thibaud, et al.
Published: (2025)
Learnable Linguistic Watermarks for Tracing Model Extraction Attacks on Large Language Models
by: Bai, Minhao, et al.
Published: (2024)
by: Bai, Minhao, et al.
Published: (2024)
Advancing Beyond Identification: Multi-bit Watermark for Large Language Models
by: Yoo, KiYoon, et al.
Published: (2023)
by: Yoo, KiYoon, et al.
Published: (2023)
Less Is More -- Until It Breaks: Security Pitfalls of Vision Token Compression in Large Vision-Language Models
by: Zhang, Xiaomei, et al.
Published: (2026)
by: Zhang, Xiaomei, et al.
Published: (2026)
A Survey of Fragile Model Watermarking
by: Gao, Zhenzhe, et al.
Published: (2024)
by: Gao, Zhenzhe, et al.
Published: (2024)
Breaking Distortion-free Watermarks in Large Language Models
by: Reynolds, Shayleen, et al.
Published: (2025)
by: Reynolds, Shayleen, et al.
Published: (2025)
Inevitable Trade-off between Watermark Strength and Speculative Sampling Efficiency for Language Models
by: Hu, Zhengmian, et al.
Published: (2024)
by: Hu, Zhengmian, et al.
Published: (2024)
Breaking the Loop: Detecting and Mitigating Denial-of-Service Vulnerabilities in Large Language Models
by: Yu, Junzhe, et al.
Published: (2025)
by: Yu, Junzhe, et al.
Published: (2025)
Watermarking Discrete Diffusion Language Models
by: Bagchi, Avi, et al.
Published: (2025)
by: Bagchi, Avi, et al.
Published: (2025)
The Coding Limits of Robust Watermarking for Generative Models
by: Francati, Danilo, et al.
Published: (2025)
by: Francati, Danilo, et al.
Published: (2025)
Duwak: Dual Watermarks in Large Language Models
by: Zhu, Chaoyi, et al.
Published: (2024)
by: Zhu, Chaoyi, et al.
Published: (2024)
Similar Items
-
BlockDoor: Blocking Backdoor Based Watermarks in Deep Neural Networks
by: Puah, Yi Hao, et al.
Published: (2024) -
Persistence of Backdoor-based Watermarks for Neural Networks: A Comprehensive Evaluation
by: Ngo, Anh Tu, et al.
Published: (2025) -
Invariant-based Robust Weights Watermark for Large Language Models
by: Guo, Qingxiao, et al.
Published: (2025) -
Cryptographic Backdoor for Neural Networks: Boon and Bane
by: Ngo, Anh Tu, et al.
Published: (2025) -
Functional Subspace Watermarking for Large Language Models
by: Ding, Zikang, et al.
Published: (2026)