Disappearing Ink: Obfuscation Breaks N-gram Code Watermarks in Theory and Practice
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Gehao, Bagdasarian, Eugene, Zhai, Juan, Ma, Shiqing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Marking Code Without Breaking It: Code Watermarking for Detecting LLM-Generated Code
von: Kim, Jungin, et al.
Veröffentlicht: (2025)
von: Kim, Jungin, et al.
Veröffentlicht: (2025)
StealthInk: A Multi-bit and Stealthy Watermark for Large Language Models
von: Jiang, Ya, et al.
Veröffentlicht: (2025)
von: Jiang, Ya, et al.
Veröffentlicht: (2025)
Network-Level Prompt and Trait Leakage in Local Research Agents
von: Jeong, Hyejun, et al.
Veröffentlicht: (2025)
von: Jeong, Hyejun, et al.
Veröffentlicht: (2025)
An Empirical Study of Code Obfuscation Practices in the Google Play Store
von: Niroshan, Akila, et al.
Veröffentlicht: (2025)
von: Niroshan, Akila, et al.
Veröffentlicht: (2025)
Backdooring Bias ($B^2$) into Stable Diffusion Models
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
Vaporizer: Breaking Watermarking Schemes for Large Language Model Outputs
von: Ng, Jonathan Hong Jin, et al.
Veröffentlicht: (2026)
von: Ng, Jonathan Hong Jin, et al.
Veröffentlicht: (2026)
Can LLMs Obfuscate Code? A Systematic Analysis of Large Language Models into Assembly Code Obfuscation
von: Mohseni, Seyedreza, et al.
Veröffentlicht: (2024)
von: Mohseni, Seyedreza, et al.
Veröffentlicht: (2024)
Can Large Language Models Really Recognize Your Name?
von: Pham, Dzung, et al.
Veröffentlicht: (2025)
von: Pham, Dzung, et al.
Veröffentlicht: (2025)
The Invisible Hand: Unveiling Provider Bias in Large Language Models for Code Generation
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
Self-interpreting Adversarial Images
von: Zhang, Tingwei, et al.
Veröffentlicht: (2024)
von: Zhang, Tingwei, et al.
Veröffentlicht: (2024)
The Coding Limits of Robust Watermarking for Generative Models
von: Francati, Danilo, et al.
Veröffentlicht: (2025)
von: Francati, Danilo, et al.
Veröffentlicht: (2025)
CODE ACROSTIC: Robust Watermarking for Code Generation
von: Lin, Li, et al.
Veröffentlicht: (2025)
von: Lin, Li, et al.
Veröffentlicht: (2025)
Throttling Web Agents Using Reasoning Gates
von: Kumar, Abhinav, et al.
Veröffentlicht: (2025)
von: Kumar, Abhinav, et al.
Veröffentlicht: (2025)
Obfuscated Memory Malware Detection
von: P, Sharmila S, et al.
Veröffentlicht: (2024)
von: P, Sharmila S, et al.
Veröffentlicht: (2024)
False Friends in the Shell: Unveiling the Emoticon Semantic Confusion in Large Language Models
von: Jiang, Weipeng, et al.
Veröffentlicht: (2026)
von: Jiang, Weipeng, et al.
Veröffentlicht: (2026)
Mechanistic Interpretability in the Presence of Architectural Obfuscation
von: Florencio, Marcos, et al.
Veröffentlicht: (2025)
von: Florencio, Marcos, et al.
Veröffentlicht: (2025)
Coward: Collision-based OOD Watermarking for Practical Proactive Federated Backdoor Detection
von: Li, Wenjie, et al.
Veröffentlicht: (2025)
von: Li, Wenjie, et al.
Veröffentlicht: (2025)
From Theory to Practice: Code Generation Using LLMs for CAPEC and CWE Frameworks
von: Shahzad, Murtuza, et al.
Veröffentlicht: (2026)
von: Shahzad, Murtuza, et al.
Veröffentlicht: (2026)
Towards Privacy-Preserving LLM Inference via Covariant Obfuscation (Technical Report)
von: Lin, Yu, et al.
Veröffentlicht: (2026)
von: Lin, Yu, et al.
Veröffentlicht: (2026)
Breaking the Code: Security Assessment of AI Code Agents Through Systematic Jailbreaking Attacks
von: Saha, Shoumik, et al.
Veröffentlicht: (2025)
von: Saha, Shoumik, et al.
Veröffentlicht: (2025)
Invariant-based Robust Weights Watermark for Large Language Models
von: Guo, Qingxiao, et al.
Veröffentlicht: (2025)
von: Guo, Qingxiao, et al.
Veröffentlicht: (2025)
Beyond A Fixed Seal: Adaptive Stealing Watermark in Large Language Models
von: Zhang, Shuhao, et al.
Veröffentlicht: (2026)
von: Zhang, Shuhao, et al.
Veröffentlicht: (2026)
Character-Level Perturbations Disrupt LLM Watermarks
von: Zhang, Zhaoxi, et al.
Veröffentlicht: (2025)
von: Zhang, Zhaoxi, et al.
Veröffentlicht: (2025)
SOFT: Selective Data Obfuscation for Protecting LLM Fine-tuning against Membership Inference Attacks
von: Zhang, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Kaiyuan, et al.
Veröffentlicht: (2025)
Trusted Machine Learning Models Unlock Private Inference for Problems Currently Infeasible with Cryptography
von: Shumailov, Ilia, et al.
Veröffentlicht: (2025)
von: Shumailov, Ilia, et al.
Veröffentlicht: (2025)
Babel: Jailbreaking Safety Attention via Obfuscation Distribution Optimized Sampling
von: Wang, Ziwei, et al.
Veröffentlicht: (2026)
von: Wang, Ziwei, et al.
Veröffentlicht: (2026)
Learning to Watermark: A Selective Watermarking Framework for Large Language Models via Multi-Objective Optimization
von: Wang, Chenrui, et al.
Veröffentlicht: (2025)
von: Wang, Chenrui, et al.
Veröffentlicht: (2025)
Can Reasoning Models Obfuscate Reasoning? Stress-Testing Chain-of-Thought Monitorability
von: Zolkowski, Artur, et al.
Veröffentlicht: (2025)
von: Zolkowski, Artur, et al.
Veröffentlicht: (2025)
FedAdOb: Privacy-Preserving Federated Deep Learning with Adaptive Obfuscation
von: Gu, Hanlin, et al.
Veröffentlicht: (2024)
von: Gu, Hanlin, et al.
Veröffentlicht: (2024)
Is Your Writing Being Mimicked by AI? Unveiling Imitation with Invisible Watermarks in Creative Writing
von: Zhang, Ziwei, et al.
Veröffentlicht: (2025)
von: Zhang, Ziwei, et al.
Veröffentlicht: (2025)
Can LLMs Deeply Detect Complex Malicious Queries? A Framework for Jailbreaking via Obfuscating Intent
von: Shang, Shang, et al.
Veröffentlicht: (2024)
von: Shang, Shang, et al.
Veröffentlicht: (2024)
Large Language Model Watermark Stealing With Mixed Integer Programming
von: Zhang, Zhaoxi, et al.
Veröffentlicht: (2024)
von: Zhang, Zhaoxi, et al.
Veröffentlicht: (2024)
TokenMark: A Modality-Agnostic Watermark for Pre-trained Transformers
von: Xu, Hengyuan, et al.
Veröffentlicht: (2024)
von: Xu, Hengyuan, et al.
Veröffentlicht: (2024)
Purify Once, Edit Freely: Breaking Image Protections under Model Mismatch
von: Zhao, Qichen, et al.
Veröffentlicht: (2026)
von: Zhao, Qichen, et al.
Veröffentlicht: (2026)
Your Semantic-Independent Watermark is Fragile: A Semantic Perturbation Attack against EaaS Watermark
von: Fei, Zekun, et al.
Veröffentlicht: (2024)
von: Fei, Zekun, et al.
Veröffentlicht: (2024)
AgentMark: Utility-Preserving Behavioral Watermarking for Agents
von: Huang, Kaibo, et al.
Veröffentlicht: (2026)
von: Huang, Kaibo, et al.
Veröffentlicht: (2026)
Detecting Benchmark Contamination Through Watermarking
von: Sander, Tom, et al.
Veröffentlicht: (2025)
von: Sander, Tom, et al.
Veröffentlicht: (2025)
Sequential Behavioral Watermarking for LLM Agents
von: An, Hyeseon, et al.
Veröffentlicht: (2026)
von: An, Hyeseon, et al.
Veröffentlicht: (2026)
Probabilistically Robust Watermarking of Neural Networks
von: Pautov, Mikhail, et al.
Veröffentlicht: (2024)
von: Pautov, Mikhail, et al.
Veröffentlicht: (2024)
A Survey of Fragile Model Watermarking
von: Gao, Zhenzhe, et al.
Veröffentlicht: (2024)
von: Gao, Zhenzhe, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Marking Code Without Breaking It: Code Watermarking for Detecting LLM-Generated Code
von: Kim, Jungin, et al.
Veröffentlicht: (2025) -
StealthInk: A Multi-bit and Stealthy Watermark for Large Language Models
von: Jiang, Ya, et al.
Veröffentlicht: (2025) -
Network-Level Prompt and Trait Leakage in Local Research Agents
von: Jeong, Hyejun, et al.
Veröffentlicht: (2025) -
An Empirical Study of Code Obfuscation Practices in the Google Play Store
von: Niroshan, Akila, et al.
Veröffentlicht: (2025) -
Backdooring Bias ($B^2$) into Stable Diffusion Models
von: Naseh, Ali, et al.
Veröffentlicht: (2024)