RAG-WM: An Efficient Black-Box Watermarking Approach for Retrieval-Augmented Generation of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lv, Peizhuo, Sun, Mengjie, Wang, Hao, Wang, Xiaofeng, Zhang, Shengzhi, Chen, Yuxuan, Chen, Kai, Sun, Limin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LoRAGuard: An Effective Black-box Watermarking Approach for LoRAs
von: Lv, Peizhuo, et al.
Veröffentlicht: (2025)
von: Lv, Peizhuo, et al.
Veröffentlicht: (2025)
SSL-WM: A Black-Box Watermarking Approach for Encoders Pre-trained by Self-supervised Learning
von: Lv, Peizhuo, et al.
Veröffentlicht: (2022)
von: Lv, Peizhuo, et al.
Veröffentlicht: (2022)
An Automated Attack Investigation Approach Leveraging Threat-Knowledge-Augmented Large Language Models
von: Dai, Rujie, et al.
Veröffentlicht: (2025)
von: Dai, Rujie, et al.
Veröffentlicht: (2025)
MEA-Defender: A Robust Watermark against Model Extraction Attack
von: Lv, Peizhuo, et al.
Veröffentlicht: (2024)
von: Lv, Peizhuo, et al.
Veröffentlicht: (2024)
Hot-Swap MarkBoard: An Efficient Black-box Watermarking Approach for Large-scale Model Distribution
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
A Model Stealing Attack Against Multi-Exit Networks
von: Pan, Li, et al.
Veröffentlicht: (2023)
von: Pan, Li, et al.
Veröffentlicht: (2023)
Black-Box Detection of Language Model Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2024)
Black-Box Opinion Manipulation Attacks to Retrieval-Augmented Generation of Large Language Models
von: Chen, Zhuo, et al.
Veröffentlicht: (2024)
von: Chen, Zhuo, et al.
Veröffentlicht: (2024)
PR-Attack: Coordinated Prompt-RAG Attacks on Retrieval-Augmented Generation in Large Language Models via Bilevel Optimization
von: Jiao, Yang, et al.
Veröffentlicht: (2025)
von: Jiao, Yang, et al.
Veröffentlicht: (2025)
The RAG Paradox: A Black-Box Attack Exploiting Unintentional Vulnerabilities in Retrieval-Augmented Generation Systems
von: Choi, Chanwoo, et al.
Veröffentlicht: (2025)
von: Choi, Chanwoo, et al.
Veröffentlicht: (2025)
ComMark: Covert and Robust Black-Box Model Watermarking with Compressed Samples
von: Yang, Yunfei, et al.
Veröffentlicht: (2025)
von: Yang, Yunfei, et al.
Veröffentlicht: (2025)
Watermarking Text Data on Large Language Models for Dataset Copyright
von: Liu, Yixin, et al.
Veröffentlicht: (2023)
von: Liu, Yixin, et al.
Veröffentlicht: (2023)
Making Theft Useless: Adulteration-Based Protection of Proprietary Knowledge Graphs in GraphRAG Systems
von: Wang, Weijie, et al.
Veröffentlicht: (2026)
von: Wang, Weijie, et al.
Veröffentlicht: (2026)
CPA-RAG:Covert Poisoning Attacks on Retrieval-Augmented Generation in Large Language Models
von: Li, Chunyang, et al.
Veröffentlicht: (2025)
von: Li, Chunyang, et al.
Veröffentlicht: (2025)
Turning Your Strength into Watermark: Watermarking Large Language Model via Knowledge Injection
von: Li, Shuai, et al.
Veröffentlicht: (2023)
von: Li, Shuai, et al.
Veröffentlicht: (2023)
A Watermark for Black-Box Language Models
von: Bahri, Dara, et al.
Veröffentlicht: (2024)
von: Bahri, Dara, et al.
Veröffentlicht: (2024)
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
von: Zou, Wei, et al.
Veröffentlicht: (2024)
von: Zou, Wei, et al.
Veröffentlicht: (2024)
TimeTravel: Real-time Timing Drift Attack on System Time Using Acoustic Waves
von: Liu, Jianshuo, et al.
Veröffentlicht: (2024)
von: Liu, Jianshuo, et al.
Veröffentlicht: (2024)
Why Neural Structural Obfuscation Can't Kill White-Box Watermarks for Good!
von: Jiang, Yanna, et al.
Veröffentlicht: (2026)
von: Jiang, Yanna, et al.
Veröffentlicht: (2026)
Tricking Retrievers with Influential Tokens: An Efficient Black-Box Corpus Poisoning Attack
von: Wang, Cheng, et al.
Veröffentlicht: (2025)
von: Wang, Cheng, et al.
Veröffentlicht: (2025)
VOW: Verifiable and Oblivious Watermark Detection for Large Language Models
von: Luan, Xiaokun, et al.
Veröffentlicht: (2026)
von: Luan, Xiaokun, et al.
Veröffentlicht: (2026)
SEW: Strengthening Robustness of Black-box DNN Watermarking via Specificity Enhancement
von: Qiu, Huming, et al.
Veröffentlicht: (2026)
von: Qiu, Huming, et al.
Veröffentlicht: (2026)
SafeRAG: Benchmarking Security in Retrieval-Augmented Generation of Large Language Model
von: Liang, Xun, et al.
Veröffentlicht: (2025)
von: Liang, Xun, et al.
Veröffentlicht: (2025)
Towards Traitor Tracing in Black-and-White-Box DNN Watermarking with Tardos-based Codes
von: Rodriguez-Lois, Elena, et al.
Veröffentlicht: (2023)
von: Rodriguez-Lois, Elena, et al.
Veröffentlicht: (2023)
EditMark: Watermarking Large Language Models based on Model Editing
von: Li, Shuai, et al.
Veröffentlicht: (2025)
von: Li, Shuai, et al.
Veröffentlicht: (2025)
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models
von: Gong, Yuyang, et al.
Veröffentlicht: (2025)
von: Gong, Yuyang, et al.
Veröffentlicht: (2025)
Efficient and Universal Watermarking for LLM-Generated Code Detection
von: Li, Boquan, et al.
Veröffentlicht: (2024)
von: Li, Boquan, et al.
Veröffentlicht: (2024)
A Certified Robust Watermark For Large Language Models
von: Feng, Xianheng, et al.
Veröffentlicht: (2024)
von: Feng, Xianheng, et al.
Veröffentlicht: (2024)
TrojanRAG: Retrieval-Augmented Generation Can Be Backdoor Driver in Large Language Models
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2024)
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2024)
SLIM: Stealthy Low-Coverage Black-Box Watermarking via Latent-Space Confusion Zones
von: Wu, Hengyu, et al.
Veröffentlicht: (2026)
von: Wu, Hengyu, et al.
Veröffentlicht: (2026)
Performance-lossless Black-box Model Watermarking
von: Zhao, Na, et al.
Veröffentlicht: (2023)
von: Zhao, Na, et al.
Veröffentlicht: (2023)
EmbTracker: Traceable Black-box Watermarking for Federated Language Models
von: Zhao, Haodong, et al.
Veröffentlicht: (2026)
von: Zhao, Haodong, et al.
Veröffentlicht: (2026)
Cross-Lingual Summarization as a Black-Box Watermark Removal Attack
von: Ganesan, Gokul
Veröffentlicht: (2025)
von: Ganesan, Gokul
Veröffentlicht: (2025)
MorphMark: Flexible Adaptive Watermarking for Large Language Models
von: Wang, Zongqi, et al.
Veröffentlicht: (2025)
von: Wang, Zongqi, et al.
Veröffentlicht: (2025)
ArmSSL: Adversarial Robust Black-Box Watermarking for Self-Supervised Learning Pre-trained Encoders
von: Jiang, Yongqi, et al.
Veröffentlicht: (2026)
von: Jiang, Yongqi, et al.
Veröffentlicht: (2026)
External Data Extraction Attacks against Retrieval-Augmented Large Language Models
von: He, Yu, et al.
Veröffentlicht: (2025)
von: He, Yu, et al.
Veröffentlicht: (2025)
MIRAGE: Misleading Retrieval-Augmented Generation via Black-box and Query-agnostic Poisoning Attacks
von: Chen, Tailun, et al.
Veröffentlicht: (2025)
von: Chen, Tailun, et al.
Veröffentlicht: (2025)
Rethinking LLM Watermark Detection in Black-Box Settings: A Non-Intrusive Third-Party Framework
von: Wang, Zhuoshang, et al.
Veröffentlicht: (2026)
von: Wang, Zhuoshang, et al.
Veröffentlicht: (2026)
HijackRAG: Hijacking Attacks against Retrieval-Augmented Large Language Models
von: Zhang, Yucheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yucheng, et al.
Veröffentlicht: (2024)
Fading the Digital Ink: A Universal Black-Box Attack Framework for 3DGS Watermarking Systems
von: Zeng, Qingyuan, et al.
Veröffentlicht: (2025)
von: Zeng, Qingyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LoRAGuard: An Effective Black-box Watermarking Approach for LoRAs
von: Lv, Peizhuo, et al.
Veröffentlicht: (2025) -
SSL-WM: A Black-Box Watermarking Approach for Encoders Pre-trained by Self-supervised Learning
von: Lv, Peizhuo, et al.
Veröffentlicht: (2022) -
An Automated Attack Investigation Approach Leveraging Threat-Knowledge-Augmented Large Language Models
von: Dai, Rujie, et al.
Veröffentlicht: (2025) -
MEA-Defender: A Robust Watermark against Model Extraction Attack
von: Lv, Peizhuo, et al.
Veröffentlicht: (2024) -
Hot-Swap MarkBoard: An Efficient Black-box Watermarking Approach for Large-scale Model Distribution
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)