Safeguarding Multimodal Knowledge Copyright in the RAG-as-a-Service Environment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Tianyu, Lou, Jian, Wang, Wenjie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Silent Leaks: Implicit Knowledge Extraction Attack on RAG Systems through Benign Queries
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
Why Safeguarded Ships Run Aground? Aligned Large Language Models' Safety Mechanisms Tend to Be Anchored in The Template Region
von: Leong, Chak Tou, et al.
Veröffentlicht: (2025)
von: Leong, Chak Tou, et al.
Veröffentlicht: (2025)
Safeguarding Text-to-Image Generative Models Against Unauthorized Knowledge Distillation
von: Gao, Yilan, et al.
Veröffentlicht: (2026)
von: Gao, Yilan, et al.
Veröffentlicht: (2026)
AdaShield: Safeguarding Multimodal Large Language Models from Structure-based Attack via Adaptive Shield Prompting
von: Wang, Yu, et al.
Veröffentlicht: (2024)
von: Wang, Yu, et al.
Veröffentlicht: (2024)
P$^2$RAG: Efficient Privacy-Preserving RAG Service Supporting Arbitrary Top-$k$ Retrieval
von: Ming, Yulong, et al.
Veröffentlicht: (2026)
von: Ming, Yulong, et al.
Veröffentlicht: (2026)
AGATE: Stealthy Black-box Watermarking for Multimodal Model Copyright Protection
von: Gao, Jianbo, et al.
Veröffentlicht: (2025)
von: Gao, Jianbo, et al.
Veröffentlicht: (2025)
CopyrightMeter: Revisiting Copyright Protection in Text-to-image Models
von: Xu, Naen, et al.
Veröffentlicht: (2024)
von: Xu, Naen, et al.
Veröffentlicht: (2024)
LLM Safeguard is a Double-Edged Sword: Exploiting False Positives for Denial-of-Service Attacks
von: Zhang, Qingzhao, et al.
Veröffentlicht: (2024)
von: Zhang, Qingzhao, et al.
Veröffentlicht: (2024)
TransLinkGuard: Safeguarding Transformer Models Against Model Stealing in Edge Deployment
von: Li, Qinfeng, et al.
Veröffentlicht: (2024)
von: Li, Qinfeng, et al.
Veröffentlicht: (2024)
A Trajectory-Based Safety Audit of Clawdbot (OpenClaw)
von: Chen, Tianyu, et al.
Veröffentlicht: (2026)
von: Chen, Tianyu, et al.
Veröffentlicht: (2026)
An AI Agent Execution Environment to Safeguard User Data
von: Stanley, Robert, et al.
Veröffentlicht: (2026)
von: Stanley, Robert, et al.
Veröffentlicht: (2026)
GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning
von: Liu, Yue, et al.
Veröffentlicht: (2025)
von: Liu, Yue, et al.
Veröffentlicht: (2025)
Privacy-Aware RAG: Secure and Isolated Knowledge Retrieval
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
On the Evidentiary Limits of Membership Inference for Copyright Auditing
von: Ertan, Murat Bilgehan, et al.
Veröffentlicht: (2026)
von: Ertan, Murat Bilgehan, et al.
Veröffentlicht: (2026)
Cordon-MAS: Defending RAG against Knowledge Poisoning via Information-Flow Control
von: Yu, Zhe, et al.
Veröffentlicht: (2026)
von: Yu, Zhe, et al.
Veröffentlicht: (2026)
On Evaluating the Durability of Safeguards for Open-Weight LLMs
von: Qi, Xiangyu, et al.
Veröffentlicht: (2024)
von: Qi, Xiangyu, et al.
Veröffentlicht: (2024)
Safeguarding Large Language Models: A Survey
von: Dong, Yi, et al.
Veröffentlicht: (2024)
von: Dong, Yi, et al.
Veröffentlicht: (2024)
PromptKeeper: Safeguarding System Prompts for LLMs
von: Jiang, Zhifeng, et al.
Veröffentlicht: (2024)
von: Jiang, Zhifeng, et al.
Veröffentlicht: (2024)
Exploring and Developing a Pre-Model Safeguard with Draft Models
von: Cai, Hongyu, et al.
Veröffentlicht: (2026)
von: Cai, Hongyu, et al.
Veröffentlicht: (2026)
RTLMarker: Protecting LLM-Generated RTL Copyright via a Hardware Watermarking Framework
von: Wang, Kun, et al.
Veröffentlicht: (2025)
von: Wang, Kun, et al.
Veröffentlicht: (2025)
Safeguarding Federated Learning-based Road Condition Classification
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
Safeguarding AI Agents: Developing and Analyzing Safety Architectures
von: Domkundwar, Ishaan, et al.
Veröffentlicht: (2024)
von: Domkundwar, Ishaan, et al.
Veröffentlicht: (2024)
Re-Triggering Safeguards within LLMs for Jailbreak Detection
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
von: Lin, Zheng, et al.
Veröffentlicht: (2026)
Do Multimodal RAG Systems Leak Data? A Comprehensive Evaluation of Membership Inference and Image Caption Retrieval Attacks
von: Al-Lawati, Ali, et al.
Veröffentlicht: (2026)
von: Al-Lawati, Ali, et al.
Veröffentlicht: (2026)
Shattering the Echo Chamber: Hidden Safeguards in Manuscripts Against the AI Takeover of Peer Review
von: Ma, Oubo, et al.
Veröffentlicht: (2026)
von: Ma, Oubo, et al.
Veröffentlicht: (2026)
PIR-RAG: A System for Private Information Retrieval in Retrieval-Augmented Generation
von: Wang, Baiqiang, et al.
Veröffentlicht: (2025)
von: Wang, Baiqiang, et al.
Veröffentlicht: (2025)
Embedding with Large Language Models for Classification of HIPAA Safeguard Compliance Rules
von: Rahman, Md Abdur, et al.
Veröffentlicht: (2024)
von: Rahman, Md Abdur, et al.
Veröffentlicht: (2024)
Do Not Merge My Model! Safeguarding Open-Source LLMs Against Unauthorized Model Merging
von: Li, Qinfeng, et al.
Veröffentlicht: (2025)
von: Li, Qinfeng, et al.
Veröffentlicht: (2025)
CoopGuard: Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Round Attacks
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
von: Lin, Lixing, et al.
Veröffentlicht: (2026)
von: Lin, Lixing, et al.
Veröffentlicht: (2026)
ME: Trigger Element Combination Backdoor Attack on Copyright Infringement
von: Yang, Feiyu, et al.
Veröffentlicht: (2025)
von: Yang, Feiyu, et al.
Veröffentlicht: (2025)
CrossGuard: Safeguarding MLLMs against Joint-Modal Implicit Malicious Attacks
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
Bridging the Copyright Gap: Do Large Vision-Language Models Recognize and Respect Copyrighted Content?
von: Xu, Naen, et al.
Veröffentlicht: (2025)
von: Xu, Naen, et al.
Veröffentlicht: (2025)
ProjLens: Unveiling the Role of Projectors in Multimodal Model Safety
von: Wang, Kun, et al.
Veröffentlicht: (2026)
von: Wang, Kun, et al.
Veröffentlicht: (2026)
Deep Learning-based Dual Watermarking for Image Copyright Protection and Authentication
von: Padhi, Sudev Kumar, et al.
Veröffentlicht: (2025)
von: Padhi, Sudev Kumar, et al.
Veröffentlicht: (2025)
MCP Guardian: A Security-First Layer for Safeguarding MCP-Based AI System
von: Kumar, Sonu, et al.
Veröffentlicht: (2025)
von: Kumar, Sonu, et al.
Veröffentlicht: (2025)
SecPE: Secure Prompt Ensembling for Private and Robust Large Language Models
von: Zhang, Jiawen, et al.
Veröffentlicht: (2025)
von: Zhang, Jiawen, et al.
Veröffentlicht: (2025)
ReliabilityRAG: Effective and Provably Robust Defense for RAG-based Web-Search
von: Shen, Zeyu, et al.
Veröffentlicht: (2025)
von: Shen, Zeyu, et al.
Veröffentlicht: (2025)
RAG-WM: An Efficient Black-Box Watermarking Approach for Retrieval-Augmented Generation of Large Language Models
von: Lv, Peizhuo, et al.
Veröffentlicht: (2025)
von: Lv, Peizhuo, et al.
Veröffentlicht: (2025)
The Stronger the Diffusion Model, the Easier the Backdoor: Data Poisoning to Induce Copyright Breaches Without Adjusting Finetuning Pipeline
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Silent Leaks: Implicit Knowledge Extraction Attack on RAG Systems through Benign Queries
von: Wang, Yuhao, et al.
Veröffentlicht: (2025) -
Why Safeguarded Ships Run Aground? Aligned Large Language Models' Safety Mechanisms Tend to Be Anchored in The Template Region
von: Leong, Chak Tou, et al.
Veröffentlicht: (2025) -
Safeguarding Text-to-Image Generative Models Against Unauthorized Knowledge Distillation
von: Gao, Yilan, et al.
Veröffentlicht: (2026) -
AdaShield: Safeguarding Multimodal Large Language Models from Structure-based Attack via Adaptive Shield Prompting
von: Wang, Yu, et al.
Veröffentlicht: (2024) -
P$^2$RAG: Efficient Privacy-Preserving RAG Service Supporting Arbitrary Top-$k$ Retrieval
von: Ming, Yulong, et al.
Veröffentlicht: (2026)