UniC-RAG: Universal Knowledge Corruption Attacks to Retrieval-Augmented Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Geng, Runpeng, Wang, Yanting, Chen, Ying, Jia, Jinyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
POISONCRAFT: Practical Poisoning of Retrieval-Augmented Generation for Large Language Models
von: Shao, Yangguang, et al.
Veröffentlicht: (2025)
von: Shao, Yangguang, et al.
Veröffentlicht: (2025)
AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
Defending against Backdoor Attacks via Module Switching
von: Li, Weijun, et al.
Veröffentlicht: (2025)
von: Li, Weijun, et al.
Veröffentlicht: (2025)
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
von: Zou, Wei, et al.
Veröffentlicht: (2024)
von: Zou, Wei, et al.
Veröffentlicht: (2024)
Evaluating the Robustness of Large Language Model Safety Guardrails Against Adversarial Attacks
von: Young, Richard J.
Veröffentlicht: (2025)
von: Young, Richard J.
Veröffentlicht: (2025)
sudoLLM: On Multi-role Alignment of Language Models
von: Saha, Soumadeep, et al.
Veröffentlicht: (2025)
von: Saha, Soumadeep, et al.
Veröffentlicht: (2025)
Operationalizing a Threat Model for Red-Teaming Large Language Models (LLMs)
von: Verma, Apurv, et al.
Veröffentlicht: (2024)
von: Verma, Apurv, et al.
Veröffentlicht: (2024)
Jailbreaking Attacks vs. Content Safety Filters: How Far Are We in the LLM Safety Arms Race?
von: Xin, Yuan, et al.
Veröffentlicht: (2025)
von: Xin, Yuan, et al.
Veröffentlicht: (2025)
Blind Spots in the Guard: How Domain-Camouflaged Injection Attacks Evade Detection in Multi-Agent LLM Systems
von: Pai, Aaditya
Veröffentlicht: (2026)
von: Pai, Aaditya
Veröffentlicht: (2026)
DWFS-Obfuscation: Dynamic Weighted Feature Selection for Robust Malware Familial Classification under Obfuscation
von: Wei, Xingyuan, et al.
Veröffentlicht: (2025)
von: Wei, Xingyuan, et al.
Veröffentlicht: (2025)
Real AI Agents with Fake Memories: Fatal Context Manipulation Attacks on Web3 Agents
von: Patlan, Atharv Singh, et al.
Veröffentlicht: (2025)
von: Patlan, Atharv Singh, et al.
Veröffentlicht: (2025)
Temporal Attack Pattern Detection in Multi-Agent AI Workflows: An Open Framework for Training Trace-Based Security Models
von: Del Rosario, Ron F.
Veröffentlicht: (2025)
von: Del Rosario, Ron F.
Veröffentlicht: (2025)
Adversarial Attacks on Large Language Models Using Regularized Relaxation
von: Chacko, Samuel Jacob, et al.
Veröffentlicht: (2024)
von: Chacko, Samuel Jacob, et al.
Veröffentlicht: (2024)
Ignore Me But Don't Replace Me: Utilizing Non-Linguistic Elements for Pretraining on the Cybersecurity Domain
von: Jang, Eugene, et al.
Veröffentlicht: (2024)
von: Jang, Eugene, et al.
Veröffentlicht: (2024)
MarkLLM: An Open-Source Toolkit for LLM Watermarking
von: Pan, Leyi, et al.
Veröffentlicht: (2024)
von: Pan, Leyi, et al.
Veröffentlicht: (2024)
Large Language Models are Advanced Anonymizers
von: Staab, Robin, et al.
Veröffentlicht: (2024)
von: Staab, Robin, et al.
Veröffentlicht: (2024)
Watermarking Degrades Alignment in Language Models: Analysis and Mitigation
von: Verma, Apurv, et al.
Veröffentlicht: (2025)
von: Verma, Apurv, et al.
Veröffentlicht: (2025)
Prompted Contextual Vectors for Spear-Phishing Detection
von: Nahmias, Daniel, et al.
Veröffentlicht: (2024)
von: Nahmias, Daniel, et al.
Veröffentlicht: (2024)
A Semantic Invariant Robust Watermark for Large Language Models
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
Can Watermarked LLMs be Identified by Users via Crafted Prompts?
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
Train to Defend: First Defense Against Cryptanalytic Neural Network Parameter Extraction Attacks
von: Kurian, Ashley, et al.
Veröffentlicht: (2025)
von: Kurian, Ashley, et al.
Veröffentlicht: (2025)
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
von: Palit, Sayon, et al.
Veröffentlicht: (2025)
von: Palit, Sayon, et al.
Veröffentlicht: (2025)
Super Suffixes: Bypassing Text Generation Alignment and Guard Models Simultaneously
von: Adiletta, Andrew, et al.
Veröffentlicht: (2025)
von: Adiletta, Andrew, et al.
Veröffentlicht: (2025)
Seeing the Forest through the Trees: Data Leakage from Partial Transformer Gradients
von: Li, Weijun, et al.
Veröffentlicht: (2024)
von: Li, Weijun, et al.
Veröffentlicht: (2024)
GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing
von: Zhang, Peiyan, et al.
Veröffentlicht: (2025)
von: Zhang, Peiyan, et al.
Veröffentlicht: (2025)
Terrarium: Revisiting the Blackboard for Multi-Agent Safety, Privacy, and Security Studies
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
von: Nakamura, Mason, et al.
Veröffentlicht: (2025)
AI Safeguards, Generative AI and the Pandora Box: AI Safety Measures to Protect Businesses and Personal Reputation
von: Kumar, Prasanna
Veröffentlicht: (2026)
von: Kumar, Prasanna
Veröffentlicht: (2026)
PromptSAM+: Malware Detection based on Prompt Segment Anything Model
von: Wei, Xingyuan, et al.
Veröffentlicht: (2024)
von: Wei, Xingyuan, et al.
Veröffentlicht: (2024)
Reducing Information Overload: Because Even Security Experts Need to Blink
von: Kuehn, Philipp, et al.
Veröffentlicht: (2022)
von: Kuehn, Philipp, et al.
Veröffentlicht: (2022)
RAR: Setting Knowledge Tripwires for Retrieval Augmented Rejection
von: Buonocore, Tommaso Mario, et al.
Veröffentlicht: (2025)
von: Buonocore, Tommaso Mario, et al.
Veröffentlicht: (2025)
The Quantum State Continuity Problem and Temporal Enforcement Against Fork Attacks
von: Ünsal, Samet
Veröffentlicht: (2025)
von: Ünsal, Samet
Veröffentlicht: (2025)
Protection Is (Nearly) All You Need: Structural Protection Dominates Scoring in Globally Capped KV Eviction
von: Garcia, Gabriel
Veröffentlicht: (2026)
von: Garcia, Gabriel
Veröffentlicht: (2026)
SecEmb: Sparsity-Aware Secure Federated Learning of On-Device Recommender System with Large Embedding
von: Mai, Peihua, et al.
Veröffentlicht: (2025)
von: Mai, Peihua, et al.
Veröffentlicht: (2025)
ConfusionPrompt: Practical Private Inference for Online Large Language Models
von: Mai, Peihua, et al.
Veröffentlicht: (2023)
von: Mai, Peihua, et al.
Veröffentlicht: (2023)
Split-and-Denoise: Protect large language model inference with local differential privacy
von: Mai, Peihua, et al.
Veröffentlicht: (2023)
von: Mai, Peihua, et al.
Veröffentlicht: (2023)
Efficient LLM Safety Evaluation through Multi-Agent Debate
von: Lin, Dachuan, et al.
Veröffentlicht: (2025)
von: Lin, Dachuan, et al.
Veröffentlicht: (2025)
Mitigating the Impact of Malware Evolution on API Sequence-based Windows Malware Detector
von: Wei, Xingyuan, et al.
Veröffentlicht: (2024)
von: Wei, Xingyuan, et al.
Veröffentlicht: (2024)
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
von: Othman, Refat
Veröffentlicht: (2026)
von: Othman, Refat
Veröffentlicht: (2026)
Benchmarking Large Language Models for IoC Recovery under Adversarial Code Obfuscation and Encryption
von: Morales, Jaime, et al.
Veröffentlicht: (2026)
von: Morales, Jaime, et al.
Veröffentlicht: (2026)
MASH: Evading Black-Box AI-Generated Text Detectors via Style Humanization
von: Gu, Yongtong, et al.
Veröffentlicht: (2026)
von: Gu, Yongtong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
POISONCRAFT: Practical Poisoning of Retrieval-Augmented Generation for Large Language Models
von: Shao, Yangguang, et al.
Veröffentlicht: (2025) -
AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption
von: Wang, Yanting, et al.
Veröffentlicht: (2025) -
Defending against Backdoor Attacks via Module Switching
von: Li, Weijun, et al.
Veröffentlicht: (2025) -
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
von: Zou, Wei, et al.
Veröffentlicht: (2024) -
Evaluating the Robustness of Large Language Model Safety Guardrails Against Adversarial Attacks
von: Young, Richard J.
Veröffentlicht: (2025)