CODE: A Contradiction-Based Deliberation Extension Framework for Overthinking Attacks on Retrieval-Augmented Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Xiaolei, Jia, Xiaojun, Chen, Liquan, Li, Songze |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SEAL-Tag: Self-Tag Evidence Aggregation with Probabilistic Circuits for PII-Safe Retrieval-Augmented Generation
von: Xie, Jin, et al.
Veröffentlicht: (2026)
von: Xie, Jin, et al.
Veröffentlicht: (2026)
TUNI: A Textual Unimodal Detector for Identity Inference in CLIP Models
von: Li, Songze, et al.
Veröffentlicht: (2024)
von: Li, Songze, et al.
Veröffentlicht: (2024)
One Shot Dominance: Knowledge Poisoning Attack on Retrieval-Augmented Generation Systems
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2025)
ReCIT: Reconstructing Full Private Data from Gradient in Parameter-Efficient Fine-Tuning of Large Language Models
von: Xie, Jin, et al.
Veröffentlicht: (2025)
von: Xie, Jin, et al.
Veröffentlicht: (2025)
Benchmarking Knowledge-Extraction Attack and Defense on Retrieval-Augmented Generation
von: Qi, Zhisheng, et al.
Veröffentlicht: (2026)
von: Qi, Zhisheng, et al.
Veröffentlicht: (2026)
Securing Retrieval-Augmented Generation: A Taxonomy of Attacks, Defenses, and Future Directions
von: Xu, Yuming, et al.
Veröffentlicht: (2026)
von: Xu, Yuming, et al.
Veröffentlicht: (2026)
CPA-RAG:Covert Poisoning Attacks on Retrieval-Augmented Generation in Large Language Models
von: Li, Chunyang, et al.
Veröffentlicht: (2025)
von: Li, Chunyang, et al.
Veröffentlicht: (2025)
Generating Is Believing: Membership Inference Attacks against Retrieval-Augmented Generation
von: Li, Yuying, et al.
Veröffentlicht: (2024)
von: Li, Yuying, et al.
Veröffentlicht: (2024)
When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack
von: Sun, Zehan, et al.
Veröffentlicht: (2026)
von: Sun, Zehan, et al.
Veröffentlicht: (2026)
Knowledge-Driven Multi-Turn Jailbreaking on Large Language Models
von: Li, Songze, et al.
Veröffentlicht: (2026)
von: Li, Songze, et al.
Veröffentlicht: (2026)
Joint-GCG: Unified Gradient-Based Poisoning Attacks on Retrieval-Augmented Generation Systems
von: Wang, Haowei, et al.
Veröffentlicht: (2025)
von: Wang, Haowei, et al.
Veröffentlicht: (2025)
MIRAGE: Misleading Retrieval-Augmented Generation via Black-box and Query-agnostic Poisoning Attacks
von: Chen, Tailun, et al.
Veröffentlicht: (2025)
von: Chen, Tailun, et al.
Veröffentlicht: (2025)
BadThink: Triggered Overthinking Attacks on Chain-of-Thought Reasoning in Large Language Models
von: Liu, Shuaitong, et al.
Veröffentlicht: (2025)
von: Liu, Shuaitong, et al.
Veröffentlicht: (2025)
JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2023)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2023)
Practical Poisoning Attacks against Retrieval-Augmented Generation
von: Zhang, Baolei, et al.
Veröffentlicht: (2025)
von: Zhang, Baolei, et al.
Veröffentlicht: (2025)
External Data Extraction Attacks against Retrieval-Augmented Large Language Models
von: He, Yu, et al.
Veröffentlicht: (2025)
von: He, Yu, et al.
Veröffentlicht: (2025)
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
von: Zou, Wei, et al.
Veröffentlicht: (2024)
von: Zou, Wei, et al.
Veröffentlicht: (2024)
Traceback of Poisoning Attacks to Retrieval-Augmented Generation
von: Zhang, Baolei, et al.
Veröffentlicht: (2025)
von: Zhang, Baolei, et al.
Veröffentlicht: (2025)
BackdoorIndicator: Leveraging OOD Data for Proactive Backdoor Detection in Federated Learning
von: Li, Songze, et al.
Veröffentlicht: (2024)
von: Li, Songze, et al.
Veröffentlicht: (2024)
Arbitrary-Threshold Fully Homomorphic Encryption with Lower Complexity
von: Chang, Yijia, et al.
Veröffentlicht: (2025)
von: Chang, Yijia, et al.
Veröffentlicht: (2025)
Beyond Explicit Refusals: Soft-Failure Attacks on Retrieval-Augmented Generation
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
Benchmarking Poisoning Attacks against Retrieval-Augmented Generation
von: Zhang, Baolei, et al.
Veröffentlicht: (2025)
von: Zhang, Baolei, et al.
Veröffentlicht: (2025)
RAG Safety: Exploring Knowledge Poisoning Attacks to Retrieval-Augmented Generation
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2025)
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2025)
Hoist with His Own Petard: Inducing Guardrails to Facilitate Denial-of-Service Attacks on Retrieval-Augmented Generation of LLMs
von: Suo, Pan, et al.
Veröffentlicht: (2025)
von: Suo, Pan, et al.
Veröffentlicht: (2025)
CIBER: A Comprehensive Benchmark for Security Evaluation of Code Interpreter Agents
von: Ba, Lei, et al.
Veröffentlicht: (2026)
von: Ba, Lei, et al.
Veröffentlicht: (2026)
Odysseus: Jailbreaking Commercial Multimodal LLM-integrated Systems via Dual Steganography
von: Li, Songze, et al.
Veröffentlicht: (2025)
von: Li, Songze, et al.
Veröffentlicht: (2025)
CODE ACROSTIC: Robust Watermarking for Code Generation
von: Lin, Li, et al.
Veröffentlicht: (2025)
von: Lin, Li, et al.
Veröffentlicht: (2025)
Data Extraction Attacks in Retrieval-Augmented Generation via Backdoors
von: Peng, Yuefeng, et al.
Veröffentlicht: (2024)
von: Peng, Yuefeng, et al.
Veröffentlicht: (2024)
Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning
von: Deng, Gelei, et al.
Veröffentlicht: (2024)
von: Deng, Gelei, et al.
Veröffentlicht: (2024)
UniC-RAG: Universal Knowledge Corruption Attacks to Retrieval-Augmented Generation
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models
von: Clop, Cody, et al.
Veröffentlicht: (2024)
von: Clop, Cody, et al.
Veröffentlicht: (2024)
Bi-Erasing: A Bidirectional Framework for Concept Removal in Diffusion Models
von: Chen, Hao, et al.
Veröffentlicht: (2025)
von: Chen, Hao, et al.
Veröffentlicht: (2025)
Noise as a Probe: Membership Inference Attacks on Diffusion Models Leveraging Initial Noise
von: Lian, Puwei, et al.
Veröffentlicht: (2026)
von: Lian, Puwei, et al.
Veröffentlicht: (2026)
Enhancing Membership Inference Attacks on Diffusion Models from a Frequency-Domain Perspective
von: Lian, Puwei, et al.
Veröffentlicht: (2025)
von: Lian, Puwei, et al.
Veröffentlicht: (2025)
BiAxisAudit: A Novel Framework to Evaluate LLM Bias Across Prompt Sensitivity and Response-Layer Divergence
von: Gan, Jialing, et al.
Veröffentlicht: (2026)
von: Gan, Jialing, et al.
Veröffentlicht: (2026)
PRAG: End-to-End Privacy-Preserving Retrieval-Augmented Generation
von: Li, Zhijun, et al.
Veröffentlicht: (2026)
von: Li, Zhijun, et al.
Veröffentlicht: (2026)
TooBadRL: Trigger Optimization to Boost Effectiveness of Backdoor Attacks on Deep Reinforcement Learning
von: Zhang, Mingxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Mingxuan, et al.
Veröffentlicht: (2025)
Hidden in the Metadata: Stealth Poisoning Attacks on Multimodal Retrieval-Augmented Generation
von: Edemacu, Kennedy, et al.
Veröffentlicht: (2026)
von: Edemacu, Kennedy, et al.
Veröffentlicht: (2026)
Knowledge Poisoning Attacks on Medical Multi-Modal Retrieval-Augmented Generation
von: Yang, Peiru, et al.
Veröffentlicht: (2026)
von: Yang, Peiru, et al.
Veröffentlicht: (2026)
Using Retriever Augmented Large Language Models for Attack Graph Generation
von: Prapty, Renascence Tarafder, et al.
Veröffentlicht: (2024)
von: Prapty, Renascence Tarafder, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SEAL-Tag: Self-Tag Evidence Aggregation with Probabilistic Circuits for PII-Safe Retrieval-Augmented Generation
von: Xie, Jin, et al.
Veröffentlicht: (2026) -
TUNI: A Textual Unimodal Detector for Identity Inference in CLIP Models
von: Li, Songze, et al.
Veröffentlicht: (2024) -
One Shot Dominance: Knowledge Poisoning Attack on Retrieval-Augmented Generation Systems
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2025) -
ReCIT: Reconstructing Full Private Data from Gradient in Parameter-Efficient Fine-Tuning of Large Language Models
von: Xie, Jin, et al.
Veröffentlicht: (2025) -
Benchmarking Knowledge-Extraction Attack and Defense on Retrieval-Augmented Generation
von: Qi, Zhisheng, et al.
Veröffentlicht: (2026)