CODE: A Contradiction-Based Deliberation Extension Framework for Overthinking Attacks on Retrieval-Augmented Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Xiaolei, Jia, Xiaojun, Chen, Liquan, Li, Songze |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SEAL-Tag: Self-Tag Evidence Aggregation with Probabilistic Circuits for PII-Safe Retrieval-Augmented Generation
por: Xie, Jin, et al.
Publicado: (2026)
por: Xie, Jin, et al.
Publicado: (2026)
TUNI: A Textual Unimodal Detector for Identity Inference in CLIP Models
por: Li, Songze, et al.
Publicado: (2024)
por: Li, Songze, et al.
Publicado: (2024)
One Shot Dominance: Knowledge Poisoning Attack on Retrieval-Augmented Generation Systems
por: Chang, Zhiyuan, et al.
Publicado: (2025)
por: Chang, Zhiyuan, et al.
Publicado: (2025)
ReCIT: Reconstructing Full Private Data from Gradient in Parameter-Efficient Fine-Tuning of Large Language Models
por: Xie, Jin, et al.
Publicado: (2025)
por: Xie, Jin, et al.
Publicado: (2025)
Benchmarking Knowledge-Extraction Attack and Defense on Retrieval-Augmented Generation
por: Qi, Zhisheng, et al.
Publicado: (2026)
por: Qi, Zhisheng, et al.
Publicado: (2026)
Securing Retrieval-Augmented Generation: A Taxonomy of Attacks, Defenses, and Future Directions
por: Xu, Yuming, et al.
Publicado: (2026)
por: Xu, Yuming, et al.
Publicado: (2026)
CPA-RAG:Covert Poisoning Attacks on Retrieval-Augmented Generation in Large Language Models
por: Li, Chunyang, et al.
Publicado: (2025)
por: Li, Chunyang, et al.
Publicado: (2025)
Generating Is Believing: Membership Inference Attacks against Retrieval-Augmented Generation
por: Li, Yuying, et al.
Publicado: (2024)
por: Li, Yuying, et al.
Publicado: (2024)
When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack
por: Sun, Zehan, et al.
Publicado: (2026)
por: Sun, Zehan, et al.
Publicado: (2026)
Knowledge-Driven Multi-Turn Jailbreaking on Large Language Models
por: Li, Songze, et al.
Publicado: (2026)
por: Li, Songze, et al.
Publicado: (2026)
Joint-GCG: Unified Gradient-Based Poisoning Attacks on Retrieval-Augmented Generation Systems
por: Wang, Haowei, et al.
Publicado: (2025)
por: Wang, Haowei, et al.
Publicado: (2025)
MIRAGE: Misleading Retrieval-Augmented Generation via Black-box and Query-agnostic Poisoning Attacks
por: Chen, Tailun, et al.
Publicado: (2025)
por: Chen, Tailun, et al.
Publicado: (2025)
BadThink: Triggered Overthinking Attacks on Chain-of-Thought Reasoning in Large Language Models
por: Liu, Shuaitong, et al.
Publicado: (2025)
por: Liu, Shuaitong, et al.
Publicado: (2025)
JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks
por: Zhang, Xiaoyu, et al.
Publicado: (2023)
por: Zhang, Xiaoyu, et al.
Publicado: (2023)
Practical Poisoning Attacks against Retrieval-Augmented Generation
por: Zhang, Baolei, et al.
Publicado: (2025)
por: Zhang, Baolei, et al.
Publicado: (2025)
External Data Extraction Attacks against Retrieval-Augmented Large Language Models
por: He, Yu, et al.
Publicado: (2025)
por: He, Yu, et al.
Publicado: (2025)
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
por: Zou, Wei, et al.
Publicado: (2024)
por: Zou, Wei, et al.
Publicado: (2024)
Traceback of Poisoning Attacks to Retrieval-Augmented Generation
por: Zhang, Baolei, et al.
Publicado: (2025)
por: Zhang, Baolei, et al.
Publicado: (2025)
BackdoorIndicator: Leveraging OOD Data for Proactive Backdoor Detection in Federated Learning
por: Li, Songze, et al.
Publicado: (2024)
por: Li, Songze, et al.
Publicado: (2024)
Arbitrary-Threshold Fully Homomorphic Encryption with Lower Complexity
por: Chang, Yijia, et al.
Publicado: (2025)
por: Chang, Yijia, et al.
Publicado: (2025)
Beyond Explicit Refusals: Soft-Failure Attacks on Retrieval-Augmented Generation
por: Zhang, Wentao, et al.
Publicado: (2026)
por: Zhang, Wentao, et al.
Publicado: (2026)
Benchmarking Poisoning Attacks against Retrieval-Augmented Generation
por: Zhang, Baolei, et al.
Publicado: (2025)
por: Zhang, Baolei, et al.
Publicado: (2025)
RAG Safety: Exploring Knowledge Poisoning Attacks to Retrieval-Augmented Generation
por: Zhao, Tianzhe, et al.
Publicado: (2025)
por: Zhao, Tianzhe, et al.
Publicado: (2025)
Hoist with His Own Petard: Inducing Guardrails to Facilitate Denial-of-Service Attacks on Retrieval-Augmented Generation of LLMs
por: Suo, Pan, et al.
Publicado: (2025)
por: Suo, Pan, et al.
Publicado: (2025)
CIBER: A Comprehensive Benchmark for Security Evaluation of Code Interpreter Agents
por: Ba, Lei, et al.
Publicado: (2026)
por: Ba, Lei, et al.
Publicado: (2026)
Odysseus: Jailbreaking Commercial Multimodal LLM-integrated Systems via Dual Steganography
por: Li, Songze, et al.
Publicado: (2025)
por: Li, Songze, et al.
Publicado: (2025)
CODE ACROSTIC: Robust Watermarking for Code Generation
por: Lin, Li, et al.
Publicado: (2025)
por: Lin, Li, et al.
Publicado: (2025)
Data Extraction Attacks in Retrieval-Augmented Generation via Backdoors
por: Peng, Yuefeng, et al.
Publicado: (2024)
por: Peng, Yuefeng, et al.
Publicado: (2024)
Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning
por: Deng, Gelei, et al.
Publicado: (2024)
por: Deng, Gelei, et al.
Publicado: (2024)
UniC-RAG: Universal Knowledge Corruption Attacks to Retrieval-Augmented Generation
por: Geng, Runpeng, et al.
Publicado: (2025)
por: Geng, Runpeng, et al.
Publicado: (2025)
Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models
por: Clop, Cody, et al.
Publicado: (2024)
por: Clop, Cody, et al.
Publicado: (2024)
Bi-Erasing: A Bidirectional Framework for Concept Removal in Diffusion Models
por: Chen, Hao, et al.
Publicado: (2025)
por: Chen, Hao, et al.
Publicado: (2025)
Noise as a Probe: Membership Inference Attacks on Diffusion Models Leveraging Initial Noise
por: Lian, Puwei, et al.
Publicado: (2026)
por: Lian, Puwei, et al.
Publicado: (2026)
Enhancing Membership Inference Attacks on Diffusion Models from a Frequency-Domain Perspective
por: Lian, Puwei, et al.
Publicado: (2025)
por: Lian, Puwei, et al.
Publicado: (2025)
BiAxisAudit: A Novel Framework to Evaluate LLM Bias Across Prompt Sensitivity and Response-Layer Divergence
por: Gan, Jialing, et al.
Publicado: (2026)
por: Gan, Jialing, et al.
Publicado: (2026)
PRAG: End-to-End Privacy-Preserving Retrieval-Augmented Generation
por: Li, Zhijun, et al.
Publicado: (2026)
por: Li, Zhijun, et al.
Publicado: (2026)
TooBadRL: Trigger Optimization to Boost Effectiveness of Backdoor Attacks on Deep Reinforcement Learning
por: Zhang, Mingxuan, et al.
Publicado: (2025)
por: Zhang, Mingxuan, et al.
Publicado: (2025)
Hidden in the Metadata: Stealth Poisoning Attacks on Multimodal Retrieval-Augmented Generation
por: Edemacu, Kennedy, et al.
Publicado: (2026)
por: Edemacu, Kennedy, et al.
Publicado: (2026)
Knowledge Poisoning Attacks on Medical Multi-Modal Retrieval-Augmented Generation
por: Yang, Peiru, et al.
Publicado: (2026)
por: Yang, Peiru, et al.
Publicado: (2026)
Using Retriever Augmented Large Language Models for Attack Graph Generation
por: Prapty, Renascence Tarafder, et al.
Publicado: (2024)
por: Prapty, Renascence Tarafder, et al.
Publicado: (2024)
Ejemplares similares
-
SEAL-Tag: Self-Tag Evidence Aggregation with Probabilistic Circuits for PII-Safe Retrieval-Augmented Generation
por: Xie, Jin, et al.
Publicado: (2026) -
TUNI: A Textual Unimodal Detector for Identity Inference in CLIP Models
por: Li, Songze, et al.
Publicado: (2024) -
One Shot Dominance: Knowledge Poisoning Attack on Retrieval-Augmented Generation Systems
por: Chang, Zhiyuan, et al.
Publicado: (2025) -
ReCIT: Reconstructing Full Private Data from Gradient in Parameter-Efficient Fine-Tuning of Large Language Models
por: Xie, Jin, et al.
Publicado: (2025) -
Benchmarking Knowledge-Extraction Attack and Defense on Retrieval-Augmented Generation
por: Qi, Zhisheng, et al.
Publicado: (2026)