Fact2Fiction: Targeted Poisoning Attack to Agentic Fact-checking System
Fuente:
arXiv
Salvato in:
| Autori principali: | He, Haorui, Li, Yupeng, Zhu, Bin Benjamin, Wen, Dacheng, Cheng, Reynold, Lau, Francis C. M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ADMIT: Few-shot Knowledge Poisoning Attacks on RAG-based Fact Checking
di: Wu, Yutao, et al.
Pubblicazione: (2025)
di: Wu, Yutao, et al.
Pubblicazione: (2025)
AI-Powered Spearphishing Cyber Attacks: Fact or Fiction?
di: Kemp, Matthew, et al.
Pubblicazione: (2025)
di: Kemp, Matthew, et al.
Pubblicazione: (2025)
Adversarial Attacks Against Automated Fact-Checking: A Survey
di: Liu, Fanzhen, et al.
Pubblicazione: (2025)
di: Liu, Fanzhen, et al.
Pubblicazione: (2025)
SEEP: Training Dynamics Grounds Latent Representation Search for Mitigating Backdoor Poisoning Attacks
di: He, Xuanli, et al.
Pubblicazione: (2024)
di: He, Xuanli, et al.
Pubblicazione: (2024)
RAG Safety: Exploring Knowledge Poisoning Attacks to Retrieval-Augmented Generation
di: Zhao, Tianzhe, et al.
Pubblicazione: (2025)
di: Zhao, Tianzhe, et al.
Pubblicazione: (2025)
Denial-of-Service Poisoning Attacks against Large Language Models
di: Gao, Kuofeng, et al.
Pubblicazione: (2024)
di: Gao, Kuofeng, et al.
Pubblicazione: (2024)
Forbidden Facts: An Investigation of Competing Objectives in Llama-2
di: Wang, Tony T., et al.
Pubblicazione: (2023)
di: Wang, Tony T., et al.
Pubblicazione: (2023)
SCOUT: A Defense Against Data Poisoning Attacks in Fine-Tuned Language Models
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
di: Afane, Mohamed, et al.
Pubblicazione: (2025)
Mitigating Backdoor Triggered and Targeted Data Poisoning Attacks in Voice Authentication Systems
di: Mohammadi, Alireza, et al.
Pubblicazione: (2025)
di: Mohammadi, Alireza, et al.
Pubblicazione: (2025)
Defending Against Weight-Poisoning Backdoor Attacks for Parameter-Efficient Fine-Tuning
di: Zhao, Shuai, et al.
Pubblicazione: (2024)
di: Zhao, Shuai, et al.
Pubblicazione: (2024)
Sharpness-Aware Data Poisoning Attack
di: He, Pengfei, et al.
Pubblicazione: (2023)
di: He, Pengfei, et al.
Pubblicazione: (2023)
Security Attacks on LLM-based Code Completion Tools
di: Cheng, Wen, et al.
Pubblicazione: (2024)
di: Cheng, Wen, et al.
Pubblicazione: (2024)
Retrieval-Augmented Review Generation for Poisoning Recommender Systems
di: Yang, Shiyi, et al.
Pubblicazione: (2025)
di: Yang, Shiyi, et al.
Pubblicazione: (2025)
LLM Anonymization Against Agentic Re-Identification
di: Li, Ziwen, et al.
Pubblicazione: (2026)
di: Li, Ziwen, et al.
Pubblicazione: (2026)
Fight Poison with Poison: Enhancing Robustness in Few-shot Machine-Generated Text Detection with Adversarial Training
di: Duan, Wenjing, et al.
Pubblicazione: (2026)
di: Duan, Wenjing, et al.
Pubblicazione: (2026)
DECEIVE-AFC: Adversarial Claim Attacks against Search-Enabled LLM-based Fact-Checking Systems
di: Ou, Haoran, et al.
Pubblicazione: (2026)
di: Ou, Haoran, et al.
Pubblicazione: (2026)
Supply-Chain Poisoning Attacks Against LLM Coding Agent Skill Ecosystems
di: Qu, Yubin, et al.
Pubblicazione: (2026)
di: Qu, Yubin, et al.
Pubblicazione: (2026)
Virus Infection Attack on LLMs: Your Poisoning Can Spread "VIA" Synthetic Data
di: Liang, Zi, et al.
Pubblicazione: (2025)
di: Liang, Zi, et al.
Pubblicazione: (2025)
IP Leakage Attacks Targeting LLM-Based Multi-Agent Systems
di: Wang, Liwen, et al.
Pubblicazione: (2025)
di: Wang, Liwen, et al.
Pubblicazione: (2025)
PoisonCatcher: Revealing and Identifying LDP Poisoning Attacks in IIoT
di: Shuai, Lisha, et al.
Pubblicazione: (2024)
di: Shuai, Lisha, et al.
Pubblicazione: (2024)
P2P: A Poison-to-Poison Remedy for Reliable Backdoor Defense in LLMs
di: Zhao, Shuai, et al.
Pubblicazione: (2025)
di: Zhao, Shuai, et al.
Pubblicazione: (2025)
PARASITE: Conditional System Prompt Poisoning to Hijack LLMs
di: Pham, Viet, et al.
Pubblicazione: (2025)
di: Pham, Viet, et al.
Pubblicazione: (2025)
TuBA: Cross-Lingual Transferability of Backdoor Attacks in LLMs with Instruction Tuning
di: He, Xuanli, et al.
Pubblicazione: (2024)
di: He, Xuanli, et al.
Pubblicazione: (2024)
Sugar-Coated Poison: Benign Generation Unlocks LLM Jailbreaking
di: Wu, Yu-Hang, et al.
Pubblicazione: (2025)
di: Wu, Yu-Hang, et al.
Pubblicazione: (2025)
Poisoning the Pixels: Revisiting Backdoor Attacks on Semantic Segmentation
di: Zhang, Guangsheng, et al.
Pubblicazione: (2026)
di: Zhang, Guangsheng, et al.
Pubblicazione: (2026)
Membership Inference Attacks Against In-Context Learning
di: Wen, Rui, et al.
Pubblicazione: (2024)
di: Wen, Rui, et al.
Pubblicazione: (2024)
Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data
di: Baumgärtner, Tim, et al.
Pubblicazione: (2024)
di: Baumgärtner, Tim, et al.
Pubblicazione: (2024)
Enhancing Prompt Injection Attacks to LLMs via Poisoning Alignment
di: Shao, Zedian, et al.
Pubblicazione: (2024)
di: Shao, Zedian, et al.
Pubblicazione: (2024)
Logit Poisoning Attack in Distillation-based Federated Learning and its Countermeasures
di: Yu, Yonghao, et al.
Pubblicazione: (2024)
di: Yu, Yonghao, et al.
Pubblicazione: (2024)
Debating Truth: Debate-driven Claim Verification with Multiple Large Language Model Agents
di: He, Haorui, et al.
Pubblicazione: (2025)
di: He, Haorui, et al.
Pubblicazione: (2025)
From Theory to Practice: Evaluating Data Poisoning Attacks and Defenses in In-Context Learning on Social Media Health Discourse
di: Jhuma, Rabeya Amin, et al.
Pubblicazione: (2025)
di: Jhuma, Rabeya Amin, et al.
Pubblicazione: (2025)
PoisonBench: Assessing Large Language Model Vulnerability to Data Poisoning
di: Fu, Tingchen, et al.
Pubblicazione: (2024)
di: Fu, Tingchen, et al.
Pubblicazione: (2024)
AdaPPA: Adaptive Position Pre-Fill Jailbreak Attack Approach Targeting LLMs
di: Lv, Lijia, et al.
Pubblicazione: (2024)
di: Lv, Lijia, et al.
Pubblicazione: (2024)
Provable Watermarking for Data Poisoning Attacks
di: Zhu, Yifan, et al.
Pubblicazione: (2025)
di: Zhu, Yifan, et al.
Pubblicazione: (2025)
Disabling Self-Correction in Retrieval-Augmented Generation via Stealthy Retriever Poisoning
di: Dai, Yanbo, et al.
Pubblicazione: (2025)
di: Dai, Yanbo, et al.
Pubblicazione: (2025)
FlexLLM: Exploring LLM Customization for Moving Target Defense on Black-Box LLMs Against Jailbreak Attacks
di: Chen, Bocheng, et al.
Pubblicazione: (2024)
di: Chen, Bocheng, et al.
Pubblicazione: (2024)
Architecture Matters: Comparing RAG Systems under Knowledge Base Poisoning
di: Korn, Samuel
Pubblicazione: (2026)
di: Korn, Samuel
Pubblicazione: (2026)
Bias Amplification in RAG: Poisoning Knowledge Retrieval to Steer LLMs
di: Wang, Linlin, et al.
Pubblicazione: (2025)
di: Wang, Linlin, et al.
Pubblicazione: (2025)
RTD-Guard: A Black-Box Textual Adversarial Detection Framework via Replacement Token Detection
di: Zhu, He, et al.
Pubblicazione: (2026)
di: Zhu, He, et al.
Pubblicazione: (2026)
Learning to Poison Large Language Models for Downstream Manipulation
di: Zhou, Xiangyu, et al.
Pubblicazione: (2024)
di: Zhou, Xiangyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ADMIT: Few-shot Knowledge Poisoning Attacks on RAG-based Fact Checking
di: Wu, Yutao, et al.
Pubblicazione: (2025) -
AI-Powered Spearphishing Cyber Attacks: Fact or Fiction?
di: Kemp, Matthew, et al.
Pubblicazione: (2025) -
Adversarial Attacks Against Automated Fact-Checking: A Survey
di: Liu, Fanzhen, et al.
Pubblicazione: (2025) -
SEEP: Training Dynamics Grounds Latent Representation Search for Mitigating Backdoor Poisoning Attacks
di: He, Xuanli, et al.
Pubblicazione: (2024) -
RAG Safety: Exploring Knowledge Poisoning Attacks to Retrieval-Augmented Generation
di: Zhao, Tianzhe, et al.
Pubblicazione: (2025)