Turning Generative Models Degenerate: The Power of Data Poisoning Attacks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Shuli, Kadhe, Swanand Ravindra, Zhou, Yi, Ahmed, Farhan, Cai, Ling, Baracaldo, Nathalie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards a Re-evaluation of Data Forging Attacks in Practice
von: Suliman, Mohamed, et al.
Veröffentlicht: (2024)
von: Suliman, Mohamed, et al.
Veröffentlicht: (2024)
In-Context Probing for Membership Inference in Fine-Tuned Language Models
von: Lu, Zhexi, et al.
Veröffentlicht: (2025)
von: Lu, Zhexi, et al.
Veröffentlicht: (2025)
Split, Unlearn, Merge: Leveraging Data Attributes for More Effective Unlearning in LLMs
von: Kadhe, Swanand Ravindra, et al.
Veröffentlicht: (2024)
von: Kadhe, Swanand Ravindra, et al.
Veröffentlicht: (2024)
Human-Imperceptible Retrieval Poisoning Attacks in LLM-Powered Applications
von: Zhang, Quan, et al.
Veröffentlicht: (2024)
von: Zhang, Quan, et al.
Veröffentlicht: (2024)
AgentSCOPE: Evaluating Contextual Privacy Across Agentic Workflows
von: Ngong, Ivoline C., et al.
Veröffentlicht: (2026)
von: Ngong, Ivoline C., et al.
Veröffentlicht: (2026)
CSC: Turning the Adversary's Poison against Itself
von: Shi, Yuchen, et al.
Veröffentlicht: (2026)
von: Shi, Yuchen, et al.
Veröffentlicht: (2026)
Vulnerabilities in AI Code Generators: Exploring Targeted Data Poisoning Attacks
von: Cotroneo, Domenico, et al.
Veröffentlicht: (2023)
von: Cotroneo, Domenico, et al.
Veröffentlicht: (2023)
System Prompt Poisoning: Persistent Attacks on Large Language Models Beyond User Injection
von: Li, Zongze, et al.
Veröffentlicht: (2025)
von: Li, Zongze, et al.
Veröffentlicht: (2025)
Poison Once, Exploit Forever: Environment-Injected Memory Poisoning Attacks on Web Agents
von: Zou, Wei, et al.
Veröffentlicht: (2026)
von: Zou, Wei, et al.
Veröffentlicht: (2026)
Nightshade: Prompt-Specific Poisoning Attacks on Text-to-Image Generative Models
von: Shan, Shawn, et al.
Veröffentlicht: (2023)
von: Shan, Shawn, et al.
Veröffentlicht: (2023)
One Shot Dominance: Knowledge Poisoning Attack on Retrieval-Augmented Generation Systems
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2025)
When and Where do Data Poisons Attack Textual Inversion?
von: Styborski, Jeremy, et al.
Veröffentlicht: (2025)
von: Styborski, Jeremy, et al.
Veröffentlicht: (2025)
CBPF: Filtering Poisoned Data Based on Composite Backdoor Attack
von: Xia, Hanfeng, et al.
Veröffentlicht: (2024)
von: Xia, Hanfeng, et al.
Veröffentlicht: (2024)
BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
Protecting Users From Themselves: Safeguarding Contextual Privacy in Interactions with Conversational Agents
von: Ngong, Ivoline, et al.
Veröffentlicht: (2025)
von: Ngong, Ivoline, et al.
Veröffentlicht: (2025)
Precision Guided Approach to Mitigate Data Poisoning Attacks in Federated Learning
von: Kumar, K Naveen, et al.
Veröffentlicht: (2024)
von: Kumar, K Naveen, et al.
Veröffentlicht: (2024)
Robustness Analysis of Machine Learning Models for IoT Intrusion Detection Under Data Poisoning Attacks
von: Wulnye, Fortunatus Aabangbio, et al.
Veröffentlicht: (2026)
von: Wulnye, Fortunatus Aabangbio, et al.
Veröffentlicht: (2026)
SafeCOMM: A Study on Safety Degradation in Fine-Tuned Telecom Large Language Models
von: Djuhera, Aladin, et al.
Veröffentlicht: (2025)
von: Djuhera, Aladin, et al.
Veröffentlicht: (2025)
PIDP-Attack: Combining Prompt Injection with Database Poisoning Attacks on Retrieval-Augmented Generation Systems
von: Wang, Haozhen, et al.
Veröffentlicht: (2026)
von: Wang, Haozhen, et al.
Veröffentlicht: (2026)
Hidden in the Metadata: Stealth Poisoning Attacks on Multimodal Retrieval-Augmented Generation
von: Edemacu, Kennedy, et al.
Veröffentlicht: (2026)
von: Edemacu, Kennedy, et al.
Veröffentlicht: (2026)
Knowledge Poisoning Attacks on Medical Multi-Modal Retrieval-Augmented Generation
von: Yang, Peiru, et al.
Veröffentlicht: (2026)
von: Yang, Peiru, et al.
Veröffentlicht: (2026)
Defending Against Beta Poisoning Attacks in Machine Learning Models
von: Gulciftci, Nilufer, et al.
Veröffentlicht: (2025)
von: Gulciftci, Nilufer, et al.
Veröffentlicht: (2025)
Great, Now Write an Article About That: The Crescendo Multi-Turn LLM Jailbreak Attack
von: Russinovich, Mark, et al.
Veröffentlicht: (2024)
von: Russinovich, Mark, et al.
Veröffentlicht: (2024)
FedCC: Robust Federated Learning against Model Poisoning Attacks
von: Jeong, Hyejun, et al.
Veröffentlicht: (2022)
von: Jeong, Hyejun, et al.
Veröffentlicht: (2022)
Joint-GCG: Unified Gradient-Based Poisoning Attacks on Retrieval-Augmented Generation Systems
von: Wang, Haowei, et al.
Veröffentlicht: (2025)
von: Wang, Haowei, et al.
Veröffentlicht: (2025)
LoopTrap: Termination Poisoning Attacks on LLM Agents
von: Xu, Huiyu, et al.
Veröffentlicht: (2026)
von: Xu, Huiyu, et al.
Veröffentlicht: (2026)
Shadowcast: Stealthy Data Poisoning Attacks Against Vision-Language Models
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
RevPRAG: Revealing Poisoning Attacks in Retrieval-Augmented Generation through LLM Activation Analysis
von: Tan, Xue, et al.
Veröffentlicht: (2024)
von: Tan, Xue, et al.
Veröffentlicht: (2024)
A Set of Generalized Components to Achieve Effective Poison-only Clean-label Backdoor Attacks with Collaborative Sample Selection and Triggers
von: Wu, Zhixiao, et al.
Veröffentlicht: (2025)
von: Wu, Zhixiao, et al.
Veröffentlicht: (2025)
Detecting Data Poisoning in Code Generation LLMs via Black-Box, Vulnerability-Oriented Scanning
von: Yan, Shenao, et al.
Veröffentlicht: (2026)
von: Yan, Shenao, et al.
Veröffentlicht: (2026)
When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents
von: Liu, Shi, et al.
Veröffentlicht: (2026)
von: Liu, Shi, et al.
Veröffentlicht: (2026)
Data Poisoning Attacks on Off-Policy Policy Evaluation Methods
von: Lobo, Elita, et al.
Veröffentlicht: (2024)
von: Lobo, Elita, et al.
Veröffentlicht: (2024)
FIDELIS: Blockchain-Enabled Protection Against Poisoning Attacks in Federated Learning
von: Carney, Jane, et al.
Veröffentlicht: (2025)
von: Carney, Jane, et al.
Veröffentlicht: (2025)
Dual Defense: Enhancing Privacy and Mitigating Poisoning Attacks in Federated Learning
von: Xu, Runhua, et al.
Veröffentlicht: (2025)
von: Xu, Runhua, et al.
Veröffentlicht: (2025)
Data-Free Model-Related Attacks: Unleashing the Potential of Generative AI
von: Ye, Dayong, et al.
Veröffentlicht: (2025)
von: Ye, Dayong, et al.
Veröffentlicht: (2025)
Bidirectional Intention Inference Enhances LLMs' Defense Against Multi-Turn Jailbreak Attacks
von: Tong, Haibo, et al.
Veröffentlicht: (2025)
von: Tong, Haibo, et al.
Veröffentlicht: (2025)
ACE: A Model Poisoning Attack on Contribution Evaluation Methods in Federated Learning
von: Xu, Zhangchen, et al.
Veröffentlicht: (2024)
von: Xu, Zhangchen, et al.
Veröffentlicht: (2024)
Reasoning-Style Poisoning of LLM Agents via Stealthy Style Transfer: Process-Level Attacks and Runtime Monitoring in RSV Space
von: Zhou, Xingfu, et al.
Veröffentlicht: (2025)
von: Zhou, Xingfu, et al.
Veröffentlicht: (2025)
Evaluating the Dynamics of Membership Privacy in Deep Learning
von: Chen, Yuetian, et al.
Veröffentlicht: (2025)
von: Chen, Yuetian, et al.
Veröffentlicht: (2025)
Scaling Trends for Data Poisoning in LLMs
von: Bowen, Dillon, et al.
Veröffentlicht: (2024)
von: Bowen, Dillon, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards a Re-evaluation of Data Forging Attacks in Practice
von: Suliman, Mohamed, et al.
Veröffentlicht: (2024) -
In-Context Probing for Membership Inference in Fine-Tuned Language Models
von: Lu, Zhexi, et al.
Veröffentlicht: (2025) -
Split, Unlearn, Merge: Leveraging Data Attributes for More Effective Unlearning in LLMs
von: Kadhe, Swanand Ravindra, et al.
Veröffentlicht: (2024) -
Human-Imperceptible Retrieval Poisoning Attacks in LLM-Powered Applications
von: Zhang, Quan, et al.
Veröffentlicht: (2024) -
AgentSCOPE: Evaluating Contextual Privacy Across Agentic Workflows
von: Ngong, Ivoline C., et al.
Veröffentlicht: (2026)