PoisonedParrot: Subtle Data Poisoning Attacks to Elicit Copyright-Infringing Content from Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Panaitescu-Liess, Michael-Andrei, Pathmanathan, Pankayaraj, Kaya, Yigitcan, Che, Zora, An, Bang, Zhu, Sicheng, Agrawal, Aakriti, Huang, Furong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Like Oil and Water: Group Robustness Methods and Poisoning Defenses May Be at Odds
por: Panaitescu-Liess, Michael-Andrei, et al.
Publicado: (2025)
por: Panaitescu-Liess, Michael-Andrei, et al.
Publicado: (2025)
RAGPart & RAGMask: Retrieval-Stage Defenses Against Corpus Poisoning in Retrieval-Augmented Generation
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2025)
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2025)
Automatic Pseudo-Harmful Prompt Generation for Evaluating False Refusals in Large Language Models
por: An, Bang, et al.
Publicado: (2024)
por: An, Bang, et al.
Publicado: (2024)
Can Watermarking Large Language Models Prevent Copyrighted Text Generation and Hide Training Data?
por: Panaitescu-Liess, Michael-Andrei, et al.
Publicado: (2024)
por: Panaitescu-Liess, Michael-Andrei, et al.
Publicado: (2024)
Is poisoning a real threat to LLM alignment? Maybe more so than you think
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2024)
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2024)
PoisonCatcher: Revealing and Identifying LDP Poisoning Attacks in IIoT
por: Shuai, Lisha, et al.
Publicado: (2024)
por: Shuai, Lisha, et al.
Publicado: (2024)
AdvBDGen: Adversarially Fortified Prompt-Specific Fuzzy Backdoor Generator Against LLM Alignment
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2024)
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2024)
Sharpness-Aware Data Poisoning Attack
por: He, Pengfei, et al.
Publicado: (2023)
por: He, Pengfei, et al.
Publicado: (2023)
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation
por: Liu, Yinuo, et al.
Publicado: (2025)
por: Liu, Yinuo, et al.
Publicado: (2025)
Transferable Availability Poisoning Attacks
por: Liu, Yiyong, et al.
Publicado: (2023)
por: Liu, Yiyong, et al.
Publicado: (2023)
Poison with Style: A Practical Poisoning Attack on Code Large Language Models
por: Tran, Khang, et al.
Publicado: (2026)
por: Tran, Khang, et al.
Publicado: (2026)
Mitigating Data Poisoning Attacks to Local Differential Privacy
por: Li, Xiaolin, et al.
Publicado: (2025)
por: Li, Xiaolin, et al.
Publicado: (2025)
Poisoning Attacks to Local Differential Privacy for Ranking Estimation
por: Zhan, Pei, et al.
Publicado: (2025)
por: Zhan, Pei, et al.
Publicado: (2025)
Poisoning the Pixels: Revisiting Backdoor Attacks on Semantic Segmentation
por: Zhang, Guangsheng, et al.
Publicado: (2026)
por: Zhang, Guangsheng, et al.
Publicado: (2026)
PACE: Poisoning Attacks on Learned Cardinality Estimation
por: Zhang, Jintao, et al.
Publicado: (2024)
por: Zhang, Jintao, et al.
Publicado: (2024)
Poison Once, Exploit Forever: Environment-Injected Memory Poisoning Attacks on Web Agents
por: Zou, Wei, et al.
Publicado: (2026)
por: Zou, Wei, et al.
Publicado: (2026)
Provable Watermarking for Data Poisoning Attacks
por: Zhu, Yifan, et al.
Publicado: (2025)
por: Zhu, Yifan, et al.
Publicado: (2025)
Shadowcast: Stealthy Data Poisoning Attacks Against Vision-Language Models
por: Xu, Yuancheng, et al.
Publicado: (2024)
por: Xu, Yuancheng, et al.
Publicado: (2024)
Effectiveness of Adversarial Benign and Malware Examples in Evasion and Poisoning Attacks
por: Kozák, Matouš, et al.
Publicado: (2025)
por: Kozák, Matouš, et al.
Publicado: (2025)
Fake Resume Attacks: Data Poisoning on Online Job Platforms
por: Yamashita, Michiharu, et al.
Publicado: (2024)
por: Yamashita, Michiharu, et al.
Publicado: (2024)
Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM
por: Yu, Lei, et al.
Publicado: (2025)
por: Yu, Lei, et al.
Publicado: (2025)
Defense against Poisoning Attacks under Shuffle-DP
por: Wang, Siyi, et al.
Publicado: (2026)
por: Wang, Siyi, et al.
Publicado: (2026)
GShield: Mitigating Poisoning Attacks in Federated Learning
por: M., Sameera K., et al.
Publicado: (2025)
por: M., Sameera K., et al.
Publicado: (2025)
Indiscriminate Data Poisoning Attacks on Neural Networks
por: Lu, Yiwei, et al.
Publicado: (2022)
por: Lu, Yiwei, et al.
Publicado: (2022)
Deep-Research Agents Can Be Poisoned via User-Generated Content
por: Zhang, Tingwei, et al.
Publicado: (2026)
por: Zhang, Tingwei, et al.
Publicado: (2026)
Blockchain Address Poisoning
por: Tsuchiya, Taro, et al.
Publicado: (2025)
por: Tsuchiya, Taro, et al.
Publicado: (2025)
Phantom: Untargeted Poisoning Attacks on Semi-Supervised Learning (Full Version)
por: Knauer, Jonathan, et al.
Publicado: (2024)
por: Knauer, Jonathan, et al.
Publicado: (2024)
Poison Attacks and Adversarial Prompts Against an Informed University Virtual Assistant
por: Fernandez, Ivan A., et al.
Publicado: (2024)
por: Fernandez, Ivan A., et al.
Publicado: (2024)
Logit Poisoning Attack in Distillation-based Federated Learning and its Countermeasures
por: Yu, Yonghao, et al.
Publicado: (2024)
por: Yu, Yonghao, et al.
Publicado: (2024)
Model Poisoning Attacks to Federated Learning via Multi-Round Consistency
por: Xie, Yueqi, et al.
Publicado: (2024)
por: Xie, Yueqi, et al.
Publicado: (2024)
On the Robustness of LDP Protocols for Numerical Attributes under Data Poisoning Attacks
por: Li, Xiaoguang, et al.
Publicado: (2024)
por: Li, Xiaoguang, et al.
Publicado: (2024)
Activation Gradient based Poisoned Sample Detection Against Backdoor Attacks
por: Yuan, Danni, et al.
Publicado: (2023)
por: Yuan, Danni, et al.
Publicado: (2023)
FedRecAttack: Model Poisoning Attack to Federated Recommendation
por: Rong, Dazhong, et al.
Publicado: (2022)
por: Rong, Dazhong, et al.
Publicado: (2022)
Hidden Poison: Machine Unlearning Enables Camouflaged Poisoning Attacks
por: Di, Jimmy Z., et al.
Publicado: (2022)
por: Di, Jimmy Z., et al.
Publicado: (2022)
Data Poisoning Attacks to Local Differential Privacy Protocols for Graphs
por: He, Xi, et al.
Publicado: (2024)
por: He, Xi, et al.
Publicado: (2024)
LoopTrap: Termination Poisoning Attacks on LLM Agents
por: Xu, Huiyu, et al.
Publicado: (2026)
por: Xu, Huiyu, et al.
Publicado: (2026)
Local Environment Poisoning Attacks on Federated Reinforcement Learning
por: Ma, Evelyn, et al.
Publicado: (2023)
por: Ma, Evelyn, et al.
Publicado: (2023)
Poisoning Attacks and Defenses in Recommender Systems: A Survey
por: Wang, Zongwei, et al.
Publicado: (2024)
por: Wang, Zongwei, et al.
Publicado: (2024)
Inverting Gradient Attacks Makes Powerful Data Poisoning
por: Bouaziz, Wassim, et al.
Publicado: (2024)
por: Bouaziz, Wassim, et al.
Publicado: (2024)
VisPoison: An Effective Backdoor Attack Framework for Tabular Data Visualization Models
por: Li, Shuaimin, et al.
Publicado: (2024)
por: Li, Shuaimin, et al.
Publicado: (2024)
Ejemplares similares
-
Like Oil and Water: Group Robustness Methods and Poisoning Defenses May Be at Odds
por: Panaitescu-Liess, Michael-Andrei, et al.
Publicado: (2025) -
RAGPart & RAGMask: Retrieval-Stage Defenses Against Corpus Poisoning in Retrieval-Augmented Generation
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2025) -
Automatic Pseudo-Harmful Prompt Generation for Evaluating False Refusals in Large Language Models
por: An, Bang, et al.
Publicado: (2024) -
Can Watermarking Large Language Models Prevent Copyrighted Text Generation and Hide Training Data?
por: Panaitescu-Liess, Michael-Andrei, et al.
Publicado: (2024) -
Is poisoning a real threat to LLM alignment? Maybe more so than you think
por: Pathmanathan, Pankayaraj, et al.
Publicado: (2024)