Can In-Context Reinforcement Learning Recover From Reward Poisoning Attacks?
Fuente:
arXiv
Saved in:
| Main Authors: | Sasnauskas, Paulius, Yalın, Yiğit, Radanović, Goran |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Independent Learning in Performative Markov Potential Games
by: Sahitaj, Rilind, et al.
Published: (2025)
by: Sahitaj, Rilind, et al.
Published: (2025)
Distributionally Robust Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2025)
by: Mandal, Debmalya, et al.
Published: (2025)
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Local Environment Poisoning Attacks on Federated Reinforcement Learning
by: Ma, Evelyn, et al.
Published: (2023)
by: Ma, Evelyn, et al.
Published: (2023)
SleeperNets: Universal Backdoor Poisoning Attacks Against Reinforcement Learning Agents
by: Rathbun, Ethan, et al.
Published: (2024)
by: Rathbun, Ethan, et al.
Published: (2024)
Online Poisoning Attack Against Reinforcement Learning under Black-box Environments
by: Li, Jianhui, et al.
Published: (2024)
by: Li, Jianhui, et al.
Published: (2024)
GShield: Mitigating Poisoning Attacks in Federated Learning
by: M., Sameera K., et al.
Published: (2025)
by: M., Sameera K., et al.
Published: (2025)
Transferable Availability Poisoning Attacks
by: Liu, Yiyong, et al.
Published: (2023)
by: Liu, Yiyong, et al.
Published: (2023)
SecureLearn -- An Attack-agnostic Defense for Multiclass Machine Learning Against Data Poisoning Attacks
by: Paracha, Anum, et al.
Published: (2025)
by: Paracha, Anum, et al.
Published: (2025)
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation
by: Liu, Yinuo, et al.
Published: (2025)
by: Liu, Yinuo, et al.
Published: (2025)
Sybil-based Virtual Data Poisoning Attacks in Federated Learning
by: Zhu, Changxun, et al.
Published: (2025)
by: Zhu, Changxun, et al.
Published: (2025)
From Theory to Practice: Evaluating Data Poisoning Attacks and Defenses in In-Context Learning on Social Media Health Discourse
by: Jhuma, Rabeya Amin, et al.
Published: (2025)
by: Jhuma, Rabeya Amin, et al.
Published: (2025)
Provable Watermarking for Data Poisoning Attacks
by: Zhu, Yifan, et al.
Published: (2025)
by: Zhu, Yifan, et al.
Published: (2025)
Poison with Style: A Practical Poisoning Attack on Code Large Language Models
by: Tran, Khang, et al.
Published: (2026)
by: Tran, Khang, et al.
Published: (2026)
FedRecAttack: Model Poisoning Attack to Federated Recommendation
by: Rong, Dazhong, et al.
Published: (2022)
by: Rong, Dazhong, et al.
Published: (2022)
Beware Untrusted Simulators -- Reward-Free Backdoor Attacks in Reinforcement Learning
by: Rathbun, Ethan, et al.
Published: (2026)
by: Rathbun, Ethan, et al.
Published: (2026)
Partner in Crime: Boosting Targeted Poisoning Attacks against Federated Learning
by: Sun, Shihua, et al.
Published: (2024)
by: Sun, Shihua, et al.
Published: (2024)
Towards Efficient and Certified Recovery from Poisoning Attacks in Federated Learning
by: Jiang, Yu, et al.
Published: (2024)
by: Jiang, Yu, et al.
Published: (2024)
Using Anomaly Detection to Detect Poisoning Attacks in Federated Learning Applications
by: Raza, Ali, et al.
Published: (2022)
by: Raza, Ali, et al.
Published: (2022)
Indiscriminate Data Poisoning Attacks on Neural Networks
by: Lu, Yiwei, et al.
Published: (2022)
by: Lu, Yiwei, et al.
Published: (2022)
Devil's Hand: Data Poisoning Attacks to Locally Private Graph Learning Protocols
by: He, Longzhu, et al.
Published: (2025)
by: He, Longzhu, et al.
Published: (2025)
Defending Against Sophisticated Poisoning Attacks with RL-based Aggregation in Federated Learning
by: Wang, Yujing, et al.
Published: (2024)
by: Wang, Yujing, et al.
Published: (2024)
Inverting Gradient Attacks Makes Powerful Data Poisoning
by: Bouaziz, Wassim, et al.
Published: (2024)
by: Bouaziz, Wassim, et al.
Published: (2024)
PoisonedParrot: Subtle Data Poisoning Attacks to Elicit Copyright-Infringing Content from Large Language Models
by: Panaitescu-Liess, Michael-Andrei, et al.
Published: (2025)
by: Panaitescu-Liess, Michael-Andrei, et al.
Published: (2025)
Poisoning Attacks to Local Differential Privacy Protocols for Trajectory Data
by: Hsu, I-Jung, et al.
Published: (2025)
by: Hsu, I-Jung, et al.
Published: (2025)
Indiscriminate Data Poisoning Attacks on Pre-trained Feature Extractors
by: Lu, Yiwei, et al.
Published: (2024)
by: Lu, Yiwei, et al.
Published: (2024)
Data Poisoning Attacks in Intelligent Transportation Systems: A Survey
by: Wang, Feilong, et al.
Published: (2024)
by: Wang, Feilong, et al.
Published: (2024)
Debiased Graph Poisoning Attack via Contrastive Surrogate Objective
by: Yoon, Kanghoon, et al.
Published: (2024)
by: Yoon, Kanghoon, et al.
Published: (2024)
Enhancing the Antidote: Improved Pointwise Certifications against Poisoning Attacks
by: Liu, Shijie, et al.
Published: (2023)
by: Liu, Shijie, et al.
Published: (2023)
FreqFed: A Frequency Analysis-Based Approach for Mitigating Poisoning Attacks in Federated Learning
by: Fereidooni, Hossein, et al.
Published: (2023)
by: Fereidooni, Hossein, et al.
Published: (2023)
FedRDF: A Robust and Dynamic Aggregation Function against Poisoning Attacks in Federated Learning
by: Campos, Enrique Mármol, et al.
Published: (2024)
by: Campos, Enrique Mármol, et al.
Published: (2024)
A Data-Driven Defense against Edge-case Model Poisoning Attacks on Federated Learning
by: Purohit, Kiran, et al.
Published: (2023)
by: Purohit, Kiran, et al.
Published: (2023)
Dataset Poisoning Attacks on Behavioral Cloning Policies
by: Kalra, Akansha, et al.
Published: (2025)
by: Kalra, Akansha, et al.
Published: (2025)
A Systematic Review of Poisoning Attacks Against Large Language Models
by: Fendley, Neil, et al.
Published: (2025)
by: Fendley, Neil, et al.
Published: (2025)
Data Poisoning Attacks to Locally Differentially Private Range Query Protocols
by: Liao, Ting-Wei, et al.
Published: (2025)
by: Liao, Ting-Wei, et al.
Published: (2025)
Hiding Backdoors within Event Sequence Data via Poisoning Attacks
by: Ermilova, Alina, et al.
Published: (2023)
by: Ermilova, Alina, et al.
Published: (2023)
The SkipSponge Attack: Sponge Weight Poisoning of Deep Neural Networks
by: Lintelo, Jona te, et al.
Published: (2024)
by: Lintelo, Jona te, et al.
Published: (2024)
How to Defend Against Large-scale Model Poisoning Attacks in Federated Learning: A Vertical Solution
by: Wang, Jinbo, et al.
Published: (2024)
by: Wang, Jinbo, et al.
Published: (2024)
XOXO: Stealthy Cross-Origin Context Poisoning Attacks against AI Coding Assistants
by: Štorek, Adam, et al.
Published: (2025)
by: Štorek, Adam, et al.
Published: (2025)
MCPTox: A Benchmark for Tool Poisoning Attack on Real-World MCP Servers
by: Wang, Zhiqiang, et al.
Published: (2025)
by: Wang, Zhiqiang, et al.
Published: (2025)
Similar Items
-
Independent Learning in Performative Markov Potential Games
by: Sahitaj, Rilind, et al.
Published: (2025) -
Distributionally Robust Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2025) -
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024) -
Local Environment Poisoning Attacks on Federated Reinforcement Learning
by: Ma, Evelyn, et al.
Published: (2023) -
SleeperNets: Universal Backdoor Poisoning Attacks Against Reinforcement Learning Agents
by: Rathbun, Ethan, et al.
Published: (2024)