Robust Thompson Sampling Algorithms Against Reward Poisoning Attacks
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Yinglun, Wang, Zhiwei, Singh, Gagandeep |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Binary Reward Labeling: Bridging Offline Preference and Reward-Based Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Learning a Pessimistic Reward Model in RLHF
by: Xu, Yinglun, et al.
Published: (2025)
by: Xu, Yinglun, et al.
Published: (2025)
RAMP: Boosting Adversarial Robustness Against Multiple $l_p$ Perturbations for Universal Robustness
by: Jiang, Enyi, et al.
Published: (2024)
by: Jiang, Enyi, et al.
Published: (2024)
Preference Poisoning Attacks on Reward Model Learning
by: Wu, Junlin, et al.
Published: (2024)
by: Wu, Junlin, et al.
Published: (2024)
Provable Robustness of (Graph) Neural Networks Against Data Poisoning and Backdoor Attacks
by: Gosch, Lukas, et al.
Published: (2024)
by: Gosch, Lukas, et al.
Published: (2024)
Poisoning Attacks on LLMs Require a Near-constant Number of Poison Samples
by: Souly, Alexandra, et al.
Published: (2025)
by: Souly, Alexandra, et al.
Published: (2025)
Adversarial Training for Defense Against Label Poisoning Attacks
by: Bal, Melis Ilayda, et al.
Published: (2025)
by: Bal, Melis Ilayda, et al.
Published: (2025)
Compression Aware Certified Training
by: Xu, Changming, et al.
Published: (2025)
by: Xu, Changming, et al.
Published: (2025)
On the Relevance of Byzantine Robust Optimization Against Data Poisoning
by: Farhadkhani, Sadegh, et al.
Published: (2024)
by: Farhadkhani, Sadegh, et al.
Published: (2024)
Thompson Sampling-like Algorithms for Stochastic Rising Bandits
by: Fiandri, Marco, et al.
Published: (2025)
by: Fiandri, Marco, et al.
Published: (2025)
Gradient Purification: Defense Against Poisoning Attack in Decentralized Federated Learning
by: Li, Bin, et al.
Published: (2025)
by: Li, Bin, et al.
Published: (2025)
When Can You Poison Rewards? A Tight Characterization of Reward Poisoning in Linear MDPs
by: Escamilla, Jose Efraim Aguilar, et al.
Published: (2026)
by: Escamilla, Jose Efraim Aguilar, et al.
Published: (2026)
Is Thompson Sampling Susceptible to Algorithmic Collusion?
by: Xiong, Yi, et al.
Published: (2024)
by: Xiong, Yi, et al.
Published: (2024)
Thompson Sampling for Repeated Newsvendor
by: Chen, Li, et al.
Published: (2025)
by: Chen, Li, et al.
Published: (2025)
Are LLM-Enhanced Graph Neural Networks Robust against Poisoning Attacks?
by: Ma, Yuhang, et al.
Published: (2026)
by: Ma, Yuhang, et al.
Published: (2026)
Can In-Context Reinforcement Learning Recover From Reward Poisoning Attacks?
by: Sasnauskas, Paulius, et al.
Published: (2025)
by: Sasnauskas, Paulius, et al.
Published: (2025)
Defending Against Sophisticated Poisoning Attacks with RL-based Aggregation in Federated Learning
by: Wang, Yujing, et al.
Published: (2024)
by: Wang, Yujing, et al.
Published: (2024)
Cross-Input Certified Training for Universal Perturbations
by: Xu, Changming, et al.
Published: (2024)
by: Xu, Changming, et al.
Published: (2024)
Shadowcast: Stealthy Data Poisoning Attacks Against Vision-Language Models
by: Xu, Yuancheng, et al.
Published: (2024)
by: Xu, Yuancheng, et al.
Published: (2024)
A Systematic Review of Poisoning Attacks Against Large Language Models
by: Fendley, Neil, et al.
Published: (2025)
by: Fendley, Neil, et al.
Published: (2025)
Defending Against Knowledge Poisoning Attacks During Retrieval-Augmented Generation
by: Edemacu, Kennedy, et al.
Published: (2025)
by: Edemacu, Kennedy, et al.
Published: (2025)
Towards Generalized Certified Robustness with Multi-Norm Training
by: Jiang, Enyi, et al.
Published: (2024)
by: Jiang, Enyi, et al.
Published: (2024)
Defending Against Poisoning Attacks in Federated Learning with Blockchain
by: Dong, Nanqing, et al.
Published: (2023)
by: Dong, Nanqing, et al.
Published: (2023)
SecureLearn -- An Attack-agnostic Defense for Multiclass Machine Learning Against Data Poisoning Attacks
by: Paracha, Anum, et al.
Published: (2025)
by: Paracha, Anum, et al.
Published: (2025)
Formal Synthesis of Certifiably Robust Neural Lyapunov-Barrier Certificates
by: Wang, Chengxiao, et al.
Published: (2026)
by: Wang, Chengxiao, et al.
Published: (2026)
Bypassing the Safety Training of Open-Source LLMs with Priming Attacks
by: Vega, Jason, et al.
Published: (2023)
by: Vega, Jason, et al.
Published: (2023)
SleeperNets: Universal Backdoor Poisoning Attacks Against Reinforcement Learning Agents
by: Rathbun, Ethan, et al.
Published: (2024)
by: Rathbun, Ethan, et al.
Published: (2024)
Online Poisoning Attack Against Reinforcement Learning under Black-box Environments
by: Li, Jianhui, et al.
Published: (2024)
by: Li, Jianhui, et al.
Published: (2024)
Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning
by: Liu, Shijie, et al.
Published: (2025)
by: Liu, Shijie, et al.
Published: (2025)
On the Robustness of Random Forest Against Untargeted Data Poisoning: An Ensemble-Based Approach
by: Anisetti, Marco, et al.
Published: (2022)
by: Anisetti, Marco, et al.
Published: (2022)
How to Defend Against Large-scale Model Poisoning Attacks in Federated Learning: A Vertical Solution
by: Wang, Jinbo, et al.
Published: (2024)
by: Wang, Jinbo, et al.
Published: (2024)
Reward Under Attack: Analyzing the Robustness and Hackability of Process Reward Models
by: Tiwari, Rishabh, et al.
Published: (2026)
by: Tiwari, Rishabh, et al.
Published: (2026)
Relational DNN Verification With Cross Executional Bound Refinement
by: Banerjee, Debangshu, et al.
Published: (2024)
by: Banerjee, Debangshu, et al.
Published: (2024)
Interactive Machine Learning: From Theory to Scale
by: Zhu, Yinglun
Published: (2025)
by: Zhu, Yinglun
Published: (2025)
Reward-Preserving Attacks For Robust Reinforcement Learning
by: Schott, Lucas, et al.
Published: (2026)
by: Schott, Lucas, et al.
Published: (2026)
Logits Poisoning Attack in Federated Distillation
by: Tang, Yuhan, et al.
Published: (2024)
by: Tang, Yuhan, et al.
Published: (2024)
Constrained Linear Thompson Sampling
by: Gangrade, Aditya, et al.
Published: (2025)
by: Gangrade, Aditya, et al.
Published: (2025)
Are Targeted Data Poisoning Attacks as Effective as We Think?
by: Xu, William, et al.
Published: (2025)
by: Xu, William, et al.
Published: (2025)
Is The Watermarking Of LLM-Generated Code Robust?
by: Suresh, Tarun, et al.
Published: (2024)
by: Suresh, Tarun, et al.
Published: (2024)
Similar Items
-
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024) -
Binary Reward Labeling: Bridging Offline Preference and Reward-Based Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024) -
Learning a Pessimistic Reward Model in RLHF
by: Xu, Yinglun, et al.
Published: (2025) -
RAMP: Boosting Adversarial Robustness Against Multiple $l_p$ Perturbations for Universal Robustness
by: Jiang, Enyi, et al.
Published: (2024) -
Preference Poisoning Attacks on Reward Model Learning
by: Wu, Junlin, et al.
Published: (2024)