Beware Untrusted Simulators -- Reward-Free Backdoor Attacks in Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Rathbun, Ethan, Lin, Wo Wei, Oprea, Alina, Amato, Christopher |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adversarial Inception Backdoor Attacks against Reinforcement Learning
by: Rathbun, Ethan, et al.
Published: (2024)
by: Rathbun, Ethan, et al.
Published: (2024)
SleeperNets: Universal Backdoor Poisoning Attacks Against Reinforcement Learning Agents
by: Rathbun, Ethan, et al.
Published: (2024)
by: Rathbun, Ethan, et al.
Published: (2024)
Backdoor Attacks in Peer-to-Peer Federated Learning
by: Syros, Georgios, et al.
Published: (2023)
by: Syros, Georgios, et al.
Published: (2023)
Thought-Transfer: Indirect Targeted Poisoning Attacks on Chain-of-Thought Reasoning Models
by: Chaudhari, Harsh, et al.
Published: (2026)
by: Chaudhari, Harsh, et al.
Published: (2026)
ARMOR: Robust Reinforcement Learning-based Control for UAVs under Physical Attacks
by: Dash, Pritam, et al.
Published: (2025)
by: Dash, Pritam, et al.
Published: (2025)
Hierarchical Multi-agent Reinforcement Learning for Cyber Network Defense
by: Singh, Aditya Vikram, et al.
Published: (2024)
by: Singh, Aditya Vikram, et al.
Published: (2024)
Dataset Poisoning Attacks on Behavioral Cloning Policies
by: Kalra, Akansha, et al.
Published: (2025)
by: Kalra, Akansha, et al.
Published: (2025)
Manipulating Trajectory Prediction with Backdoors
by: Messaoud, Kaouther, et al.
Published: (2023)
by: Messaoud, Kaouther, et al.
Published: (2023)
Phantom: General Backdoor Attacks on Retrieval Augmented Language Generation
by: Chaudhari, Harsh, et al.
Published: (2024)
by: Chaudhari, Harsh, et al.
Published: (2024)
Black-Box Privacy Attacks on Shared Representations in Multitask Learning
by: Abascal, John, et al.
Published: (2025)
by: Abascal, John, et al.
Published: (2025)
Learning-based Detection of GPS Spoofing Attack for Quadrotors
by: Wang, Pengyu, et al.
Published: (2025)
by: Wang, Pengyu, et al.
Published: (2025)
How Vulnerable Is My Learned Policy? Universal Adversarial Perturbation Attacks On Modern Behavior Cloning Policies
by: Kalra, Akansha, et al.
Published: (2025)
by: Kalra, Akansha, et al.
Published: (2025)
Hiding Backdoors within Event Sequence Data via Poisoning Attacks
by: Ermilova, Alina, et al.
Published: (2023)
by: Ermilova, Alina, et al.
Published: (2023)
TooBadRL: Trigger Optimization to Boost Effectiveness of Backdoor Attacks on Deep Reinforcement Learning
by: Zhang, Mingxuan, et al.
Published: (2025)
by: Zhang, Mingxuan, et al.
Published: (2025)
Cooperative Decentralized Backdoor Attacks on Vertical Federated Learning
by: Lee, Seohyun, et al.
Published: (2025)
by: Lee, Seohyun, et al.
Published: (2025)
Attacks and Mitigations for Distributed Governance of Agentic AI under Byzantine Adversaries
by: Laws, Matthew D., et al.
Published: (2026)
by: Laws, Matthew D., et al.
Published: (2026)
ARBoids: Adaptive Residual Reinforcement Learning With Boids Model for Cooperative Multi-USV Target Defense
by: Tao, Jiyue, et al.
Published: (2025)
by: Tao, Jiyue, et al.
Published: (2025)
RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors
by: Bai, Fengshuo, et al.
Published: (2024)
by: Bai, Fengshuo, et al.
Published: (2024)
Secure Control Systems for Autonomous Quadrotors against Cyber-Attacks
by: Belkadi, Samuel
Published: (2024)
by: Belkadi, Samuel
Published: (2024)
Reconstruction of Personally Identifiable Information from Supervised Finetuned Models
by: Furukawa, Sae, et al.
Published: (2026)
by: Furukawa, Sae, et al.
Published: (2026)
Persistent Backdoor Attacks in Continual Learning
by: Guo, Zhen, et al.
Published: (2024)
by: Guo, Zhen, et al.
Published: (2024)
User Inference Attacks on Large Language Models
by: Kandpal, Nikhil, et al.
Published: (2023)
by: Kandpal, Nikhil, et al.
Published: (2023)
Can In-Context Reinforcement Learning Recover From Reward Poisoning Attacks?
by: Sasnauskas, Paulius, et al.
Published: (2025)
by: Sasnauskas, Paulius, et al.
Published: (2025)
On the Out-of-Distribution Backdoor Attack for Federated Learning
by: Xu, Jiahao, et al.
Published: (2025)
by: Xu, Jiahao, et al.
Published: (2025)
TMI! Finetuned Models Leak Private Information from their Pretraining Data
by: Abascal, John, et al.
Published: (2023)
by: Abascal, John, et al.
Published: (2023)
PoolFlip: A Multi-Agent Reinforcement Learning Security Environment for Cyber Defense
by: Cadet, Xavier, et al.
Published: (2025)
by: Cadet, Xavier, et al.
Published: (2025)
UTrace: Poisoning Forensics for Private Collaborative Learning
by: Rose, Evan, et al.
Published: (2024)
by: Rose, Evan, et al.
Published: (2024)
Lurking in the shadows: Unveiling Stealthy Backdoor Attacks against Personalized Federated Learning
by: Lyu, Xiaoting, et al.
Published: (2024)
by: Lyu, Xiaoting, et al.
Published: (2024)
Defending against Backdoor Attack on Deep Neural Networks
by: Cheng, Hao, et al.
Published: (2020)
by: Cheng, Hao, et al.
Published: (2020)
Rethinking Graph Backdoor Attacks: A Distribution-Preserving Perspective
by: Zhang, Zhiwei, et al.
Published: (2024)
by: Zhang, Zhiwei, et al.
Published: (2024)
Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem
by: Lin, Shuyi, et al.
Published: (2025)
by: Lin, Shuyi, et al.
Published: (2025)
Imperio: Language-Guided Backdoor Attacks for Arbitrary Model Control
by: Chow, Ka-Ho, et al.
Published: (2024)
by: Chow, Ka-Ho, et al.
Published: (2024)
Rogue Cell: Adversarial Attack and Defense in Untrusted O-RAN Setup Exploiting the Traffic Steering xApp
by: Aizikovich, Eran, et al.
Published: (2025)
by: Aizikovich, Eran, et al.
Published: (2025)
Syntax- and Compilation-Preserving Evasion of LLM Vulnerability Detectors
by: Sun, Luze, et al.
Published: (2026)
by: Sun, Luze, et al.
Published: (2026)
Backdoor Attacks on Decentralised Post-Training
by: Ersoy, Oğuzhan, et al.
Published: (2026)
by: Ersoy, Oğuzhan, et al.
Published: (2026)
Detecting Backdoor Attacks via Similarity in Semantic Communication Systems
by: Wei, Ziyang, et al.
Published: (2025)
by: Wei, Ziyang, et al.
Published: (2025)
Let's Focus: Focused Backdoor Attack against Federated Transfer Learning
by: Arazzi, Marco, et al.
Published: (2024)
by: Arazzi, Marco, et al.
Published: (2024)
Structure-Aware Distributed Backdoor Attacks in Federated Learning
by: Jian, Wang, et al.
Published: (2026)
by: Jian, Wang, et al.
Published: (2026)
EmInspector: Combating Backdoor Attacks in Federated Self-Supervised Learning Through Embedding Inspection
by: Qian, Yuwen, et al.
Published: (2024)
by: Qian, Yuwen, et al.
Published: (2024)
Dynamic Free-Rider Detection in Federated Learning via Simulated Attack Patterns
by: Nakamura, Motoki
Published: (2026)
by: Nakamura, Motoki
Published: (2026)
Similar Items
-
Adversarial Inception Backdoor Attacks against Reinforcement Learning
by: Rathbun, Ethan, et al.
Published: (2024) -
SleeperNets: Universal Backdoor Poisoning Attacks Against Reinforcement Learning Agents
by: Rathbun, Ethan, et al.
Published: (2024) -
Backdoor Attacks in Peer-to-Peer Federated Learning
by: Syros, Georgios, et al.
Published: (2023) -
Thought-Transfer: Indirect Targeted Poisoning Attacks on Chain-of-Thought Reasoning Models
by: Chaudhari, Harsh, et al.
Published: (2026) -
ARMOR: Robust Reinforcement Learning-based Control for UAVs under Physical Attacks
by: Dash, Pritam, et al.
Published: (2025)