Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nie, Buqing, Fu, Yangqing, Ji, Jingtian, Gao, Yue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Select before Act: Spatially Decoupled Action Repetition for Continuous Control
von: Nie, Buqing, et al.
Veröffentlicht: (2025)
von: Nie, Buqing, et al.
Veröffentlicht: (2025)
Adversarial Constrained Policy Optimization: Improving Constrained Reinforcement Learning by Adapting Budgets
von: Ma, Jianmina, et al.
Veröffentlicht: (2024)
von: Ma, Jianmina, et al.
Veröffentlicht: (2024)
Disturbance-Aware Adaptive Compensation in Hybrid Force-Position Locomotion Policy for Legged Robots
von: Zhang, Yang, et al.
Veröffentlicht: (2025)
von: Zhang, Yang, et al.
Veröffentlicht: (2025)
Can Tabular Foundation Models Guide Exploration in Robot Policy Learning?
von: Ou, Buqing, et al.
Veröffentlicht: (2026)
von: Ou, Buqing, et al.
Veröffentlicht: (2026)
On the Tension Between Optimality and Adversarial Robustness in Policy Optimization
von: Li, Haoran, et al.
Veröffentlicht: (2025)
von: Li, Haoran, et al.
Veröffentlicht: (2025)
Toward Evaluating Robustness of Reinforcement Learning with Adversarial Policy
von: Zheng, Xiang, et al.
Veröffentlicht: (2023)
von: Zheng, Xiang, et al.
Veröffentlicht: (2023)
Actor-Accelerated Policy Dual Averaging for Reinforcement Learning in Continuous Action Spaces
von: Gao, Ji, et al.
Veröffentlicht: (2026)
von: Gao, Ji, et al.
Veröffentlicht: (2026)
Preferred-Action-Optimized Diffusion Policies for Offline Reinforcement Learning
von: Zhang, Tianle, et al.
Veröffentlicht: (2024)
von: Zhang, Tianle, et al.
Veröffentlicht: (2024)
Towards Optimal Adversarial Robust Reinforcement Learning with Infinity Measurement Error
von: Li, Haoran, et al.
Veröffentlicht: (2025)
von: Li, Haoran, et al.
Veröffentlicht: (2025)
Fairness Aware Reinforcement Learning via Proximal Policy Optimization
von: La Malfa, Gabriele, et al.
Veröffentlicht: (2025)
von: La Malfa, Gabriele, et al.
Veröffentlicht: (2025)
Robust Deep Reinforcement Learning with Adaptive Adversarial Perturbations in Action Space
von: Liu, Qianmei, et al.
Veröffentlicht: (2024)
von: Liu, Qianmei, et al.
Veröffentlicht: (2024)
Extreme Value Policy Optimization for Safe Reinforcement Learning
von: Gao, Shiqing, et al.
Veröffentlicht: (2026)
von: Gao, Shiqing, et al.
Veröffentlicht: (2026)
Adversarial Policy Optimization for Offline Preference-based Reinforcement Learning
von: Kang, Hyungkyu, et al.
Veröffentlicht: (2025)
von: Kang, Hyungkyu, et al.
Veröffentlicht: (2025)
Reinforcing Language Agents via Policy Optimization with Action Decomposition
von: Wen, Muning, et al.
Veröffentlicht: (2024)
von: Wen, Muning, et al.
Veröffentlicht: (2024)
Dual Action Policy for Robust Sim-to-Real Reinforcement Learning
von: Terence, Ng Wen Zheng, et al.
Veröffentlicht: (2024)
von: Terence, Ng Wen Zheng, et al.
Veröffentlicht: (2024)
ORVIT: Near-Optimal Online Distributionally Robust Reinforcement Learning
von: Ghosh, Debamita, et al.
Veröffentlicht: (2025)
von: Ghosh, Debamita, et al.
Veröffentlicht: (2025)
Incentivizing Safer Actions in Policy Optimization for Constrained Reinforcement Learning
von: Hazra, Somnath, et al.
Veröffentlicht: (2025)
von: Hazra, Somnath, et al.
Veröffentlicht: (2025)
Rethinking Adversarial Attacks in Reinforcement Learning from Policy Distribution Perspective
von: Duan, Tianyang, et al.
Veröffentlicht: (2025)
von: Duan, Tianyang, et al.
Veröffentlicht: (2025)
Improving Policy Exploitation in Online Reinforcement Learning with Instant Retrospect Action
von: Gao, Gong, et al.
Veröffentlicht: (2026)
von: Gao, Gong, et al.
Veröffentlicht: (2026)
Robust Adversarial Policy Optimization Under Dynamics Uncertainty
von: Kim, Mintae, et al.
Veröffentlicht: (2026)
von: Kim, Mintae, et al.
Veröffentlicht: (2026)
ISEP: Implicit Support Expansion for Offline Reinforcement Learning via Stochastic Policy Optimization
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
Policy Optimization via Adv2: Adversarial Learning on Advantage Functions
von: Jonckheere, Matthieu, et al.
Veröffentlicht: (2023)
von: Jonckheere, Matthieu, et al.
Veröffentlicht: (2023)
Towards Robust Policy: Enhancing Offline Reinforcement Learning with Adversarial Attacks and Defenses
von: Nguyen, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thanh, et al.
Veröffentlicht: (2024)
Adversarial Diffusion for Robust Reinforcement Learning
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
Action-Graph Policies: Learning Action Co-dependencies in Multi-Agent Reinforcement Learning
von: Gupta, Nikunj, et al.
Veröffentlicht: (2026)
von: Gupta, Nikunj, et al.
Veröffentlicht: (2026)
Doubly Optimal Policy Evaluation for Reinforcement Learning
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024)
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024)
Viability of Future Actions: Robust Safety in Reinforcement Learning via Entropy Regularization
von: Massiani, Pierre-François, et al.
Veröffentlicht: (2025)
von: Massiani, Pierre-François, et al.
Veröffentlicht: (2025)
Robust Deep Reinforcement Learning in Robotics via Adaptive Gradient-Masked Adversarial Attacks
von: Zhang, Zongyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Zongyuan, et al.
Veröffentlicht: (2025)
Constraint-Aware Reinforcement Learning via Adaptive Action Scaling
von: Dawood, Murad, et al.
Veröffentlicht: (2025)
von: Dawood, Murad, et al.
Veröffentlicht: (2025)
Vulnerability-Aware Robust Multimodal Adversarial Training
von: Zhang, Junrui, et al.
Veröffentlicht: (2025)
von: Zhang, Junrui, et al.
Veröffentlicht: (2025)
Robust off-policy Reinforcement Learning via Soft Constrained Adversary
von: Nakanishi, Kosuke, et al.
Veröffentlicht: (2024)
von: Nakanishi, Kosuke, et al.
Veröffentlicht: (2024)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
Robust Adversarial Quantification via Conflict-Aware Evidential Deep Learning
von: Barker, Charmaine, et al.
Veröffentlicht: (2025)
von: Barker, Charmaine, et al.
Veröffentlicht: (2025)
Belief-Based Offline Reinforcement Learning for Delay-Robust Policy Optimization
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2025)
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2025)
Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization
von: Dus, Mathias
Veröffentlicht: (2026)
von: Dus, Mathias
Veröffentlicht: (2026)
Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions
von: Kong, Lingkai, et al.
Veröffentlicht: (2026)
von: Kong, Lingkai, et al.
Veröffentlicht: (2026)
Regularization for Adversarial Robust Learning
von: Wang, Jie, et al.
Veröffentlicht: (2024)
von: Wang, Jie, et al.
Veröffentlicht: (2024)
Rectified Robust Policy Optimization for Model-Uncertain Constrained Reinforcement Learning without Strong Duality
von: Ma, Shaocong, et al.
Veröffentlicht: (2025)
von: Ma, Shaocong, et al.
Veröffentlicht: (2025)
Causal-Aware Generative Adversarial Networks with Reinforcement Learning
von: Nguyen, Tu Anh Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Tu Anh Hoang, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning using Action Projection: Safeguard the Policy or the Environment?
von: Markgraf, Hannah, et al.
Veröffentlicht: (2025)
von: Markgraf, Hannah, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Select before Act: Spatially Decoupled Action Repetition for Continuous Control
von: Nie, Buqing, et al.
Veröffentlicht: (2025) -
Adversarial Constrained Policy Optimization: Improving Constrained Reinforcement Learning by Adapting Budgets
von: Ma, Jianmina, et al.
Veröffentlicht: (2024) -
Disturbance-Aware Adaptive Compensation in Hybrid Force-Position Locomotion Policy for Legged Robots
von: Zhang, Yang, et al.
Veröffentlicht: (2025) -
Can Tabular Foundation Models Guide Exploration in Robot Policy Learning?
von: Ou, Buqing, et al.
Veröffentlicht: (2026) -
On the Tension Between Optimality and Adversarial Robustness in Policy Optimization
von: Li, Haoran, et al.
Veröffentlicht: (2025)