Towards Optimal Adversarial Robust Reinforcement Learning with Infinity Measurement Error
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Haoran, Zhang, Zicheng, Luo, Wang, Han, Congying, Lv, Jiayu, Guo, Tiande, Hu, Yudong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Optimal Adversarial Robust Q-learning with Bellman Infinity-error
von: Li, Haoran, et al.
Veröffentlicht: (2024)
von: Li, Haoran, et al.
Veröffentlicht: (2024)
On the Tension Between Optimality and Adversarial Robustness in Policy Optimization
von: Li, Haoran, et al.
Veröffentlicht: (2025)
von: Li, Haoran, et al.
Veröffentlicht: (2025)
Mitigating Distribution Shift in Model-based Offline RL via Shifts-aware Reward Learning
von: Luo, Wang, et al.
Veröffentlicht: (2024)
von: Luo, Wang, et al.
Veröffentlicht: (2024)
Dual Alignment Maximin Optimization for Offline Model-based RL
von: Zhou, Chi, et al.
Veröffentlicht: (2025)
von: Zhou, Chi, et al.
Veröffentlicht: (2025)
Purity Law for Generalizable Neural TSP Solvers
von: Liu, Wenzhao, et al.
Veröffentlicht: (2025)
von: Liu, Wenzhao, et al.
Veröffentlicht: (2025)
A Fast Anti-Jamming Cognitive Radar Deployment Algorithm Based on Reinforcement Learning
von: Cai, Wencheng, et al.
Veröffentlicht: (2025)
von: Cai, Wencheng, et al.
Veröffentlicht: (2025)
Understanding Oversmoothing in Diffusion-Based GNNs From the Perspective of Operator Semigroup Theory
von: Zhao, Weichen, et al.
Veröffentlicht: (2024)
von: Zhao, Weichen, et al.
Veröffentlicht: (2024)
A-PSRO: A Unified Strategy Learning Method with Advantage Function for Normal-form Games
von: Hu, Yudong, et al.
Veröffentlicht: (2023)
von: Hu, Yudong, et al.
Veröffentlicht: (2023)
Applying Opponent Modeling for Automatic Bidding in Online Repeated Auctions
von: Hu, Yudong, et al.
Veröffentlicht: (2022)
von: Hu, Yudong, et al.
Veröffentlicht: (2022)
Toward Evaluating Robustness of Reinforcement Learning with Adversarial Policy
von: Zheng, Xiang, et al.
Veröffentlicht: (2023)
von: Zheng, Xiang, et al.
Veröffentlicht: (2023)
A Near-optimal, Scalable and Parallelizable Framework for Stochastic Bandits Robust to Adversarial Corruptions and Beyond
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
Towards Robust Model-Based Reinforcement Learning Against Adversarial Corruption
von: Ye, Chenlu, et al.
Veröffentlicht: (2024)
von: Ye, Chenlu, et al.
Veröffentlicht: (2024)
Preference-based opponent shaping in differentiable games
von: Qiao, Xinyu, et al.
Veröffentlicht: (2024)
von: Qiao, Xinyu, et al.
Veröffentlicht: (2024)
DR-BFR: Degradation Representation with Diffusion Models for Blind Face Restoration
von: Qiu, Xinmin, et al.
Veröffentlicht: (2024)
von: Qiu, Xinmin, et al.
Veröffentlicht: (2024)
On Achieving Optimal Adversarial Test Error
von: Li, Justin D., et al.
Veröffentlicht: (2023)
von: Li, Justin D., et al.
Veröffentlicht: (2023)
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization
von: Nie, Buqing, et al.
Veröffentlicht: (2025)
von: Nie, Buqing, et al.
Veröffentlicht: (2025)
Resolving Endpoint Underfitting in Diffusion Bridges via Noise Alignment
von: Gao, Yurong, et al.
Veröffentlicht: (2026)
von: Gao, Yurong, et al.
Veröffentlicht: (2026)
Towards Minimax Optimality of Model-based Robust Reinforcement Learning
von: Clavier, Pierre, et al.
Veröffentlicht: (2023)
von: Clavier, Pierre, et al.
Veröffentlicht: (2023)
Dynamic Adversarial Reinforcement Learning for Robust Multimodal Large Language Models
von: Bao, Yicheng, et al.
Veröffentlicht: (2026)
von: Bao, Yicheng, et al.
Veröffentlicht: (2026)
Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2026)
Robust Accelerated Adaptive Search: High-Probability Complexity Bounds under Bounded-Moment Stochastic Oracles
von: Zhang, Shunzhi, et al.
Veröffentlicht: (2026)
von: Zhang, Shunzhi, et al.
Veröffentlicht: (2026)
Robust Learning with Optimal Error
von: Blanc, Guy
Veröffentlicht: (2026)
von: Blanc, Guy
Veröffentlicht: (2026)
Hierarchical Refinement: Optimal Transport to Infinity and Beyond
von: Halmos, Peter, et al.
Veröffentlicht: (2025)
von: Halmos, Peter, et al.
Veröffentlicht: (2025)
On the optimal pivot path of simplex method for linear programming based on reinforcement learning
von: Li, Anqi, et al.
Veröffentlicht: (2022)
von: Li, Anqi, et al.
Veröffentlicht: (2022)
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
von: Xu, Haoran, et al.
Veröffentlicht: (2025)
von: Xu, Haoran, et al.
Veröffentlicht: (2025)
Relating Checkpoint Update Probabilities to Momentum Parameters in Single-Loop Variance Reduction Methods
von: Liu, Hai, et al.
Veröffentlicht: (2026)
von: Liu, Hai, et al.
Veröffentlicht: (2026)
Minimax-Optimal Multi-Agent Robust Reinforcement Learning
von: Jiao, Yuchen, et al.
Veröffentlicht: (2024)
von: Jiao, Yuchen, et al.
Veröffentlicht: (2024)
Adversarial Diffusion for Robust Reinforcement Learning
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
Towards Robust Offline Reinforcement Learning under Diverse Data Corruption
von: Yang, Rui, et al.
Veröffentlicht: (2023)
von: Yang, Rui, et al.
Veröffentlicht: (2023)
Distributionally Robust Reinforcement Learning with Interactive Data Collection: Fundamental Hardness and Near-Optimal Algorithm
von: Lu, Miao, et al.
Veröffentlicht: (2024)
von: Lu, Miao, et al.
Veröffentlicht: (2024)
Discovery of Optimal Quantum Error Correcting Codes via Reinforcement Learning
von: Su, Vincent Paul, et al.
Veröffentlicht: (2023)
von: Su, Vincent Paul, et al.
Veröffentlicht: (2023)
Robust Decentralized Multi-armed Bandits: From Corruption-Resilience to Byzantine-Resilience
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
von: Hu, Zicheng, et al.
Veröffentlicht: (2025)
Towards Robust Policy: Enhancing Offline Reinforcement Learning with Adversarial Attacks and Defenses
von: Nguyen, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thanh, et al.
Veröffentlicht: (2024)
UACER: An Uncertainty-Adaptive Critic Ensemble Framework for Robust Adversarial Reinforcement Learning
von: Wu, Jiaxi, et al.
Veröffentlicht: (2025)
von: Wu, Jiaxi, et al.
Veröffentlicht: (2025)
ORVIT: Near-Optimal Online Distributionally Robust Reinforcement Learning
von: Ghosh, Debamita, et al.
Veröffentlicht: (2025)
von: Ghosh, Debamita, et al.
Veröffentlicht: (2025)
Towards Robust Zero-Shot Reinforcement Learning
von: Zheng, Kexin, et al.
Veröffentlicht: (2025)
von: Zheng, Kexin, et al.
Veröffentlicht: (2025)
Enhancing Robustness of Offline Reinforcement Learning Under Data Corruption via Sharpness-Aware Minimization
von: Xu, Le, et al.
Veröffentlicht: (2025)
von: Xu, Le, et al.
Veröffentlicht: (2025)
Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning
von: Li, Zhenghao, et al.
Veröffentlicht: (2025)
von: Li, Zhenghao, et al.
Veröffentlicht: (2025)
Intriguing Frequency Interpretation of Adversarial Robustness for CNNs and ViTs
von: Chen, Lu, et al.
Veröffentlicht: (2025)
von: Chen, Lu, et al.
Veröffentlicht: (2025)
Interaction-Breaking Adversarial Learning Framework for Robust Multi-Agent Reinforcement Learning
von: Lee, Sunwoo, et al.
Veröffentlicht: (2026)
von: Lee, Sunwoo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Optimal Adversarial Robust Q-learning with Bellman Infinity-error
von: Li, Haoran, et al.
Veröffentlicht: (2024) -
On the Tension Between Optimality and Adversarial Robustness in Policy Optimization
von: Li, Haoran, et al.
Veröffentlicht: (2025) -
Mitigating Distribution Shift in Model-based Offline RL via Shifts-aware Reward Learning
von: Luo, Wang, et al.
Veröffentlicht: (2024) -
Dual Alignment Maximin Optimization for Offline Model-based RL
von: Zhou, Chi, et al.
Veröffentlicht: (2025) -
Purity Law for Generalizable Neural TSP Solvers
von: Liu, Wenzhao, et al.
Veröffentlicht: (2025)