On the Tension Between Optimality and Adversarial Robustness in Policy Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Haoran, Lv, Jiayu, Han, Congying, Zhang, Zicheng, Li, Anqi, Liu, Yan, Guo, Tiande, Jiang, Nan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Optimal Adversarial Robust Reinforcement Learning with Infinity Measurement Error
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
Towards Optimal Adversarial Robust Q-learning with Bellman Infinity-error
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
Purity Law for Generalizable Neural TSP Solvers
by: Liu, Wenzhao, et al.
Published: (2025)
by: Liu, Wenzhao, et al.
Published: (2025)
Mitigating Distribution Shift in Model-based Offline RL via Shifts-aware Reward Learning
by: Luo, Wang, et al.
Published: (2024)
by: Luo, Wang, et al.
Published: (2024)
Dual Alignment Maximin Optimization for Offline Model-based RL
by: Zhou, Chi, et al.
Published: (2025)
by: Zhou, Chi, et al.
Published: (2025)
A Fast Anti-Jamming Cognitive Radar Deployment Algorithm Based on Reinforcement Learning
by: Cai, Wencheng, et al.
Published: (2025)
by: Cai, Wencheng, et al.
Published: (2025)
On the optimal pivot path of simplex method for linear programming based on reinforcement learning
by: Li, Anqi, et al.
Published: (2022)
by: Li, Anqi, et al.
Published: (2022)
Understanding Oversmoothing in Diffusion-Based GNNs From the Perspective of Operator Semigroup Theory
by: Zhao, Weichen, et al.
Published: (2024)
by: Zhao, Weichen, et al.
Published: (2024)
DR-BFR: Degradation Representation with Diffusion Models for Blind Face Restoration
by: Qiu, Xinmin, et al.
Published: (2024)
by: Qiu, Xinmin, et al.
Published: (2024)
Action Robust Reinforcement Learning via Optimal Adversary Aware Policy Optimization
by: Nie, Buqing, et al.
Published: (2025)
by: Nie, Buqing, et al.
Published: (2025)
Distributionally Robust Policy Evaluation under General Covariate Shift in Contextual Bandits
by: Guo, Yihong, et al.
Published: (2024)
by: Guo, Yihong, et al.
Published: (2024)
Beyond State-Wise Mirror Descent: Offline Policy Optimization with Parametric Policies
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
Relating Checkpoint Update Probabilities to Momentum Parameters in Single-Loop Variance Reduction Methods
by: Liu, Hai, et al.
Published: (2026)
by: Liu, Hai, et al.
Published: (2026)
Adversarially-Robust Inference on Trees via Belief Propagation
by: Hopkins, Samuel B., et al.
Published: (2024)
by: Hopkins, Samuel B., et al.
Published: (2024)
Resolving Endpoint Underfitting in Diffusion Bridges via Noise Alignment
by: Gao, Yurong, et al.
Published: (2026)
by: Gao, Yurong, et al.
Published: (2026)
Robust Accelerated Adaptive Search: High-Probability Complexity Bounds under Bounded-Moment Stochastic Oracles
by: Zhang, Shunzhi, et al.
Published: (2026)
by: Zhang, Shunzhi, et al.
Published: (2026)
A Near-optimal, Scalable and Parallelizable Framework for Stochastic Bandits Robust to Adversarial Corruptions and Beyond
by: Hu, Zicheng, et al.
Published: (2025)
by: Hu, Zicheng, et al.
Published: (2025)
Robust Adversarial Policy Optimization Under Dynamics Uncertainty
by: Kim, Mintae, et al.
Published: (2026)
by: Kim, Mintae, et al.
Published: (2026)
ANO: A Principled Approach to Robust Policy Optimization
by: Zhang, Yiheng, et al.
Published: (2026)
by: Zhang, Yiheng, et al.
Published: (2026)
A-PSRO: A Unified Strategy Learning Method with Advantage Function for Normal-form Games
by: Hu, Yudong, et al.
Published: (2023)
by: Hu, Yudong, et al.
Published: (2023)
Constrained Online Two-stage Stochastic Optimization: Near Optimal Algorithms via Adversarial Learning
by: Jiang, Jiashuo
Published: (2023)
by: Jiang, Jiashuo
Published: (2023)
DARD: Dice Adversarial Robustness Distillation against Adversarial Attacks
by: Zou, Jing, et al.
Published: (2025)
by: Zou, Jing, et al.
Published: (2025)
Adversarial Adaptive Sampling: Unify PINN and Optimal Transport for the Approximation of PDEs
by: Tang, Kejun, et al.
Published: (2023)
by: Tang, Kejun, et al.
Published: (2023)
Q-MMR: Off-Policy Evaluation via Recursive Reweighting and Moment Matching
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
BlazeBVD: Make Scale-Time Equalization Great Again for Blind Video Deflickering
by: Qiu, Xinmin, et al.
Published: (2024)
by: Qiu, Xinmin, et al.
Published: (2024)
StyO: Stylize Your Face in Only One-shot
by: Li, Bonan, et al.
Published: (2023)
by: Li, Bonan, et al.
Published: (2023)
ETR: Outcome-Guided Elastic Trust Regions for Policy Optimization
by: Zhang, Shijie, et al.
Published: (2026)
by: Zhang, Shijie, et al.
Published: (2026)
Optimal Transport Regularized Divergences: Application to Adversarial Robustness
by: Birrell, Jeremiah, et al.
Published: (2023)
by: Birrell, Jeremiah, et al.
Published: (2023)
Toward Evaluating Robustness of Reinforcement Learning with Adversarial Policy
by: Zheng, Xiang, et al.
Published: (2023)
by: Zheng, Xiang, et al.
Published: (2023)
On Achieving Optimal Adversarial Test Error
by: Li, Justin D., et al.
Published: (2023)
by: Li, Justin D., et al.
Published: (2023)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
by: Ganesh, Swetha, et al.
Published: (2024)
by: Ganesh, Swetha, et al.
Published: (2024)
Proactive Constrained Policy Optimization with Preemptive Penalty
by: Yang, Ning, et al.
Published: (2025)
by: Yang, Ning, et al.
Published: (2025)
Analyzing Generalization in Policy Networks: A Case Study with the Double-Integrator System
by: Zhang, Ruining, et al.
Published: (2023)
by: Zhang, Ruining, et al.
Published: (2023)
OTPTO: Joint Product Selection and Inventory Optimization in Fresh E-commerce Front-End Warehouses
by: Zhang, Zheming, et al.
Published: (2025)
by: Zhang, Zheming, et al.
Published: (2025)
Robust Graph Fine-Tuning with Adversarial Graph Prompting
by: Zhang, Ziyan, et al.
Published: (2026)
by: Zhang, Ziyan, et al.
Published: (2026)
Towards Trustworthy Multimodal Moderation via Policy-Aligned Reasoning and Hierarchical Labeling
by: Li, Anqi, et al.
Published: (2025)
by: Li, Anqi, et al.
Published: (2025)
DGPO: Distribution Guided Policy Optimization for Fine Grained Credit Assignment
by: Jin, Hongbo, et al.
Published: (2026)
by: Jin, Hongbo, et al.
Published: (2026)
Density-Guided Robust Counterfactual Explanations on Tabular Data under Model Multiplicity
by: Tan, Jun, et al.
Published: (2026)
by: Tan, Jun, et al.
Published: (2026)
CANDOR: Counterfactual ANnotated DOubly Robust Off-Policy Evaluation
by: Mandyam, Aishwarya, et al.
Published: (2024)
by: Mandyam, Aishwarya, et al.
Published: (2024)
Distributionally Robust Optimization with Adversarial Data Contamination
by: Li, Shuyao, et al.
Published: (2025)
by: Li, Shuyao, et al.
Published: (2025)
Similar Items
-
Towards Optimal Adversarial Robust Reinforcement Learning with Infinity Measurement Error
by: Li, Haoran, et al.
Published: (2025) -
Towards Optimal Adversarial Robust Q-learning with Bellman Infinity-error
by: Li, Haoran, et al.
Published: (2024) -
Purity Law for Generalizable Neural TSP Solvers
by: Liu, Wenzhao, et al.
Published: (2025) -
Mitigating Distribution Shift in Model-based Offline RL via Shifts-aware Reward Learning
by: Luo, Wang, et al.
Published: (2024) -
Dual Alignment Maximin Optimization for Offline Model-based RL
by: Zhou, Chi, et al.
Published: (2025)