Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yikai, Liu, Shang, Blanchet, Jose |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Wasserstein Distributionally Robust Regret Optimization
by: Fiechtner, Lukas-Benedikt, et al.
Published: (2025)
by: Fiechtner, Lukas-Benedikt, et al.
Published: (2025)
On the Foundation of Distributionally Robust Reinforcement Learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Reinforcement Learning from Human Feedback with Active Queries
by: Ji, Kaixuan, et al.
Published: (2024)
by: Ji, Kaixuan, et al.
Published: (2024)
Unifying Distributionally Robust Optimization via Optimal Transport Theory
by: Blanchet, Jose, et al.
Published: (2023)
by: Blanchet, Jose, et al.
Published: (2023)
Wasserstein Distributionally Robust Optimization: Theory and Applications in Machine Learning
by: Kuhn, Daniel, et al.
Published: (2019)
by: Kuhn, Daniel, et al.
Published: (2019)
Distributed Online Bandit Nonconvex Optimization with One-Point Residual Feedback via Dynamic Regret
by: Hua, Youqing, et al.
Published: (2024)
by: Hua, Youqing, et al.
Published: (2024)
Tail Distribution of Regret in Optimistic Reinforcement Learning
by: Khodadadian, Sajad, et al.
Published: (2025)
by: Khodadadian, Sajad, et al.
Published: (2025)
Wasserstein Distributionally Robust Online Learning
by: Chen, Guixian, et al.
Published: (2026)
by: Chen, Guixian, et al.
Published: (2026)
On Generalization and Regularization via Wasserstein Distributionally Robust Optimization
by: Wu, Qinyu, et al.
Published: (2022)
by: Wu, Qinyu, et al.
Published: (2022)
Sample Complexity of Variance-reduced Distributionally Robust Q-learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Enhancing Distributional Robustness in Principal Component Analysis by Wasserstein Distances
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
Causal LLM Routing: End-to-End Regret Minimization from Observational Data
by: Tsiourvas, Asterios, et al.
Published: (2025)
by: Tsiourvas, Asterios, et al.
Published: (2025)
Robust Out-of-Distribution Stochastic Optimization
by: Li, Xianyu, et al.
Published: (2026)
by: Li, Xianyu, et al.
Published: (2026)
Reinforcement Learning and Regret Bounds for Admission Control
by: Weber, Lucas, et al.
Published: (2024)
by: Weber, Lucas, et al.
Published: (2024)
Stability Evaluation via Distributional Perturbation Analysis
by: Blanchet, Jose, et al.
Published: (2024)
by: Blanchet, Jose, et al.
Published: (2024)
Distributionally Robust Regret Optimal LQR with Common Stage-Law Ambiguity
by: Fiechtner, Lukas-Benedikt, et al.
Published: (2026)
by: Fiechtner, Lukas-Benedikt, et al.
Published: (2026)
Wasserstein Distributionally Robust Shallow Convex Neural Networks
by: Pallage, Julien, et al.
Published: (2024)
by: Pallage, Julien, et al.
Published: (2024)
Distributional Surgery for Language Model Activations
by: Nguyen, Bao, et al.
Published: (2025)
by: Nguyen, Bao, et al.
Published: (2025)
Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training
by: Liu, Hong, et al.
Published: (2023)
by: Liu, Hong, et al.
Published: (2023)
Robust Assortment Optimization from Observational Data
by: Lu, Miao, et al.
Published: (2026)
by: Lu, Miao, et al.
Published: (2026)
Distributionally-Robust Learning to Optimize
by: Ranjan, Vinit, et al.
Published: (2026)
by: Ranjan, Vinit, et al.
Published: (2026)
Adapprox: Adaptive Approximation in Adam Optimization via Randomized Low-Rank Matrices
by: Zhao, Pengxiang, et al.
Published: (2024)
by: Zhao, Pengxiang, et al.
Published: (2024)
Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization
by: Dus, Mathias
Published: (2026)
by: Dus, Mathias
Published: (2026)
Contextual Optimization under Covariate Shift: A Robust Approach by Intersecting Wasserstein Balls
by: Wang, Tianyu, et al.
Published: (2024)
by: Wang, Tianyu, et al.
Published: (2024)
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
by: Chen, Zijun, et al.
Published: (2025)
by: Chen, Zijun, et al.
Published: (2025)
Distributionally and Adversarially Robust Logistic Regression via Intersecting Wasserstein Balls
by: Selvi, Aras, et al.
Published: (2024)
by: Selvi, Aras, et al.
Published: (2024)
Tight Robustness Certificates and Wasserstein Distributional Attacks for Deep Neural Networks
by: Le, Bach C., et al.
Published: (2025)
by: Le, Bach C., et al.
Published: (2025)
Distributionally Robust Regret Optimal Control Under Moment-Based Ambiguity Sets
by: Taha, Feras Al, et al.
Published: (2025)
by: Taha, Feras Al, et al.
Published: (2025)
Duality and Policy Evaluation in Distributionally Robust Bayesian Diffusion Control
by: Blanchet, Jose, et al.
Published: (2025)
by: Blanchet, Jose, et al.
Published: (2025)
Sign-Based Optimizers Are Effective Under Heavy-Tailed Noise
by: Yu, Dingzhi, et al.
Published: (2026)
by: Yu, Dingzhi, et al.
Published: (2026)
Wasserstein Distributionally Robust Nash Equilibrium Seeking with Heterogeneous Data: A Lagrangian Approach
by: Wang, Zifan, et al.
Published: (2025)
by: Wang, Zifan, et al.
Published: (2025)
Doubly Optimal No-Regret Online Learning in Strongly Monotone Games with Bandit Feedback
by: Ba, Wenjia, et al.
Published: (2021)
by: Ba, Wenjia, et al.
Published: (2021)
Sinkhorn Distributionally Robust Optimization
by: Wang, Jie, et al.
Published: (2021)
by: Wang, Jie, et al.
Published: (2021)
Convex Dominance in Deep Learning I: A Scaling Law of Loss and Learning Rate
by: Bu, Zhiqi, et al.
Published: (2026)
by: Bu, Zhiqi, et al.
Published: (2026)
StablePCA: Distributionally Robust Learning of Shared Representations from Multi-Source Data
by: Wang, Zhenyu, et al.
Published: (2025)
by: Wang, Zhenyu, et al.
Published: (2025)
Accelerated Distributed Optimization with Compression and Error Feedback
by: Gao, Yuan, et al.
Published: (2025)
by: Gao, Yuan, et al.
Published: (2025)
A Short and General Duality Proof for Wasserstein Distributionally Robust Optimization
by: Zhang, Luhao, et al.
Published: (2022)
by: Zhang, Luhao, et al.
Published: (2022)
FOCUS: First Order Concentrated Updating Scheme
by: Liu, Yizhou, et al.
Published: (2025)
by: Liu, Yizhou, et al.
Published: (2025)
Distributionally Robust Optimization
by: Kuhn, Daniel, et al.
Published: (2024)
by: Kuhn, Daniel, et al.
Published: (2024)
Adaptive, Doubly Optimal No-Regret Learning in Strongly Monotone and Exp-Concave Games with Gradient Feedback
by: Jordan, Michael I., et al.
Published: (2023)
by: Jordan, Michael I., et al.
Published: (2023)
Similar Items
-
Wasserstein Distributionally Robust Regret Optimization
by: Fiechtner, Lukas-Benedikt, et al.
Published: (2025) -
On the Foundation of Distributionally Robust Reinforcement Learning
by: Wang, Shengbo, et al.
Published: (2023) -
Reinforcement Learning from Human Feedback with Active Queries
by: Ji, Kaixuan, et al.
Published: (2024) -
Unifying Distributionally Robust Optimization via Optimal Transport Theory
by: Blanchet, Jose, et al.
Published: (2023) -
Wasserstein Distributionally Robust Optimization: Theory and Applications in Machine Learning
by: Kuhn, Daniel, et al.
Published: (2019)