Wasserstein Distributionally Robust Policy Evaluation and Learning for Contextual Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Yi, Xu, Pan, Zavlanos, Michael M. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Group Distributionally Robust Machine Learning under Group Level Distributional Uncertainty
by: Konti, Xenia, et al.
Published: (2025)
by: Konti, Xenia, et al.
Published: (2025)
Distributionally Robust Clustered Federated Learning: A Case Study in Healthcare
by: Konti, Xenia, et al.
Published: (2024)
by: Konti, Xenia, et al.
Published: (2024)
Distributionally Robust Policy Evaluation under General Covariate Shift in Contextual Bandits
by: Guo, Yihong, et al.
Published: (2024)
by: Guo, Yihong, et al.
Published: (2024)
Wasserstein Distributionally Robust Nash Equilibrium Seeking with Heterogeneous Data: A Lagrangian Approach
by: Wang, Zifan, et al.
Published: (2025)
by: Wang, Zifan, et al.
Published: (2025)
Distributionally Robust Federated Learning with Outlier Resilience
by: Wang, Zifan, et al.
Published: (2025)
by: Wang, Zifan, et al.
Published: (2025)
Context-Action Embedding Learning for Off-Policy Evaluation in Contextual Bandits
by: Chandak, Kushagra, et al.
Published: (2025)
by: Chandak, Kushagra, et al.
Published: (2025)
Contextual Bandits for Unbounded Context Distributions
by: Zhao, Puning, et al.
Published: (2024)
by: Zhao, Puning, et al.
Published: (2024)
Effective Off-Policy Evaluation and Learning in Contextual Combinatorial Bandits
by: Shimizu, Tatsuhiro, et al.
Published: (2024)
by: Shimizu, Tatsuhiro, et al.
Published: (2024)
Distributionally Robust Multi-Agent Reinforcement Learning for Dynamic Chute Mapping
by: Liu, Guangyi, et al.
Published: (2025)
by: Liu, Guangyi, et al.
Published: (2025)
IBCB: Efficient Inverse Batched Contextual Bandit for Behavioral Evolution History
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
Pessimistic Risk-Aware Policy Learning in Contextual Bandits
by: Wan, Yilong, et al.
Published: (2026)
by: Wan, Yilong, et al.
Published: (2026)
Optimizing Warfarin Dosing Using Contextual Bandit: An Offline Policy Learning and Evaluation Method
by: Huang, Yong, et al.
Published: (2024)
by: Huang, Yong, et al.
Published: (2024)
PAC Off-Policy Prediction of Contextual Bandits
by: Wan, Yilong, et al.
Published: (2025)
by: Wan, Yilong, et al.
Published: (2025)
Optimal Regret for Policy Optimization in Contextual Bandits
by: Levy, Orin, et al.
Published: (2026)
by: Levy, Orin, et al.
Published: (2026)
Cramming Contextual Bandits for On-policy Statistical Evaluation
by: Jia, Zeyang, et al.
Published: (2024)
by: Jia, Zeyang, et al.
Published: (2024)
Risk-averse Learning with Non-Stationary Distributions
by: Wang, Siyi, et al.
Published: (2024)
by: Wang, Siyi, et al.
Published: (2024)
Catoni Contextual Bandits are Robust to Heavy-tailed Rewards
by: Ye, Chenlu, et al.
Published: (2025)
by: Ye, Chenlu, et al.
Published: (2025)
Off-Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits
by: Zhan, Ruohan, et al.
Published: (2021)
by: Zhan, Ruohan, et al.
Published: (2021)
Wasserstein Distributionally Robust Online Learning
by: Chen, Guixian, et al.
Published: (2026)
by: Chen, Guixian, et al.
Published: (2026)
Provable Anytime Ensemble Sampling Algorithms in Nonlinear Contextual Bandits
by: Sun, Jiazheng, et al.
Published: (2025)
by: Sun, Jiazheng, et al.
Published: (2025)
Bayesian Inference of Contextual Bandit Policies via Empirical Likelihood
by: Ouyang, Jiangrong, et al.
Published: (2026)
by: Ouyang, Jiangrong, et al.
Published: (2026)
Linear Contextual Bandits with Interference
by: Xu, Yang, et al.
Published: (2024)
by: Xu, Yang, et al.
Published: (2024)
High Probability Bound for Cross-Learning Contextual Bandits with Unknown Context Distributions
by: Huang, Ruiyuan, et al.
Published: (2024)
by: Huang, Ruiyuan, et al.
Published: (2024)
Exploration via Feature Perturbation in Contextual Bandits
by: Yi, Seouh-won, et al.
Published: (2025)
by: Yi, Seouh-won, et al.
Published: (2025)
Optimal Baseline Corrections for Off-Policy Contextual Bandits
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Adjusted Wasserstein Distributionally Robust Estimator in Statistical Learning
by: Xie, Yiling, et al.
Published: (2023)
by: Xie, Yiling, et al.
Published: (2023)
Wasserstein Distributionally Robust Multiclass Support Vector Machine
by: Ibrahim, Michael, et al.
Published: (2024)
by: Ibrahim, Michael, et al.
Published: (2024)
Harnessing the Power of Federated Learning in Federated Contextual Bandits
by: Shi, Chengshuai, et al.
Published: (2023)
by: Shi, Chengshuai, et al.
Published: (2023)
Active Learning for Stochastic Contextual Linear Bandits
by: Brunskill, Emma, et al.
Published: (2026)
by: Brunskill, Emma, et al.
Published: (2026)
Deep Proxy Causal Learning and its Application to Confounded Bandit Policy Evaluation
by: Xu, Liyuan, et al.
Published: (2021)
by: Xu, Liyuan, et al.
Published: (2021)
Inverse Reinforcement Learning from Non-Stationary Learning Agents
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
by: Sivakumar, Kavinayan P., et al.
Published: (2024)
Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions
by: Kang, Yue, et al.
Published: (2025)
by: Kang, Yue, et al.
Published: (2025)
Learning When to Trust in Contextual Bandits
by: Ghasemi, Majid, et al.
Published: (2026)
by: Ghasemi, Majid, et al.
Published: (2026)
Efficient Generalized Low-Rank Tensor Contextual Bandits
by: Yi, Qianxin, et al.
Published: (2023)
by: Yi, Qianxin, et al.
Published: (2023)
Sparse Nonparametric Contextual Bandits
by: Flynn, Hamish, et al.
Published: (2025)
by: Flynn, Hamish, et al.
Published: (2025)
Evaluating and Learning Robust Bandit Policies Under Uncertain Causal Mechanisms
by: Avery, Katherine, et al.
Published: (2025)
by: Avery, Katherine, et al.
Published: (2025)
When is Off-Policy Evaluation (Reward Modeling) Useful in Contextual Bandits? A Data-Centric Perspective
by: Sun, Hao, et al.
Published: (2023)
by: Sun, Hao, et al.
Published: (2023)
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
by: Goyal, Tanmay, et al.
Published: (2025)
by: Goyal, Tanmay, et al.
Published: (2025)
Strategic Linear Contextual Bandits
by: Buening, Thomas Kleine, et al.
Published: (2024)
by: Buening, Thomas Kleine, et al.
Published: (2024)
Transfer Learning for Contextual Multi-armed Bandits
by: Cai, Changxiao, et al.
Published: (2022)
by: Cai, Changxiao, et al.
Published: (2022)
Similar Items
-
Group Distributionally Robust Machine Learning under Group Level Distributional Uncertainty
by: Konti, Xenia, et al.
Published: (2025) -
Distributionally Robust Clustered Federated Learning: A Case Study in Healthcare
by: Konti, Xenia, et al.
Published: (2024) -
Distributionally Robust Policy Evaluation under General Covariate Shift in Contextual Bandits
by: Guo, Yihong, et al.
Published: (2024) -
Wasserstein Distributionally Robust Nash Equilibrium Seeking with Heterogeneous Data: A Lagrangian Approach
by: Wang, Zifan, et al.
Published: (2025) -
Distributionally Robust Federated Learning with Outlier Resilience
by: Wang, Zifan, et al.
Published: (2025)