Distributionally Robust Policy Evaluation under General Covariate Shift in Contextual Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Yihong, Liu, Hao, Yue, Yisong, Liu, Anqi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Group-Sensitive Offline Contextual Bandits
by: Guo, Yihong, et al.
Published: (2025)
by: Guo, Yihong, et al.
Published: (2025)
Wasserstein Distributionally Robust Policy Evaluation and Learning for Contextual Bandits
by: Shen, Yi, et al.
Published: (2023)
by: Shen, Yi, et al.
Published: (2023)
Learning Calibrated Uncertainties for Domain Shift: A Distributionally Robust Learning Approach
by: Wang, Haoxuan, et al.
Published: (2020)
by: Wang, Haoxuan, et al.
Published: (2020)
Optimal Policy Adaptation under Covariate Shift
by: Liu, Xueqing, et al.
Published: (2025)
by: Liu, Xueqing, et al.
Published: (2025)
Density-Softmax: Efficient Test-time Model for Uncertainty Estimation and Robustness under Distribution Shifts
by: Bui, Ha Manh, et al.
Published: (2023)
by: Bui, Ha Manh, et al.
Published: (2023)
Transfer Learning in Latent Contextual Bandits with Covariate Shift Through Causal Transportability
by: Deng, Mingwei, et al.
Published: (2025)
by: Deng, Mingwei, et al.
Published: (2025)
ODD: Overlap-aware Estimation of Model Performance under Distribution Shift
by: Mishra, Aayush, et al.
Published: (2025)
by: Mishra, Aayush, et al.
Published: (2025)
Distributionally Robust Coreset Selection under Covariate Shift
by: Tanaka, Tomonari, et al.
Published: (2025)
by: Tanaka, Tomonari, et al.
Published: (2025)
Distributionally Robust Safe Sample Elimination under Covariate Shift
by: Hanada, Hiroyuki, et al.
Published: (2024)
by: Hanada, Hiroyuki, et al.
Published: (2024)
Safe Distributionally Robust Feature Selection under Covariate Shift
by: Hanada, Hiroyuki, et al.
Published: (2026)
by: Hanada, Hiroyuki, et al.
Published: (2026)
Variance-Aware Linear UCB with Deep Representation for Neural Contextual Bandits
by: Bui, Ha Manh, et al.
Published: (2024)
by: Bui, Ha Manh, et al.
Published: (2024)
Context-Action Embedding Learning for Off-Policy Evaluation in Contextual Bandits
by: Chandak, Kushagra, et al.
Published: (2025)
by: Chandak, Kushagra, et al.
Published: (2025)
Beyond Numeric Rewards: In-Context Dueling Bandits with LLM Agents
by: Xia, Fanzeng, et al.
Published: (2024)
by: Xia, Fanzeng, et al.
Published: (2024)
Density-Regression: Efficient and Distance-Aware Deep Regressor for Uncertainty Estimation under Distribution Shifts
by: Bui, Ha Manh, et al.
Published: (2024)
by: Bui, Ha Manh, et al.
Published: (2024)
Contextual Optimization under Covariate Shift: A Robust Approach by Intersecting Wasserstein Balls
by: Wang, Tianyu, et al.
Published: (2024)
by: Wang, Tianyu, et al.
Published: (2024)
Spectral Algorithms in Misspecified Regression: Convergence under Covariate Shift
by: Liu, Ren-Rui, et al.
Published: (2025)
by: Liu, Ren-Rui, et al.
Published: (2025)
Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions
by: Kang, Yue, et al.
Published: (2025)
by: Kang, Yue, et al.
Published: (2025)
Distributionally Robust Constrained Reinforcement Learning under Strong Duality
by: Zhang, Zhengfei, et al.
Published: (2024)
by: Zhang, Zhengfei, et al.
Published: (2024)
Contextual Bandits for Unbounded Context Distributions
by: Zhao, Puning, et al.
Published: (2024)
by: Zhao, Puning, et al.
Published: (2024)
Effective Off-Policy Evaluation and Learning in Contextual Combinatorial Bandits
by: Shimizu, Tatsuhiro, et al.
Published: (2024)
by: Shimizu, Tatsuhiro, et al.
Published: (2024)
PAC Off-Policy Prediction of Contextual Bandits
by: Wan, Yilong, et al.
Published: (2025)
by: Wan, Yilong, et al.
Published: (2025)
Optimal Regret for Policy Optimization in Contextual Bandits
by: Levy, Orin, et al.
Published: (2026)
by: Levy, Orin, et al.
Published: (2026)
Learning with Incomplete Context: Linear Contextual Bandits with Pretrained Imputation
by: Yan, Hao, et al.
Published: (2025)
by: Yan, Hao, et al.
Published: (2025)
Cross-Domain Energy-Guided Diffusion Generation for Off-Dynamics Reinforcement Learning
by: Yang, Yu, et al.
Published: (2026)
by: Yang, Yu, et al.
Published: (2026)
Efficient Generalized Low-Rank Tensor Contextual Bandits
by: Yi, Qianxin, et al.
Published: (2023)
by: Yi, Qianxin, et al.
Published: (2023)
Weight Clipping for Robust Conformal Inference under Unbounded Covariate Shifts
by: Wang, James, et al.
Published: (2026)
by: Wang, James, et al.
Published: (2026)
Online Continuous Hyperparameter Optimization for Generalized Linear Contextual Bandits
by: Kang, Yue, et al.
Published: (2023)
by: Kang, Yue, et al.
Published: (2023)
Generalized Low-Rank Matrix Contextual Bandits with Graph Information
by: Wang, Yao, et al.
Published: (2025)
by: Wang, Yao, et al.
Published: (2025)
Pessimistic Risk-Aware Policy Learning in Contextual Bandits
by: Wan, Yilong, et al.
Published: (2026)
by: Wan, Yilong, et al.
Published: (2026)
Uncertainty of Joint Neural Contextual Bandit
by: Guo, Hongbo, et al.
Published: (2024)
by: Guo, Hongbo, et al.
Published: (2024)
Optimal Algorithms in Linear Regression under Covariate Shift: On the Importance of Precondition
by: Liu, Yuanshi, et al.
Published: (2025)
by: Liu, Yuanshi, et al.
Published: (2025)
When is Off-Policy Evaluation (Reward Modeling) Useful in Contextual Bandits? A Data-Centric Perspective
by: Sun, Hao, et al.
Published: (2023)
by: Sun, Hao, et al.
Published: (2023)
Catoni Contextual Bandits are Robust to Heavy-tailed Rewards
by: Ye, Chenlu, et al.
Published: (2025)
by: Ye, Chenlu, et al.
Published: (2025)
Off-Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits
by: Zhan, Ruohan, et al.
Published: (2021)
by: Zhan, Ruohan, et al.
Published: (2021)
Episodic Contextual Bandits with Knapsacks under Conversion Models
by: Cheung, Wang Chi, et al.
Published: (2025)
by: Cheung, Wang Chi, et al.
Published: (2025)
Robust and Computationally Efficient Linear Contextual Bandits under Adversarial Corruption and Heavy-Tailed Noise
by: Tani, Naoto, et al.
Published: (2026)
by: Tani, Naoto, et al.
Published: (2026)
Bayesian Inference of Contextual Bandit Policies via Empirical Likelihood
by: Ouyang, Jiangrong, et al.
Published: (2026)
by: Ouyang, Jiangrong, et al.
Published: (2026)
Certifiably Robust Model Evaluation in Federated Learning under Meta-Distributional Shifts
by: Najafi, Amir, et al.
Published: (2024)
by: Najafi, Amir, et al.
Published: (2024)
Wasserstein-regularized Conformal Prediction under General Distribution Shift
by: Xu, Rui, et al.
Published: (2025)
by: Xu, Rui, et al.
Published: (2025)
Optimal Baseline Corrections for Off-Policy Contextual Bandits
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Similar Items
-
Group-Sensitive Offline Contextual Bandits
by: Guo, Yihong, et al.
Published: (2025) -
Wasserstein Distributionally Robust Policy Evaluation and Learning for Contextual Bandits
by: Shen, Yi, et al.
Published: (2023) -
Learning Calibrated Uncertainties for Domain Shift: A Distributionally Robust Learning Approach
by: Wang, Haoxuan, et al.
Published: (2020) -
Optimal Policy Adaptation under Covariate Shift
by: Liu, Xueqing, et al.
Published: (2025) -
Density-Softmax: Efficient Test-time Model for Uncertainty Estimation and Robustness under Distribution Shifts
by: Bui, Ha Manh, et al.
Published: (2023)