Off-policy Evaluation in Doubly Inhomogeneous Environments
Fuente:
arXiv
Guardado en:
| Autores principales: | Bian, Zeyu, Shi, Chengchun, Qi, Zhengling, Wang, Lan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Double Fairness Policy Learning: Integrating Action Fairness and Outcome Fairness in Decision-making
por: Bian, Zeyu, et al.
Publicado: (2026)
por: Bian, Zeyu, et al.
Publicado: (2026)
Two-way Deconfounder for Off-policy Evaluation in Causal Reinforcement Learning
por: Yu, Shuguang, et al.
Publicado: (2024)
por: Yu, Shuguang, et al.
Publicado: (2024)
A Tale of Two Cities: Pessimism and Opportunism in Offline Dynamic Pricing
por: Bian, Zeyu, et al.
Publicado: (2024)
por: Bian, Zeyu, et al.
Publicado: (2024)
STEEL: Singularity-aware Reinforcement Learning
por: Chen, Xiaohong, et al.
Publicado: (2023)
por: Chen, Xiaohong, et al.
Publicado: (2023)
Doubly Inhomogeneous Reinforcement Learning
por: Hu, Liyuan, et al.
Publicado: (2022)
por: Hu, Liyuan, et al.
Publicado: (2022)
Distributional Off-Policy Evaluation with Deep Quantile Process Regression
por: Kuang, Qi, et al.
Publicado: (2026)
por: Kuang, Qi, et al.
Publicado: (2026)
Combining Experimental and Historical Data for Policy Evaluation
por: Li, Ting, et al.
Publicado: (2024)
por: Li, Ting, et al.
Publicado: (2024)
Distributional Off-policy Evaluation with Bellman Residual Minimization
por: Hong, Sungee, et al.
Publicado: (2024)
por: Hong, Sungee, et al.
Publicado: (2024)
Reinforcement Learning with Continuous Actions Under Unmeasured Confounding
por: Li, Yuhan, et al.
Publicado: (2025)
por: Li, Yuhan, et al.
Publicado: (2025)
POLAR: A Pessimistic Model-based Policy Learning Algorithm for Dynamic Treatment Regimes
por: Zhang, Ruijia, et al.
Publicado: (2025)
por: Zhang, Ruijia, et al.
Publicado: (2025)
DR-VIDAL -- Doubly Robust Variational Information-theoretic Deep Adversarial Learning for Counterfactual Prediction and Treatment Effect Estimation on Real World Data
por: Ghosh, Shantanu, et al.
Publicado: (2023)
por: Ghosh, Shantanu, et al.
Publicado: (2023)
A Graphical Approach to State Variable Selection in Off-policy Learning
por: Andersen, Joakim Blach, et al.
Publicado: (2025)
por: Andersen, Joakim Blach, et al.
Publicado: (2025)
Designing Time Series Experiments in A/B Testing with Transformer Reinforcement Learning
por: Wu, Xiangkun, et al.
Publicado: (2026)
por: Wu, Xiangkun, et al.
Publicado: (2026)
Doubly Robust Proximal Causal Learning for Continuous Treatments
por: Wu, Yong, et al.
Publicado: (2023)
por: Wu, Yong, et al.
Publicado: (2023)
Doubly-Regressing Approach for Subgroup Fairness
por: Kim, Kunwoong, et al.
Publicado: (2025)
por: Kim, Kunwoong, et al.
Publicado: (2025)
Learning Robust Treatment Rules for Censored Data
por: Cui, Yifan, et al.
Publicado: (2024)
por: Cui, Yifan, et al.
Publicado: (2024)
Doubly robust identification of treatment effects from multiple environments
por: De Bartolomeis, Piersilvio, et al.
Publicado: (2025)
por: De Bartolomeis, Piersilvio, et al.
Publicado: (2025)
Doubly Robust Fusion of Many Treatments for Policy Learning
por: Zhu, Ke, et al.
Publicado: (2025)
por: Zhu, Ke, et al.
Publicado: (2025)
Detecting and Localizing Anomalous Cliques in Inhomogeneous Networks using Egonets
por: Bhadra, Subhankar, et al.
Publicado: (2018)
por: Bhadra, Subhankar, et al.
Publicado: (2018)
Doubly Robust Interval Estimation for Optimal Policy Evaluation in Online Learning
por: Shen, Ye, et al.
Publicado: (2021)
por: Shen, Ye, et al.
Publicado: (2021)
Doubly Robust Conformalized Survival Analysis with Right-Censored Data
por: Sesia, Matteo, et al.
Publicado: (2024)
por: Sesia, Matteo, et al.
Publicado: (2024)
Doubly Robust Conditional Independence Testing with Generative Neural Networks
por: Zhang, Yi, et al.
Publicado: (2024)
por: Zhang, Yi, et al.
Publicado: (2024)
Beyond Demand Estimation: Consumer Surplus Evaluation via Cumulative Propensity Weights
por: Bian, Zeyu, et al.
Publicado: (2026)
por: Bian, Zeyu, et al.
Publicado: (2026)
Proximal Projection for Doubly Sparse Regularized Models
por: He, Jia Wei, et al.
Publicado: (2026)
por: He, Jia Wei, et al.
Publicado: (2026)
Counterfactually Fair Reinforcement Learning via Sequential Data Preprocessing
por: Wang, Jitao, et al.
Publicado: (2025)
por: Wang, Jitao, et al.
Publicado: (2025)
2D Stability Selection: Design Jittering for Doubly Stable Feature Selection
por: Nouraie, Mahdi, et al.
Publicado: (2026)
por: Nouraie, Mahdi, et al.
Publicado: (2026)
Beyond Differences: Doubly Robust Meta-Learners for Ratio-Based Treatment Effects
por: Fuchs, Michael, et al.
Publicado: (2026)
por: Fuchs, Michael, et al.
Publicado: (2026)
Doubly Robust Inference in Causal Latent Factor Models
por: Abadie, Alberto, et al.
Publicado: (2024)
por: Abadie, Alberto, et al.
Publicado: (2024)
Benchmarking Estimators for Natural Experiments: A Novel Dataset and a Doubly Robust Algorithm
por: Witter, R. Teal, et al.
Publicado: (2024)
por: Witter, R. Teal, et al.
Publicado: (2024)
Penalized Empirical Likelihood for Doubly Robust Causal Inference under Contamination in High Dimensions
por: Lee, Byeonghee, et al.
Publicado: (2025)
por: Lee, Byeonghee, et al.
Publicado: (2025)
Doubly robust inference via calibration
por: van der Laan, Lars, et al.
Publicado: (2024)
por: van der Laan, Lars, et al.
Publicado: (2024)
Learning covariate importance for matching in policy-relevant observational research
por: Zhang, Hongzhe, et al.
Publicado: (2024)
por: Zhang, Hongzhe, et al.
Publicado: (2024)
A Doubly Robust Machine Learning Approach for Disentangling Treatment Effect Heterogeneity with Functional Outcomes
por: Salmaso, Filippo, et al.
Publicado: (2026)
por: Salmaso, Filippo, et al.
Publicado: (2026)
Confidence Diagram of Nonparametric Ranking for Uncertainty Assessment in Large Language Models Evaluation
por: Wang, Zebin, et al.
Publicado: (2024)
por: Wang, Zebin, et al.
Publicado: (2024)
Off-Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits
por: Zhan, Ruohan, et al.
Publicado: (2021)
por: Zhan, Ruohan, et al.
Publicado: (2021)
Cramming Contextual Bandits for On-policy Statistical Evaluation
por: Jia, Zeyang, et al.
Publicado: (2024)
por: Jia, Zeyang, et al.
Publicado: (2024)
Automatic Doubly Robust Forests
por: Chen, Zhaomeng, et al.
Publicado: (2024)
por: Chen, Zhaomeng, et al.
Publicado: (2024)
Position: Stop Chasing the C-index when Evaluating Survival Analysis Models
por: Lillelund, Christian Marius, et al.
Publicado: (2025)
por: Lillelund, Christian Marius, et al.
Publicado: (2025)
Joint modeling for learning decision-making dynamics in behavioral experiments
por: Bian, Yuan, et al.
Publicado: (2025)
por: Bian, Yuan, et al.
Publicado: (2025)
Off-Policy Evaluation and Learning for Survival Outcomes under Censoring
por: Kubota, Kohsuke, et al.
Publicado: (2026)
por: Kubota, Kohsuke, et al.
Publicado: (2026)
Ejemplares similares
-
Double Fairness Policy Learning: Integrating Action Fairness and Outcome Fairness in Decision-making
por: Bian, Zeyu, et al.
Publicado: (2026) -
Two-way Deconfounder for Off-policy Evaluation in Causal Reinforcement Learning
por: Yu, Shuguang, et al.
Publicado: (2024) -
A Tale of Two Cities: Pessimism and Opportunism in Offline Dynamic Pricing
por: Bian, Zeyu, et al.
Publicado: (2024) -
STEEL: Singularity-aware Reinforcement Learning
por: Chen, Xiaohong, et al.
Publicado: (2023) -
Doubly Inhomogeneous Reinforcement Learning
por: Hu, Liyuan, et al.
Publicado: (2022)