POLAR: A Pessimistic Model-based Policy Learning Algorithm for Dynamic Treatment Regimes
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Ruijia, Zhang, Xiangyu, Qi, Zhengling, Wu, Yue, Xu, Yanxun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MoMA: Model-based Mirror Ascent for Offline Reinforcement Learning
by: Hong, Mao, et al.
Published: (2024)
by: Hong, Mao, et al.
Published: (2024)
Learning Robust Treatment Rules for Censored Data
by: Cui, Yifan, et al.
Published: (2024)
by: Cui, Yifan, et al.
Published: (2024)
Policy Learning for Optimal Dynamic Treatment Regimes with Observational Data
by: Sakaguchi, Shosei
Published: (2024)
by: Sakaguchi, Shosei
Published: (2024)
On Multiple Robustness of Proximal Dynamic Treatment Regimes
by: Gao, Yuanshan, et al.
Published: (2025)
by: Gao, Yuanshan, et al.
Published: (2025)
Medical Knowledge Integration into Reinforcement Learning Algorithms for Dynamic Treatment Regimes
by: Yazzourh, Sophia, et al.
Published: (2024)
by: Yazzourh, Sophia, et al.
Published: (2024)
STEEL: Singularity-aware Reinforcement Learning
by: Chen, Xiaohong, et al.
Published: (2023)
by: Chen, Xiaohong, et al.
Published: (2023)
TransformerLSR: Attentive Joint Model of Longitudinal Data, Survival, and Recurrent Events with Concurrent Latent Structure
by: Zhang, Zhiyue, et al.
Published: (2024)
by: Zhang, Zhiyue, et al.
Published: (2024)
Spectral Ranking Inferences based on General Multiway Comparisons
by: Fan, Jianqing, et al.
Published: (2023)
by: Fan, Jianqing, et al.
Published: (2023)
High-dimensional Clustering and Signal Recovery under Block Signals
by: Su, Wu, et al.
Published: (2025)
by: Su, Wu, et al.
Published: (2025)
Modular Learning of Deep Causal Generative Models for High-dimensional Causal Inference
by: Rahman, Md Musfiqur, et al.
Published: (2024)
by: Rahman, Md Musfiqur, et al.
Published: (2024)
Unifying Summary Statistic Selection for Approximate Bayesian Computation
by: Hoffmann, Till, et al.
Published: (2022)
by: Hoffmann, Till, et al.
Published: (2022)
Bayesian model-averaging stochastic item selection for adaptive testing
by: Su, Tina, et al.
Published: (2025)
by: Su, Tina, et al.
Published: (2025)
Locally Private Parametric Methods for Change-Point Detection
by: Yadav, Anuj Kumar, et al.
Published: (2026)
by: Yadav, Anuj Kumar, et al.
Published: (2026)
Amortized Vine Copulas for High-Dimensional Density and Information Estimation
by: Safaai, Houman
Published: (2026)
by: Safaai, Houman
Published: (2026)
An efficient search-and-score algorithm for ancestral graphs using multivariate information scores
by: Lagrange, Nikita, et al.
Published: (2024)
by: Lagrange, Nikita, et al.
Published: (2024)
Maximum entropy based testing in network models: ERGMs and constrained optimization
by: Ghosh, Subhro, et al.
Published: (2026)
by: Ghosh, Subhro, et al.
Published: (2026)
Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals
by: Dong, Zihan, et al.
Published: (2026)
by: Dong, Zihan, et al.
Published: (2026)
Spectrum-Aware Debiasing: A Modern Inference Framework with Applications to Principal Components Regression
by: Li, Yufan, et al.
Published: (2023)
by: Li, Yufan, et al.
Published: (2023)
Evaluating and Learning Optimal Dynamic Treatment Regimes under Truncation by Death
by: Park, Sihyung, et al.
Published: (2025)
by: Park, Sihyung, et al.
Published: (2025)
Off-policy Evaluation in Doubly Inhomogeneous Environments
by: Bian, Zeyu, et al.
Published: (2023)
by: Bian, Zeyu, et al.
Published: (2023)
Adaptive Learn-then-Test: Statistically Valid and Efficient Hyperparameter Selection
by: Zecchin, Matteo, et al.
Published: (2024)
by: Zecchin, Matteo, et al.
Published: (2024)
Inference for Heteroskedastic PCA with Missing Data
by: Yan, Yuling, et al.
Published: (2021)
by: Yan, Yuling, et al.
Published: (2021)
Time-Uniform Confidence Spheres for Means of Random Vectors
by: Chugg, Ben, et al.
Published: (2023)
by: Chugg, Ben, et al.
Published: (2023)
Stability of a Generalized Debiased Lasso with Applications to Resampling-Based Variable Selection
by: Liu, Jingbo
Published: (2024)
by: Liu, Jingbo
Published: (2024)
Consistent model selection in the spiked Wigner model via AIC-type criteria
by: Mukherjee, Soumendu Sundar
Published: (2023)
by: Mukherjee, Soumendu Sundar
Published: (2023)
Deflated HeteroPCA: Overcoming the curse of ill-conditioning in heteroskedastic PCA
by: Zhou, Yuchen, et al.
Published: (2023)
by: Zhou, Yuchen, et al.
Published: (2023)
Optimal Differentially Private PCA and Estimation for Spiked Covariance Matrices
by: Cai, T. Tony, et al.
Published: (2024)
by: Cai, T. Tony, et al.
Published: (2024)
Decentralized Conformal Novelty Detection via Quantized Model Exchange
by: Loh, Kyle, et al.
Published: (2026)
by: Loh, Kyle, et al.
Published: (2026)
Distributional Treatment Effect Estimation across Heterogeneous Sites via Optimal Transport
by: Bateni, Borna, et al.
Published: (2025)
by: Bateni, Borna, et al.
Published: (2025)
Gradient descent inference in empirical risk minimization
by: Han, Qiyang, et al.
Published: (2024)
by: Han, Qiyang, et al.
Published: (2024)
Optimal No-regret Learning in Repeated First-price Auctions
by: Han, Yanjun, et al.
Published: (2020)
by: Han, Yanjun, et al.
Published: (2020)
Computationally efficient reductions between some statistical models
by: Lou, Mengqi, et al.
Published: (2024)
by: Lou, Mengqi, et al.
Published: (2024)
Reinforcement Learning with Continuous Actions Under Unmeasured Confounding
by: Li, Yuhan, et al.
Published: (2025)
by: Li, Yuhan, et al.
Published: (2025)
Double Machine Learning of Continuous Treatment Effects with General Instrumental Variables
by: Chen, Shuyuan, et al.
Published: (2026)
by: Chen, Shuyuan, et al.
Published: (2026)
Bridging the Gap between Empirical Welfare Maximization and Conditional Average Treatment Effect Estimation in Policy Learning
by: Kato, Masahiro
Published: (2025)
by: Kato, Masahiro
Published: (2025)
Treatment Effects in Extreme Regimes
by: Aloui, Ahmed, et al.
Published: (2023)
by: Aloui, Ahmed, et al.
Published: (2023)
Robustness Against Weak or Invalid Instruments: Exploring Nonlinear Treatment Models with Machine Learning
by: Guo, Zijian, et al.
Published: (2022)
by: Guo, Zijian, et al.
Published: (2022)
Probing the Information Theoretical Roots of Spatial Dependence Measures
by: Wang, Zhangyu, et al.
Published: (2024)
by: Wang, Zhangyu, et al.
Published: (2024)
CURATE: Scaling-up Differentially Private Causal Graph Discovery
by: Bhattacharjee, Payel, et al.
Published: (2024)
by: Bhattacharjee, Payel, et al.
Published: (2024)
Meta Off-Policy Estimation
by: Jeunen, Olivier
Published: (2025)
by: Jeunen, Olivier
Published: (2025)
Similar Items
-
MoMA: Model-based Mirror Ascent for Offline Reinforcement Learning
by: Hong, Mao, et al.
Published: (2024) -
Learning Robust Treatment Rules for Censored Data
by: Cui, Yifan, et al.
Published: (2024) -
Policy Learning for Optimal Dynamic Treatment Regimes with Observational Data
by: Sakaguchi, Shosei
Published: (2024) -
On Multiple Robustness of Proximal Dynamic Treatment Regimes
by: Gao, Yuanshan, et al.
Published: (2025) -
Medical Knowledge Integration into Reinforcement Learning Algorithms for Dynamic Treatment Regimes
by: Yazzourh, Sophia, et al.
Published: (2024)