Reinforcement Learning in Dynamic Treatment Regimes Needs Critical Reexamination
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Zhiyao, Pan, Yangchen, Watkinson, Peter, Zhu, Tingting |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DTR-Bench: An in silico Environment and Benchmark Platform for Reinforcement Learning Based Dynamic Treatment Regime
by: Luo, Zhiyao, et al.
Published: (2024)
by: Luo, Zhiyao, et al.
Published: (2024)
ProtoEHR: Hierarchical Prototype Learning for EHR-based Healthcare Predictions
by: Cai, Zi, et al.
Published: (2025)
by: Cai, Zi, et al.
Published: (2025)
Bridging Data Gaps of Rare Conditions in ICU: A Multi-Disease Adaptation Approach for Clinical Prediction
by: Zhu, Mingcheng, et al.
Published: (2025)
by: Zhu, Mingcheng, et al.
Published: (2025)
Cross-Representation Benchmarking in Time-Series Electronic Health Records for Clinical Outcome Prediction
by: Chen, Tianyi, et al.
Published: (2025)
by: Chen, Tianyi, et al.
Published: (2025)
Kernel-Based Distributed Q-Learning: A Scalable Reinforcement Learning Approach for Dynamic Treatment Regimes
by: Wang, Di, et al.
Published: (2023)
by: Wang, Di, et al.
Published: (2023)
Measures of Variability for Risk-averse Policy Gradient
by: Luo, Yudong, et al.
Published: (2025)
by: Luo, Yudong, et al.
Published: (2025)
An MRP Formulation for Supervised Learning: Generalized Temporal Difference Learning Models
by: Pan, Yangchen, et al.
Published: (2024)
by: Pan, Yangchen, et al.
Published: (2024)
Censoring-Aware Tree-Based Reinforcement Learning for Estimating Dynamic Treatment Regimes with Censored Outcomes
by: Paul, Animesh Kumar, et al.
Published: (2025)
by: Paul, Animesh Kumar, et al.
Published: (2025)
Gradient Residual Connections
by: Pan, Yangchen, et al.
Published: (2026)
by: Pan, Yangchen, et al.
Published: (2026)
The Three Regimes of Offline-to-Online Reinforcement Learning
by: Li, Lu, et al.
Published: (2025)
by: Li, Lu, et al.
Published: (2025)
Are Large Language Models Dynamic Treatment Planners? An In Silico Study from a Prior Knowledge Injection Angle
by: Luo, Zhiyao, et al.
Published: (2025)
by: Luo, Zhiyao, et al.
Published: (2025)
Robust Exploration in Directed Controller Synthesis via Reinforcement Learning with Soft Mixture-of-Experts
by: Ubukata, Toshihide, et al.
Published: (2026)
by: Ubukata, Toshihide, et al.
Published: (2026)
Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning
by: Liu, Zhishuai, et al.
Published: (2024)
by: Liu, Zhishuai, et al.
Published: (2024)
Study of the Proper NNUE Dataset
by: Tan, Daniel, et al.
Published: (2024)
by: Tan, Daniel, et al.
Published: (2024)
Diffusion Models for Reinforcement Learning: A Survey
by: Zhu, Zhengbang, et al.
Published: (2023)
by: Zhu, Zhengbang, et al.
Published: (2023)
Easy Samples Are All You Need: Self-Evolving LLMs via Data-Efficient Reinforcement Learning
by: Yu, Zhiyin, et al.
Published: (2026)
by: Yu, Zhiyin, et al.
Published: (2026)
UAV-MARL: Multi-Agent Reinforcement Learning for Time-Critical and Dynamic Medical Supply Delivery
by: Guven, Islam, et al.
Published: (2026)
by: Guven, Islam, et al.
Published: (2026)
Flow Actor-Critic for Offline Reinforcement Learning
by: Chae, Jongseong, et al.
Published: (2026)
by: Chae, Jongseong, et al.
Published: (2026)
Probabilistic Constraint for Safety-Critical Reinforcement Learning
by: Chen, Weiqin, et al.
Published: (2023)
by: Chen, Weiqin, et al.
Published: (2023)
DynaGraph: Interpretable Multi-Label Prediction from EHRs via Dynamic Graph Learning and Contrastive Augmentation
by: Mesinovic, Munib, et al.
Published: (2025)
by: Mesinovic, Munib, et al.
Published: (2025)
Localized Dynamics-Aware Domain Adaption for Off-Dynamics Offline Reinforcement Learning
by: Xia, Zhangjie, et al.
Published: (2026)
by: Xia, Zhangjie, et al.
Published: (2026)
Provably Efficient Action-Manipulation Attack Against Continuous Reinforcement Learning
by: Luo, Zhi, et al.
Published: (2024)
by: Luo, Zhi, et al.
Published: (2024)
Adaptive Insurance Reserving with CVaR-Constrained Reinforcement Learning under Macroeconomic Regimes
by: Dong, Stella C.
Published: (2025)
by: Dong, Stella C.
Published: (2025)
A Survey of Explainable Reinforcement Learning: Targets, Methods and Needs
by: Saulières, Léo
Published: (2025)
by: Saulières, Léo
Published: (2025)
Skill-Critic: Refining Learned Skills for Hierarchical Reinforcement Learning
by: Hao, Ce, et al.
Published: (2023)
by: Hao, Ce, et al.
Published: (2023)
CAE: Repurposing the Critic as an Explorer in Deep Reinforcement Learning
by: Li, Yexin
Published: (2025)
by: Li, Yexin
Published: (2025)
Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning
by: Zhuang, Yuan, et al.
Published: (2026)
by: Zhuang, Yuan, et al.
Published: (2026)
MOBODY: Model Based Off-Dynamics Offline Reinforcement Learning
by: Guo, Yihong, et al.
Published: (2025)
by: Guo, Yihong, et al.
Published: (2025)
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
by: Thalagala, Shiron, et al.
Published: (2024)
by: Thalagala, Shiron, et al.
Published: (2024)
Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning
by: Garcin, Samuel, et al.
Published: (2025)
by: Garcin, Samuel, et al.
Published: (2025)
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
by: Küçükoğlu, Burcu, et al.
Published: (2025)
by: Küçükoğlu, Burcu, et al.
Published: (2025)
Treatment Stitching with Schrödinger Bridge for Enhancing Offline Reinforcement Learning in Adaptive Treatment Strategies
by: Shin, Dong-Hee, et al.
Published: (2025)
by: Shin, Dong-Hee, et al.
Published: (2025)
Return Augmented Decision Transformer for Off-Dynamics Reinforcement Learning
by: Wang, Ruhan, et al.
Published: (2024)
by: Wang, Ruhan, et al.
Published: (2024)
Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic
by: Vo, Thanh Vinh, et al.
Published: (2025)
by: Vo, Thanh Vinh, et al.
Published: (2025)
Is Exploration All You Need? Effective Exploration Characteristics for Transfer in Reinforcement Learning
by: Balloch, Jonathan C., et al.
Published: (2024)
by: Balloch, Jonathan C., et al.
Published: (2024)
Sample-Efficiency in Multi-Batch Reinforcement Learning: The Need for Dimension-Dependent Adaptivity
by: Johnson, Emmeran, et al.
Published: (2023)
by: Johnson, Emmeran, et al.
Published: (2023)
Overconfident Errors Need Stronger Correction: Asymmetric Confidence Penalties for Reinforcement Learning
by: Xu, Yuanda, et al.
Published: (2026)
by: Xu, Yuanda, et al.
Published: (2026)
A Survey of Few-Shot Learning for Biomedical Time Series
by: Li, Chenqi, et al.
Published: (2024)
by: Li, Chenqi, et al.
Published: (2024)
HALO: Hierarchical Reinforcement Learning for Large-Scale Adaptive Traffic Signal Control
by: Zhu, Yaqiao, et al.
Published: (2025)
by: Zhu, Yaqiao, et al.
Published: (2025)
DySurv: dynamic deep learning model for survival analysis with conditional variational inference
by: Mesinovic, Munib, et al.
Published: (2023)
by: Mesinovic, Munib, et al.
Published: (2023)
Similar Items
-
DTR-Bench: An in silico Environment and Benchmark Platform for Reinforcement Learning Based Dynamic Treatment Regime
by: Luo, Zhiyao, et al.
Published: (2024) -
ProtoEHR: Hierarchical Prototype Learning for EHR-based Healthcare Predictions
by: Cai, Zi, et al.
Published: (2025) -
Bridging Data Gaps of Rare Conditions in ICU: A Multi-Disease Adaptation Approach for Clinical Prediction
by: Zhu, Mingcheng, et al.
Published: (2025) -
Cross-Representation Benchmarking in Time-Series Electronic Health Records for Clinical Outcome Prediction
by: Chen, Tianyi, et al.
Published: (2025) -
Kernel-Based Distributed Q-Learning: A Scalable Reinforcement Learning Approach for Dynamic Treatment Regimes
by: Wang, Di, et al.
Published: (2023)