Distributionally Robust Policy Evaluation and Learning for Continuous Treatment with Observational Data
Fuente:
arXiv
Guardado en:
| Autores principales: | Leung, Cheuk Hang, Huang, Yiyan, Li, Yijun, Wu, Qi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Unveiling the Potential of Robustness in Selecting Conditional Average Treatment Effect Estimators
por: Huang, Yiyan, et al.
Publicado: (2024)
por: Huang, Yiyan, et al.
Publicado: (2024)
Probabilistic Learning of Multivariate Time Series with Temporal Irregularity
por: Li, Yijun, et al.
Publicado: (2023)
por: Li, Yijun, et al.
Publicado: (2023)
Latent Conditional Diffusion-based Data Augmentation for Continuous-Time Dynamic Graph Model
por: Tian, Yuxing, et al.
Publicado: (2024)
por: Tian, Yuxing, et al.
Publicado: (2024)
Task-Distributionally Robust Data-Free Meta-Learning
por: Hu, Zixuan, et al.
Publicado: (2023)
por: Hu, Zixuan, et al.
Publicado: (2023)
Belief-Based Offline Reinforcement Learning for Delay-Robust Policy Optimization
por: Zhan, Simon Sinong, et al.
Publicado: (2025)
por: Zhan, Simon Sinong, et al.
Publicado: (2025)
A Two-Stage Feature Selection Approach for Robust Evaluation of Treatment Effects in High-Dimensional Observational Data
por: Islam, Md Saiful, et al.
Publicado: (2021)
por: Islam, Md Saiful, et al.
Publicado: (2021)
Guided Learning: Lubricating End-to-End Modeling for Multi-stage Decision-making
por: Guo, Jian, et al.
Publicado: (2024)
por: Guo, Jian, et al.
Publicado: (2024)
NUM2EVENT: Interpretable Event Reasoning from Numerical time-series
por: Feng, Ninghui, et al.
Publicado: (2025)
por: Feng, Ninghui, et al.
Publicado: (2025)
Ranking Policy Learning via Marketplace Expected Value Estimation From Observational Data
por: Ebrahimzadeh, Ehsan, et al.
Publicado: (2024)
por: Ebrahimzadeh, Ehsan, et al.
Publicado: (2024)
Off-Policy Actor-Critic for Adversarial Observation Robustness: Virtual Alternative Training via Symmetric Policy Evaluation
por: Nakanishi, Kosuke, et al.
Publicado: (2025)
por: Nakanishi, Kosuke, et al.
Publicado: (2025)
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
por: Zhu, Yuanyang, et al.
Publicado: (2024)
por: Zhu, Yuanyang, et al.
Publicado: (2024)
Pareto-Optimal Estimation and Policy Learning on Short-term and Long-term Treatment Effects
por: Wang, Yingrong, et al.
Publicado: (2024)
por: Wang, Yingrong, et al.
Publicado: (2024)
Offline Policy Evaluation for Reinforcement Learning with Adaptively Collected Data
por: Madhow, Sunil, et al.
Publicado: (2023)
por: Madhow, Sunil, et al.
Publicado: (2023)
Wasserstein Distributionally Robust Bayesian Optimization with Continuous Context
por: Micheli, Francesco, et al.
Publicado: (2025)
por: Micheli, Francesco, et al.
Publicado: (2025)
Learning Discriminative and Generalizable Anomaly Detector for Dynamic Graph with Limited Supervision
por: Tian, Yuxing, et al.
Publicado: (2026)
por: Tian, Yuxing, et al.
Publicado: (2026)
Optimal Policy Learning with Observational Data in Multi-Action Scenarios: Estimation, Risk Preference, and Potential Failures
por: Cerulli, Giovanni
Publicado: (2024)
por: Cerulli, Giovanni
Publicado: (2024)
Categorical Policies: Multimodal Policy Learning and Exploration in Continuous Control
por: Islam, SM Mazharul, et al.
Publicado: (2025)
por: Islam, SM Mazharul, et al.
Publicado: (2025)
Reinforcement Learning with Euclidean Data Augmentation for State-Based Continuous Control
por: Luo, Jinzhu, et al.
Publicado: (2024)
por: Luo, Jinzhu, et al.
Publicado: (2024)
Breaking the Curse of Repulsion: Optimistic Distributionally Robust Policy Optimization for Off-Policy Generative Recommendation
por: Jiang, Jie, et al.
Publicado: (2026)
por: Jiang, Jie, et al.
Publicado: (2026)
Energy-Structured Low-Rank Adaptation for Continual Learning
por: Li, Longhua, et al.
Publicado: (2026)
por: Li, Longhua, et al.
Publicado: (2026)
Continual Task Learning through Adaptive Policy Self-Composition
por: Hu, Shengchao, et al.
Publicado: (2024)
por: Hu, Shengchao, et al.
Publicado: (2024)
Maintaining Adversarial Robustness in Continuous Learning
por: Ru, Xiaolei, et al.
Publicado: (2024)
por: Ru, Xiaolei, et al.
Publicado: (2024)
Better Generative Replay for Continual Federated Learning
por: Qi, Daiqing, et al.
Publicado: (2023)
por: Qi, Daiqing, et al.
Publicado: (2023)
Robust Policy Expansion for Offline-to-Online RL under Diverse Data Corruption
por: He, Longxiang, et al.
Publicado: (2025)
por: He, Longxiang, et al.
Publicado: (2025)
Rethinking Adversarial Attacks in Reinforcement Learning from Policy Distribution Perspective
por: Duan, Tianyang, et al.
Publicado: (2025)
por: Duan, Tianyang, et al.
Publicado: (2025)
Policy Regularized Distributionally Robust Markov Decision Processes with Linear Function Approximation
por: Gu, Jingwen, et al.
Publicado: (2025)
por: Gu, Jingwen, et al.
Publicado: (2025)
Distribution-Level Memory Recall for Continual Learning: Preserving Knowledge and Avoiding Confusion
por: Cheng, Shaoxu, et al.
Publicado: (2024)
por: Cheng, Shaoxu, et al.
Publicado: (2024)
SCENT: Robust Spatiotemporal Learning for Continuous Scientific Data via Scalable Conditioned Neural Fields
por: Park, David Keetae, et al.
Publicado: (2025)
por: Park, David Keetae, et al.
Publicado: (2025)
Low-redundancy Distillation for Continual Learning
por: Liu, RuiQi, et al.
Publicado: (2023)
por: Liu, RuiQi, et al.
Publicado: (2023)
MetaCLBench: Meta Continual Learning Benchmark on Resource-Constrained Edge Devices
por: Li, Sijia, et al.
Publicado: (2025)
por: Li, Sijia, et al.
Publicado: (2025)
Efficient and Sharp Off-Policy Evaluation in Robust Markov Decision Processes
por: Bennett, Andrew, et al.
Publicado: (2024)
por: Bennett, Andrew, et al.
Publicado: (2024)
A Closer Look at the Application of Causal Inference in Graph Representation Learning
por: Gao, Hang, et al.
Publicado: (2026)
por: Gao, Hang, et al.
Publicado: (2026)
Blending Imitation and Reinforcement Learning for Robust Policy Improvement
por: Liu, Xuefeng, et al.
Publicado: (2023)
por: Liu, Xuefeng, et al.
Publicado: (2023)
Uncertainty-based Offline Variational Bayesian Reinforcement Learning for Robustness under Diverse Data Corruptions
por: Yang, Rui, et al.
Publicado: (2024)
por: Yang, Rui, et al.
Publicado: (2024)
Information-Theoretic Dual Memory System for Continual Learning
por: Wu, RunQing, et al.
Publicado: (2025)
por: Wu, RunQing, et al.
Publicado: (2025)
Optimizing Warfarin Dosing Using Contextual Bandit: An Offline Policy Learning and Evaluation Method
por: Huang, Yong, et al.
Publicado: (2024)
por: Huang, Yong, et al.
Publicado: (2024)
Single-Trajectory Distributionally Robust Reinforcement Learning
por: Liang, Zhipeng, et al.
Publicado: (2023)
por: Liang, Zhipeng, et al.
Publicado: (2023)
MLLM Is a Strong Reranker: Advancing Multimodal Retrieval-augmented Generation via Knowledge-enhanced Reranking and Noise-injected Training
por: Chen, Zhanpeng, et al.
Publicado: (2024)
por: Chen, Zhanpeng, et al.
Publicado: (2024)
Learning Action Embeddings for Off-Policy Evaluation
por: Cief, Matej, et al.
Publicado: (2023)
por: Cief, Matej, et al.
Publicado: (2023)
Learning to Compress Graphs via Dual Agents for Consistent Topological Robustness Evaluation
por: Chai, Qisen, et al.
Publicado: (2025)
por: Chai, Qisen, et al.
Publicado: (2025)
Ejemplares similares
-
Unveiling the Potential of Robustness in Selecting Conditional Average Treatment Effect Estimators
por: Huang, Yiyan, et al.
Publicado: (2024) -
Probabilistic Learning of Multivariate Time Series with Temporal Irregularity
por: Li, Yijun, et al.
Publicado: (2023) -
Latent Conditional Diffusion-based Data Augmentation for Continuous-Time Dynamic Graph Model
por: Tian, Yuxing, et al.
Publicado: (2024) -
Task-Distributionally Robust Data-Free Meta-Learning
por: Hu, Zixuan, et al.
Publicado: (2023) -
Belief-Based Offline Reinforcement Learning for Delay-Robust Policy Optimization
por: Zhan, Simon Sinong, et al.
Publicado: (2025)