SNPL: Simultaneous Policy Learning and Evaluation for Safe Multi-Objective Policy Improvement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cho, Brian, Pop, Ana-Roxana, Evnine, Ariel, Kallus, Nathan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CSPI-MT: Calibrated Safe Policy Improvement with Multiple Testing for Threshold Policies
von: Cho, Brian M, et al.
Veröffentlicht: (2024)
von: Cho, Brian M, et al.
Veröffentlicht: (2024)
Robust and Agnostic Learning of Conditional Distributional Treatment Effects
von: Kallus, Nathan, et al.
Veröffentlicht: (2022)
von: Kallus, Nathan, et al.
Veröffentlicht: (2022)
Bellman Calibration for $V$-Learning in Offline Reinforcement Learning
von: van der Laan, Lars, et al.
Veröffentlicht: (2025)
von: van der Laan, Lars, et al.
Veröffentlicht: (2025)
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
von: Kallus, Nathan
Veröffentlicht: (2025)
von: Kallus, Nathan
Veröffentlicht: (2025)
Demistifying Inference after Adaptive Experiments
von: Bibaut, Aurélien, et al.
Veröffentlicht: (2024)
von: Bibaut, Aurélien, et al.
Veröffentlicht: (2024)
Inverse Reinforcement Learning with Just Classification and a Few Regressions
von: van der Laan, Lars, et al.
Veröffentlicht: (2025)
von: van der Laan, Lars, et al.
Veröffentlicht: (2025)
Long-term Causal Inference Under Persistent Confounding via Data Combination
von: Imbens, Guido, et al.
Veröffentlicht: (2022)
von: Imbens, Guido, et al.
Veröffentlicht: (2022)
Policy Learning with Abstention
von: Sawarni, Ayush, et al.
Veröffentlicht: (2025)
von: Sawarni, Ayush, et al.
Veröffentlicht: (2025)
Policy Learning with Competing Agents
von: Sahoo, Roshni, et al.
Veröffentlicht: (2022)
von: Sahoo, Roshni, et al.
Veröffentlicht: (2022)
Near-Optimal Non-Parametric Sequential Tests and Confidence Sequences with Possibly Dependent Observations
von: Bibaut, Aurelien, et al.
Veröffentlicht: (2022)
von: Bibaut, Aurelien, et al.
Veröffentlicht: (2022)
Applied Causal Inference Powered by ML and AI
von: Chernozhukov, Victor, et al.
Veröffentlicht: (2024)
von: Chernozhukov, Victor, et al.
Veröffentlicht: (2024)
Individualized Policy Evaluation and Learning under Clustered Network Interference
von: Zhang, Yi, et al.
Veröffentlicht: (2023)
von: Zhang, Yi, et al.
Veröffentlicht: (2023)
Quantile-Optimal Policy Learning under Unmeasured Confounding
von: Chen, Zhongren, et al.
Veröffentlicht: (2025)
von: Chen, Zhongren, et al.
Veröffentlicht: (2025)
Profit-Aligned CATE Estimation: Reconciling Policy Learning and Inference
von: Timoshenko, Artem, et al.
Veröffentlicht: (2025)
von: Timoshenko, Artem, et al.
Veröffentlicht: (2025)
Externally Valid Policy Choice
von: Adjaho, Christopher, et al.
Veröffentlicht: (2022)
von: Adjaho, Christopher, et al.
Veröffentlicht: (2022)
Data-Automated Policy Learning for Nonlinear Welfare
von: Ai, Chunrong, et al.
Veröffentlicht: (2026)
von: Ai, Chunrong, et al.
Veröffentlicht: (2026)
Robustness of Refugee-Matching Gains to Off-Policy Evaluation Choices
von: Bansak, Kirk, et al.
Veröffentlicht: (2026)
von: Bansak, Kirk, et al.
Veröffentlicht: (2026)
Causal-Policy Forest for End-to-End Policy Learning
von: Kato, Masahiro
Veröffentlicht: (2025)
von: Kato, Masahiro
Veröffentlicht: (2025)
Off-Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits
von: Zhan, Ruohan, et al.
Veröffentlicht: (2021)
von: Zhan, Ruohan, et al.
Veröffentlicht: (2021)
Beating the Winner's Curse via Inference-Aware Policy Optimization
von: Bastani, Hamsa, et al.
Veröffentlicht: (2025)
von: Bastani, Hamsa, et al.
Veröffentlicht: (2025)
Policy design in experiments with unknown interference
von: Viviano, Davide, et al.
Veröffentlicht: (2020)
von: Viviano, Davide, et al.
Veröffentlicht: (2020)
Nonparametric Instrumental Variable Analysis Without Structural Equations: Debiased Inference on Functionals of Inverse Problems with No Solutions
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
General Bayesian Policy Learning
von: Kato, Masahiro
Veröffentlicht: (2026)
von: Kato, Masahiro
Veröffentlicht: (2026)
Policy Learning with Distributional Welfare
von: Cui, Yifan, et al.
Veröffentlicht: (2023)
von: Cui, Yifan, et al.
Veröffentlicht: (2023)
Policy-Oriented Binary Classification: Improving (KD-)CART Final Splits for Subpopulation Targeting
von: Wang, Lei Bill, et al.
Veröffentlicht: (2025)
von: Wang, Lei Bill, et al.
Veröffentlicht: (2025)
Adaptive Experimental Design for Policy Learning
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
Semiparametric Off-Policy Inference for Optimal Policy Values under Possible Non-Uniqueness
von: Wei, Haoyu
Veröffentlicht: (2025)
von: Wei, Haoyu
Veröffentlicht: (2025)
Policy Learning with Observational Data: The Case of Hepatitis C Treatment for HIV/HCV Co-Infected Patients
von: Langevin, Raphaël
Veröffentlicht: (2026)
von: Langevin, Raphaël
Veröffentlicht: (2026)
Federated Offline Policy Learning
von: Carranza, Aldo Gael, et al.
Veröffentlicht: (2023)
von: Carranza, Aldo Gael, et al.
Veröffentlicht: (2023)
Estimation of Optimal Dynamic Treatment Assignment Rules under Policy Constraints
von: Sakaguchi, Shosei
Veröffentlicht: (2021)
von: Sakaguchi, Shosei
Veröffentlicht: (2021)
Econometric vs. Causal Structure-Learning for Time-Series Policy Decisions: Evidence from the UK COVID-19 Policies
von: Petrungaro, Bruno, et al.
Veröffentlicht: (2026)
von: Petrungaro, Bruno, et al.
Veröffentlicht: (2026)
Policy Learning for Optimal Dynamic Treatment Regimes with Observational Data
von: Sakaguchi, Shosei
Veröffentlicht: (2024)
von: Sakaguchi, Shosei
Veröffentlicht: (2024)
Identification and Estimation of Simultaneous Equation Models Using Higher-Order Cumulant Restrictions
von: Jiang, Ziyu
Veröffentlicht: (2025)
von: Jiang, Ziyu
Veröffentlicht: (2025)
Causal Multi-Task Demand Learning
von: Gupta, Varun, et al.
Veröffentlicht: (2026)
von: Gupta, Varun, et al.
Veröffentlicht: (2026)
ScoreMatchingRiesz: Score Matching for Debiased Machine Learning and Policy Path Estimation
von: Kato, Masahiro
Veröffentlicht: (2025)
von: Kato, Masahiro
Veröffentlicht: (2025)
Bridging the Gap between Empirical Welfare Maximization and Conditional Average Treatment Effect Estimation in Policy Learning
von: Kato, Masahiro
Veröffentlicht: (2025)
von: Kato, Masahiro
Veröffentlicht: (2025)
Machine Learning and Econometric Approaches to Fiscal Policies: Understanding Industrial Investment Dynamics in Uruguay (1974-2010)
von: Vallarino, Diego
Veröffentlicht: (2024)
von: Vallarino, Diego
Veröffentlicht: (2024)
Policy Targeting under Network Interference
von: Viviano, Davide
Veröffentlicht: (2019)
von: Viviano, Davide
Veröffentlicht: (2019)
Multi-Agent Reinforcement Learning for Dynamic Pricing in Supply Chains: Benchmarking Strategic Agent Behaviours under Realistically Simulated Market Conditions
von: Hazenberg, Thomas, et al.
Veröffentlicht: (2025)
von: Hazenberg, Thomas, et al.
Veröffentlicht: (2025)
Inference on Optimal Policy Values and Other Irregular Functionals via Softmax Smoothing
von: Whitehouse, Justin, et al.
Veröffentlicht: (2025)
von: Whitehouse, Justin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CSPI-MT: Calibrated Safe Policy Improvement with Multiple Testing for Threshold Policies
von: Cho, Brian M, et al.
Veröffentlicht: (2024) -
Robust and Agnostic Learning of Conditional Distributional Treatment Effects
von: Kallus, Nathan, et al.
Veröffentlicht: (2022) -
Bellman Calibration for $V$-Learning in Offline Reinforcement Learning
von: van der Laan, Lars, et al.
Veröffentlicht: (2025) -
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
von: Kallus, Nathan
Veröffentlicht: (2025) -
Demistifying Inference after Adaptive Experiments
von: Bibaut, Aurélien, et al.
Veröffentlicht: (2024)