Bellman Calibration for $V$-Learning in Offline Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | van der Laan, Lars, Kallus, Nathan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Inverse Reinforcement Learning with Just Classification and a Few Regressions
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Robust and Agnostic Learning of Conditional Distributional Treatment Effects
by: Kallus, Nathan, et al.
Published: (2022)
by: Kallus, Nathan, et al.
Published: (2022)
Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
by: Kallus, Nathan
Published: (2025)
by: Kallus, Nathan
Published: (2025)
SNPL: Simultaneous Policy Learning and Evaluation for Safe Multi-Objective Policy Improvement
by: Cho, Brian, et al.
Published: (2025)
by: Cho, Brian, et al.
Published: (2025)
Calibeating Prediction-Powered Inference
by: van der Laan, Lars, et al.
Published: (2026)
by: van der Laan, Lars, et al.
Published: (2026)
Demistifying Inference after Adaptive Experiments
by: Bibaut, Aurélien, et al.
Published: (2024)
by: Bibaut, Aurélien, et al.
Published: (2024)
Long-term Causal Inference Under Persistent Confounding via Data Combination
by: Imbens, Guido, et al.
Published: (2022)
by: Imbens, Guido, et al.
Published: (2022)
Near-Optimal Non-Parametric Sequential Tests and Confidence Sequences with Possibly Dependent Observations
by: Bibaut, Aurelien, et al.
Published: (2022)
by: Bibaut, Aurelien, et al.
Published: (2022)
Applied Causal Inference Powered by ML and AI
by: Chernozhukov, Victor, et al.
Published: (2024)
by: Chernozhukov, Victor, et al.
Published: (2024)
Efficient Inference for Inverse Reinforcement Learning and Dynamic Discrete Choice Models
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
A Lecture Note on Offline RL and IRL, Part II: Foundations of Inverse Reinforcement Learning and Dynamic Discrete Choice Models
by: Kang, Enoch Hyunwook
Published: (2026)
by: Kang, Enoch Hyunwook
Published: (2026)
Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Post Reinforcement Learning Inference
by: Syrgkanis, Vasilis, et al.
Published: (2023)
by: Syrgkanis, Vasilis, et al.
Published: (2023)
Multi-Agent Reinforcement Learning for Dynamic Pricing in Supply Chains: Benchmarking Strategic Agent Behaviours under Realistically Simulated Market Conditions
by: Hazenberg, Thomas, et al.
Published: (2025)
by: Hazenberg, Thomas, et al.
Published: (2025)
Reward Transfer from Inverse Reinforcement Learning: A Coupled Minimax Approach
by: Hao, Guang-Yuan, et al.
Published: (2026)
by: Hao, Guang-Yuan, et al.
Published: (2026)
Incorporating Cognitive Biases into Reinforcement Learning for Financial Decision-Making
by: He, Liu
Published: (2026)
by: He, Liu
Published: (2026)
Nonparametric Instrumental Variable Analysis Without Structural Equations: Debiased Inference on Functionals of Inverse Problems with No Solutions
by: Shen, Zikai, et al.
Published: (2026)
by: Shen, Zikai, et al.
Published: (2026)
Federated Offline Policy Learning
by: Carranza, Aldo Gael, et al.
Published: (2023)
by: Carranza, Aldo Gael, et al.
Published: (2023)
Improving the Finite Sample Estimation of Average Treatment Effects using Double/Debiased Machine Learning with Propensity Score Calibration
by: Ballinari, Daniele, et al.
Published: (2024)
by: Ballinari, Daniele, et al.
Published: (2024)
STEEL: Singularity-aware Reinforcement Learning
by: Chen, Xiaohong, et al.
Published: (2023)
by: Chen, Xiaohong, et al.
Published: (2023)
Semiparametric Double Reinforcement Learning with Applications to Long-Term Causal Inference
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Calibrating doubly-robust estimators with unbalanced treatment assignment
by: Ballinari, Daniele
Published: (2024)
by: Ballinari, Daniele
Published: (2024)
Policy Learning with Abstention
by: Sawarni, Ayush, et al.
Published: (2025)
by: Sawarni, Ayush, et al.
Published: (2025)
Macroeconomic Forecasting and Machine Learning
by: Chi, Ta-Chung, et al.
Published: (2025)
by: Chi, Ta-Chung, et al.
Published: (2025)
Policy Learning with Competing Agents
by: Sahoo, Roshni, et al.
Published: (2022)
by: Sahoo, Roshni, et al.
Published: (2022)
Asymptotic Theory for IV-Based Reinforcement Learning with Potential Endogeneity
by: Li, Jin, et al.
Published: (2021)
by: Li, Jin, et al.
Published: (2021)
Dual Interpretation of Machine Learning Forecasts
by: Coulombe, Philippe Goulet, et al.
Published: (2024)
by: Coulombe, Philippe Goulet, et al.
Published: (2024)
Causal Multi-Task Demand Learning
by: Gupta, Varun, et al.
Published: (2026)
by: Gupta, Varun, et al.
Published: (2026)
Model Averaging and Double Machine Learning
by: Ahrens, Achim, et al.
Published: (2024)
by: Ahrens, Achim, et al.
Published: (2024)
Causal Machine Learning for Moderation Effects
by: Bearth, Nora, et al.
Published: (2024)
by: Bearth, Nora, et al.
Published: (2024)
Simulation-Based Benchmarking of Reinforcement Learning Agents for Personalized Retail Promotions
by: Xia, Yu, et al.
Published: (2024)
by: Xia, Yu, et al.
Published: (2024)
The Uncertainty of Machine Learning Predictions in Asset Pricing
by: Liao, Yuan, et al.
Published: (2025)
by: Liao, Yuan, et al.
Published: (2025)
Automatic Debiased Machine Learning for Smooth Functionals of Nonparametric M-Estimands
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
A Hybrid Framework for Reinsurance Optimization: Integrating Generative Models and Reinforcement Learning
by: Dong, Stella C.
Published: (2025)
by: Dong, Stella C.
Published: (2025)
Unemployment Dynamics Forecasting with Machine Learning Regression Models
by: Kim, Kyungsu
Published: (2025)
by: Kim, Kyungsu
Published: (2025)
Quantile-Optimal Policy Learning under Unmeasured Confounding
by: Chen, Zhongren, et al.
Published: (2025)
by: Chen, Zhongren, et al.
Published: (2025)
Learning Correlated Reward Models: Statistical Barriers and Opportunities
by: Cherapanamjeri, Yeshwanth, et al.
Published: (2025)
by: Cherapanamjeri, Yeshwanth, et al.
Published: (2025)
Inference for Regression with Variables Generated by AI or Machine Learning
by: Battaglia, Laura, et al.
Published: (2024)
by: Battaglia, Laura, et al.
Published: (2024)
Profit-Aligned CATE Estimation: Reconciling Policy Learning and Inference
by: Timoshenko, Artem, et al.
Published: (2025)
by: Timoshenko, Artem, et al.
Published: (2025)
Similar Items
-
Inverse Reinforcement Learning with Just Classification and a Few Regressions
by: van der Laan, Lars, et al.
Published: (2025) -
Robust and Agnostic Learning of Conditional Distributional Treatment Effects
by: Kallus, Nathan, et al.
Published: (2022) -
Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting
by: van der Laan, Lars, et al.
Published: (2025) -
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
by: Kallus, Nathan
Published: (2025) -
SNPL: Simultaneous Policy Learning and Evaluation for Safe Multi-Objective Policy Improvement
by: Cho, Brian, et al.
Published: (2025)