Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting
Fuente:
arXiv
Saved in:
| Main Authors: | van der Laan, Lars, Kallus, Nathan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Bellman Calibration for $V$-Learning in Offline Reinforcement Learning
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Efficient Inference for Inverse Reinforcement Learning and Dynamic Discrete Choice Models
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Nonparametric Instrumental Variable Inference with Many Weak Instruments
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Inverse Reinforcement Learning with Just Classification and a Few Regressions
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Reward Transfer from Inverse Reinforcement Learning: A Coupled Minimax Approach
by: Hao, Guang-Yuan, et al.
Published: (2026)
by: Hao, Guang-Yuan, et al.
Published: (2026)
Automatic Debiased Machine Learning for Smooth Functionals of Nonparametric M-Estimands
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Semiparametric Double Reinforcement Learning with Applications to Long-Term Causal Inference
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Stabilized Inverse Probability Weighting via Isotonic Calibration
by: van der Laan, Lars, et al.
Published: (2024)
by: van der Laan, Lars, et al.
Published: (2024)
A Researcher's Guide to Empirical Risk Minimization
by: van der Laan, Lars
Published: (2026)
by: van der Laan, Lars
Published: (2026)
Deterministic Exploration via Stationary Bellman Error Maximization
by: Griesbach, Sebastian, et al.
Published: (2024)
by: Griesbach, Sebastian, et al.
Published: (2024)
Generalized Venn and Venn-Abers Calibration with Applications in Conformal Prediction
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Smooth Non-Stationary Bandits
by: Jia, Su, et al.
Published: (2023)
by: Jia, Su, et al.
Published: (2023)
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
by: Kallus, Nathan
Published: (2025)
by: Kallus, Nathan
Published: (2025)
Adaptive-TMLE for the Average Treatment Effect based on Randomized Controlled Trial Augmented with Real-World Data
by: van der Laan, Mark, et al.
Published: (2024)
by: van der Laan, Mark, et al.
Published: (2024)
Doubly robust inference via calibration
by: van der Laan, Lars, et al.
Published: (2024)
by: van der Laan, Lars, et al.
Published: (2024)
Adaptive debiased machine learning using data-driven model selection techniques
by: van der Laan, Lars, et al.
Published: (2023)
by: van der Laan, Lars, et al.
Published: (2023)
ShiQ: Bringing back Bellman to LLMs
by: Clavier, Pierre, et al.
Published: (2025)
by: Clavier, Pierre, et al.
Published: (2025)
Self-Calibrating Conformal Prediction
by: van der Laan, Lars, et al.
Published: (2024)
by: van der Laan, Lars, et al.
Published: (2024)
Calibeating Prediction-Powered Inference
by: van der Laan, Lars, et al.
Published: (2026)
by: van der Laan, Lars, et al.
Published: (2026)
Hybrid Meta-learners for Estimating Heterogeneous Treatment Effects
by: Liang, Zhongyuan, et al.
Published: (2025)
by: Liang, Zhongyuan, et al.
Published: (2025)
Combining T-learning and DR-learning: a framework for oracle-efficient estimation of causal contrasts
by: van der Laan, Lars, et al.
Published: (2024)
by: van der Laan, Lars, et al.
Published: (2024)
Estimating Heterogeneous Treatment Effects by Combining Weak Instruments and Observational Data
by: Oprescu, Miruna, et al.
Published: (2024)
by: Oprescu, Miruna, et al.
Published: (2024)
On the role of surrogates in the efficient estimation of treatment effects with limited outcome data
by: Kallus, Nathan, et al.
Published: (2020)
by: Kallus, Nathan, et al.
Published: (2020)
Robust and Agnostic Learning of Conditional Distributional Treatment Effects
by: Kallus, Nathan, et al.
Published: (2022)
by: Kallus, Nathan, et al.
Published: (2022)
Variation Due to Regularization Tractably Recovers Bayesian Deep Learning
by: McInerney, James, et al.
Published: (2024)
by: McInerney, James, et al.
Published: (2024)
Does Weighting Improve Matrix Factorization for Recommender Systems?
by: Ayoub, Alex, et al.
Published: (2025)
by: Ayoub, Alex, et al.
Published: (2025)
Demistifying Inference after Adaptive Experiments
by: Bibaut, Aurélien, et al.
Published: (2024)
by: Bibaut, Aurélien, et al.
Published: (2024)
Multi-Armed Bandits with Interference
by: Jia, Su, et al.
Published: (2024)
by: Jia, Su, et al.
Published: (2024)
SNPL: Simultaneous Policy Learning and Evaluation for Safe Multi-Objective Policy Improvement
by: Cho, Brian, et al.
Published: (2025)
by: Cho, Brian, et al.
Published: (2025)
Exploration in the Limit
by: Cho, Brian M., et al.
Published: (2025)
by: Cho, Brian M., et al.
Published: (2025)
Peeking with PEAK: Sequential, Nonparametric Composite Hypothesis Tests for Means of Multiple Data Streams
by: Cho, Brian, et al.
Published: (2024)
by: Cho, Brian, et al.
Published: (2024)
Long-term Causal Inference Under Persistent Confounding via Data Combination
by: Imbens, Guido, et al.
Published: (2022)
by: Imbens, Guido, et al.
Published: (2022)
Towards Optimal Adversarial Robust Q-learning with Bellman Infinity-error
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
The Context Gathering Decision Process: A POMDP Framework for Agentic Search
by: Kausik, Chinmaya, et al.
Published: (2026)
by: Kausik, Chinmaya, et al.
Published: (2026)
Low-Rank MDPs with Continuous Action Spaces
by: Bennett, Andrew, et al.
Published: (2023)
by: Bennett, Andrew, et al.
Published: (2023)
The Central Role of the Loss Function in Reinforcement Learning
by: Wang, Kaiwen, et al.
Published: (2024)
by: Wang, Kaiwen, et al.
Published: (2024)
Is Cosine-Similarity of Embeddings Really About Similarity?
by: Steck, Harald, et al.
Published: (2024)
by: Steck, Harald, et al.
Published: (2024)
Nonparametric Instrumental Variable Analysis Without Structural Equations: Debiased Inference on Functionals of Inverse Problems with No Solutions
by: Shen, Zikai, et al.
Published: (2026)
by: Shen, Zikai, et al.
Published: (2026)
Efficient Adaptive Experimentation with Noncompliance
by: Oprescu, Miruna, et al.
Published: (2025)
by: Oprescu, Miruna, et al.
Published: (2025)
Similar Items
-
Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration
by: van der Laan, Lars, et al.
Published: (2025) -
Bellman Calibration for $V$-Learning in Offline Reinforcement Learning
by: van der Laan, Lars, et al.
Published: (2025) -
Efficient Inference for Inverse Reinforcement Learning and Dynamic Discrete Choice Models
by: van der Laan, Lars, et al.
Published: (2025) -
Nonparametric Instrumental Variable Inference with Many Weak Instruments
by: van der Laan, Lars, et al.
Published: (2025) -
Inverse Reinforcement Learning with Just Classification and a Few Regressions
by: van der Laan, Lars, et al.
Published: (2025)