Inverse Reinforcement Learning with Just Classification and a Few Regressions
Fuente:
arXiv
Guardado en:
| Autores principales: | van der Laan, Lars, Kallus, Nathan, Bibaut, Aurelien |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Bellman Calibration for $V$-Learning in Offline Reinforcement Learning
por: van der Laan, Lars, et al.
Publicado: (2025)
por: van der Laan, Lars, et al.
Publicado: (2025)
Demistifying Inference after Adaptive Experiments
por: Bibaut, Aurélien, et al.
Publicado: (2024)
por: Bibaut, Aurélien, et al.
Publicado: (2024)
Efficient Inference for Inverse Reinforcement Learning and Dynamic Discrete Choice Models
por: van der Laan, Lars, et al.
Publicado: (2025)
por: van der Laan, Lars, et al.
Publicado: (2025)
Near-Optimal Non-Parametric Sequential Tests and Confidence Sequences with Possibly Dependent Observations
por: Bibaut, Aurelien, et al.
Publicado: (2022)
por: Bibaut, Aurelien, et al.
Publicado: (2022)
Asymptotic Theory for IV-Based Reinforcement Learning with Potential Endogeneity
por: Li, Jin, et al.
Publicado: (2021)
por: Li, Jin, et al.
Publicado: (2021)
Nonparametric Instrumental Variable Analysis Without Structural Equations: Debiased Inference on Functionals of Inverse Problems with No Solutions
por: Shen, Zikai, et al.
Publicado: (2026)
por: Shen, Zikai, et al.
Publicado: (2026)
Stochastic Optimization Algorithms for Instrumental Variable Regression with Streaming Data
por: Chen, Xuxing, et al.
Publicado: (2024)
por: Chen, Xuxing, et al.
Publicado: (2024)
Reward Transfer from Inverse Reinforcement Learning: A Coupled Minimax Approach
por: Hao, Guang-Yuan, et al.
Publicado: (2026)
por: Hao, Guang-Yuan, et al.
Publicado: (2026)
Nonparametric Instrumental Variable Inference with Many Weak Instruments
por: van der Laan, Lars, et al.
Publicado: (2025)
por: van der Laan, Lars, et al.
Publicado: (2025)
PPI-SVRG: Unifying Prediction-Powered Inference and Variance Reduction for Semi-Supervised Optimization
por: Ao, Ruicheng, et al.
Publicado: (2026)
por: Ao, Ruicheng, et al.
Publicado: (2026)
A Nonparametric Approach with Marginals for Modeling Consumer Choice
por: Ruan, Yanqiu, et al.
Publicado: (2022)
por: Ruan, Yanqiu, et al.
Publicado: (2022)
On Sinkhorn's Algorithm and Choice Modeling
por: Qu, Zhaonan, et al.
Publicado: (2023)
por: Qu, Zhaonan, et al.
Publicado: (2023)
Semiparametric Double Reinforcement Learning with Applications to Long-Term Causal Inference
por: van der Laan, Lars, et al.
Publicado: (2025)
por: van der Laan, Lars, et al.
Publicado: (2025)
Automatic Debiased Machine Learning for Smooth Functionals of Nonparametric M-Estimands
por: van der Laan, Lars, et al.
Publicado: (2025)
por: van der Laan, Lars, et al.
Publicado: (2025)
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
por: Kallus, Nathan
Publicado: (2025)
por: Kallus, Nathan
Publicado: (2025)
Robust and Agnostic Learning of Conditional Distributional Treatment Effects
por: Kallus, Nathan, et al.
Publicado: (2022)
por: Kallus, Nathan, et al.
Publicado: (2022)
Distributionally Robust Instrumental Variables Estimation
por: Qu, Zhaonan, et al.
Publicado: (2024)
por: Qu, Zhaonan, et al.
Publicado: (2024)
What Is the Alignment Tax?
por: Young, Robin
Publicado: (2026)
por: Young, Robin
Publicado: (2026)
SNPL: Simultaneous Policy Learning and Evaluation for Safe Multi-Objective Policy Improvement
por: Cho, Brian, et al.
Publicado: (2025)
por: Cho, Brian, et al.
Publicado: (2025)
Calibeating Prediction-Powered Inference
por: van der Laan, Lars, et al.
Publicado: (2026)
por: van der Laan, Lars, et al.
Publicado: (2026)
Portfolio Optimization with Robust Covariance and Conditional Value-at-Risk Constraints
por: Zhou, Qiqin
Publicado: (2024)
por: Zhou, Qiqin
Publicado: (2024)
Classification and Treatment Learning with Constraints via Composite Heaviside Optimization: a Progressive MIP Method
por: Fang, Yue, et al.
Publicado: (2024)
por: Fang, Yue, et al.
Publicado: (2024)
On the Estimation of Multinomial Logit and Nested Logit Models: A Conic Optimization Approach
por: Pham, Hoang Giang, et al.
Publicado: (2025)
por: Pham, Hoang Giang, et al.
Publicado: (2025)
Joint Location and Cost Planning in Maximum Capture Facility Location under Multiplicative Random Utility Maximization
por: Duong, Ngan Ha, et al.
Publicado: (2022)
por: Duong, Ngan Ha, et al.
Publicado: (2022)
Reducing Marketplace Interference Bias Via Shadow Prices
por: Bright, Ido, et al.
Publicado: (2022)
por: Bright, Ido, et al.
Publicado: (2022)
Load Asymptotics and Dynamic Speed Optimization for the Greenest Path Problem: A Comprehensive Analysis
por: Moradi, Poulad, et al.
Publicado: (2023)
por: Moradi, Poulad, et al.
Publicado: (2023)
The Nonstationary Newsvendor with (and without) Predictions
por: An, Lin, et al.
Publicado: (2023)
por: An, Lin, et al.
Publicado: (2023)
Competitive Facility Location with Market Expansion and Customer-centric Objective
por: Le, Cuong, et al.
Publicado: (2024)
por: Le, Cuong, et al.
Publicado: (2024)
Long-term Causal Inference Under Persistent Confounding via Data Combination
por: Imbens, Guido, et al.
Publicado: (2022)
por: Imbens, Guido, et al.
Publicado: (2022)
IISE PG&E Energy Analytics Challenge 2025: Hourly-Binned Regression Models Beat Transformers in Load Forecasting
por: Roy, Millend, et al.
Publicado: (2025)
por: Roy, Millend, et al.
Publicado: (2025)
Applied Causal Inference Powered by ML and AI
por: Chernozhukov, Victor, et al.
Publicado: (2024)
por: Chernozhukov, Victor, et al.
Publicado: (2024)
Optimising pandemic response through vaccination strategies using neural networks
por: Zhai, Chang, et al.
Publicado: (2025)
por: Zhai, Chang, et al.
Publicado: (2025)
Posterior and Likelihood Sensitivity in Bayesian Distributionally Robust Optimization
por: Gotoh, Jun-ya, et al.
Publicado: (2026)
por: Gotoh, Jun-ya, et al.
Publicado: (2026)
Offline Reinforcement Learning via Inverse Optimization
por: Dimanidis, Ioannis, et al.
Publicado: (2025)
por: Dimanidis, Ioannis, et al.
Publicado: (2025)
Revealed Information
por: Doval, Laura, et al.
Publicado: (2024)
por: Doval, Laura, et al.
Publicado: (2024)
Contrasting the optimal resource allocation to cybersecurity and cyber insurance using prospect theory versus expected utility theory
por: Joshi, Chaitanya, et al.
Publicado: (2024)
por: Joshi, Chaitanya, et al.
Publicado: (2024)
Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration
por: van der Laan, Lars, et al.
Publicado: (2025)
por: van der Laan, Lars, et al.
Publicado: (2025)
Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting
por: van der Laan, Lars, et al.
Publicado: (2025)
por: van der Laan, Lars, et al.
Publicado: (2025)
Offline Hierarchical Reinforcement Learning via Inverse Optimization
por: Schmidt, Carolin, et al.
Publicado: (2024)
por: Schmidt, Carolin, et al.
Publicado: (2024)
Mitigating Covariate Shift in Misspecified Regression with Applications to Reinforcement Learning
por: Amortila, Philip, et al.
Publicado: (2024)
por: Amortila, Philip, et al.
Publicado: (2024)
Ejemplares similares
-
Bellman Calibration for $V$-Learning in Offline Reinforcement Learning
por: van der Laan, Lars, et al.
Publicado: (2025) -
Demistifying Inference after Adaptive Experiments
por: Bibaut, Aurélien, et al.
Publicado: (2024) -
Efficient Inference for Inverse Reinforcement Learning and Dynamic Discrete Choice Models
por: van der Laan, Lars, et al.
Publicado: (2025) -
Near-Optimal Non-Parametric Sequential Tests and Confidence Sequences with Possibly Dependent Observations
por: Bibaut, Aurelien, et al.
Publicado: (2022) -
Asymptotic Theory for IV-Based Reinforcement Learning with Potential Endogeneity
por: Li, Jin, et al.
Publicado: (2021)