Structural Estimation of Markov Decision Processes in High-Dimensional State Space with Finite-Time Guarantees
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zeng, Siliang, Hong, Mingyi, Garcia, Alfredo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Econometric vs. Causal Structure-Learning for Time-Series Policy Decisions: Evidence from the UK COVID-19 Policies
von: Petrungaro, Bruno, et al.
Veröffentlicht: (2026)
von: Petrungaro, Bruno, et al.
Veröffentlicht: (2026)
Statistical Tests for Replacing Human Decision Makers with Algorithms
von: Feng, Kai, et al.
Veröffentlicht: (2023)
von: Feng, Kai, et al.
Veröffentlicht: (2023)
Management Decisions in Manufacturing using Causal Machine Learning -- To Rework, or not to Rework?
von: Schwarz, Philipp, et al.
Veröffentlicht: (2024)
von: Schwarz, Philipp, et al.
Veröffentlicht: (2024)
Unveiling the Potential of Robustness in Selecting Conditional Average Treatment Effect Estimators
von: Huang, Yiyan, et al.
Veröffentlicht: (2024)
von: Huang, Yiyan, et al.
Veröffentlicht: (2024)
When Demonstrations Meet Generative World Models: A Maximum Likelihood Framework for Offline Inverse Reinforcement Learning
von: Zeng, Siliang, et al.
Veröffentlicht: (2023)
von: Zeng, Siliang, et al.
Veröffentlicht: (2023)
DeXposure-FM: A Time-series, Graph Foundation Model for Credit Exposures and Stability on Decentralized Financial Networks
von: Shu, Aijie, et al.
Veröffentlicht: (2026)
von: Shu, Aijie, et al.
Veröffentlicht: (2026)
Multi-Band Variable-Lag Granger Causality: A Unified Framework for Causal Time Series Inference across Frequencies
von: Sookkongwaree, Chakattrai, et al.
Veröffentlicht: (2025)
von: Sookkongwaree, Chakattrai, et al.
Veröffentlicht: (2025)
LLM Personas as a Substitute for Field Experiments in Method Benchmarking
von: Kang, Enoch Hyunwook
Veröffentlicht: (2025)
von: Kang, Enoch Hyunwook
Veröffentlicht: (2025)
Foundation Priors
von: Misra, Sanjog
Veröffentlicht: (2025)
von: Misra, Sanjog
Veröffentlicht: (2025)
Causal EpiNets: Precision-corrected Bounds on Individual Treatment Effects using Epistemic Neural Networks
von: Patil, Gandharv, et al.
Veröffentlicht: (2026)
von: Patil, Gandharv, et al.
Veröffentlicht: (2026)
Simulation-Based Benchmarking of Reinforcement Learning Agents for Personalized Retail Promotions
von: Xia, Yu, et al.
Veröffentlicht: (2024)
von: Xia, Yu, et al.
Veröffentlicht: (2024)
Auditing and Fixing Economic Validity in Tabular Foundation Models for Discrete Choice
von: Wang, Yingshuo, et al.
Veröffentlicht: (2026)
von: Wang, Yingshuo, et al.
Veröffentlicht: (2026)
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
von: Kallus, Nathan
Veröffentlicht: (2025)
von: Kallus, Nathan
Veröffentlicht: (2025)
GDP nowcasting with artificial neural networks: How much does long-term memory matter?
von: Németh, Kristóf, et al.
Veröffentlicht: (2023)
von: Németh, Kristóf, et al.
Veröffentlicht: (2023)
Generating density nowcasts for U.S. GDP growth with deep learning: Bayes by Backprop and Monte Carlo dropout
von: Németh, Kristóf, et al.
Veröffentlicht: (2024)
von: Németh, Kristóf, et al.
Veröffentlicht: (2024)
A Hybrid Framework for Reinsurance Optimization: Integrating Generative Models and Reinforcement Learning
von: Dong, Stella C.
Veröffentlicht: (2025)
von: Dong, Stella C.
Veröffentlicht: (2025)
An Empirical Risk Minimization Approach for Offline Inverse RL and Dynamic Discrete Choice Model
von: Kang, Enoch H., et al.
Veröffentlicht: (2025)
von: Kang, Enoch H., et al.
Veröffentlicht: (2025)
How Well Do LLMs Predict Human Behavior? A Measure of their Pretrained Knowledge
von: Gao, Wayne, et al.
Veröffentlicht: (2026)
von: Gao, Wayne, et al.
Veröffentlicht: (2026)
Model-Estimation-Free, Dense, and High Dimensional Consistent Precision Matrix Estimators
von: Stojnic, Mehmet Caner Agostino Capponi Mihailo
Veröffentlicht: (2025)
von: Stojnic, Mehmet Caner Agostino Capponi Mihailo
Veröffentlicht: (2025)
$ρ$-GNF: A Copula-based Sensitivity Analysis to Unobserved Confounding Using Normalizing Flows
von: Balgi, Sourabh, et al.
Veröffentlicht: (2022)
von: Balgi, Sourabh, et al.
Veröffentlicht: (2022)
Towards Generalizing Inferences from Trials to Target Populations
von: Huang, Melody Y, et al.
Veröffentlicht: (2024)
von: Huang, Melody Y, et al.
Veröffentlicht: (2024)
Adaptive Experimental Design for Policy Learning
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
von: Kato, Masahiro, et al.
Veröffentlicht: (2024)
Learning Causal Representations from General Environments: Identifiability and Intrinsic Ambiguity
von: Jin, Jikai, et al.
Veröffentlicht: (2023)
von: Jin, Jikai, et al.
Veröffentlicht: (2023)
Selective Reviews of Bandit Problems in AI via a Statistical View
von: Zhou, Pengjie, et al.
Veröffentlicht: (2024)
von: Zhou, Pengjie, et al.
Veröffentlicht: (2024)
Global Ease of Living Index: a machine learning framework for longitudinal analysis of major economies
von: Panat, Tanay, et al.
Veröffentlicht: (2025)
von: Panat, Tanay, et al.
Veröffentlicht: (2025)
Learning from Double Positive and Unlabeled Data for Potential-Customer Identification
von: Kato, Masahiro, et al.
Veröffentlicht: (2025)
von: Kato, Masahiro, et al.
Veröffentlicht: (2025)
From What Ifs to Insights: Counterfactuals in Causal Inference vs. Explainable AI
von: Shmueli, Galit, et al.
Veröffentlicht: (2025)
von: Shmueli, Galit, et al.
Veröffentlicht: (2025)
A Double Machine Learning Approach to Combining Experimental and Observational Data
von: Parikh, Harsh, et al.
Veröffentlicht: (2023)
von: Parikh, Harsh, et al.
Veröffentlicht: (2023)
Generalized Neyman Allocation for Locally Minimax Optimal Best-Arm Identification
von: Kato, Masahiro
Veröffentlicht: (2024)
von: Kato, Masahiro
Veröffentlicht: (2024)
Enhancing Preference-based Linear Bandits via Human Response Time
von: Li, Shen, et al.
Veröffentlicht: (2024)
von: Li, Shen, et al.
Veröffentlicht: (2024)
On LASSO for High Dimensional Predictive Regression
von: Mei, Ziwei, et al.
Veröffentlicht: (2022)
von: Mei, Ziwei, et al.
Veröffentlicht: (2022)
High-Dimensional Tail Index Regression
von: Sasaki, Yuya, et al.
Veröffentlicht: (2024)
von: Sasaki, Yuya, et al.
Veröffentlicht: (2024)
Causality Elicitation from Large Language Models
von: Kameyama, Takashi, et al.
Veröffentlicht: (2026)
von: Kameyama, Takashi, et al.
Veröffentlicht: (2026)
Transformers Handle Endogeneity in In-Context Linear Regression
von: Liang, Haodong, et al.
Veröffentlicht: (2024)
von: Liang, Haodong, et al.
Veröffentlicht: (2024)
Differentially Private Two-Stage Gradient Descent for Instrumental Variable Regression
von: Liang, Haodong, et al.
Veröffentlicht: (2025)
von: Liang, Haodong, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Monetary Policy Under Macroeconomic Uncertainty: Analyzing Tabular and Function Approximation Methods
von: Wang, Tony, et al.
Veröffentlicht: (2025)
von: Wang, Tony, et al.
Veröffentlicht: (2025)
What Is the Alignment Tax?
von: Young, Robin
Veröffentlicht: (2026)
von: Young, Robin
Veröffentlicht: (2026)
Estimating the Value of Evidence-Based Decision Making
von: Abadie, Alberto, et al.
Veröffentlicht: (2023)
von: Abadie, Alberto, et al.
Veröffentlicht: (2023)
Bias-Reduced Estimation of Finite Mixtures: An Application to Latent Group Structures in Panel Data
von: Langevin, Raphaël
Veröffentlicht: (2026)
von: Langevin, Raphaël
Veröffentlicht: (2026)
Testing Effect Homogeneity and Confounding in High-Dimensional Experimental and Observational Studies
von: Armendariz, Ana, et al.
Veröffentlicht: (2026)
von: Armendariz, Ana, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Econometric vs. Causal Structure-Learning for Time-Series Policy Decisions: Evidence from the UK COVID-19 Policies
von: Petrungaro, Bruno, et al.
Veröffentlicht: (2026) -
Statistical Tests for Replacing Human Decision Makers with Algorithms
von: Feng, Kai, et al.
Veröffentlicht: (2023) -
Management Decisions in Manufacturing using Causal Machine Learning -- To Rework, or not to Rework?
von: Schwarz, Philipp, et al.
Veröffentlicht: (2024) -
Unveiling the Potential of Robustness in Selecting Conditional Average Treatment Effect Estimators
von: Huang, Yiyan, et al.
Veröffentlicht: (2024) -
When Demonstrations Meet Generative World Models: A Maximum Likelihood Framework for Offline Inverse Reinforcement Learning
von: Zeng, Siliang, et al.
Veröffentlicht: (2023)