Distributional Offline Policy Evaluation with Predictive Error Guarantees
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Runzhe, Uehara, Masatoshi, Sun, Wen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Functional Graphical Models: Structure Enables Offline Data-Driven Optimization
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2024)
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2024)
Making RL with Preference-based Feedback Efficient via Randomization
von: Wu, Runzhe, et al.
Veröffentlicht: (2023)
von: Wu, Runzhe, et al.
Veröffentlicht: (2023)
Distorted Distributional Policy Evaluation for Offline Reinforcement Learning
von: Iwaki, Ryo, et al.
Veröffentlicht: (2026)
von: Iwaki, Ryo, et al.
Veröffentlicht: (2026)
Provable Reward-Agnostic Preference-Based Reinforcement Learning
von: Zhan, Wenhao, et al.
Veröffentlicht: (2023)
von: Zhan, Wenhao, et al.
Veröffentlicht: (2023)
Classification with Reject Option: Distribution-free Error Guarantees via Conformal Prediction
von: Szabadváry, Johan Hallberg, et al.
Veröffentlicht: (2025)
von: Szabadváry, Johan Hallberg, et al.
Veröffentlicht: (2025)
Robust Offline Reinforcement learning with Heavy-Tailed Rewards
von: Zhu, Jin, et al.
Veröffentlicht: (2023)
von: Zhu, Jin, et al.
Veröffentlicht: (2023)
Computationally Efficient RL under Linear Bellman Completeness for Deterministic Dynamics
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
Diffusing States and Matching Scores: A New Framework for Imitation Learning
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
Statistical Guarantees for Offline Domain Randomization
von: Fickinger, Arnaud, et al.
Veröffentlicht: (2025)
von: Fickinger, Arnaud, et al.
Veröffentlicht: (2025)
Efficient Policy Evaluation with Offline Data Informed Behavior Policy Design
von: Liu, Shuze, et al.
Veröffentlicht: (2023)
von: Liu, Shuze, et al.
Veröffentlicht: (2023)
Mildly Constrained Evaluation Policy for Offline Reinforcement Learning
von: Xu, Linjie, et al.
Veröffentlicht: (2023)
von: Xu, Linjie, et al.
Veröffentlicht: (2023)
Algorithmic Guarantees for Distilling Supervised and Offline RL Datasets
von: Gupta, Aaryan, et al.
Veröffentlicht: (2025)
von: Gupta, Aaryan, et al.
Veröffentlicht: (2025)
Coverage-Guaranteed Prediction Sets for Out-of-Distribution Data
von: Zou, Xin, et al.
Veröffentlicht: (2024)
von: Zou, Xin, et al.
Veröffentlicht: (2024)
Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
Evaluation-Time Policy Switching for Offline Reinforcement Learning
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2025)
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2025)
Conformal Prediction Beyond the Horizon: Distribution-Free Inference for Policy Evaluation
von: Gan, Feichen, et al.
Veröffentlicht: (2025)
von: Gan, Feichen, et al.
Veröffentlicht: (2025)
Offline Policy Evaluation for Reinforcement Learning with Adaptively Collected Data
von: Madhow, Sunil, et al.
Veröffentlicht: (2023)
von: Madhow, Sunil, et al.
Veröffentlicht: (2023)
Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees
von: Jiang, Nan, et al.
Veröffentlicht: (2025)
von: Jiang, Nan, et al.
Veröffentlicht: (2025)
Regularized DeepIV with Model Selection
von: Li, Zihao, et al.
Veröffentlicht: (2024)
von: Li, Zihao, et al.
Veröffentlicht: (2024)
Calibrated Inference for the Conditional Average Treatment Effect in the Few-Placebo Regime via Gaussian Processes
von: Uehara, Eichi
Veröffentlicht: (2026)
von: Uehara, Eichi
Veröffentlicht: (2026)
Stop Suppressing the Tail: Causal Inference for Extreme Events
von: Uehara, Eichi
Veröffentlicht: (2026)
von: Uehara, Eichi
Veröffentlicht: (2026)
Efficient Evaluation of LLM Performance with Statistical Guarantees
von: Wu, Skyler, et al.
Veröffentlicht: (2026)
von: Wu, Skyler, et al.
Veröffentlicht: (2026)
Fat-to-Thin Policy Optimization: Offline RL with Sparse Policies
von: Zhu, Lingwei, et al.
Veröffentlicht: (2025)
von: Zhu, Lingwei, et al.
Veröffentlicht: (2025)
Efficient Federated Conformal Prediction with Group-Conditional Guarantees
von: Wen, Haifeng, et al.
Veröffentlicht: (2026)
von: Wen, Haifeng, et al.
Veröffentlicht: (2026)
Enhancing Generative Auto-bidding with Offline Reward Evaluation and Policy Search
von: Mou, Zhiyu, et al.
Veröffentlicht: (2025)
von: Mou, Zhiyu, et al.
Veröffentlicht: (2025)
Frequentist Guarantees of Distributed (Non)-Bayesian Inference
von: Wu, Bohan, et al.
Veröffentlicht: (2023)
von: Wu, Bohan, et al.
Veröffentlicht: (2023)
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling
von: He, Shenghong
Veröffentlicht: (2025)
von: He, Shenghong
Veröffentlicht: (2025)
Policy-regularized Offline Multi-objective Reinforcement Learning
von: Lin, Qian, et al.
Veröffentlicht: (2024)
von: Lin, Qian, et al.
Veröffentlicht: (2024)
Cross-Domain Offline Policy Adaptation via Selective Transition Correction
von: Yan, Mengbei, et al.
Veröffentlicht: (2026)
von: Yan, Mengbei, et al.
Veröffentlicht: (2026)
Online Policy Learning from Offline Preferences
von: Zhang, Guoxi, et al.
Veröffentlicht: (2024)
von: Zhang, Guoxi, et al.
Veröffentlicht: (2024)
Towards Establishing Guaranteed Error for Learned Database Operations
von: Zeighami, Sepanta, et al.
Veröffentlicht: (2024)
von: Zeighami, Sepanta, et al.
Veröffentlicht: (2024)
OPERA: Automatic Offline Policy Evaluation with Re-weighted Aggregates of Multiple Estimators
von: Nie, Allen, et al.
Veröffentlicht: (2024)
von: Nie, Allen, et al.
Veröffentlicht: (2024)
Policy Constraint by Only Support Constraint for Offline Reinforcement Learning
von: Gao, Yunkai, et al.
Veröffentlicht: (2025)
von: Gao, Yunkai, et al.
Veröffentlicht: (2025)
Theoretically Guaranteed Distribution Adaptable Learning
von: Xu, Chao, et al.
Veröffentlicht: (2024)
von: Xu, Chao, et al.
Veröffentlicht: (2024)
Diffusion Policies for Out-of-Distribution Generalization in Offline Reinforcement Learning
von: Ada, Suzan Ece, et al.
Veröffentlicht: (2023)
von: Ada, Suzan Ece, et al.
Veröffentlicht: (2023)
Optimistic Model Rollouts for Pessimistic Offline Policy Optimization
von: Zhai, Yuanzhao, et al.
Veröffentlicht: (2024)
von: Zhai, Yuanzhao, et al.
Veröffentlicht: (2024)
Active Reinforcement Learning Strategies for Offline Policy Improvement
von: Dukkipati, Ambedkar, et al.
Veröffentlicht: (2024)
von: Dukkipati, Ambedkar, et al.
Veröffentlicht: (2024)
Hypercube Policy Regularization Framework for Offline Reinforcement Learning
von: Shen, Yi, et al.
Veröffentlicht: (2024)
von: Shen, Yi, et al.
Veröffentlicht: (2024)
Adaptive Scaling of Policy Constraints for Offline Reinforcement Learning
von: Jing, Tan, et al.
Veröffentlicht: (2025)
von: Jing, Tan, et al.
Veröffentlicht: (2025)
Hierarchical Subspaces of Policies for Continual Offline Reinforcement Learning
von: Kobanda, Anthony, et al.
Veröffentlicht: (2024)
von: Kobanda, Anthony, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Functional Graphical Models: Structure Enables Offline Data-Driven Optimization
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2024) -
Making RL with Preference-based Feedback Efficient via Randomization
von: Wu, Runzhe, et al.
Veröffentlicht: (2023) -
Distorted Distributional Policy Evaluation for Offline Reinforcement Learning
von: Iwaki, Ryo, et al.
Veröffentlicht: (2026) -
Provable Reward-Agnostic Preference-Based Reinforcement Learning
von: Zhan, Wenhao, et al.
Veröffentlicht: (2023) -
Classification with Reject Option: Distribution-free Error Guarantees via Conformal Prediction
von: Szabadváry, Johan Hallberg, et al.
Veröffentlicht: (2025)