Distributional Offline Policy Evaluation with Predictive Error Guarantees
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Runzhe, Uehara, Masatoshi, Sun, Wen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Functional Graphical Models: Structure Enables Offline Data-Driven Optimization
por: Kuba, Jakub Grudzien, et al.
Publicado: (2024)
por: Kuba, Jakub Grudzien, et al.
Publicado: (2024)
Making RL with Preference-based Feedback Efficient via Randomization
por: Wu, Runzhe, et al.
Publicado: (2023)
por: Wu, Runzhe, et al.
Publicado: (2023)
Distorted Distributional Policy Evaluation for Offline Reinforcement Learning
por: Iwaki, Ryo, et al.
Publicado: (2026)
por: Iwaki, Ryo, et al.
Publicado: (2026)
Provable Reward-Agnostic Preference-Based Reinforcement Learning
por: Zhan, Wenhao, et al.
Publicado: (2023)
por: Zhan, Wenhao, et al.
Publicado: (2023)
Classification with Reject Option: Distribution-free Error Guarantees via Conformal Prediction
por: Szabadváry, Johan Hallberg, et al.
Publicado: (2025)
por: Szabadváry, Johan Hallberg, et al.
Publicado: (2025)
Robust Offline Reinforcement learning with Heavy-Tailed Rewards
por: Zhu, Jin, et al.
Publicado: (2023)
por: Zhu, Jin, et al.
Publicado: (2023)
Computationally Efficient RL under Linear Bellman Completeness for Deterministic Dynamics
por: Wu, Runzhe, et al.
Publicado: (2024)
por: Wu, Runzhe, et al.
Publicado: (2024)
Diffusing States and Matching Scores: A New Framework for Imitation Learning
por: Wu, Runzhe, et al.
Publicado: (2024)
por: Wu, Runzhe, et al.
Publicado: (2024)
Statistical Guarantees for Offline Domain Randomization
por: Fickinger, Arnaud, et al.
Publicado: (2025)
por: Fickinger, Arnaud, et al.
Publicado: (2025)
Efficient Policy Evaluation with Offline Data Informed Behavior Policy Design
por: Liu, Shuze, et al.
Publicado: (2023)
por: Liu, Shuze, et al.
Publicado: (2023)
Mildly Constrained Evaluation Policy for Offline Reinforcement Learning
por: Xu, Linjie, et al.
Publicado: (2023)
por: Xu, Linjie, et al.
Publicado: (2023)
Algorithmic Guarantees for Distilling Supervised and Offline RL Datasets
por: Gupta, Aaryan, et al.
Publicado: (2025)
por: Gupta, Aaryan, et al.
Publicado: (2025)
Coverage-Guaranteed Prediction Sets for Out-of-Distribution Data
por: Zou, Xin, et al.
Publicado: (2024)
por: Zou, Xin, et al.
Publicado: (2024)
Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
por: Uehara, Masatoshi, et al.
Publicado: (2024)
por: Uehara, Masatoshi, et al.
Publicado: (2024)
Evaluation-Time Policy Switching for Offline Reinforcement Learning
por: Neggatu, Natinael Solomon, et al.
Publicado: (2025)
por: Neggatu, Natinael Solomon, et al.
Publicado: (2025)
Conformal Prediction Beyond the Horizon: Distribution-Free Inference for Policy Evaluation
por: Gan, Feichen, et al.
Publicado: (2025)
por: Gan, Feichen, et al.
Publicado: (2025)
Offline Policy Evaluation for Reinforcement Learning with Adaptively Collected Data
por: Madhow, Sunil, et al.
Publicado: (2023)
por: Madhow, Sunil, et al.
Publicado: (2023)
Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees
por: Jiang, Nan, et al.
Publicado: (2025)
por: Jiang, Nan, et al.
Publicado: (2025)
Regularized DeepIV with Model Selection
por: Li, Zihao, et al.
Publicado: (2024)
por: Li, Zihao, et al.
Publicado: (2024)
Calibrated Inference for the Conditional Average Treatment Effect in the Few-Placebo Regime via Gaussian Processes
por: Uehara, Eichi
Publicado: (2026)
por: Uehara, Eichi
Publicado: (2026)
Stop Suppressing the Tail: Causal Inference for Extreme Events
por: Uehara, Eichi
Publicado: (2026)
por: Uehara, Eichi
Publicado: (2026)
Efficient Evaluation of LLM Performance with Statistical Guarantees
por: Wu, Skyler, et al.
Publicado: (2026)
por: Wu, Skyler, et al.
Publicado: (2026)
Fat-to-Thin Policy Optimization: Offline RL with Sparse Policies
por: Zhu, Lingwei, et al.
Publicado: (2025)
por: Zhu, Lingwei, et al.
Publicado: (2025)
Efficient Federated Conformal Prediction with Group-Conditional Guarantees
por: Wen, Haifeng, et al.
Publicado: (2026)
por: Wen, Haifeng, et al.
Publicado: (2026)
Enhancing Generative Auto-bidding with Offline Reward Evaluation and Policy Search
por: Mou, Zhiyu, et al.
Publicado: (2025)
por: Mou, Zhiyu, et al.
Publicado: (2025)
Frequentist Guarantees of Distributed (Non)-Bayesian Inference
por: Wu, Bohan, et al.
Publicado: (2023)
por: Wu, Bohan, et al.
Publicado: (2023)
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling
por: He, Shenghong
Publicado: (2025)
por: He, Shenghong
Publicado: (2025)
Policy-regularized Offline Multi-objective Reinforcement Learning
por: Lin, Qian, et al.
Publicado: (2024)
por: Lin, Qian, et al.
Publicado: (2024)
Cross-Domain Offline Policy Adaptation via Selective Transition Correction
por: Yan, Mengbei, et al.
Publicado: (2026)
por: Yan, Mengbei, et al.
Publicado: (2026)
Online Policy Learning from Offline Preferences
por: Zhang, Guoxi, et al.
Publicado: (2024)
por: Zhang, Guoxi, et al.
Publicado: (2024)
Towards Establishing Guaranteed Error for Learned Database Operations
por: Zeighami, Sepanta, et al.
Publicado: (2024)
por: Zeighami, Sepanta, et al.
Publicado: (2024)
OPERA: Automatic Offline Policy Evaluation with Re-weighted Aggregates of Multiple Estimators
por: Nie, Allen, et al.
Publicado: (2024)
por: Nie, Allen, et al.
Publicado: (2024)
Policy Constraint by Only Support Constraint for Offline Reinforcement Learning
por: Gao, Yunkai, et al.
Publicado: (2025)
por: Gao, Yunkai, et al.
Publicado: (2025)
Theoretically Guaranteed Distribution Adaptable Learning
por: Xu, Chao, et al.
Publicado: (2024)
por: Xu, Chao, et al.
Publicado: (2024)
Diffusion Policies for Out-of-Distribution Generalization in Offline Reinforcement Learning
por: Ada, Suzan Ece, et al.
Publicado: (2023)
por: Ada, Suzan Ece, et al.
Publicado: (2023)
Optimistic Model Rollouts for Pessimistic Offline Policy Optimization
por: Zhai, Yuanzhao, et al.
Publicado: (2024)
por: Zhai, Yuanzhao, et al.
Publicado: (2024)
Active Reinforcement Learning Strategies for Offline Policy Improvement
por: Dukkipati, Ambedkar, et al.
Publicado: (2024)
por: Dukkipati, Ambedkar, et al.
Publicado: (2024)
Hypercube Policy Regularization Framework for Offline Reinforcement Learning
por: Shen, Yi, et al.
Publicado: (2024)
por: Shen, Yi, et al.
Publicado: (2024)
Adaptive Scaling of Policy Constraints for Offline Reinforcement Learning
por: Jing, Tan, et al.
Publicado: (2025)
por: Jing, Tan, et al.
Publicado: (2025)
Hierarchical Subspaces of Policies for Continual Offline Reinforcement Learning
por: Kobanda, Anthony, et al.
Publicado: (2024)
por: Kobanda, Anthony, et al.
Publicado: (2024)
Ejemplares similares
-
Functional Graphical Models: Structure Enables Offline Data-Driven Optimization
por: Kuba, Jakub Grudzien, et al.
Publicado: (2024) -
Making RL with Preference-based Feedback Efficient via Randomization
por: Wu, Runzhe, et al.
Publicado: (2023) -
Distorted Distributional Policy Evaluation for Offline Reinforcement Learning
por: Iwaki, Ryo, et al.
Publicado: (2026) -
Provable Reward-Agnostic Preference-Based Reinforcement Learning
por: Zhan, Wenhao, et al.
Publicado: (2023) -
Classification with Reject Option: Distribution-free Error Guarantees via Conformal Prediction
por: Szabadváry, Johan Hallberg, et al.
Publicado: (2025)