Gespeichert in:
| Hauptverfasser: | Shi, C., Zhang, S., Lu, W., Song, R. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2020
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2001.04515 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Infinite-Horizon Reinforcement Learning with Multinomial Logistic Function Approximation
von: Park, Jaehyun, et al.
Veröffentlicht: (2024)
von: Park, Jaehyun, et al.
Veröffentlicht: (2024)
Statistical Inference in Reinforcement Learning: A Selective Survey
von: Shi, Chengchun
Veröffentlicht: (2025)
von: Shi, Chengchun
Veröffentlicht: (2025)
Provably Efficient Infinite-Horizon Average-Reward Reinforcement Learning with Linear Function Approximation
von: Chae, Woojin, et al.
Veröffentlicht: (2024)
von: Chae, Woojin, et al.
Veröffentlicht: (2024)
Optimistically Optimistic Exploration for Provably Efficient Infinite-Horizon Reinforcement and Imitation Learning
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
Policy Zooming: Adaptive Discretization-based Infinite-Horizon Average-Reward Reinforcement Learning
von: Kar, Avik, et al.
Veröffentlicht: (2024)
von: Kar, Avik, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Infinite-Horizon Average-Reward Linear MDPs via Approximation by Discounted-Reward MDPs
von: Hong, Kihyuk, et al.
Veröffentlicht: (2024)
von: Hong, Kihyuk, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning for Infinite Horizon Mean Field Problems in Continuous Spaces
von: Angiuli, Andrea, et al.
Veröffentlicht: (2023)
von: Angiuli, Andrea, et al.
Veröffentlicht: (2023)
Nonparametric Additive Value Functions: Interpretable Reinforcement Learning with an Application to Surgical Recovery
von: Emedom-Nnamdi, Patrick, et al.
Veröffentlicht: (2023)
von: Emedom-Nnamdi, Patrick, et al.
Veröffentlicht: (2023)
Infinite-Horizon Reach-Avoid Zero-Sum Games via Deep Reinforcement Learning
von: Li, Jingqi, et al.
Veröffentlicht: (2022)
von: Li, Jingqi, et al.
Veröffentlicht: (2022)
Learning Infinite-Horizon Average-Reward Linear Mixture MDPs of Bounded Span
von: Chae, Woojin, et al.
Veröffentlicht: (2024)
von: Chae, Woojin, et al.
Veröffentlicht: (2024)
Statistical Inference for Temporal Difference Learning with Linear Function Approximation
von: Wu, Weichen, et al.
Veröffentlicht: (2024)
von: Wu, Weichen, et al.
Veröffentlicht: (2024)
Conformal Prediction Beyond the Horizon: Distribution-Free Inference for Policy Evaluation
von: Gan, Feichen, et al.
Veröffentlicht: (2025)
von: Gan, Feichen, et al.
Veröffentlicht: (2025)
Horizon Generalization in Reinforcement Learning
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
Exchange Policy Optimization Algorithm for Semi-Infinite Safe Reinforcement Learning
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
Reinforcement Learning from Human Feedback: A Statistical Perspective
von: Liu, Pangpang, et al.
Veröffentlicht: (2026)
von: Liu, Pangpang, et al.
Veröffentlicht: (2026)
Active Value Querying to Minimize Additive Error in Subadditive Set Function Learning
von: Černý, Martin, et al.
Veröffentlicht: (2026)
von: Černý, Martin, et al.
Veröffentlicht: (2026)
Statistical Inference for Fuzzy Clustering
von: Wu, Qiuyi, et al.
Veröffentlicht: (2026)
von: Wu, Qiuyi, et al.
Veröffentlicht: (2026)
Learning Set Functions with Implicit Differentiation
von: Özcan, Gözde, et al.
Veröffentlicht: (2024)
von: Özcan, Gözde, et al.
Veröffentlicht: (2024)
UDQL: Bridging The Gap between MSE Loss and The Optimal Value Function in Offline Reinforcement Learning
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
On the Importance of Multistability for Horizon Generalization in Reinforcement Learning
von: Bakija, Asad, et al.
Veröffentlicht: (2026)
von: Bakija, Asad, et al.
Veröffentlicht: (2026)
Inverse Reinforcement Learning with Multiple Planning Horizons
von: Yao, Jiayu, et al.
Veröffentlicht: (2024)
von: Yao, Jiayu, et al.
Veröffentlicht: (2024)
Stochastic Decision Horizons for Constrained Reinforcement Learning
von: Milosevic, Nikola, et al.
Veröffentlicht: (2026)
von: Milosevic, Nikola, et al.
Veröffentlicht: (2026)
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
von: Vamplew, Peter, et al.
Veröffentlicht: (2024)
von: Vamplew, Peter, et al.
Veröffentlicht: (2024)
Statistical Learning from Attribution Sets
von: Applebaum, Lorne, et al.
Veröffentlicht: (2026)
von: Applebaum, Lorne, et al.
Veröffentlicht: (2026)
Factored Value Functions for Graph-Based Multi-Agent Reinforcement Learning
von: Rashwan, Ahmed, et al.
Veröffentlicht: (2026)
von: Rashwan, Ahmed, et al.
Veröffentlicht: (2026)
Fixing Incomplete Value Function Decomposition for Multi-Agent Reinforcement Learning
von: Baisero, Andrea, et al.
Veröffentlicht: (2025)
von: Baisero, Andrea, et al.
Veröffentlicht: (2025)
On the Effective Horizon of Inverse Reinforcement Learning
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
Online Learning with Set-Valued Feedback
von: Raman, Vinod, et al.
Veröffentlicht: (2023)
von: Raman, Vinod, et al.
Veröffentlicht: (2023)
Estimation and Inference in Distributional Reinforcement Learning
von: Zhang, Liangyu, et al.
Veröffentlicht: (2023)
von: Zhang, Liangyu, et al.
Veröffentlicht: (2023)
What Makes Value Learning Efficient in Residual Reinforcement Learning?
von: Ma, Guozheng, et al.
Veröffentlicht: (2026)
von: Ma, Guozheng, et al.
Veröffentlicht: (2026)
A Physics-Informed Learning Framework to Solve the Infinite-Horizon Optimal Control Problem
von: Fotiadis, Filippos, et al.
Veröffentlicht: (2025)
von: Fotiadis, Filippos, et al.
Veröffentlicht: (2025)
Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive Recipe
von: Wu, Xixi, et al.
Veröffentlicht: (2026)
von: Wu, Xixi, et al.
Veröffentlicht: (2026)
VIPO: Value Function Inconsistency Penalized Offline Reinforcement Learning
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
Set-Valued Policy Learning
von: Fuentes-Vicente, Laura, et al.
Veröffentlicht: (2026)
von: Fuentes-Vicente, Laura, et al.
Veröffentlicht: (2026)
Learning U-Statistics with Active Inference
von: Wang, Xiaoning, et al.
Veröffentlicht: (2026)
von: Wang, Xiaoning, et al.
Veröffentlicht: (2026)
Horizon Reduction as Information Loss in Offline Reinforcement Learning
von: Nidadala, Uday Kumar, et al.
Veröffentlicht: (2025)
von: Nidadala, Uday Kumar, et al.
Veröffentlicht: (2025)
Conformal Calibration of Statistical Confidence Sets
von: Cabezas, Luben M. C., et al.
Veröffentlicht: (2024)
von: Cabezas, Luben M. C., et al.
Veröffentlicht: (2024)
Reinforcement Learning with Random Time Horizons
von: Borrell, Enric Ribera, et al.
Veröffentlicht: (2025)
von: Borrell, Enric Ribera, et al.
Veröffentlicht: (2025)
Investigating Lagrangian Neural Networks for Infinite Horizon Planning in Quadrupedal Locomotion
von: Kotecha, Prakrut, et al.
Veröffentlicht: (2025)
von: Kotecha, Prakrut, et al.
Veröffentlicht: (2025)
MARPLE: A Benchmark for Long-Horizon Inference
von: Jin, Emily, et al.
Veröffentlicht: (2024)
von: Jin, Emily, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Infinite-Horizon Reinforcement Learning with Multinomial Logistic Function Approximation
von: Park, Jaehyun, et al.
Veröffentlicht: (2024) -
Statistical Inference in Reinforcement Learning: A Selective Survey
von: Shi, Chengchun
Veröffentlicht: (2025) -
Provably Efficient Infinite-Horizon Average-Reward Reinforcement Learning with Linear Function Approximation
von: Chae, Woojin, et al.
Veröffentlicht: (2024) -
Optimistically Optimistic Exploration for Provably Efficient Infinite-Horizon Reinforcement and Imitation Learning
von: Moulin, Antoine, et al.
Veröffentlicht: (2025) -
Policy Zooming: Adaptive Discretization-based Infinite-Horizon Average-Reward Reinforcement Learning
von: Kar, Avik, et al.
Veröffentlicht: (2024)