Guo, J., Ho, C. W., & Singh, S. S. (2025). Bayesian learning of the optimal action-value function in a Markov decision process.
Chicago-Zitierstil (17. Ausg.)Guo, Jiaqi, Chon Wai Ho, und Sumeetpal S. Singh. Bayesian Learning of the Optimal Action-value Function in a Markov Decision Process. 2025.
MLA-Zitierstil (9. Ausg.)Guo, Jiaqi, et al. Bayesian Learning of the Optimal Action-value Function in a Markov Decision Process. 2025.
Achtung: Diese Zitate sind unter Umständen nicht zu 100% korrekt.