Mukherjee, D., & Kalyanakrishnan, S. (2025). Howard's Policy Iteration is Subexponential for Deterministic Markov Decision Problems with Rewards of Fixed Bit-size and Arbitrary Discount Factor.
Chicago-Zitierstil (17. Ausg.)Mukherjee, Dibyangshu, und Shivaram Kalyanakrishnan. Howard's Policy Iteration Is Subexponential for Deterministic Markov Decision Problems with Rewards of Fixed Bit-size and Arbitrary Discount Factor. 2025.
MLA-Zitierstil (9. Ausg.)Mukherjee, Dibyangshu, und Shivaram Kalyanakrishnan. Howard's Policy Iteration Is Subexponential for Deterministic Markov Decision Problems with Rewards of Fixed Bit-size and Arbitrary Discount Factor. 2025.
Achtung: Diese Zitate sind unter Umständen nicht zu 100% korrekt.