Lee, D. N., & Kosorok, M. R. (2024). Off-Policy Reinforcement Learning with High Dimensional Reward.
Chicago-Zitierstil (17. Ausg.)Lee, Dong Neuck, und Michael R. Kosorok. Off-Policy Reinforcement Learning with High Dimensional Reward. 2024.
MLA-Zitierstil (9. Ausg.)Lee, Dong Neuck, und Michael R. Kosorok. Off-Policy Reinforcement Learning with High Dimensional Reward. 2024.
Achtung: Diese Zitate sind unter Umständen nicht zu 100% korrekt.