Style de citation APA (7e éd.)

Kikkawa, N., & Ohno, H. (2024). Unified theory of upper confidence bound policies for bandit problems targeting total reward, maximal reward, and more.

Style de citation Chicago (17e éd.)

Kikkawa, Nobuaki, et Hiroshi Ohno. Unified Theory of Upper Confidence Bound Policies for Bandit Problems Targeting Total Reward, Maximal Reward, and More. 2024.

Style de citation MLA (9e éd.)

Kikkawa, Nobuaki, et Hiroshi Ohno. Unified Theory of Upper Confidence Bound Policies for Bandit Problems Targeting Total Reward, Maximal Reward, and More. 2024.

Attention : ces citations peuvent ne pas être correctes à 100%.