Monea, G., Bosselut, A., Brantley, K., & Artzi, Y. (2024). LLMs Are In-Context Bandit Reinforcement Learners.
Chicago-Zitierstil (17. Ausg.)Monea, Giovanni, Antoine Bosselut, Kianté Brantley, und Yoav Artzi. LLMs Are In-Context Bandit Reinforcement Learners. 2024.
MLA-Zitierstil (9. Ausg.)Monea, Giovanni, et al. LLMs Are In-Context Bandit Reinforcement Learners. 2024.
Achtung: Diese Zitate sind unter Umständen nicht zu 100% korrekt.