Aggarwal, P., & Welleck, S. (2025). L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.
Chicago-Zitierstil (17. Ausg.)Aggarwal, Pranjal, und Sean Welleck. L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning. 2025.
MLA-Zitierstil (9. Ausg.)Aggarwal, Pranjal, und Sean Welleck. L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning. 2025.
Achtung: Diese Zitate sind unter Umständen nicht zu 100% korrekt.