Du, H., Dong, Y., & Ning, X. (2025). Latent Thinking Optimization: Your Latent Reasoning Language Model Secretly Encodes Reward Signals in Its Latent Thoughts.
Style de citation Chicago (17e éd.)Du, Hanwen, Yuxin Dong, et Xia Ning. Latent Thinking Optimization: Your Latent Reasoning Language Model Secretly Encodes Reward Signals in Its Latent Thoughts. 2025.
Style de citation MLA (9e éd.)Du, Hanwen, et al. Latent Thinking Optimization: Your Latent Reasoning Language Model Secretly Encodes Reward Signals in Its Latent Thoughts. 2025.
Attention : ces citations peuvent ne pas être correctes à 100%.