Ramos, M. M., Alves, D. M., & Martins, A. F. T. (2026). Combining On-Policy Optimization and Distillation for Long-Context Reasoning in Large Language Models.
Chicago Style (17th ed.) CitationRamos, Miguel Moura, Duarte M. Alves, and André F. T. Martins. Combining On-Policy Optimization and Distillation for Long-Context Reasoning in Large Language Models. 2026.
MLA (9th ed.) CitationRamos, Miguel Moura, et al. Combining On-Policy Optimization and Distillation for Long-Context Reasoning in Large Language Models. 2026.
Warning: These citations may not always be 100% accurate.