Zhu, Y., Steck, H., Liang, D., He, Y., Ostuni, V., Li, J., & Kallus, N. (2025). Rank-GRPO: Training LLM-based Conversational Recommender Systems with Reinforcement Learning.
Chicago Style (17th ed.) CitationZhu, Yaochen, Harald Steck, Dawen Liang, Yinhan He, Vito Ostuni, Jundong Li, and Nathan Kallus. Rank-GRPO: Training LLM-based Conversational Recommender Systems with Reinforcement Learning. 2025.
MLA (9th ed.) CitationZhu, Yaochen, et al. Rank-GRPO: Training LLM-based Conversational Recommender Systems with Reinforcement Learning. 2025.
Warning: These citations may not always be 100% accurate.