Yang, R., Ding, R., Lin, Y., Zhang, H., & Zhang, T. (2024). Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs.
Chicago Style (17th ed.) CitationYang, Rui, Ruomeng Ding, Yong Lin, Huan Zhang, and Tong Zhang. Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs. 2024.
MLA (9th ed.) CitationYang, Rui, et al. Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs. 2024.
Warning: These citations may not always be 100% accurate.