APA (7th ed.) Citation

Ye, C., Xiong, W., Zhang, Y., Dong, H., Jiang, N., & Zhang, T. (2024). Online Iterative Reinforcement Learning from Human Feedback with General Preference Model.

Chicago Style (17th ed.) Citation

Ye, Chenlu, Wei Xiong, Yuheng Zhang, Hanze Dong, Nan Jiang, and Tong Zhang. Online Iterative Reinforcement Learning from Human Feedback with General Preference Model. 2024.

MLA (9th ed.) Citation

Ye, Chenlu, et al. Online Iterative Reinforcement Learning from Human Feedback with General Preference Model. 2024.

Warning: These citations may not always be 100% accurate.