Wu, R., & Sun, W. (2023). Making RL with Preference-based Feedback Efficient via Randomization.
Chicago Style (17th ed.) CitationWu, Runzhe, and Wen Sun. Making RL with Preference-based Feedback Efficient via Randomization. 2023.
MLA (9th ed.) CitationWu, Runzhe, and Wen Sun. Making RL with Preference-based Feedback Efficient via Randomization. 2023.
Warning: These citations may not always be 100% accurate.