Wu, X. (2024). From Reward Shaping to Q-Shaping: Achieving Unbiased Learning with LLM-Guided Knowledge.
Chicago Style (17th ed.) CitationWu, Xiefeng. From Reward Shaping to Q-Shaping: Achieving Unbiased Learning with LLM-Guided Knowledge. 2024.
MLA (9th ed.) CitationWu, Xiefeng. From Reward Shaping to Q-Shaping: Achieving Unbiased Learning with LLM-Guided Knowledge. 2024.
Warning: These citations may not always be 100% accurate.