Gu, Z., Chen, X., Shi, X., Wang, T., Zheng, S., Li, T., . . . Xiao, Y. (2025). GAPO: Learning Preferential Prompt through Generative Adversarial Policy Optimization.
Chicago Style (17th ed.) CitationGu, Zhouhong, Xingzhou Chen, Xiaoran Shi, Tao Wang, Suhang Zheng, Tianyu Li, Hongwei Feng, and Yanghua Xiao. GAPO: Learning Preferential Prompt Through Generative Adversarial Policy Optimization. 2025.
MLA (9th ed.) CitationGu, Zhouhong, et al. GAPO: Learning Preferential Prompt Through Generative Adversarial Policy Optimization. 2025.
Warning: These citations may not always be 100% accurate.