Shen, H., Yang, H., Gu, Z., & Han, W. (2026). ScholarGym: Benchmarking Large Language Model Capabilities in the Information-Gathering Stage of Deep Research.
Chicago Style (17th ed.) CitationShen, Hao, Hang Yang, Zhouhong Gu, and Weili Han. ScholarGym: Benchmarking Large Language Model Capabilities in the Information-Gathering Stage of Deep Research. 2026.
MLA (9th ed.) CitationShen, Hao, et al. ScholarGym: Benchmarking Large Language Model Capabilities in the Information-Gathering Stage of Deep Research. 2026.
Warning: These citations may not always be 100% accurate.