APA (7th ed.) Citation

Fan, M., Han, W., Wang, D., Chen, C., Zhang, Z., & Zhou, J. (2026). Evaluating Interactive Reasoning in Large Language Models: A Hierarchical Benchmark with Executable Games.

Chicago Style (17th ed.) Citation

Fan, Mingyuan, Weiguang Han, Daixin Wang, Cen Chen, Zhiqiang Zhang, and Jun Zhou. Evaluating Interactive Reasoning in Large Language Models: A Hierarchical Benchmark with Executable Games. 2026.

MLA (9th ed.) Citation

Fan, Mingyuan, et al. Evaluating Interactive Reasoning in Large Language Models: A Hierarchical Benchmark with Executable Games. 2026.

Warning: These citations may not always be 100% accurate.