Yan, J., Luo, Y., & Zhang, Y. (2025). RefuteBench 2.0 -- Agentic Benchmark for Dynamic Evaluation of LLM Responses to Refutation Instruction.
Chicago Style (17th ed.) CitationYan, Jianhao, Yun Luo, and Yue Zhang. RefuteBench 2.0 -- Agentic Benchmark for Dynamic Evaluation of LLM Responses to Refutation Instruction. 2025.
MLA (9th ed.) CitationYan, Jianhao, et al. RefuteBench 2.0 -- Agentic Benchmark for Dynamic Evaluation of LLM Responses to Refutation Instruction. 2025.
Warning: These citations may not always be 100% accurate.