Treviño, E., Contant, H., Ngai, J., Neubig, G., & Wang, Z. Z. (2025). Benchmarking Failures in Tool-Augmented Language Models.
Chicago Style (17th ed.) CitationTreviño, Eduardo, Hugo Contant, James Ngai, Graham Neubig, and Zora Zhiruo Wang. Benchmarking Failures in Tool-Augmented Language Models. 2025.
MLA (9th ed.) CitationTreviño, Eduardo, et al. Benchmarking Failures in Tool-Augmented Language Models. 2025.
Warning: These citations may not always be 100% accurate.