Bai, H., & Ma, Y. (2024). Improving Neuron-level Interpretability with White-box Language Models.
Chicago Style (17th ed.) CitationBai, Hao, and Yi Ma. Improving Neuron-level Interpretability with White-box Language Models. 2024.
MLA (9th ed.) CitationBai, Hao, and Yi Ma. Improving Neuron-level Interpretability with White-box Language Models. 2024.
Warning: These citations may not always be 100% accurate.