Ye, M., Suzuki, J., Inaba, T., & Kuribayashi, T. (2025). Transformer Key-Value Memories Are Nearly as Interpretable as Sparse Autoencoders.
Chicago Style (17th ed.) CitationYe, Mengyu, Jun Suzuki, Tatsuro Inaba, and Tatsuki Kuribayashi. Transformer Key-Value Memories Are Nearly as Interpretable as Sparse Autoencoders. 2025.
MLA (9th ed.) CitationYe, Mengyu, et al. Transformer Key-Value Memories Are Nearly as Interpretable as Sparse Autoencoders. 2025.
Warning: These citations may not always be 100% accurate.