Song, J., Heo, B., Gu, G., Choo, J., Han, D., & Yun, S. (2026). Learning to See What You Need: Gaze Attention for Multimodal Large Language Models.
Chicago Style (17th ed.) CitationSong, Junha, Byeongho Heo, Geonmo Gu, Jaegul Choo, Dongyoon Han, and Sangdoo Yun. Learning to See What You Need: Gaze Attention for Multimodal Large Language Models. 2026.
MLA (9th ed.) CitationSong, Junha, et al. Learning to See What You Need: Gaze Attention for Multimodal Large Language Models. 2026.
Warning: These citations may not always be 100% accurate.