Kang, S., Kim, J., Kim, J., & Hwang, S. J. (2025). Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding.
Chicago Style (17th ed.) CitationKang, Seil, Jinyeong Kim, Junhyeok Kim, and Seong Jae Hwang. Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding. 2025.
MLA (9th ed.) CitationKang, Seil, et al. Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding. 2025.
Warning: These citations may not always be 100% accurate.