Yuan, L., Wang, J., Sun, H., Zhang, Y., & Lin, Y. (2025). Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding.
Chicago Style (17th ed.) CitationYuan, Liping, Jiawei Wang, Haomiao Sun, Yuchen Zhang, and Yuan Lin. Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding. 2025.
MLA (9th ed.) CitationYuan, Liping, et al. Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding. 2025.
Warning: These citations may not always be 100% accurate.