Wang, S., Guo, W., Chen, Z., Hu, X., & Xiong, H. (2026). Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding.
Cita Chicago Style (17a ed.)Wang, Shaoguang, Weiyu Guo, Ziyang Chen, Xuming Hu, y Hui Xiong. Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding. 2026.
Cita MLA (9a ed.)Wang, Shaoguang, et al. Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding. 2026.
Precaución: Estas citas no son 100% exactas.