Wang, W., Duan, C., Peng, Z., Liu, Y., & Zhou, B. (2025). Embodied Scene Understanding for Vision Language Models via MetaVQA.
Citazione stile Chigago Style (17a edizione)Wang, Weizhen, Chenda Duan, Zhenghao Peng, Yuxin Liu, e Bolei Zhou. Embodied Scene Understanding for Vision Language Models via MetaVQA. 2025.
Citatione MLA (9a ed.)Wang, Weizhen, et al. Embodied Scene Understanding for Vision Language Models via MetaVQA. 2025.
Attenzione: Queste citazioni potrebbero non essere precise al 100%.