Cai, M., Liu, H., Park, D., Mustikovela, S. K., Meyer, G. P., Chai, Y., & Lee, Y. J. (2023). ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts.
Cita Chicago Style (17a ed.)Cai, Mu, Haotian Liu, Dennis Park, Siva Karthik Mustikovela, Gregory P. Meyer, Yuning Chai, y Yong Jae Lee. ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts. 2023.
Cita MLA (9a ed.)Cai, Mu, et al. ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts. 2023.
Precaución: Estas citas no son 100% exactas.