Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Luo, Gen, Yang, Xue, Dou, Wenhan, Wang, Zhaokai, Liu, Jiawen, Dai, Jifeng, Qiao, Yu, Zhu, Xizhou
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!