Yu, Y., Liao, M., Zhang, J., & Wu, J. (2024). TextHawk2: A Large Vision-Language Model Excels in Bilingual OCR and Grounding with 16x Fewer Tokens.
Style de citation Chicago (17e éd.)Yu, Ya-Qi, Minghui Liao, Jiwen Zhang, et Jihao Wu. TextHawk2: A Large Vision-Language Model Excels in Bilingual OCR and Grounding with 16x Fewer Tokens. 2024.
Style de citation MLA (9e éd.)Yu, Ya-Qi, et al. TextHawk2: A Large Vision-Language Model Excels in Bilingual OCR and Grounding with 16x Fewer Tokens. 2024.
Attention : ces citations peuvent ne pas être correctes à 100%.