VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Liu, Qing'an, Feng, Juntong, Wang, Yuhao, Han, Xinzhe, Cheng, Yujie, Zhu, Yue, Diao, Haiwen, Zhuge, Yunzhi, Lu, Huchuan
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!