Reading $\neq$ Seeing: Diagnosing and Closing the Typography Gap in Vision-Language Models

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Zhou, Heng, Yu, Ao, Kang, Li, Fan, Yuchen, Fan, Yutao, Song, Xiufeng, Geng, Hejia, Qin, Yiran
Format: Preprint
Veröffentlicht: 2026
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!