VTCBench: Can Vision-Language Models Understand Long Context with Vision-Text Compression?

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhao, Hongbo, Wang, Meng, Zhu, Fei, Liu, Wenzhuo, Ni, Bolin, Zeng, Fanhu, Meng, Gaofeng, Zhang, Zhaoxiang
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!