KidVis: Do Multimodal Large Language Models Possess the Visual Perceptual Capabilities of a 6-Year-Old?

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wang, Xianfeng, Zhang, Kaiwei, Jia, Qi, Chen, Zijian, Zhai, Guangtao, Min, Xiongkuo
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!