Detect2Interact: Localizing Object Key Field in Visual Question Answering (VQA) with LLMs

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wang, Jialou, Zhu, Manli, Li, Yulei, Li, Honglei, Yang, Longzhi, Woo, Wai Lok
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!