Enhancing Multimodal Understanding with CLIP-Based Image-to-Text Transformation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Che, Chang, Lin, Qunwei, Zhao, Xinyu, Huang, Jiaxin, Yu, Liqiang
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!

Similar Items