Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Huang, Shaofei, Ling, Rui, Li, Hongyu, Hui, Tianrui, Tang, Zongheng, Wei, Xiaoming, Han, Jizhong, Liu, Si
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!