Fewer Tokens and Fewer Videos: Extending Video Understanding Abilities in Large Vision-Language Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Chen, Shimin, Yuan, Yitian, Chen, Shaoxiang, Jie, Zequn, Ma, Lin
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!