QMAVIS: Long Video-Audio Understanding using Fusion of Large Multimodal Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Lin, Zixing, Wang, Jiale, Ng, Gee Wah, Mak, Lee Onn, Jeriel, Chan Zhi Yang, Lee, Jun Yang, Li, Yaohao
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!