Crab$^{+}$: A Scalable and Unified Audio-Visual Scene Understanding Model with Explicit Cooperation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Cai, Dongnuan, Du, Henghui, Zhou, Chang, Chen, Xi, Guo, Dan, Zhang, Hongyuan, Li, Xuelong, Hu, Di
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!