VATSA: Video, Audio, Text, Sensory, Action - A Unified Five-Modality Architecture for Human-Level Perception and Action

Fuente: Zenodo
Saved in:
Bibliographic Details
Main Author: K V (Kengeri Vijaya Kumar), Vinay Kumar
Format: Recurso digital
Language:English
Published: Zenodo 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!