SpeechCLIP+: Self-supervised multi-task representation learning for speech via CLIP and speech-image data

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wang, Hsuan-Fu, Shih, Yi-Jen, Chang, Heng-Jui, Berry, Layne, Peng, Puyuan, Lee, Hung-yi, Wang, Hsin-Min, Harwath, David
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!