Integrating Self-supervised Speech Model with Pseudo Word-level Targets from Visually-grounded Speech Model

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Fang, Hung-Chieh, Ye, Nai-Xuan, Shih, Yi-Jen, Peng, Puyuan, Wang, Hsuan-Fu, Berry, Layne, Lee, Hung-yi, Harwath, David
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!