Bridging the Modality Gap: Softly Discretizing Audio Representation for LLM-based Automatic Speech Recognition

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Yang, Mu, Chen, Szu-Jui, Xie, Jiamin, Hansen, John
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!