Low-rank Adaptation of Large Language Model Rescoring for Parameter-Efficient Speech Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Yu, Yang, Chao-Han Huck, Kolehmainen, Jari, Shivakumar, Prashanth G., Gu, Yile, Ryu, Sungho, Ren, Roger, Luo, Qi, Gourav, Aditya, Chen, I-Fan, Liu, Yi-Chieh, Dinh, Tuan, Gandhe, Ankur, Filimonov, Denis, Ghosh, Shalini, Stolcke, Andreas, Rastow, Ariya, Bulyko, Ivan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Speech Recognition Rescoring with Large Speech-Text Foundation Models
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2024)
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2024)
Investigating Training Strategies and Model Robustness of Low-Rank Adaptation for Language Modeling in Speech Recognition
von: Yu, Yu, et al.
Veröffentlicht: (2024)
von: Yu, Yu, et al.
Veröffentlicht: (2024)
Multi-Modal Retrieval For Large Language Model Based Speech Recognition
von: Kolehmainen, Jari, et al.
Veröffentlicht: (2024)
von: Kolehmainen, Jari, et al.
Veröffentlicht: (2024)
Towards ASR Robust Spoken Language Understanding Through In-Context Learning With Word Confusion Networks
von: Everson, Kevin, et al.
Veröffentlicht: (2024)
von: Everson, Kevin, et al.
Veröffentlicht: (2024)
Group Relative Policy Optimization for Speech Recognition
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2025)
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2025)
Paralinguistics-Enhanced Large Language Modeling of Spoken Dialogue
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2023)
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2023)
Align-SLM: Textless Spoken Language Models with Reinforcement Learning from AI Feedback
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2024)
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2024)
Streaming Speech-to-Confusion Network Speech Recognition
von: Filimonov, Denis, et al.
Veröffentlicht: (2023)
von: Filimonov, Denis, et al.
Veröffentlicht: (2023)
Generative Speech Recognition Error Correction with Large Language Models and Task-Activating Prompting
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2023)
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2023)
PROCTER: PROnunciation-aware ConTextual adaptER for personalized speech recognition in neural transducers
von: Pandey, Rahul, et al.
Veröffentlicht: (2023)
von: Pandey, Rahul, et al.
Veröffentlicht: (2023)
Incentivizing Consistent, Effective and Scalable Reasoning Capability in Audio LLMs via Reasoning Process Rewards
von: Fan, Jiajun, et al.
Veröffentlicht: (2025)
von: Fan, Jiajun, et al.
Veröffentlicht: (2025)
Cross-utterance ASR Rescoring with Graph-based Label Propagation
von: Tankasala, Srinath, et al.
Veröffentlicht: (2023)
von: Tankasala, Srinath, et al.
Veröffentlicht: (2023)
Holomorphic mappings of maximal rank into projective spaces
von: Huynh, Dinh Tuan
Veröffentlicht: (2023)
von: Huynh, Dinh Tuan
Veröffentlicht: (2023)
Task Oriented Dialogue as a Catalyst for Self-Supervised Automatic Speech Recognition
von: Chan, David M., et al.
Veröffentlicht: (2024)
von: Chan, David M., et al.
Veröffentlicht: (2024)
HDEE: Heterogeneous Domain Expert Ensemble
von: Ersoy, Oğuzhan, et al.
Veröffentlicht: (2025)
von: Ersoy, Oğuzhan, et al.
Veröffentlicht: (2025)
Metal‐Free Organocatalytic Formylation by CO 2 ‐Masked Carbene Functionalized Graphene Oxide Nanosheets
von: Swarbhanu Ghosh, et al.
Veröffentlicht: (2025)
von: Swarbhanu Ghosh, et al.
Veröffentlicht: (2025)
Spoken Conversational Agents with Large Language Models
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2025)
von: Yang, Chao-Han Huck, et al.
Veröffentlicht: (2025)
NoLoCo: No-all-reduce Low Communication Training Method for Large Models
von: Kolehmainen, Jari, et al.
Veröffentlicht: (2025)
von: Kolehmainen, Jari, et al.
Veröffentlicht: (2025)
The inverse conductivity problem with an imperfectly known boundary / Ville Kolehmainen, Matti Lassas, Petri Ola
von: Kolehmainen, Ville
Veröffentlicht: (2005)
von: Kolehmainen, Ville
Veröffentlicht: (2005)
On the structure of the seminal receptacle in cyclopids (Copepoda, Cyclopoida). [Translation from: Informatsionnyi Byulleten Biologiya Vnutrennikh Vod (6) 26-31, 1970.]
von: Filimonov, L. A.
Veröffentlicht: (1974)
von: Filimonov, L. A.
Veröffentlicht: (1974)
Phone Duration Modeling for Speaker Age Estimation in Children
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2021)
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2021)
Reinforcement Learning for Secrecy Optimization in Underwater Energy Harvesting Relay Network
von: Tripathi, Shalini, et al.
Veröffentlicht: (2026)
von: Tripathi, Shalini, et al.
Veröffentlicht: (2026)
La influencia de la esclavitud en la estructura doméstica y la familia en Jamaica, Cuba y Brasil / Verena Stolcke
von: Stolcke, Verena
Veröffentlicht: (1970)
von: Stolcke, Verena
Veröffentlicht: (1970)
La influencia de la esclavitud en la estructura doméstica y la familia en Jamaica, Cuba y Brasil
von: Verena Stolcke
Veröffentlicht: (2003)
von: Verena Stolcke
Veröffentlicht: (2003)
ACTO HOMENAJE A JOHN V. MURRA EN EL INSTITUT D'ESTUDIS CATALANS, BARCELONA, 20 DE FEBRERO DE 2007
von: Verena Stolcke
Veröffentlicht: (2010)
von: Verena Stolcke
Veröffentlicht: (2010)
¿Es el sexo para el género lo que la raza para la etnicidad... y la naturaleza para la sociedad?
von: Verena Stolcke
Veröffentlicht: (2000)
von: Verena Stolcke
Veröffentlicht: (2000)
Los mestizos no nacen sino que se hacen
von: Verena Stolcke
Veröffentlicht: (2009)
von: Verena Stolcke
Veröffentlicht: (2009)
Las nuevas tecnologías reproductivas, la vieja paternidad
von: Verena Stolcke
Veröffentlicht: (2018)
von: Verena Stolcke
Veröffentlicht: (2018)
Minimum mean-squared error estimation with bandit feedback
von: Ghosh, Ayon, et al.
Veröffentlicht: (2022)
von: Ghosh, Ayon, et al.
Veröffentlicht: (2022)
Bridging Structural Causal Inference and Machine Learning The S-DIDML Estimator for Heterogeneous Treatment Effects
von: Yu, Yile, et al.
Veröffentlicht: (2025)
von: Yu, Yile, et al.
Veröffentlicht: (2025)
CRACI: A Cloud-Native Reference Architecture for the Industrial Compute Continuum
von: Dinh-Tuan, Hai
Veröffentlicht: (2025)
von: Dinh-Tuan, Hai
Veröffentlicht: (2025)
Unikernels vs. Containers: A Runtime-Level Performance Comparison for Resource-Constrained Edge Workloads
von: Dinh-Tuan, Hai
Veröffentlicht: (2025)
von: Dinh-Tuan, Hai
Veröffentlicht: (2025)
Degeneracy of holomorphic mappings into or avoiding Fermat type hypersurfaces
von: Huynh, Dinh Tuan
Veröffentlicht: (2024)
von: Huynh, Dinh Tuan
Veröffentlicht: (2024)
SHADOW: Seamless Handoff And Zero-Downtime Orchestrated Workload Migration for Stateful Microservices
von: Dinh-Tuan, Hai
Veröffentlicht: (2026)
von: Dinh-Tuan, Hai
Veröffentlicht: (2026)
Serverless5GC: Private 5G Core Deployment via a Procedure-as-a-Function Architecture
von: Dinh-Tuan, Hai
Veröffentlicht: (2026)
von: Dinh-Tuan, Hai
Veröffentlicht: (2026)
A $p$-adic Second Main Theorem
von: Huynh, Dinh Tuan
Veröffentlicht: (2024)
von: Huynh, Dinh Tuan
Veröffentlicht: (2024)
On the Gauss maps of complete minimal surfaces in $\mathbb{R}^n$
von: Huynh, Dinh Tuan
Veröffentlicht: (2024)
von: Huynh, Dinh Tuan
Veröffentlicht: (2024)
Some variants of the generalized Borel Theorem and applications
von: Huynh, Dinh Tuan
Veröffentlicht: (2024)
von: Huynh, Dinh Tuan
Veröffentlicht: (2024)
Phonetically-Augmented Discriminative Rescoring for Voice Search Error Correction
von: Van Gysel, Christophe, et al.
Veröffentlicht: (2025)
von: Van Gysel, Christophe, et al.
Veröffentlicht: (2025)
Qualitative Evaluation of Language Model Rescoring in Automatic Speech Recognition
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Speech Recognition Rescoring with Large Speech-Text Foundation Models
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2024) -
Investigating Training Strategies and Model Robustness of Low-Rank Adaptation for Language Modeling in Speech Recognition
von: Yu, Yu, et al.
Veröffentlicht: (2024) -
Multi-Modal Retrieval For Large Language Model Based Speech Recognition
von: Kolehmainen, Jari, et al.
Veröffentlicht: (2024) -
Towards ASR Robust Spoken Language Understanding Through In-Context Learning With Word Confusion Networks
von: Everson, Kevin, et al.
Veröffentlicht: (2024) -
Group Relative Policy Optimization for Speech Recognition
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2025)