PROCTER: PROnunciation-aware ConTextual adaptER for personalized speech recognition in neural transducers
Fuente:
arXiv
Saved in:
| Main Authors: | Pandey, Rahul, Ren, Roger, Luo, Qi, Liu, Jing, Rastrow, Ariya, Gandhe, Ankur, Filimonov, Denis, Strimel, Grant, Stolcke, Andreas, Bulyko, Ivan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Streaming Speech-to-Confusion Network Speech Recognition
by: Filimonov, Denis, et al.
Published: (2023)
by: Filimonov, Denis, et al.
Published: (2023)
Multi-Modal Retrieval For Large Language Model Based Speech Recognition
by: Kolehmainen, Jari, et al.
Published: (2024)
by: Kolehmainen, Jari, et al.
Published: (2024)
Speech Recognition Rescoring with Large Speech-Text Foundation Models
by: Shivakumar, Prashanth Gurunath, et al.
Published: (2024)
by: Shivakumar, Prashanth Gurunath, et al.
Published: (2024)
Towards ASR Robust Spoken Language Understanding Through In-Context Learning With Word Confusion Networks
by: Everson, Kevin, et al.
Published: (2024)
by: Everson, Kevin, et al.
Published: (2024)
Investigating Training Strategies and Model Robustness of Low-Rank Adaptation for Language Modeling in Speech Recognition
by: Yu, Yu, et al.
Published: (2024)
by: Yu, Yu, et al.
Published: (2024)
Group Relative Policy Optimization for Speech Recognition
by: Shivakumar, Prashanth Gurunath, et al.
Published: (2025)
by: Shivakumar, Prashanth Gurunath, et al.
Published: (2025)
Self-consistent context aware conformer transducer for speech recognition
by: Kolokolov, Konstantin, et al.
Published: (2024)
by: Kolokolov, Konstantin, et al.
Published: (2024)
Dial-MAE: ConTextual Masked Auto-Encoder for Retrieval-based Dialogue Systems
by: Su, Zhenpeng, et al.
Published: (2023)
by: Su, Zhenpeng, et al.
Published: (2023)
ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models
by: Wadhawan, Rohan, et al.
Published: (2024)
by: Wadhawan, Rohan, et al.
Published: (2024)
ConTextual: Improving Clinical Text Summarization in LLMs with Context-preserving Token Filtering and Knowledge Graphs
by: Piya, Fahmida Liza, et al.
Published: (2025)
by: Piya, Fahmida Liza, et al.
Published: (2025)
Incentivizing Consistent, Effective and Scalable Reasoning Capability in Audio LLMs via Reasoning Process Rewards
by: Fan, Jiajun, et al.
Published: (2025)
by: Fan, Jiajun, et al.
Published: (2025)
Low-rank Adaptation of Large Language Model Rescoring for Parameter-Efficient Speech Recognition
by: Yu, Yu, et al.
Published: (2023)
by: Yu, Yu, et al.
Published: (2023)
Paralinguistics-Enhanced Large Language Modeling of Spoken Dialogue
by: Lin, Guan-Ting, et al.
Published: (2023)
by: Lin, Guan-Ting, et al.
Published: (2023)
Prominence-aware automatic speech recognition for conversational speech
by: Linke, Julian, et al.
Published: (2025)
by: Linke, Julian, et al.
Published: (2025)
Align-SLM: Textless Spoken Language Models with Reinforcement Learning from AI Feedback
by: Lin, Guan-Ting, et al.
Published: (2024)
by: Lin, Guan-Ting, et al.
Published: (2024)
Task Oriented Dialogue as a Catalyst for Self-Supervised Automatic Speech Recognition
by: Chan, David M., et al.
Published: (2024)
by: Chan, David M., et al.
Published: (2024)
Universal Semantic Disentangled Privacy-preserving Speech Representation Learning
by: Vecino, Biel Tura, et al.
Published: (2025)
by: Vecino, Biel Tura, et al.
Published: (2025)
An Efficient Self-Learning Framework For Interactive Spoken Dialog Systems
by: Tulsiani, Hitesh, et al.
Published: (2024)
by: Tulsiani, Hitesh, et al.
Published: (2024)
Keyword spotting using convolutional neural network for speech recognition in Hindi
by: Bharti, Saru, et al.
Published: (2026)
by: Bharti, Saru, et al.
Published: (2026)
Introduction to speech recognition
by: Dauphin, Gabriel
Published: (2024)
by: Dauphin, Gabriel
Published: (2024)
Teaching the Teachers: Boosting unsupervised domain adaptation in speech recognition by ensemble update
by: Ahmad, Rehan, et al.
Published: (2026)
by: Ahmad, Rehan, et al.
Published: (2026)
CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing
by: Lu, Yen-Ju, et al.
Published: (2024)
by: Lu, Yen-Ju, et al.
Published: (2024)
Enhancing Conversational TTS with Cascaded Prompting and ICL-Based Online Reinforcement Learning
by: Ouyang, Zhicheng, et al.
Published: (2026)
by: Ouyang, Zhicheng, et al.
Published: (2026)
Generative Speech Recognition Error Correction with Large Language Models and Task-Activating Prompting
by: Yang, Chao-Han Huck, et al.
Published: (2023)
by: Yang, Chao-Han Huck, et al.
Published: (2023)
Improving child speech recognition with augmented child-like speech
by: Zhang, Yuanyuan, et al.
Published: (2024)
by: Zhang, Yuanyuan, et al.
Published: (2024)
Acoustic and linguistic effects in synthesized speech augmentation for speech recognition
by: Yohan Lim, et al.
Published: (2025)
by: Yohan Lim, et al.
Published: (2025)
Automatic speech recognition for the Nepali language using CNN, bidirectional LSTM and ResNet
by: Dhakal, Manish, et al.
Published: (2024)
by: Dhakal, Manish, et al.
Published: (2024)
On the structure of the seminal receptacle in cyclopids (Copepoda, Cyclopoida). [Translation from: Informatsionnyi Byulleten Biologiya Vnutrennikh Vod (6) 26-31, 1970.]
by: Filimonov, L. A.
Published: (1974)
by: Filimonov, L. A.
Published: (1974)
Dialect-based automatic speech recognition in Tamil
by: Saranya S, et al.
Published: (2025)
by: Saranya S, et al.
Published: (2025)
Automatic recognition and detection of aphasic natural speech
by: Barberis, Mara, et al.
Published: (2024)
by: Barberis, Mara, et al.
Published: (2024)
Frequency compression and speech recognition in elderly people
by: Amanda Dal Piva Gresele
Published: (2014)
by: Amanda Dal Piva Gresele
Published: (2014)
La influencia de la esclavitud en la estructura doméstica y la familia en Jamaica, Cuba y Brasil / Verena Stolcke
by: Stolcke, Verena
Published: (1970)
by: Stolcke, Verena
Published: (1970)
La influencia de la esclavitud en la estructura doméstica y la familia en Jamaica, Cuba y Brasil
by: Verena Stolcke
Published: (2003)
by: Verena Stolcke
Published: (2003)
ACTO HOMENAJE A JOHN V. MURRA EN EL INSTITUT D'ESTUDIS CATALANS, BARCELONA, 20 DE FEBRERO DE 2007
by: Verena Stolcke
Published: (2010)
by: Verena Stolcke
Published: (2010)
¿Es el sexo para el género lo que la raza para la etnicidad... y la naturaleza para la sociedad?
by: Verena Stolcke
Published: (2000)
by: Verena Stolcke
Published: (2000)
Los mestizos no nacen sino que se hacen
by: Verena Stolcke
Published: (2009)
by: Verena Stolcke
Published: (2009)
Las nuevas tecnologías reproductivas, la vieja paternidad
by: Verena Stolcke
Published: (2018)
by: Verena Stolcke
Published: (2018)
SEMIÓTICA DA MARCA DOS PRODUTOS PROCTER & GAMBLE NO FILME “MINHA MÃE É UMA PEÇA”
by: Pablo Moreno Fernandes Viana.
Published: (2014)
by: Pablo Moreno Fernandes Viana.
Published: (2014)
SIFT-50M: A Large-Scale Multilingual Dataset for Speech Instruction Fine-Tuning
by: Pandey, Prabhat, et al.
Published: (2025)
by: Pandey, Prabhat, et al.
Published: (2025)
Development and multi-center evaluation of domain-adapted speech recognition for human-AI teaming in real-world gastrointestinal endoscopy
by: Yang, Ruijie, et al.
Published: (2026)
by: Yang, Ruijie, et al.
Published: (2026)
Similar Items
-
Streaming Speech-to-Confusion Network Speech Recognition
by: Filimonov, Denis, et al.
Published: (2023) -
Multi-Modal Retrieval For Large Language Model Based Speech Recognition
by: Kolehmainen, Jari, et al.
Published: (2024) -
Speech Recognition Rescoring with Large Speech-Text Foundation Models
by: Shivakumar, Prashanth Gurunath, et al.
Published: (2024) -
Towards ASR Robust Spoken Language Understanding Through In-Context Learning With Word Confusion Networks
by: Everson, Kevin, et al.
Published: (2024) -
Investigating Training Strategies and Model Robustness of Low-Rank Adaptation for Language Modeling in Speech Recognition
by: Yu, Yu, et al.
Published: (2024)