InterBiasing: Boost Unseen Word Recognition through Biasing Intermediate Predictions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nakagome, Yu, Hentschel, Michael |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WCTC-Biasing: Retraining-free Contextual Biasing ASR with Wildcard CTC-based Keyword Spotting and Inter-layer Biasing
von: Nakagome, Yu, et al.
Veröffentlicht: (2025)
von: Nakagome, Yu, et al.
Veröffentlicht: (2025)
Contextualized End-to-end Automatic Speech Recognition with Intermediate Biasing Loss
von: Shakeel, Muhammad, et al.
Veröffentlicht: (2024)
von: Shakeel, Muhammad, et al.
Veröffentlicht: (2024)
Minimising Biasing Word Errors for Contextual ASR with the Tree-Constrained Pointer Generator
von: Sun, Guangzhi, et al.
Veröffentlicht: (2022)
von: Sun, Guangzhi, et al.
Veröffentlicht: (2022)
Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation
von: Huang, Ruizhe, et al.
Veröffentlicht: (2024)
von: Huang, Ruizhe, et al.
Veröffentlicht: (2024)
Post-decoder Biasing for End-to-End Speech Recognition of Multi-turn Medical Interview
von: Liu, Heyang, et al.
Veröffentlicht: (2024)
von: Liu, Heyang, et al.
Veröffentlicht: (2024)
Lightweight Prompt Biasing for Contextualized End-to-End ASR Systems
von: Ren, Bo, et al.
Veröffentlicht: (2025)
von: Ren, Bo, et al.
Veröffentlicht: (2025)
Improving ASR Contextual Biasing with Guided Attention
von: Tang, Jiyang, et al.
Veröffentlicht: (2024)
von: Tang, Jiyang, et al.
Veröffentlicht: (2024)
ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
von: Kumar, Subham, et al.
Veröffentlicht: (2025)
von: Kumar, Subham, et al.
Veröffentlicht: (2025)
Automatic Text Pronunciation Correlation Generation and Application for Contextual Biasing
von: Cheng, Gaofeng, et al.
Veröffentlicht: (2025)
von: Cheng, Gaofeng, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition Biases in Newcastle English: an Error Analysis
von: Serditova, Dana, et al.
Veröffentlicht: (2025)
von: Serditova, Dana, et al.
Veröffentlicht: (2025)
Unveiling Biases while Embracing Sustainability: Assessing the Dual Challenges of Automatic Speech Recognition Systems
von: Kulkarni, Ajinkya, et al.
Veröffentlicht: (2025)
von: Kulkarni, Ajinkya, et al.
Veröffentlicht: (2025)
OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary
von: Sudo, Yui, et al.
Veröffentlicht: (2025)
von: Sudo, Yui, et al.
Veröffentlicht: (2025)
Zero-shot Context Biasing with Trie-based Decoding using Synthetic Multi-Pronunciation
von: Liu, Changsong, et al.
Veröffentlicht: (2025)
von: Liu, Changsong, et al.
Veröffentlicht: (2025)
MauBERT: Universal Phonetic Inductive Biases for Few-Shot Acoustic Units Discovery
von: Tandazo, Angelo Ortiz, et al.
Veröffentlicht: (2025)
von: Tandazo, Angelo Ortiz, et al.
Veröffentlicht: (2025)
Lost in Transcription: Identifying and Quantifying the Accuracy Biases of Automatic Speech Recognition Systems Against Disfluent Speech
von: Mujtaba, Dena, et al.
Veröffentlicht: (2024)
von: Mujtaba, Dena, et al.
Veröffentlicht: (2024)
Contextual Biasing for ASR in Speech LLM with Common Word Cues and Bias Word Position Prediction
von: Novitasari, Sashi, et al.
Veröffentlicht: (2026)
von: Novitasari, Sashi, et al.
Veröffentlicht: (2026)
Unifying Global and Near-Context Biasing in a Single Trie Pass
von: Thorbecke, Iuliia, et al.
Veröffentlicht: (2024)
von: Thorbecke, Iuliia, et al.
Veröffentlicht: (2024)
Investigating the Impact of Word Informativeness on Speech Emotion Recognition
von: Kakouros, Sofoklis
Veröffentlicht: (2025)
von: Kakouros, Sofoklis
Veröffentlicht: (2025)
TurboBias: Universal ASR Context-Biasing powered by GPU-accelerated Phrase-Boosting Tree
von: Andrusenko, Andrei, et al.
Veröffentlicht: (2025)
von: Andrusenko, Andrei, et al.
Veröffentlicht: (2025)
Enhancing Dysarthric Speech Recognition for Unseen Speakers via Prototype-Based Adaptation
von: Wang, Shiyao, et al.
Veröffentlicht: (2024)
von: Wang, Shiyao, et al.
Veröffentlicht: (2024)
Understanding Zero-shot Rare Word Recognition Improvements Through LLM Integration
von: Wang, Haoxuan
Veröffentlicht: (2025)
von: Wang, Haoxuan
Veröffentlicht: (2025)
PHRASED: Phrase Dictionary Biasing for Speech Translation
von: Wang, Peidong, et al.
Veröffentlicht: (2025)
von: Wang, Peidong, et al.
Veröffentlicht: (2025)
Contextual Biasing for Streaming ASR via CTC-based Word Spotting
von: Tsai, Kai-Chen, et al.
Veröffentlicht: (2026)
von: Tsai, Kai-Chen, et al.
Veröffentlicht: (2026)
Enhancing the Robustness of Contextual ASR to Varying Biasing Information Volumes Through Purified Semantic Correlation Joint Modeling
von: Gu, Yue, et al.
Veröffentlicht: (2025)
von: Gu, Yue, et al.
Veröffentlicht: (2025)
Boosting CTC-Based ASR Using LLM-Based Intermediate Loss Regularization
von: Altinok, Duygu
Veröffentlicht: (2025)
von: Altinok, Duygu
Veröffentlicht: (2025)
Fast Context-Biasing for CTC and Transducer ASR models with CTC-based Word Spotter
von: Andrusenko, Andrei, et al.
Veröffentlicht: (2024)
von: Andrusenko, Andrei, et al.
Veröffentlicht: (2024)
RAG-Boost: Retrieval-Augmented Generation Enhanced LLM-based Speech Recognition
von: Wang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Wang, Pengcheng, et al.
Veröffentlicht: (2025)
Keep Decoding Parallel with Effective Knowledge Distillation from Language Models to End-to-end Speech Recognisers
von: Hentschel, Michael, et al.
Veröffentlicht: (2024)
von: Hentschel, Michael, et al.
Veröffentlicht: (2024)
Word Level Timestamp Generation for Automatic Speech Recognition and Translation
von: Hu, Ke, et al.
Veröffentlicht: (2025)
von: Hu, Ke, et al.
Veröffentlicht: (2025)
Enhancing Multilingual ASR for Unseen Languages via Language Embedding Modeling
von: Huang, Shao-Syuan, et al.
Veröffentlicht: (2024)
von: Huang, Shao-Syuan, et al.
Veröffentlicht: (2024)
Deepfake Word Detection by Next-token Prediction using Fine-tuned Whisper
von: Tran, Hoan My, et al.
Veröffentlicht: (2026)
von: Tran, Hoan My, et al.
Veröffentlicht: (2026)
Automatic Speech Recognition System-Independent Word Error Rate Estimation
von: Park, Chanho, et al.
Veröffentlicht: (2024)
von: Park, Chanho, et al.
Veröffentlicht: (2024)
A Neural Model for Contextual Biasing Score Learning and Filtering
von: Huang, Wanting, et al.
Veröffentlicht: (2025)
von: Huang, Wanting, et al.
Veröffentlicht: (2025)
Enhancing Large Language Model-based Speech Recognition by Contextualization for Rare and Ambiguous Words
von: Nozawa, Kento, et al.
Veröffentlicht: (2024)
von: Nozawa, Kento, et al.
Veröffentlicht: (2024)
Whisper Has an Internal Word Aligner
von: Yeh, Sung-Lin, et al.
Veröffentlicht: (2025)
von: Yeh, Sung-Lin, et al.
Veröffentlicht: (2025)
Layer-Wise Analysis of Self-Supervised Acoustic Word Embeddings: A Study on Speech Emotion Recognition
von: Saliba, Alexandra, et al.
Veröffentlicht: (2024)
von: Saliba, Alexandra, et al.
Veröffentlicht: (2024)
Contextualized Automatic Speech Recognition with Attention-Based Bias Phrase Boosted Beam Search
von: Sudo, Yui, et al.
Veröffentlicht: (2024)
von: Sudo, Yui, et al.
Veröffentlicht: (2024)
Improving Acoustic Word Embeddings through Correspondence Training of Self-supervised Speech Representations
von: Meghanani, Amit, et al.
Veröffentlicht: (2024)
von: Meghanani, Amit, et al.
Veröffentlicht: (2024)
In-Context Learning Boosts Speech Recognition via Human-like Adaptation to Speakers and Language Varieties
von: Roll, Nathan, et al.
Veröffentlicht: (2025)
von: Roll, Nathan, et al.
Veröffentlicht: (2025)
Children's Speech Recognition through Discrete Token Enhancement
von: Sukhadia, Vrunda N., et al.
Veröffentlicht: (2024)
von: Sukhadia, Vrunda N., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
WCTC-Biasing: Retraining-free Contextual Biasing ASR with Wildcard CTC-based Keyword Spotting and Inter-layer Biasing
von: Nakagome, Yu, et al.
Veröffentlicht: (2025) -
Contextualized End-to-end Automatic Speech Recognition with Intermediate Biasing Loss
von: Shakeel, Muhammad, et al.
Veröffentlicht: (2024) -
Minimising Biasing Word Errors for Contextual ASR with the Tree-Constrained Pointer Generator
von: Sun, Guangzhi, et al.
Veröffentlicht: (2022) -
Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation
von: Huang, Ruizhe, et al.
Veröffentlicht: (2024) -
Post-decoder Biasing for End-to-End Speech Recognition of Multi-turn Medical Interview
von: Liu, Heyang, et al.
Veröffentlicht: (2024)