CTC-DID: CTC-Based Arabic dialect identification for streaming applications
Fuente:
arXiv
Salvato in:
| Autori principali: | Farooq, Muhammad Umar, Saz, Oscar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SAGE: Spliced-Audio Generated Data for Enhancing Foundational Models in Low-Resource Arabic-English Code-Switched Speech Recognition
di: Farooq, Muhammad Umar, et al.
Pubblicazione: (2025)
di: Farooq, Muhammad Umar, et al.
Pubblicazione: (2025)
CTC-GMM: CTC guided modality matching for fast and accurate streaming speech translation
di: Zhao, Rui, et al.
Pubblicazione: (2024)
di: Zhao, Rui, et al.
Pubblicazione: (2024)
Speaker-Distinguishable CTC: Learning Speaker Distinction Using CTC for Multi-Talker Speech Recognition
di: Sakuma, Asahi, et al.
Pubblicazione: (2025)
di: Sakuma, Asahi, et al.
Pubblicazione: (2025)
FlexCTC: GPU-powered CTC Beam Decoding With Advanced Contextual Abilities
di: Grigoryan, Lilit, et al.
Pubblicazione: (2025)
di: Grigoryan, Lilit, et al.
Pubblicazione: (2025)
CTC-Assisted LLM-Based Contextual ASR
di: Yang, Guanrou, et al.
Pubblicazione: (2024)
di: Yang, Guanrou, et al.
Pubblicazione: (2024)
Fast Context-Biasing for CTC and Transducer ASR models with CTC-based Word Spotter
di: Andrusenko, Andrei, et al.
Pubblicazione: (2024)
di: Andrusenko, Andrei, et al.
Pubblicazione: (2024)
Joint Beam Search Integrating CTC, Attention, and Transducer Decoders
di: Sudo, Yui, et al.
Pubblicazione: (2024)
di: Sudo, Yui, et al.
Pubblicazione: (2024)
Unimodal Aggregation for CTC-based Speech Recognition
di: Fang, Ying, et al.
Pubblicazione: (2023)
di: Fang, Ying, et al.
Pubblicazione: (2023)
Improvement in Sign Language Translation Using Text CTC Alignment
di: Tan, Sihan, et al.
Pubblicazione: (2024)
di: Tan, Sihan, et al.
Pubblicazione: (2024)
Mixture-of-Experts with Intermediate CTC Supervision for Accented Speech Recognition
di: Lee, Wonjun, et al.
Pubblicazione: (2026)
di: Lee, Wonjun, et al.
Pubblicazione: (2026)
Boosting CTC-Based ASR Using LLM-Based Intermediate Loss Regularization
di: Altinok, Duygu
Pubblicazione: (2025)
di: Altinok, Duygu
Pubblicazione: (2025)
Automatic Speech Recognition with BERT and CTC Transformers: A Review
di: Djeffal, Noussaiba, et al.
Pubblicazione: (2024)
di: Djeffal, Noussaiba, et al.
Pubblicazione: (2024)
OWSM-CTC: An Open Encoder-Only Speech Foundation Model for Speech Recognition, Translation, and Language Identification
di: Peng, Yifan, et al.
Pubblicazione: (2024)
di: Peng, Yifan, et al.
Pubblicazione: (2024)
Label-Context-Dependent Internal Language Model Estimation for CTC
di: Yang, Zijian, et al.
Pubblicazione: (2025)
di: Yang, Zijian, et al.
Pubblicazione: (2025)
LegoSLM: Connecting LLM with Speech Encoder using CTC Posteriors
di: Ma, Rao, et al.
Pubblicazione: (2025)
di: Ma, Rao, et al.
Pubblicazione: (2025)
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration
di: Wen, Zhuofan, et al.
Pubblicazione: (2024)
di: Wen, Zhuofan, et al.
Pubblicazione: (2024)
CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment
di: Liu, Hanwen, et al.
Pubblicazione: (2026)
di: Liu, Hanwen, et al.
Pubblicazione: (2026)
CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition
di: Bartelds, Martijn, et al.
Pubblicazione: (2025)
di: Bartelds, Martijn, et al.
Pubblicazione: (2025)
Guiding Frame-Level CTC Alignments Using Self-knowledge Distillation
di: Kim, Eungbeom, et al.
Pubblicazione: (2024)
di: Kim, Eungbeom, et al.
Pubblicazione: (2024)
Improving Non-autoregressive Translation Quality with Pretrained Language Model, Embedding Distillation and Upsampling Strategy for CTC
di: Syu, Shen-sian, et al.
Pubblicazione: (2023)
di: Syu, Shen-sian, et al.
Pubblicazione: (2023)
CTC-based Non-autoregressive Textless Speech-to-Speech Translation
di: Fang, Qingkai, et al.
Pubblicazione: (2024)
di: Fang, Qingkai, et al.
Pubblicazione: (2024)
Less Peaky and More Accurate CTC Forced Alignment by Label Priors
di: Huang, Ruizhe, et al.
Pubblicazione: (2024)
di: Huang, Ruizhe, et al.
Pubblicazione: (2024)
JurisCTC: Enhancing Legal Judgment Prediction via Cross-Domain Transfer and Contrastive Learning
di: Kang, Zhaolu, et al.
Pubblicazione: (2025)
di: Kang, Zhaolu, et al.
Pubblicazione: (2025)
A Language-Agnostic Hierarchical LoRA-MoE Architecture for CTC-based Multilingual ASR
di: Zheng, Yuang, et al.
Pubblicazione: (2026)
di: Zheng, Yuang, et al.
Pubblicazione: (2026)
Align With Purpose: Optimize Desired Properties in CTC Models with a General Plug-and-Play Framework
di: Segev, Eliya, et al.
Pubblicazione: (2023)
di: Segev, Eliya, et al.
Pubblicazione: (2023)
Enhancing CTC-Based Visual Speech Recognition
di: Laux, Hendrik, et al.
Pubblicazione: (2024)
di: Laux, Hendrik, et al.
Pubblicazione: (2024)
Improving Zero-Shot Chinese-English Code-Switching ASR with kNN-CTC and Gated Monolingual Datastores
di: Zhou, Jiaming, et al.
Pubblicazione: (2024)
di: Zhou, Jiaming, et al.
Pubblicazione: (2024)
Enhancing Code-Switching ASR Leveraging Non-Peaky CTC Loss and Deep Language Posterior Injection
di: Yang, Tzu-Ting, et al.
Pubblicazione: (2024)
di: Yang, Tzu-Ting, et al.
Pubblicazione: (2024)
Estimating the Level of Dialectness Predicts Interannotator Agreement in Multi-dialect Arabic Datasets
di: Keleg, Amr, et al.
Pubblicazione: (2024)
di: Keleg, Amr, et al.
Pubblicazione: (2024)
WCTC-Biasing: Retraining-free Contextual Biasing ASR with Wildcard CTC-based Keyword Spotting and Inter-layer Biasing
di: Nakagome, Yu, et al.
Pubblicazione: (2025)
di: Nakagome, Yu, et al.
Pubblicazione: (2025)
Improving Multilingual Speech Models on ML-SUPERB 2.0: Fine-tuning with Data Augmentation and LID-Aware CTC
di: Wang, Qingzheng, et al.
Pubblicazione: (2025)
di: Wang, Qingzheng, et al.
Pubblicazione: (2025)
Sentiment Analysis Dataset in Moroccan Dialect: Bridging the Gap Between Arabic and Latin Scripted dialect
di: Jbel, Mouad, et al.
Pubblicazione: (2023)
di: Jbel, Mouad, et al.
Pubblicazione: (2023)
CR-CTC: Consistency regularization on CTC for improved speech recognition
di: Yao, Zengwei, et al.
Pubblicazione: (2024)
di: Yao, Zengwei, et al.
Pubblicazione: (2024)
LV-CTC: Non-autoregressive ASR with CTC and latent variable models
di: Fujita, Yuya, et al.
Pubblicazione: (2024)
di: Fujita, Yuya, et al.
Pubblicazione: (2024)
kNN-CTC: Enhancing ASR via Retrieval of CTC Pseudo Labels
di: Zhou, Jiaming, et al.
Pubblicazione: (2023)
di: Zhou, Jiaming, et al.
Pubblicazione: (2023)
CTC Blank Triggered Dynamic Layer-Skipping for Efficient CTC-based Speech Recognition
di: Hou, Junfeng, et al.
Pubblicazione: (2024)
di: Hou, Junfeng, et al.
Pubblicazione: (2024)
Unification of Balti and trans-border sister dialects in the essence of LLMs and AI Technology
di: Sharif, Muhammad, et al.
Pubblicazione: (2024)
di: Sharif, Muhammad, et al.
Pubblicazione: (2024)
Open Universal Arabic ASR Leaderboard
di: Wang, Yingzhi, et al.
Pubblicazione: (2024)
di: Wang, Yingzhi, et al.
Pubblicazione: (2024)
Supporting dataset for panCTC
di: Ye, Bin
Pubblicazione: (2026)
di: Ye, Bin
Pubblicazione: (2026)
Classifying several dialectal Nawatl varieties
di: Guzmán-Landa, Juan-José, et al.
Pubblicazione: (2026)
di: Guzmán-Landa, Juan-José, et al.
Pubblicazione: (2026)
Documenti analoghi
-
SAGE: Spliced-Audio Generated Data for Enhancing Foundational Models in Low-Resource Arabic-English Code-Switched Speech Recognition
di: Farooq, Muhammad Umar, et al.
Pubblicazione: (2025) -
CTC-GMM: CTC guided modality matching for fast and accurate streaming speech translation
di: Zhao, Rui, et al.
Pubblicazione: (2024) -
Speaker-Distinguishable CTC: Learning Speaker Distinction Using CTC for Multi-Talker Speech Recognition
di: Sakuma, Asahi, et al.
Pubblicazione: (2025) -
FlexCTC: GPU-powered CTC Beam Decoding With Advanced Contextual Abilities
di: Grigoryan, Lilit, et al.
Pubblicazione: (2025) -
CTC-Assisted LLM-Based Contextual ASR
di: Yang, Guanrou, et al.
Pubblicazione: (2024)