The USTC-NERCSLIP Systems for The ICMC-ASR Challenge
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wu, Minghui, Xu, Luzhen, Zhang, Jie, Tang, Haitao, Yue, Yanyan, Liao, Ruizhi, Zhao, Jintao, Zhang, Zhengzhe, Wang, Yichi, Yan, Haoyin, Yu, Hongliang, Ma, Tongle, Liu, Jiachen, Wu, Chongliang, Li, Yongchao, Zhang, Yanyong, Fang, Xin, Zhang, Yue |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The USTC-NERCSLIP Systems for the CHiME-8 NOTSOFAR-1 Challenge
par: Niu, Shutong, et autres
Publié: (2024)
par: Niu, Shutong, et autres
Publié: (2024)
The USTC-NERCSLIP Systems for the CHiME-9 MCoRec Challenge
par: Jiang, Ya, et autres
Publié: (2026)
par: Jiang, Ya, et autres
Publié: (2026)
The USTC-NERCSLIP Systems for the CHiME-8 MMCSG Challenge
par: Jiang, Ya, et autres
Publié: (2024)
par: Jiang, Ya, et autres
Publié: (2024)
ICMC-ASR: The ICASSP 2024 In-Car Multi-Channel Automatic Speech Recognition Challenge
par: Wang, He, et autres
Publié: (2024)
par: Wang, He, et autres
Publié: (2024)
USTC-KXDIGIT System Description for ASVspoof5 Challenge
par: Chen, Yihao, et autres
Publié: (2024)
par: Chen, Yihao, et autres
Publié: (2024)
A Composite Predictive-Generative Approach to Monaural Universal Speech Enhancement
par: Zhang, Jie, et autres
Publié: (2025)
par: Zhang, Jie, et autres
Publié: (2025)
LABNet: A Lightweight Attentive Beamforming Network for Ad-hoc Multichannel Microphone Invariant Real-Time Speech Enhancement
par: Yan, Haoyin, et autres
Publié: (2025)
par: Yan, Haoyin, et autres
Publié: (2025)
DARS: Dysarthria-Aware Rhythm-Style Synthesis for ASR Enhancement
par: Wu, Minghui, et autres
Publié: (2026)
par: Wu, Minghui, et autres
Publié: (2026)
Overlap-Adaptive Hybrid Speaker Diarization and ASR-Aware Observation Addition for MISP 2025 Challenge
par: Huang, Shangkun, et autres
Publié: (2025)
par: Huang, Shangkun, et autres
Publié: (2025)
CaSNet: Compress-and-Send Network Based Multi-Device Speech Enhancement Model for Distributed Microphone Arrays
par: Jiang, Chengqian, et autres
Publié: (2026)
par: Jiang, Chengqian, et autres
Publié: (2026)
End-to-End Simultaneous Dysarthric Speech Reconstruction with Frame-Level Adaptor and Multiple Wait-k Knowledge Distillation
par: Wu, Minghui, et autres
Publié: (2026)
par: Wu, Minghui, et autres
Publié: (2026)
Speech Recognition on TV Series with Video-guided Post-ASR Correction
par: Yang, Haoyuan, et autres
Publié: (2025)
par: Yang, Haoyuan, et autres
Publié: (2025)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
par: Yan, Haoyin, et autres
Publié: (2024)
par: Yan, Haoyin, et autres
Publié: (2024)
DNCASR: End-to-End Training for Speaker-Attributed ASR
par: Zheng, Xianrui, et autres
Publié: (2025)
par: Zheng, Xianrui, et autres
Publié: (2025)
The THUEE System Description for the IARPA OpenASR21 Challenge
par: Zhao, Jing, et autres
Publié: (2022)
par: Zhao, Jing, et autres
Publié: (2022)
Enhancing the Robustness of Contextual ASR to Varying Biasing Information Volumes Through Purified Semantic Correlation Joint Modeling
par: Gu, Yue, et autres
Publié: (2025)
par: Gu, Yue, et autres
Publié: (2025)
MaLa-ASR: Multimedia-Assisted LLM-Based ASR
par: Yang, Guanrou, et autres
Publié: (2024)
par: Yang, Guanrou, et autres
Publié: (2024)
SOT Triggered Neural Clustering for Speaker Attributed ASR
par: Zheng, Xianrui, et autres
Publié: (2024)
par: Zheng, Xianrui, et autres
Publié: (2024)
Metis: A Foundation Speech Generation Model with Masked Generative Pre-training
par: Wang, Yuancheng, et autres
Publié: (2025)
par: Wang, Yuancheng, et autres
Publié: (2025)
NTU Speechlab LLM-Based Multilingual ASR System for Interspeech MLC-SLM Challenge 2025
par: Peng, Yizhou, et autres
Publié: (2025)
par: Peng, Yizhou, et autres
Publié: (2025)
NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR
par: Xie, Yuan, et autres
Publié: (2026)
par: Xie, Yuan, et autres
Publié: (2026)
Index-ASR Technical Report
par: Song, Zheshu, et autres
Publié: (2025)
par: Song, Zheshu, et autres
Publié: (2025)
ASR for Affective Speech: Investigating Impact of Emotion and Speech Generative Strategy
par: Wu, Ya-Tse, et autres
Publié: (2026)
par: Wu, Ya-Tse, et autres
Publié: (2026)
Adaptive Differential Denoising for Respiratory Sounds Classification
par: Dong, Gaoyang, et autres
Publié: (2025)
par: Dong, Gaoyang, et autres
Publié: (2025)
Speaker Adaptation for Quantised End-to-End ASR Models
par: Zhao, Qiuming, et autres
Publié: (2024)
par: Zhao, Qiuming, et autres
Publié: (2024)
Alternating Weak Triphone/BPE Alignment Supervision from Hybrid Model Improves End-to-End ASR
par: Jiang, Jintao, et autres
Publié: (2024)
par: Jiang, Jintao, et autres
Publié: (2024)
Contextual Biasing for LLM-Based ASR with Hotword Retrieval and Reinforcement Learning
par: Kong, YuXiang, et autres
Publié: (2025)
par: Kong, YuXiang, et autres
Publié: (2025)
CHSER: A Dataset and Case Study on Generative Speech Error Correction for Child ASR
par: Shankar, Natarajan Balaji, et autres
Publié: (2025)
par: Shankar, Natarajan Balaji, et autres
Publié: (2025)
Comparing Unsupervised and Supervised Semantic Speech Tokens: A Case Study of Child ASR
par: Shi, Mohan, et autres
Publié: (2025)
par: Shi, Mohan, et autres
Publié: (2025)
Advancing Multi-talker ASR Performance with Large Language Models
par: Shi, Mohan, et autres
Publié: (2024)
par: Shi, Mohan, et autres
Publié: (2024)
MedASR: An Open-Source Model for High-Accuracy Medical Dictation
par: Wu, Ke, et autres
Publié: (2026)
par: Wu, Ke, et autres
Publié: (2026)
Qwen3-ASR Technical Report
par: Shi, Xian, et autres
Publié: (2026)
par: Shi, Xian, et autres
Publié: (2026)
SA-SOT: Speaker-Aware Serialized Output Training for Multi-Talker ASR
par: Fan, Zhiyun, et autres
Publié: (2024)
par: Fan, Zhiyun, et autres
Publié: (2024)
Are Transformers in Pre-trained LM A Good ASR Encoder? An Empirical Study
par: An, Keyu, et autres
Publié: (2024)
par: An, Keyu, et autres
Publié: (2024)
Multi-Channel Multi-Speaker ASR Using Target Speaker's Solo Segment
par: Shao, Yiwen, et autres
Publié: (2024)
par: Shao, Yiwen, et autres
Publié: (2024)
Loss Masking Is Not Needed in Decoder-only Transformer for Discrete-token-based ASR
par: Chen, Qian, et autres
Publié: (2023)
par: Chen, Qian, et autres
Publié: (2023)
NTT Multi-Speaker ASR System for the DASR Task of CHiME-8 Challenge
par: Kamo, Naoyuki, et autres
Publié: (2024)
par: Kamo, Naoyuki, et autres
Publié: (2024)
SAML: Speaker Adaptive Mixture of LoRA Experts for End-to-End ASR
par: Zhao, Qiuming, et autres
Publié: (2024)
par: Zhao, Qiuming, et autres
Publié: (2024)
TagSpeech: End-to-End Multi-Speaker ASR and Diarization with Fine-Grained Temporal Grounding
par: Huo, Mingyue, et autres
Publié: (2026)
par: Huo, Mingyue, et autres
Publié: (2026)
Retrieval Augmented Generation based context discovery for ASR
par: Siskos, Dimitrios, et autres
Publié: (2025)
par: Siskos, Dimitrios, et autres
Publié: (2025)
Documents similaires
-
The USTC-NERCSLIP Systems for the CHiME-8 NOTSOFAR-1 Challenge
par: Niu, Shutong, et autres
Publié: (2024) -
The USTC-NERCSLIP Systems for the CHiME-9 MCoRec Challenge
par: Jiang, Ya, et autres
Publié: (2026) -
The USTC-NERCSLIP Systems for the CHiME-8 MMCSG Challenge
par: Jiang, Ya, et autres
Publié: (2024) -
ICMC-ASR: The ICASSP 2024 In-Car Multi-Channel Automatic Speech Recognition Challenge
par: Wang, He, et autres
Publié: (2024) -
USTC-KXDIGIT System Description for ASVspoof5 Challenge
par: Chen, Yihao, et autres
Publié: (2024)