Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Pu, Van hamme, Hugo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Using Adapters to Overcome Catastrophic Forgetting in End-to-End Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2022)
por: Eeckt, Steven Vander, et al.
Publicado: (2022)
Continual Learning for Monolingual End-to-End Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2021)
por: Eeckt, Steven Vander, et al.
Publicado: (2021)
Unsupervised Online Continual Learning for Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2024)
por: Eeckt, Steven Vander, et al.
Publicado: (2024)
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
por: Poncelet, Jakob, et al.
Publicado: (2025)
por: Poncelet, Jakob, et al.
Publicado: (2025)
Rehearsal-Free Online Continual Learning for Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2023)
por: Eeckt, Steven Vander, et al.
Publicado: (2023)
Multitask Learning with Capsule Networks for Speech-to-Intent Applications
por: Poncelet, Jakob, et al.
Publicado: (2020)
por: Poncelet, Jakob, et al.
Publicado: (2020)
Weight Averaging: A Simple Yet Effective Method to Overcome Catastrophic Forgetting in Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2022)
por: Eeckt, Steven Vander, et al.
Publicado: (2022)
Breaking Walls: Pioneering Automatic Speech Recognition for Central Kurdish: End-to-End Transformer Paradigm
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2024)
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2024)
SSVD-O: Parameter-Efficient Fine-Tuning with Structured SVD for Speech Recognition
por: Wang, Pu, et al.
Publicado: (2026)
por: Wang, Pu, et al.
Publicado: (2026)
Unsupervised Accent Adaptation Through Masked Language Model Correction Of Discrete Self-Supervised Speech Units
por: Poncelet, Jakob, et al.
Publicado: (2023)
por: Poncelet, Jakob, et al.
Publicado: (2023)
Comparison of Self-Supervised Speech Pre-Training Methods on Flemish Dutch
por: Poncelet, Jakob, et al.
Publicado: (2021)
por: Poncelet, Jakob, et al.
Publicado: (2021)
Efficient Extraction of Noise-Robust Discrete Units from Self-Supervised Speech Models
por: Poncelet, Jakob, et al.
Publicado: (2024)
por: Poncelet, Jakob, et al.
Publicado: (2024)
End-to-End Speech Recognition with Pre-trained Masked Language Model
por: Higuchi, Yosuke, et al.
Publicado: (2024)
por: Higuchi, Yosuke, et al.
Publicado: (2024)
End-to-End Transformer-based Automatic Speech Recognition for Northern Kurdish: A Pioneering Approach
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2024)
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2024)
Lightweight and Robust Multi-Channel End-to-End Speech Recognition with Spherical Harmonic Transform
por: Kong, Xiangzhu, et al.
Publicado: (2025)
por: Kong, Xiangzhu, et al.
Publicado: (2025)
IKFST: IOO and KOO Algorithms for Accelerated and Precise WFST-based End-to-End Automatic Speech Recognition
por: Zhuang, Zhuoran, et al.
Publicado: (2026)
por: Zhuang, Zhuoran, et al.
Publicado: (2026)
Code-Switching in End-to-End Automatic Speech Recognition: A Systematic Literature Review
por: Agro, Maha Tufail, et al.
Publicado: (2025)
por: Agro, Maha Tufail, et al.
Publicado: (2025)
End-to-End Target Speaker Speech Recognition Using Context-Aware Attention Mechanisms for Challenging Enrollment Scenario
por: Ghane, Mohsen, et al.
Publicado: (2025)
por: Ghane, Mohsen, et al.
Publicado: (2025)
CUSIDE-array: A Streaming Multi-Channel End-to-End Speech Recognition System with Realistic Evaluations
por: Kong, Xiangzhu, et al.
Publicado: (2024)
por: Kong, Xiangzhu, et al.
Publicado: (2024)
Continual Test-time Adaptation for End-to-end Speech Recognition on Noisy Speech
por: Lin, Guan-Ting, et al.
Publicado: (2024)
por: Lin, Guan-Ting, et al.
Publicado: (2024)
Contextualized End-to-end Automatic Speech Recognition with Intermediate Biasing Loss
por: Shakeel, Muhammad, et al.
Publicado: (2024)
por: Shakeel, Muhammad, et al.
Publicado: (2024)
Central Kurdish Text-to-Speech Synthesis with Novel End-to-End Transformer Training
por: Ahmad, Hawraz A., et al.
Publicado: (2024)
por: Ahmad, Hawraz A., et al.
Publicado: (2024)
Enhancing Fully Formatted End-to-End Speech Recognition with Knowledge Distillation via Multi-Codebook Vector Quantization
por: You, Jian, et al.
Publicado: (2025)
por: You, Jian, et al.
Publicado: (2025)
CosyEdit: Unlocking End-to-End Speech Editing Capability from Zero-Shot Text-to-Speech Models
por: Chen, Junyang, et al.
Publicado: (2026)
por: Chen, Junyang, et al.
Publicado: (2026)
Survey of End-to-End Multi-Speaker Automatic Speech Recognition for Monaural Audio
por: He, Xinlu, et al.
Publicado: (2025)
por: He, Xinlu, et al.
Publicado: (2025)
SSVD: Structured SVD for Parameter-Efficient Fine-Tuning and Benchmarking under Domain Shift in ASR
por: Wang, Pu, et al.
Publicado: (2025)
por: Wang, Pu, et al.
Publicado: (2025)
End-to-End DOA-Guided Speech Extraction in Noisy Multi-Talker Scenarios
por: Jing, Kangqi, et al.
Publicado: (2025)
por: Jing, Kangqi, et al.
Publicado: (2025)
Fairness of Automatic Speech Recognition in Cleft Lip and Palate Speech
por: Bhattacharjee, Susmita, et al.
Publicado: (2025)
por: Bhattacharjee, Susmita, et al.
Publicado: (2025)
Reference Channel Selection by Multi-Channel Masking for End-to-End Multi-Channel Speech Enhancement
por: Dai, Wang, et al.
Publicado: (2024)
por: Dai, Wang, et al.
Publicado: (2024)
Data Augmentation for End-to-end Code-switching Speech Recognition
por: Du, Chenpeng, et al.
Publicado: (2020)
por: Du, Chenpeng, et al.
Publicado: (2020)
Post-decoder Biasing for End-to-End Speech Recognition of Multi-turn Medical Interview
por: Liu, Heyang, et al.
Publicado: (2024)
por: Liu, Heyang, et al.
Publicado: (2024)
Speech-to-See: End-to-End Speech-Driven Open-Set Object Detection
por: Lu, Wenhuan, et al.
Publicado: (2025)
por: Lu, Wenhuan, et al.
Publicado: (2025)
Inverse-Hessian Regularization for Continual Learning in ASR
por: Eeckt, Steven Vander, et al.
Publicado: (2026)
por: Eeckt, Steven Vander, et al.
Publicado: (2026)
Towards End-to-End Training of Automatic Speech Recognition for Nigerian Pidgin
por: Rufai, Amina Mardiyyah, et al.
Publicado: (2020)
por: Rufai, Amina Mardiyyah, et al.
Publicado: (2020)
End-to-End Integration of Speech Emotion Recognition with Voice Activity Detection using Self-Supervised Learning Features
por: Yamashita, Natsuo, et al.
Publicado: (2024)
por: Yamashita, Natsuo, et al.
Publicado: (2024)
Speaker Attributed Automatic Speech Recognition Using Speech Aware LLMS
por: Aronowitz, Hagai, et al.
Publicado: (2026)
por: Aronowitz, Hagai, et al.
Publicado: (2026)
SoulX-Transcriber: A Robust End-to-End Framework for Multi-Speaker Speech Transcription
por: Dai, Yuhang, et al.
Publicado: (2026)
por: Dai, Yuhang, et al.
Publicado: (2026)
Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition
por: Hono, Yukiya, et al.
Publicado: (2023)
por: Hono, Yukiya, et al.
Publicado: (2023)
Gammatonegram Representation for End-to-End Dysarthric Speech Processing Tasks: Speech Recognition, Speaker Identification, and Intelligibility Assessment
por: Farhadipour, Aref, et al.
Publicado: (2023)
por: Farhadipour, Aref, et al.
Publicado: (2023)
DEFORMER: Coupling Deformed Localized Patterns with Global Context for Robust End-to-end Speech Recognition
por: Xie, Jiamin, et al.
Publicado: (2022)
por: Xie, Jiamin, et al.
Publicado: (2022)
Ejemplares similares
-
Using Adapters to Overcome Catastrophic Forgetting in End-to-End Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2022) -
Continual Learning for Monolingual End-to-End Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2021) -
Unsupervised Online Continual Learning for Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2024) -
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
por: Poncelet, Jakob, et al.
Publicado: (2025) -
Rehearsal-Free Online Continual Learning for Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2023)