AC-Mix: Self-Supervised Adaptation for Low-Resource Automatic Speech Recognition using Agnostic Contrastive Mixup
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Carvalho, Carlos, Abad, Alberto |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition
par: Farhadipour, Aref, et autres
Publié: (2024)
par: Farhadipour, Aref, et autres
Publié: (2024)
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
par: Nespoli, Francesco, et autres
Publié: (2024)
par: Nespoli, Francesco, et autres
Publié: (2024)
Towards Automatic Assessment of Self-Supervised Speech Models using Rank
par: Aldeneh, Zakaria, et autres
Publié: (2024)
par: Aldeneh, Zakaria, et autres
Publié: (2024)
The CHiME-8 DASR Challenge for Generalizable and Array Agnostic Distant Automatic Speech Recognition and Diarization
par: Cornell, Samuele, et autres
Publié: (2024)
par: Cornell, Samuele, et autres
Publié: (2024)
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition
par: Wang, Yujin, et autres
Publié: (2022)
par: Wang, Yujin, et autres
Publié: (2022)
GigaAM: Efficient Self-Supervised Learner for Speech Recognition
par: Kutsakov, Aleksandr, et autres
Publié: (2025)
par: Kutsakov, Aleksandr, et autres
Publié: (2025)
SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech Recognition
par: Hsu, Ming-Hao, et autres
Publié: (2024)
par: Hsu, Ming-Hao, et autres
Publié: (2024)
Fusion of Discrete Representations and Self-Augmented Representations for Multilingual Automatic Speech Recognition
par: Wang, Shih-heng, et autres
Publié: (2024)
par: Wang, Shih-heng, et autres
Publié: (2024)
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
par: Tian, Jingguang, et autres
Publié: (2024)
par: Tian, Jingguang, et autres
Publié: (2024)
Keyword-Guided Adaptation of Automatic Speech Recognition
par: Shamsian, Aviv, et autres
Publié: (2024)
par: Shamsian, Aviv, et autres
Publié: (2024)
AGADIR: Towards Array-Geometry Agnostic Directional Speech Recognition
par: Lin, Ju, et autres
Publié: (2024)
par: Lin, Ju, et autres
Publié: (2024)
Emotion-Aware Contrastive Adaptation Network for Source-Free Cross-Corpus Speech Emotion Recognition
par: Zhao, Yan, et autres
Publié: (2024)
par: Zhao, Yan, et autres
Publié: (2024)
Crab: Multi Layer Contrastive Supervision to Improve Speech Emotion Recognition Under Both Acted and Natural Speech Condition
par: Ueda, Lucas H., et autres
Publié: (2026)
par: Ueda, Lucas H., et autres
Publié: (2026)
Automatic Speech Recognition for African Low-Resource Languages: Challenges and Future Directions
par: Imam, Sukairaj Hafiz, et autres
Publié: (2025)
par: Imam, Sukairaj Hafiz, et autres
Publié: (2025)
Low-Resource Self-Supervised Learning with SSL-Enhanced TTS
par: Hsu, Po-chun, et autres
Publié: (2023)
par: Hsu, Po-chun, et autres
Publié: (2023)
Windowed SummaryMixing: An Efficient Fine-Tuning of Self-Supervised Learning Models for Low-resource Speech Recognition
par: Menon, Aditya Srinivas, et autres
Publié: (2026)
par: Menon, Aditya Srinivas, et autres
Publié: (2026)
End-to-End Integration of Speech Emotion Recognition with Voice Activity Detection using Self-Supervised Learning Features
par: Yamashita, Natsuo, et autres
Publié: (2024)
par: Yamashita, Natsuo, et autres
Publié: (2024)
Emotion-Coherent Speech Data Augmentation and Self-Supervised Contrastive Style Training for Enhancing Kids's Story Speech Synthesis
par: Chung, Raymond
Publié: (2026)
par: Chung, Raymond
Publié: (2026)
Can you Remove the Downstream Model for Speaker Recognition with Self-Supervised Speech Features?
par: Aldeneh, Zakaria, et autres
Publié: (2024)
par: Aldeneh, Zakaria, et autres
Publié: (2024)
On the Relevance of Clinical Assessment Tasks for the Automatic Detection of Parkinson's Disease Medication State from Speech
par: Gimeno-Gómez, David, et autres
Publié: (2025)
par: Gimeno-Gómez, David, et autres
Publié: (2025)
SOA: Reducing Domain Mismatch in SSL Pipeline by Speech Only Adaptation for Low Resource ASR
par: Shankar, Natarajan Balaji, et autres
Publié: (2024)
par: Shankar, Natarajan Balaji, et autres
Publié: (2024)
Zero-Shot Recognition of Dysarthric Speech Using Commercial Automatic Speech Recognition and Multimodal Large Language Models
par: Alsayegh, Ali, et autres
Publié: (2025)
par: Alsayegh, Ali, et autres
Publié: (2025)
Transliterated Zero-Shot Domain Adaptation for Automatic Speech Recognition
par: Zhu, Han, et autres
Publié: (2024)
par: Zhu, Han, et autres
Publié: (2024)
Unsupervised Accent Adaptation Through Masked Language Model Correction Of Discrete Self-Supervised Speech Units
par: Poncelet, Jakob, et autres
Publié: (2023)
par: Poncelet, Jakob, et autres
Publié: (2023)
Exploring Local Interpretable Model-Agnostic Explanations for Speech Emotion Recognition with Distribution-Shift
par: Hjuler, Maja J., et autres
Publié: (2025)
par: Hjuler, Maja J., et autres
Publié: (2025)
Refining Self-Supervised Learnt Speech Representation using Brain Activations
par: Li, Hengyu, et autres
Publié: (2024)
par: Li, Hengyu, et autres
Publié: (2024)
Attention-Guided Adaptation for Code-Switching Speech Recognition
par: Aditya, Bobbi, et autres
Publié: (2023)
par: Aditya, Bobbi, et autres
Publié: (2023)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
par: Bondaruk, Łukasz, et autres
Publié: (2024)
par: Bondaruk, Łukasz, et autres
Publié: (2024)
Rehearsal-Free Online Continual Learning for Automatic Speech Recognition
par: Eeckt, Steven Vander, et autres
Publié: (2023)
par: Eeckt, Steven Vander, et autres
Publié: (2023)
Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis
par: Leung, Wing-Zin, et autres
Publié: (2024)
par: Leung, Wing-Zin, et autres
Publié: (2024)
Continual Test-time Adaptation for End-to-end Speech Recognition on Noisy Speech
par: Lin, Guan-Ting, et autres
Publié: (2024)
par: Lin, Guan-Ting, et autres
Publié: (2024)
Speech as a Biomarker for Disease Detection
par: Botelho, Catarina, et autres
Publié: (2024)
par: Botelho, Catarina, et autres
Publié: (2024)
Evaluating Parkinson's Disease Detection in Anonymized Speech: A Performance and Acoustic Analysis
par: Franzreb, Carlos, et autres
Publié: (2026)
par: Franzreb, Carlos, et autres
Publié: (2026)
Streaming Decoder-Only Automatic Speech Recognition with Discrete Speech Units: A Pilot Study
par: Chen, Peikun, et autres
Publié: (2024)
par: Chen, Peikun, et autres
Publié: (2024)
Investigation of Deep Neural Network Acoustic Modelling Approaches for Low Resource Accented Mandarin Speech Recognition
par: Xie, Xurong, et autres
Publié: (2022)
par: Xie, Xurong, et autres
Publié: (2022)
Latent-Level Enhancement with Flow Matching for Robust Automatic Speech Recognition
par: Yang, Da-Hee, et autres
Publié: (2026)
par: Yang, Da-Hee, et autres
Publié: (2026)
Leveraging Broadcast Media Subtitle Transcripts for Automatic Speech Recognition and Subtitling
par: Poncelet, Jakob, et autres
Publié: (2025)
par: Poncelet, Jakob, et autres
Publié: (2025)
Analysis of Self-Supervised Speech Models on Children's Speech and Infant Vocalizations
par: Li, Jialu, et autres
Publié: (2024)
par: Li, Jialu, et autres
Publié: (2024)
Multilingual Zero Resource Speech Recognition Base on Self-Supervise Pre-Trained Acoustic Models
par: Wang, Haoyu, et autres
Publié: (2022)
par: Wang, Haoyu, et autres
Publié: (2022)
Fine-Tuning Automatic Speech Recognition for People with Parkinson's: An Effective Strategy for Enhancing Speech Technology Accessibility
par: Zheng, Xiuwen, et autres
Publié: (2024)
par: Zheng, Xiuwen, et autres
Publié: (2024)
Documents similaires
-
Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition
par: Farhadipour, Aref, et autres
Publié: (2024) -
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
par: Nespoli, Francesco, et autres
Publié: (2024) -
Towards Automatic Assessment of Self-Supervised Speech Models using Rank
par: Aldeneh, Zakaria, et autres
Publié: (2024) -
The CHiME-8 DASR Challenge for Generalizable and Array Agnostic Distant Automatic Speech Recognition and Diarization
par: Cornell, Samuele, et autres
Publié: (2024) -
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition
par: Wang, Yujin, et autres
Publié: (2022)