Saved in:
| Main Authors: | Hernandez, Abner, Arias-Vergara, Tomás, Liu, Daiqi, Maier, Andreas, Pérez-Toro, Paula Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.25596 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Confidence-Guided Error Correction for Disordered Speech Recognition
by: Hernandez, Abner, et al.
Published: (2025)
by: Hernandez, Abner, et al.
Published: (2025)
Audio-Vision Contrastive Learning for Phonological Class Recognition
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
Adapting Self-Supervised Speech Representations for Cross-lingual Dysarthria Detection in Parkinson's Disease
by: Hernandez, Abner, et al.
Published: (2026)
by: Hernandez, Abner, et al.
Published: (2026)
SpeechCT-CLIP: Distilling Text-Image Knowledge to Speech for Voice-Native Multimodal CT Analysis
by: Buess, Lukas, et al.
Published: (2025)
by: Buess, Lukas, et al.
Published: (2025)
SIREM: Speech-Informed MRI Reconstruction with Learned Sampling
by: Hasan, Md, et al.
Published: (2026)
by: Hasan, Md, et al.
Published: (2026)
Perception of Phonological Assimilation by Neural Speech Recognition Models
by: Pouw, Charlotte, et al.
Published: (2024)
by: Pouw, Charlotte, et al.
Published: (2024)
VocSegMRI: Multimodal Learning for Precise Vocal Tract Segmentation in Real-time MRI
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
SNOBERT: A Benchmark for clinical notes entity linking in the SNOMED CT clinical terminology
by: Kulyabin, Mikhail, et al.
Published: (2024)
by: Kulyabin, Mikhail, et al.
Published: (2024)
Probing for Phonology in Self-Supervised Speech Representations: A Case Study on Accent Perception
by: Venkateswaran, Nitin, et al.
Published: (2025)
by: Venkateswaran, Nitin, et al.
Published: (2025)
Multilingual Zero Resource Speech Recognition Base on Self-Supervise Pre-Trained Acoustic Models
by: Wang, Haoyu, et al.
Published: (2022)
by: Wang, Haoyu, et al.
Published: (2022)
Speech-Guided Multimodal Learning for Vocal Tract Segmentation in Real-Time MRI
by: Liu, Daiqi, et al.
Published: (2026)
by: Liu, Daiqi, et al.
Published: (2026)
Leveraging Audio-Visual Data to Reduce the Multilingual Gap in Self-Supervised Speech Models
by: Blandón, María Andrea Cruz, et al.
Published: (2025)
by: Blandón, María Andrea Cruz, et al.
Published: (2025)
IPA-CHILDES & G2P+: Feature-Rich Resources for Cross-Lingual Phonology and Phonemic Language Modeling
by: Goriely, Zébulon, et al.
Published: (2025)
by: Goriely, Zébulon, et al.
Published: (2025)
Training-Free Cross-Lingual Dysarthria Severity Assessment via Phonological Subspace Analysis in Self-Supervised Speech Representations
by: Muller, Bernard, et al.
Published: (2026)
by: Muller, Bernard, et al.
Published: (2026)
Language Complexity and Speech Recognition Accuracy: Orthographic Complexity Hurts, Phonological Complexity Doesn't
by: Taguchi, Chihiro, et al.
Published: (2024)
by: Taguchi, Chihiro, et al.
Published: (2024)
Phonology Recognition in American Sign Language
by: Tavella, Federico, et al.
Published: (2021)
by: Tavella, Federico, et al.
Published: (2021)
Building a Non-native Speech Corpus Featuring Chinese-English Bilingual Children: Compilation and Rationale
by: Hung, Hiuchung, et al.
Published: (2023)
by: Hung, Hiuchung, et al.
Published: (2023)
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition
by: Wang, Yujin, et al.
Published: (2022)
by: Wang, Yujin, et al.
Published: (2022)
A Category-Fragment Segmentation Framework for Pelvic Fracture Segmentation in X-ray Images
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
Phonology-Guided Speech-to-Speech Translation for African Languages
by: Ochieng, Peter, et al.
Published: (2024)
by: Ochieng, Peter, et al.
Published: (2024)
[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic
by: Choi, Kwanghee, et al.
Published: (2026)
by: Choi, Kwanghee, et al.
Published: (2026)
SSHR: Leveraging Self-supervised Hierarchical Representations for Multilingual Automatic Speech Recognition
by: Xue, Hongfei, et al.
Published: (2023)
by: Xue, Hongfei, et al.
Published: (2023)
Novel-view X-ray Projection Synthesis through Geometry-Integrated Deep Learning
by: Liu, Daiqi, et al.
Published: (2025)
by: Liu, Daiqi, et al.
Published: (2025)
Simulating Articulatory Trajectories with Phonological Feature Interpolation
by: Tandazo, Angelo Ortiz, et al.
Published: (2024)
by: Tandazo, Angelo Ortiz, et al.
Published: (2024)
Whistle: Data-Efficient Multilingual and Crosslingual Speech Recognition via Weakly Phonetic Supervision
by: Yusuyin, Saierdaer, et al.
Published: (2024)
by: Yusuyin, Saierdaer, et al.
Published: (2024)
Speaker- and Text-Independent Estimation of Articulatory Movements and Phoneme Alignments from Speech
by: Weise, Tobias, et al.
Published: (2024)
by: Weise, Tobias, et al.
Published: (2024)
Quantifying Speaker Embedding Phonological Rule Interactions in Accented Speech Synthesis
by: Lertpetchpun, Thanathai, et al.
Published: (2026)
by: Lertpetchpun, Thanathai, et al.
Published: (2026)
BabyLM's First Words: Word Segmentation as a Phonological Probing Task
by: Goriely, Zébulon, et al.
Published: (2025)
by: Goriely, Zébulon, et al.
Published: (2025)
Interpreting Speaker Characteristics in the Dimensions of Self-Supervised Speech Features
by: van Rensburg, Kyle Janse, et al.
Published: (2026)
by: van Rensburg, Kyle Janse, et al.
Published: (2026)
WST: Weakly Supervised Transducer for Automatic Speech Recognition
by: Gao, Dongji, et al.
Published: (2025)
by: Gao, Dongji, et al.
Published: (2025)
Zipper-LoRA: Dynamic Parameter Decoupling for Speech-LLM based Multilingual Speech Recognition
by: Mei, Yuxiang, et al.
Published: (2026)
by: Mei, Yuxiang, et al.
Published: (2026)
Multilingual Extraction and Recognition of Implicit Discourse Relations in Speech and Text
by: Ruby, Ahmed, et al.
Published: (2026)
by: Ruby, Ahmed, et al.
Published: (2026)
Cross-Lingual Speech Emotion Recognition: Humans vs. Self-Supervised Models
by: Han, Zhichen, et al.
Published: (2024)
by: Han, Zhichen, et al.
Published: (2024)
LoGSAM: Parameter-Efficient Cross-Modal Grounding for MRI Segmentation
by: Bhuiyan, Mohammad Robaitul Islam, et al.
Published: (2026)
by: Bhuiyan, Mohammad Robaitul Islam, et al.
Published: (2026)
Twists, Humps, and Pebbles: Multilingual Speech Recognition Models Exhibit Gender Performance Gaps
by: Attanasio, Giuseppe, et al.
Published: (2024)
by: Attanasio, Giuseppe, et al.
Published: (2024)
Weighted Cross-entropy for Low-Resource Languages in Multilingual Speech Recognition
by: Piñeiro-Martín, Andrés, et al.
Published: (2024)
by: Piñeiro-Martín, Andrés, et al.
Published: (2024)
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
by: Adila, Aulia, et al.
Published: (2024)
by: Adila, Aulia, et al.
Published: (2024)
Learning More with Less: Self-Supervised Approaches for Low-Resource Speech Emotion Recognition
by: Gong, Ziwei, et al.
Published: (2025)
by: Gong, Ziwei, et al.
Published: (2025)
Low-Resourced Speech Recognition for Iu Mien Language via Weakly-Supervised Phoneme-based Multilingual Pre-training
by: Dong, Lukuan, et al.
Published: (2024)
by: Dong, Lukuan, et al.
Published: (2024)
An Exploration of Mamba for Speech Self-Supervised Models
by: Lin, Tzu-Quan, et al.
Published: (2025)
by: Lin, Tzu-Quan, et al.
Published: (2025)
Similar Items
-
Confidence-Guided Error Correction for Disordered Speech Recognition
by: Hernandez, Abner, et al.
Published: (2025) -
Audio-Vision Contrastive Learning for Phonological Class Recognition
by: Liu, Daiqi, et al.
Published: (2025) -
Adapting Self-Supervised Speech Representations for Cross-lingual Dysarthria Detection in Parkinson's Disease
by: Hernandez, Abner, et al.
Published: (2026) -
SpeechCT-CLIP: Distilling Text-Image Knowledge to Speech for Voice-Native Multimodal CT Analysis
by: Buess, Lukas, et al.
Published: (2025) -
SIREM: Speech-Informed MRI Reconstruction with Learned Sampling
by: Hasan, Md, et al.
Published: (2026)