An Effective Automated Speaking Assessment Approach to Mitigating Data Scarcity and Imbalanced Distribution
Fuente:
arXiv
Guardado en:
| Autores principales: | Lo, Tien-Hong, Chao, Fu-An, Wu, Tzu-I, Sung, Yao-Ting, Chen, Berlin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
An Effective Strategy for Modeling Score Ordinality and Non-uniform Intervals in Automated Speaking Assessment
por: Lo, Tien-Hong, et al.
Publicado: (2025)
por: Lo, Tien-Hong, et al.
Publicado: (2025)
Mitigating Data Imbalance in Automated Speaking Assessment
por: Tsai, Fong-Chun, et al.
Publicado: (2025)
por: Tsai, Fong-Chun, et al.
Publicado: (2025)
A Novel Data Augmentation Approach for Automatic Speaking Assessment on Opinion Expressions
por: Wang, Chung-Chun, et al.
Publicado: (2025)
por: Wang, Chung-Chun, et al.
Publicado: (2025)
Advancing Automated Speaking Assessment Leveraging Multifaceted Relevance and Grammar Information
por: Lu, Hao-Chien, et al.
Publicado: (2025)
por: Lu, Hao-Chien, et al.
Publicado: (2025)
An Effective Mixture-Of-Experts Approach For Code-Switching Speech Recognition Leveraging Encoder Disentanglement
por: Yang, Tzu-Ting, et al.
Publicado: (2024)
por: Yang, Tzu-Ting, et al.
Publicado: (2024)
ConPCO: Preserving Phoneme Characteristics for Automatic Pronunciation Assessment Leveraging Contrastive Ordinal Regularization
por: Yan, Bi-Cheng, et al.
Publicado: (2024)
por: Yan, Bi-Cheng, et al.
Publicado: (2024)
Zero-Shot Text-to-Speech as Golden Speech Generator: A Systematic Framework and its Applicability in Automatic Pronunciation Assessment
por: Lo, Tien-Hong, et al.
Publicado: (2024)
por: Lo, Tien-Hong, et al.
Publicado: (2024)
Probing the Hidden Talent of ASR Foundation Models for L2 English Oral Assessment
por: Chao, Fu-An, et al.
Publicado: (2025)
por: Chao, Fu-An, et al.
Publicado: (2025)
Improved Remixing Process for Domain Adaptation-Based Speech Enhancement by Mitigating Data Imbalance in Signal-to-Noise Ratio
por: Li, Li, et al.
Publicado: (2024)
por: Li, Li, et al.
Publicado: (2024)
Mitigating Category Imbalance: Fosafer System for the Multimodal Emotion and Intent Joint Understanding Challenge
por: Wang, Honghong, et al.
Publicado: (2025)
por: Wang, Honghong, et al.
Publicado: (2025)
The NTNU System at the S&I Challenge 2025 SLA Open Track
por: Lin, Hong-Yun, et al.
Publicado: (2025)
por: Lin, Hong-Yun, et al.
Publicado: (2025)
Automating Urban Soundscape Enhancements with AI: In-situ Assessment of Quality and Restorativeness in Traffic-Exposed Residential Areas
por: Lam, Bhan, et al.
Publicado: (2024)
por: Lam, Bhan, et al.
Publicado: (2024)
Speech-Aware Neural Diarization with Encoder-Decoder Attractor Guided by Attention Constraints
por: Lee, PeiYing, et al.
Publicado: (2024)
por: Lee, PeiYing, et al.
Publicado: (2024)
ConSep: a Noise- and Reverberation-Robust Speech Separation Framework by Magnitude Conditioning
por: Ho, Kuan-Hsun, et al.
Publicado: (2024)
por: Ho, Kuan-Hsun, et al.
Publicado: (2024)
What do neural networks listen to? Exploring the crucial bands in Speech Enhancement using Sinc-convolution
por: Ho, Kuan-Hsun, et al.
Publicado: (2024)
por: Ho, Kuan-Hsun, et al.
Publicado: (2024)
Optimizing Automatic Speech Assessment: W-RankSim Regularization and Hybrid Feature Fusion Strategies
por: Wu, Chung-Wen, et al.
Publicado: (2024)
por: Wu, Chung-Wen, et al.
Publicado: (2024)
Enhancing Code-Switching ASR Leveraging Non-Peaky CTC Loss and Deep Language Posterior Injection
por: Yang, Tzu-Ting, et al.
Publicado: (2024)
por: Yang, Tzu-Ting, et al.
Publicado: (2024)
HiPPO: Exploring A Novel Hierarchical Pronunciation Assessment Approach for Spoken Languages
por: Yan, Bi-Cheng, et al.
Publicado: (2025)
por: Yan, Bi-Cheng, et al.
Publicado: (2025)
A Study on Synthesizing Expressive Violin Performances: Approaches and Comparisons
por: Hung, Tzu-Yun, et al.
Publicado: (2024)
por: Hung, Tzu-Yun, et al.
Publicado: (2024)
Efficient Dialect-Aware Modeling and Conditioning for Low-Resource Taiwanese Hakka Speech Processing
por: Peng, An-Ci, et al.
Publicado: (2026)
por: Peng, An-Ci, et al.
Publicado: (2026)
Building Tailored Speech Recognizers for Japanese Speaking Assessment
por: Kubo, Yotaro, et al.
Publicado: (2025)
por: Kubo, Yotaro, et al.
Publicado: (2025)
Restorative Speech Enhancement: A Progressive Approach Using SE and Codec Modules
por: Chiang, Hsin-Tien, et al.
Publicado: (2024)
por: Chiang, Hsin-Tien, et al.
Publicado: (2024)
An Effective Context-Balanced Adaptation Approach for Long-Tailed Speech Recognition
por: Wang, Yi-Cheng, et al.
Publicado: (2024)
por: Wang, Yi-Cheng, et al.
Publicado: (2024)
SoundCollage: Automated Discovery of New Classes in Audio Datasets
por: Choi, Ryuhaerang, et al.
Publicado: (2024)
por: Choi, Ryuhaerang, et al.
Publicado: (2024)
Beyond Modality Limitations: A Unified MLLM Approach to Automated Speaking Assessment with Effective Curriculum Learning
por: Fang, Yu-Hsuan, et al.
Publicado: (2025)
por: Fang, Yu-Hsuan, et al.
Publicado: (2025)
Speaking from Coarse to Fine: Improving Neural Codec Language Model via Multi-Scale Speech Coding and Generation
por: Guo, Haohan, et al.
Publicado: (2024)
por: Guo, Haohan, et al.
Publicado: (2024)
Voice Cloning for Dysarthric Speech Synthesis: Addressing Data Scarcity in Speech-Language Pathology
por: Moell, Birger, et al.
Publicado: (2025)
por: Moell, Birger, et al.
Publicado: (2025)
Overcoming Data Scarcity in Multi-Dialectal Arabic ASR via Whisper Fine-Tuning
por: Özyilmaz, Ömer Tarik, et al.
Publicado: (2025)
por: Özyilmaz, Ömer Tarik, et al.
Publicado: (2025)
Joint Multi-scale Cross-lingual Speaking Style Transfer with Bidirectional Attention Mechanism for Automatic Dubbing
por: Li, Jingbei, et al.
Publicado: (2023)
por: Li, Jingbei, et al.
Publicado: (2023)
JIS: A Speech Corpus of Japanese Idol Speakers with Various Speaking Styles
por: Kondo, Yuto, et al.
Publicado: (2025)
por: Kondo, Yuto, et al.
Publicado: (2025)
ValSub: Subsampling Validation Data to Mitigate Forgetting during ASR Personalization
por: Mehmood, Haaris, et al.
Publicado: (2025)
por: Mehmood, Haaris, et al.
Publicado: (2025)
1st Place Solution to Odyssey Emotion Recognition Challenge Task1: Tackling Class Imbalance Problem
por: Chen, Mingjie, et al.
Publicado: (2024)
por: Chen, Mingjie, et al.
Publicado: (2024)
Investigating Effective Speaker Property Privacy Protection in Federated Learning for Speech Emotion Recognition
por: Tan, Chao, et al.
Publicado: (2024)
por: Tan, Chao, et al.
Publicado: (2024)
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval
por: Sun, Haoqin, et al.
Publicado: (2025)
por: Sun, Haoqin, et al.
Publicado: (2025)
Rethinking Mean Opinion Scores in Speech Quality Assessment: Aggregation through Quantized Distribution Fitting
por: Kondo, Yuto, et al.
Publicado: (2025)
por: Kondo, Yuto, et al.
Publicado: (2025)
Leveraging Mamba with Full-Face Vision for Audio-Visual Speech Enhancement
por: Chao, Rong, et al.
Publicado: (2025)
por: Chao, Rong, et al.
Publicado: (2025)
Multimodal Assessment of Speech Impairment in ALS Using Audio-Visual and Machine Learning Approaches
por: Pierotti, Francesco, et al.
Publicado: (2025)
por: Pierotti, Francesco, et al.
Publicado: (2025)
Serial-Parallel Dual-Path Architecture for Speaking Style Recognition
por: Li, Guojian, et al.
Publicado: (2025)
por: Li, Guojian, et al.
Publicado: (2025)
A Study on Incorporating Whisper for Robust Speech Assessment
por: Zezario, Ryandhimas E., et al.
Publicado: (2023)
por: Zezario, Ryandhimas E., et al.
Publicado: (2023)
ATRI: Mitigating Multilingual Audio Text Retrieval Inconsistencies by Reducing Data Distribution Errors
por: Yin, Yuguo, et al.
Publicado: (2025)
por: Yin, Yuguo, et al.
Publicado: (2025)
Ejemplares similares
-
An Effective Strategy for Modeling Score Ordinality and Non-uniform Intervals in Automated Speaking Assessment
por: Lo, Tien-Hong, et al.
Publicado: (2025) -
Mitigating Data Imbalance in Automated Speaking Assessment
por: Tsai, Fong-Chun, et al.
Publicado: (2025) -
A Novel Data Augmentation Approach for Automatic Speaking Assessment on Opinion Expressions
por: Wang, Chung-Chun, et al.
Publicado: (2025) -
Advancing Automated Speaking Assessment Leveraging Multifaceted Relevance and Grammar Information
por: Lu, Hao-Chien, et al.
Publicado: (2025) -
An Effective Mixture-Of-Experts Approach For Code-Switching Speech Recognition Leveraging Encoder Disentanglement
por: Yang, Tzu-Ting, et al.
Publicado: (2024)