A Novel Data Augmentation Approach for Automatic Speaking Assessment on Opinion Expressions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Chung-Chun, Lin, Jhen-Ke, Lu, Hao-Chien, Lin, Hong-Yun, Chen, Berlin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Advancing Automated Speaking Assessment Leveraging Multifaceted Relevance and Grammar Information
von: Lu, Hao-Chien, et al.
Veröffentlicht: (2025)
von: Lu, Hao-Chien, et al.
Veröffentlicht: (2025)
Acoustically Precise Hesitation Tagging Is Essential for End-to-End Verbatim Transcription Systems
von: Lin, Jhen-Ke, et al.
Veröffentlicht: (2025)
von: Lin, Jhen-Ke, et al.
Veröffentlicht: (2025)
The NTNU System at the S&I Challenge 2025 SLA Open Track
von: Lin, Hong-Yun, et al.
Veröffentlicht: (2025)
von: Lin, Hong-Yun, et al.
Veröffentlicht: (2025)
Optimizing Automatic Speech Assessment: W-RankSim Regularization and Hybrid Feature Fusion Strategies
von: Wu, Chung-Wen, et al.
Veröffentlicht: (2024)
von: Wu, Chung-Wen, et al.
Veröffentlicht: (2024)
An Effective Automated Speaking Assessment Approach to Mitigating Data Scarcity and Imbalanced Distribution
von: Lo, Tien-Hong, et al.
Veröffentlicht: (2024)
von: Lo, Tien-Hong, et al.
Veröffentlicht: (2024)
Fine-Tuning Large Multimodal Models for Automatic Pronunciation Assessment
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
Building Tailored Speech Recognizers for Japanese Speaking Assessment
von: Kubo, Yotaro, et al.
Veröffentlicht: (2025)
von: Kubo, Yotaro, et al.
Veröffentlicht: (2025)
ConPCO: Preserving Phoneme Characteristics for Automatic Pronunciation Assessment Leveraging Contrastive Ordinal Regularization
von: Yan, Bi-Cheng, et al.
Veröffentlicht: (2024)
von: Yan, Bi-Cheng, et al.
Veröffentlicht: (2024)
TG-ASR: Translation-Guided Learning with Parallel Gated Cross Attention for Low-Resource Automatic Speech Recognition
von: Yang, Cheng-Yeh, et al.
Veröffentlicht: (2026)
von: Yang, Cheng-Yeh, et al.
Veröffentlicht: (2026)
An Effective Strategy for Modeling Score Ordinality and Non-uniform Intervals in Automated Speaking Assessment
von: Lo, Tien-Hong, et al.
Veröffentlicht: (2025)
von: Lo, Tien-Hong, et al.
Veröffentlicht: (2025)
Effective Noise-aware Data Simulation for Domain-adaptive Speech Enhancement Leveraging Dynamic Stochastic Perturbation
von: Wang, Chien-Chun, et al.
Veröffentlicht: (2024)
von: Wang, Chien-Chun, et al.
Veröffentlicht: (2024)
Universal Robust Speech Adaptation for Cross-Domain Speech Recognition and Enhancement
von: Wang, Chien-Chun, et al.
Veröffentlicht: (2026)
von: Wang, Chien-Chun, et al.
Veröffentlicht: (2026)
CIF-T: A Novel CIF-based Transducer Architecture for Automatic Speech Recognition
von: Zhang, Tian-Hao, et al.
Veröffentlicht: (2023)
von: Zhang, Tian-Hao, et al.
Veröffentlicht: (2023)
Retrieval-Augmented Speech Recognition Approach for Domain Challenges
von: Shen, Peng, et al.
Veröffentlicht: (2025)
von: Shen, Peng, et al.
Veröffentlicht: (2025)
Dynamic Data Pruning for Automatic Speech Recognition
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
Automatic Proficiency Assessment in L2 English Learners
von: Mohammadi, Armita, et al.
Veröffentlicht: (2025)
von: Mohammadi, Armita, et al.
Veröffentlicht: (2025)
AV2Wav: Diffusion-Based Re-synthesis from Continuous Self-supervised Features for Audio-Visual Speech Enhancement
von: Chou, Ju-Chieh, et al.
Veröffentlicht: (2023)
von: Chou, Ju-Chieh, et al.
Veröffentlicht: (2023)
Fusion of Discrete Representations and Self-Augmented Representations for Multilingual Automatic Speech Recognition
von: Wang, Shih-heng, et al.
Veröffentlicht: (2024)
von: Wang, Shih-heng, et al.
Veröffentlicht: (2024)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary Prediction and Activation
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
LLM-Driven Multimodal Opinion Expression Identification
von: Jia, Bonian, et al.
Veröffentlicht: (2024)
von: Jia, Bonian, et al.
Veröffentlicht: (2024)
Identifying Speaker Information in Feed-Forward Layers of Self-Supervised Speech Transformers
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2025)
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2025)
Factor-Conditioned Speaking-Style Captioning
von: Ando, Atsushi, et al.
Veröffentlicht: (2024)
von: Ando, Atsushi, et al.
Veröffentlicht: (2024)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
von: Shi, Hao, et al.
Veröffentlicht: (2024)
von: Shi, Hao, et al.
Veröffentlicht: (2024)
Speech-Aware Neural Diarization with Encoder-Decoder Attractor Guided by Attention Constraints
von: Lee, PeiYing, et al.
Veröffentlicht: (2024)
von: Lee, PeiYing, et al.
Veröffentlicht: (2024)
Probing the Hidden Talent of ASR Foundation Models for L2 English Oral Assessment
von: Chao, Fu-An, et al.
Veröffentlicht: (2025)
von: Chao, Fu-An, et al.
Veröffentlicht: (2025)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
Automatic Speech Recognition for Biomedical Data in Bengali Language
von: Kabir, Shariar, et al.
Veröffentlicht: (2024)
von: Kabir, Shariar, et al.
Veröffentlicht: (2024)
An Effective Mixture-Of-Experts Approach For Code-Switching Speech Recognition Leveraging Encoder Disentanglement
von: Yang, Tzu-Ting, et al.
Veröffentlicht: (2024)
von: Yang, Tzu-Ting, et al.
Veröffentlicht: (2024)
Channel-Aware Domain-Adaptive Generative Adversarial Network for Robust Speech Recognition
von: Wang, Chien-Chun, et al.
Veröffentlicht: (2024)
von: Wang, Chien-Chun, et al.
Veröffentlicht: (2024)
Exploring Procedural Data Generation for Automatic Acoustic Guitar Fingerpicking Transcription
von: Murgul, Sebastian, et al.
Veröffentlicht: (2025)
von: Murgul, Sebastian, et al.
Veröffentlicht: (2025)
oboVox Far Field Speaker Recognition: A Novel Data Augmentation Approach with Pretrained Models
von: Dip, Muhammad Sudipto Siam, et al.
Veröffentlicht: (2024)
von: Dip, Muhammad Sudipto Siam, et al.
Veröffentlicht: (2024)
Word Level Timestamp Generation for Automatic Speech Recognition and Translation
von: Hu, Ke, et al.
Veröffentlicht: (2025)
von: Hu, Ke, et al.
Veröffentlicht: (2025)
Mitigating Data Imbalance in Automated Speaking Assessment
von: Tsai, Fong-Chun, et al.
Veröffentlicht: (2025)
von: Tsai, Fong-Chun, et al.
Veröffentlicht: (2025)
SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech Recognition
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2024)
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2024)
Emotion-Coherent Speech Data Augmentation and Self-Supervised Contrastive Style Training for Enhancing Kids's Story Speech Synthesis
von: Chung, Raymond
Veröffentlicht: (2026)
von: Chung, Raymond
Veröffentlicht: (2026)
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style
von: Kang, Wonjune, et al.
Veröffentlicht: (2025)
von: Kang, Wonjune, et al.
Veröffentlicht: (2025)
Automatic Assessment of Oral Reading Accuracy for Reading Diagnostics
von: Molenaar, Bo, et al.
Veröffentlicht: (2023)
von: Molenaar, Bo, et al.
Veröffentlicht: (2023)
An Effective Context-Balanced Adaptation Approach for Long-Tailed Speech Recognition
von: Wang, Yi-Cheng, et al.
Veröffentlicht: (2024)
von: Wang, Yi-Cheng, et al.
Veröffentlicht: (2024)
Can Large Audio-Language Models Truly Hear? Tackling Hallucinations with Multi-Task Assessment and Stepwise Audio Reasoning
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2024)
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2024)
Voice Conversion for Lombard Speaking Style with Implicit and Explicit Acoustic Feature Conditioning
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2025)
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Advancing Automated Speaking Assessment Leveraging Multifaceted Relevance and Grammar Information
von: Lu, Hao-Chien, et al.
Veröffentlicht: (2025) -
Acoustically Precise Hesitation Tagging Is Essential for End-to-End Verbatim Transcription Systems
von: Lin, Jhen-Ke, et al.
Veröffentlicht: (2025) -
The NTNU System at the S&I Challenge 2025 SLA Open Track
von: Lin, Hong-Yun, et al.
Veröffentlicht: (2025) -
Optimizing Automatic Speech Assessment: W-RankSim Regularization and Hybrid Feature Fusion Strategies
von: Wu, Chung-Wen, et al.
Veröffentlicht: (2024) -
An Effective Automated Speaking Assessment Approach to Mitigating Data Scarcity and Imbalanced Distribution
von: Lo, Tien-Hong, et al.
Veröffentlicht: (2024)