Leveraging Multiple Speech Enhancers for Non-Intrusive Intelligibility Prediction for Hearing-Impaired Listeners
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cao, Boxuan, Li, Linkai, Yu, Hanlin, Mo, Changgeng, Zhou, Haoshuai, Wang, Shan Xiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unveiling the Best Practices for Applying Speech Foundation Models to Speech Intelligibility Prediction for Hearing-Impaired People
von: Zhou, Haoshuai, et al.
Veröffentlicht: (2025)
von: Zhou, Haoshuai, et al.
Veröffentlicht: (2025)
No Audiogram: Leveraging Existing Scores for Personalized Speech Intelligibility Prediction
von: Zhou, Haoshuai, et al.
Veröffentlicht: (2025)
von: Zhou, Haoshuai, et al.
Veröffentlicht: (2025)
Reference-aware SFM layers for intrusive intelligibility prediction
von: Yu, Hanlin, et al.
Veröffentlicht: (2025)
von: Yu, Hanlin, et al.
Veröffentlicht: (2025)
Non-Intrusive Binaural Speech Intelligibility Prediction Using Mamba for Hearing-Impaired Listeners
von: Yamamoto, Katsuhiko, et al.
Veröffentlicht: (2025)
von: Yamamoto, Katsuhiko, et al.
Veröffentlicht: (2025)
A Multi-stage Low-latency Enhancement System for Hearing Aids
von: Ouyang, Chengwei, et al.
Veröffentlicht: (2025)
von: Ouyang, Chengwei, et al.
Veröffentlicht: (2025)
Feature Importance across Domains for Improving Non-Intrusive Speech Intelligibility Prediction in Hearing Aids
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2025)
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2025)
Non-Intrusive Speech Intelligibility Prediction for Hearing-Impaired Users using Intermediate ASR Features and Human Memory Models
von: Mogridge, Rhiannon, et al.
Veröffentlicht: (2024)
von: Mogridge, Rhiannon, et al.
Veröffentlicht: (2024)
Frame-Aligned Fusion of Canary and WavLM for Non-Intrusive Intelligibility Prediction of Hearing-Aid-Processed Speech
von: Nakazawa, Kazushi
Veröffentlicht: (2026)
von: Nakazawa, Kazushi
Veröffentlicht: (2026)
Non-Intrusive Intelligibility Prediction for Hearing Aids: Recent Advances, Trends, and Challenges
von: Zezario, Ryandhimas E.
Veröffentlicht: (2025)
von: Zezario, Ryandhimas E.
Veröffentlicht: (2025)
Non-Intrusive Speech Intelligibility Prediction for Hearing Aids using Whisper and Metadata
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
A Study on Zero-Shot Non-Intrusive Speech Intelligibility for Hearing Aids Using Large Language Models
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2025)
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2025)
LIWhiz: A Non-Intrusive Lyric Intelligibility Prediction System for the Cadenza Challenge
von: Shekar, Ram C. M. C., et al.
Veröffentlicht: (2025)
von: Shekar, Ram C. M. C., et al.
Veröffentlicht: (2025)
Neural Vocoders as Speech Enhancers
von: Li, Andong, et al.
Veröffentlicht: (2025)
von: Li, Andong, et al.
Veröffentlicht: (2025)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
von: Chung, Soo-Whan, et al.
Veröffentlicht: (2025)
von: Chung, Soo-Whan, et al.
Veröffentlicht: (2025)
WhiSQA: Non-Intrusive Speech Quality Prediction Using Whisper Encoder Features
von: Close, George, et al.
Veröffentlicht: (2025)
von: Close, George, et al.
Veröffentlicht: (2025)
MAGE: A Coarse-to-Fine Speech Enhancer with Masked Generative Model
von: Pham, The Hieu, et al.
Veröffentlicht: (2025)
von: Pham, The Hieu, et al.
Veröffentlicht: (2025)
ZipEnhancer: Dual-Path Down-Up Sampling-based Zipformer for Monaural Speech Enhancement
von: Wang, Haoxu, et al.
Veröffentlicht: (2025)
von: Wang, Haoxu, et al.
Veröffentlicht: (2025)
Listen First, Then Answer: Timestamp-Grounded Speech Reasoning
von: Jeong, Jihoon, et al.
Veröffentlicht: (2026)
von: Jeong, Jihoon, et al.
Veröffentlicht: (2026)
Evaluating Speech Enhancement Systems Through Listening Effort
von: Gelderblom, Femke B., et al.
Veröffentlicht: (2024)
von: Gelderblom, Femke B., et al.
Veröffentlicht: (2024)
GESI: Gammachirp Envelope Similarity Index for Predicting Intelligibility of Simulated Hearing Loss Sounds
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2023)
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2023)
Multi-Task Pseudo-Label Learning for Non-Intrusive Speech Quality Assessment Model
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
EventTrojan: Manipulating Non-Intrusive Speech Quality Assessment via Imperceptible Events
von: Ren, Ying, et al.
Veröffentlicht: (2023)
von: Ren, Ying, et al.
Veröffentlicht: (2023)
Using Speech Foundational Models in Loss Functions for Hearing Aid Speech Enhancement
von: Sutherland, Robert, et al.
Veröffentlicht: (2024)
von: Sutherland, Robert, et al.
Veröffentlicht: (2024)
Unifying Listener Scoring Scales: Comparison Learning Framework for Speech Quality Assessment and Continuous Speech Emotion Recognition
von: Hu, Cheng-Hung, et al.
Veröffentlicht: (2025)
von: Hu, Cheng-Hung, et al.
Veröffentlicht: (2025)
Deep Learning-based Non-Intrusive Multi-Objective Speech Assessment Model with Cross-Domain Features
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2021)
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2021)
Fish-Speech: Leveraging Large Language Models for Advanced Multilingual Text-to-Speech Synthesis
von: Liao, Shijia, et al.
Veröffentlicht: (2024)
von: Liao, Shijia, et al.
Veröffentlicht: (2024)
SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing
von: Zhang, Hanlin, et al.
Veröffentlicht: (2026)
von: Zhang, Hanlin, et al.
Veröffentlicht: (2026)
Generating Speakers by Prompting Listener Impressions for Pre-trained Multi-Speaker Text-to-Speech Systems
von: Chen, Zhengyang, et al.
Veröffentlicht: (2024)
von: Chen, Zhengyang, et al.
Veröffentlicht: (2024)
Developing an AI-Guided Assistant Device for the Deaf and Hearing Impaired
von: Jiayu, et al.
Veröffentlicht: (2025)
von: Jiayu, et al.
Veröffentlicht: (2025)
Creating Personalized Synthetic Voices from Articulation Impaired Speech Using Augmented Reconstruction Loss
von: Tian, Yusheng, et al.
Veröffentlicht: (2024)
von: Tian, Yusheng, et al.
Veröffentlicht: (2024)
Towards Environmental Preference Based Speech Enhancement For Individualised Multi-Modal Hearing Aids
von: Kirton-Wingate, Jasper, et al.
Veröffentlicht: (2024)
von: Kirton-Wingate, Jasper, et al.
Veröffentlicht: (2024)
A Scalable Pipeline for Enabling Non-Verbal Speech Generation and Understanding
von: Ye, Runchuan, et al.
Veröffentlicht: (2025)
von: Ye, Runchuan, et al.
Veröffentlicht: (2025)
Hearing Health in Home Healthcare: Leveraging LLMs for Illness Scoring and ALMs for Vocal Biomarker Extraction
von: Chen, Yu-Wen, et al.
Veröffentlicht: (2025)
von: Chen, Yu-Wen, et al.
Veröffentlicht: (2025)
Reading to Listen at the Cocktail Party: Multi-Modal Speech Separation
von: Rahimi, Akam, et al.
Veröffentlicht: (2025)
von: Rahimi, Akam, et al.
Veröffentlicht: (2025)
Listen, Think, and Understand
von: Gong, Yuan, et al.
Veröffentlicht: (2023)
von: Gong, Yuan, et al.
Veröffentlicht: (2023)
A Perception-Based L2 Speech Intelligibility Indicator: Leveraging a Rater's Shadowing and Sequence-to-sequence Voice Conversion
von: Geng, Haopeng, et al.
Veröffentlicht: (2025)
von: Geng, Haopeng, et al.
Veröffentlicht: (2025)
Embedding-Based Intrusive Evaluation Metrics for Musical Source Separation Using MERT Representations
von: Bereuter, Paul A., et al.
Veröffentlicht: (2026)
von: Bereuter, Paul A., et al.
Veröffentlicht: (2026)
Listening Between the Lines: Synthetic Speech Detection Disregarding Verbal Content
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
von: Salvi, Davide, et al.
Veröffentlicht: (2024)
Look Once to Hear: Target Speech Hearing with Noisy Examples
von: Veluri, Bandhav, et al.
Veröffentlicht: (2024)
von: Veluri, Bandhav, et al.
Veröffentlicht: (2024)
SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics
von: Saeki, Takaaki, et al.
Veröffentlicht: (2024)
von: Saeki, Takaaki, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Unveiling the Best Practices for Applying Speech Foundation Models to Speech Intelligibility Prediction for Hearing-Impaired People
von: Zhou, Haoshuai, et al.
Veröffentlicht: (2025) -
No Audiogram: Leveraging Existing Scores for Personalized Speech Intelligibility Prediction
von: Zhou, Haoshuai, et al.
Veröffentlicht: (2025) -
Reference-aware SFM layers for intrusive intelligibility prediction
von: Yu, Hanlin, et al.
Veröffentlicht: (2025) -
Non-Intrusive Binaural Speech Intelligibility Prediction Using Mamba for Hearing-Impaired Listeners
von: Yamamoto, Katsuhiko, et al.
Veröffentlicht: (2025) -
A Multi-stage Low-latency Enhancement System for Hearing Aids
von: Ouyang, Chengwei, et al.
Veröffentlicht: (2025)