Quantifying the effect of speech pathology on automatic and human speaker verification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Halpern, Bence Mark, Tienkamp, Thomas, Huang, Wen-Chin, Violeta, Lester Phillip, Rebernik, Teja, de Visscher, Sebastiaan, Witjes, Max, Wieling, Martijn, Abur, Defne, Toda, Tomoki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards explainable reference-free speech intelligibility evaluation of people with pathological speech
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2026)
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2026)
Relationship between objective and subjective perceptual measures of speech in individuals with head and neck cancer
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025)
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025)
XPPG-PCA: Reference-free automatic speech severity evaluation with principal components
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025)
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025)
PathBench: Speech Intelligibility Benchmark for Automatic Pathological Speech Assessment
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2026)
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2026)
Articulatory–kinematic changes in speech following surgical treatment for oral or oropharyngeal cancer: A systematic review
von: Thomas B. Tienkamp, et al.
Veröffentlicht: (2024)
von: Thomas B. Tienkamp, et al.
Veröffentlicht: (2024)
Articulatory clarity and variability before and after surgery for tongue cancer
von: Tienkamp, Thomas, et al.
Veröffentlicht: (2025)
von: Tienkamp, Thomas, et al.
Veröffentlicht: (2025)
Reference-free automatic speech severity evaluation using acoustic unit language modelling
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025)
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025)
Improving Automatic Speech Recognition for Speakers Treated for Oral Cancer using Data Augmentation and LLM Error Correction
von: Folkertsma, Hidde, et al.
Veröffentlicht: (2026)
von: Folkertsma, Hidde, et al.
Veröffentlicht: (2026)
Serenade: A Singing Style Conversion Framework Based On Audio Infilling
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2025)
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2025)
Multi-speaker Text-to-speech Training with Speaker Anonymized Data
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2024)
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2024)
Electrolaryngeal Speech Intelligibility Enhancement Through Robust Linguistic Encoders
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2023)
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2023)
An Extensive Analysis of the Singing Voice Conversion Challenge 2025 Evaluation Results
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2025)
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2025)
Tandem spoofing-robust automatic speaker verification based on time-domain embeddings
von: Weizman, Avishai, et al.
Veröffentlicht: (2024)
von: Weizman, Avishai, et al.
Veröffentlicht: (2024)
Maxillectomy patients' speech and performance of contemporary speaker‐independent automatic speech recognition platforms in Japanese
von: Ahmed Sameir Mohamed Ali, et al.
Veröffentlicht: (2024)
von: Ahmed Sameir Mohamed Ali, et al.
Veröffentlicht: (2024)
Perceptual implications of automatic anonymization in pathological speech
von: Arasteh, Soroosh Tayebi, et al.
Veröffentlicht: (2025)
von: Arasteh, Soroosh Tayebi, et al.
Veröffentlicht: (2025)
An explainable approach to detect case law on housing and eviction issues within the HUDOC database
von: Mohammadi, Mohammad, et al.
Veröffentlicht: (2024)
von: Mohammadi, Mohammad, et al.
Veröffentlicht: (2024)
Advancing Electrolaryngeal Speech Enhancement Through Speech-Text Representation Learning
von: Ma, Ding, et al.
Veröffentlicht: (2026)
von: Ma, Ding, et al.
Veröffentlicht: (2026)
Text adaptation for speaker verification with speaker-text factorized embeddings
von: Yang, Yexin, et al.
Veröffentlicht: (2025)
von: Yang, Yexin, et al.
Veröffentlicht: (2025)
Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
A Preliminary Investigation on Flexible Singing Voice Synthesis Through Decomposed Framework with Inferrable Features
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2024)
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2024)
Thinking in cocktail party: Chain-of-Thought and reinforcement learning for target speaker automatic speech recognition
von: Zhang, Yiru, et al.
Veröffentlicht: (2025)
von: Zhang, Yiru, et al.
Veröffentlicht: (2025)
Triage knowledge distillation for speaker verification
von: Kim, Ju-ho, et al.
Veröffentlicht: (2026)
von: Kim, Ju-ho, et al.
Veröffentlicht: (2026)
MOS-Bench: Benchmarking Generalization Abilities of Subjective Speech Quality Assessment Models
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2024)
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2024)
SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2025)
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2025)
Curriculum learning for self-supervised speaker verification
von: Heo, Hee-Soo, et al.
Veröffentlicht: (2022)
von: Heo, Hee-Soo, et al.
Veröffentlicht: (2022)
Variability compensation for speaker verification with short utterances
von: Flavio J. Reyes-Díaz
Veröffentlicht: (2016)
von: Flavio J. Reyes-Díaz
Veröffentlicht: (2016)
An insight to the automatic categorization of speakers according to sex and its application to the detection of voice pathologies: A comparative study
von: Jorge Andrés Gómez-García
Veröffentlicht: (2016)
von: Jorge Andrés Gómez-García
Veröffentlicht: (2016)
Improving speaker verification robustness with synthetic emotional utterances
von: Koditala, Nikhil Kumar, et al.
Veröffentlicht: (2024)
von: Koditala, Nikhil Kumar, et al.
Veröffentlicht: (2024)
On the influence of language similarity in non-target speaker verification trials
von: Reuter, Paul M., et al.
Veröffentlicht: (2025)
von: Reuter, Paul M., et al.
Veröffentlicht: (2025)
Music Similarity Representation Learning Focusing on Individual Instruments with Source Separation and Human Preference
von: Imamura, Takehiro, et al.
Veröffentlicht: (2025)
von: Imamura, Takehiro, et al.
Veröffentlicht: (2025)
Analysis and Extension of Noisy-target Training for Unsupervised Target Signal Enhancement
von: Fujimura, Takuya, et al.
Veröffentlicht: (2025)
von: Fujimura, Takuya, et al.
Veröffentlicht: (2025)
Investigation of perceptual music similarity focusing on each instrumental part
von: Hashizume, Yuka, et al.
Veröffentlicht: (2025)
von: Hashizume, Yuka, et al.
Veröffentlicht: (2025)
PMF-CEC: Phoneme-augmented Multimodal Fusion for Context-aware ASR Error Correction with Error-specific Selective Decoding
von: He, Jiajun, et al.
Veröffentlicht: (2025)
von: He, Jiajun, et al.
Veröffentlicht: (2025)
QHARMA-GAN: Quasi-Harmonic Neural Vocoder based on Autoregressive Moving Average Model
von: Chen, Shaowen, et al.
Veröffentlicht: (2025)
von: Chen, Shaowen, et al.
Veröffentlicht: (2025)
Automatic design optimization of preference-based subjective evaluation with online learning in crowdsourcing environment
von: Yasuda, Yusuke, et al.
Veröffentlicht: (2024)
von: Yasuda, Yusuke, et al.
Veröffentlicht: (2024)
2DP-2MRC: 2-Dimensional Pointer-based Machine Reading Comprehension Method for Multimodal Moment Retrieval
von: He, Jiajun, et al.
Veröffentlicht: (2024)
von: He, Jiajun, et al.
Veröffentlicht: (2024)
Online speaker diarization of meetings guided by speech separation
von: Gruttadauria, Elio, et al.
Veröffentlicht: (2024)
von: Gruttadauria, Elio, et al.
Veröffentlicht: (2024)
Multi-channel multi-speaker transformer for speech recognition
von: Yifan, Guo, et al.
Veröffentlicht: (2026)
von: Yifan, Guo, et al.
Veröffentlicht: (2026)
Prominence-aware automatic speech recognition for conversational speech
von: Linke, Julian, et al.
Veröffentlicht: (2025)
von: Linke, Julian, et al.
Veröffentlicht: (2025)
Gradient weighting for speaker verification in extremely low Signal-to-Noise Ratio
von: Ma, Yi, et al.
Veröffentlicht: (2024)
von: Ma, Yi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards explainable reference-free speech intelligibility evaluation of people with pathological speech
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2026) -
Relationship between objective and subjective perceptual measures of speech in individuals with head and neck cancer
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025) -
XPPG-PCA: Reference-free automatic speech severity evaluation with principal components
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2025) -
PathBench: Speech Intelligibility Benchmark for Automatic Pathological Speech Assessment
von: Halpern, Bence Mark, et al.
Veröffentlicht: (2026) -
Articulatory–kinematic changes in speech following surgical treatment for oral or oropharyngeal cancer: A systematic review
von: Thomas B. Tienkamp, et al.
Veröffentlicht: (2024)