Multivariate Probabilistic Assessment of Speech Quality
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cumlin, Fredrik, Liang, Xinyu, Ungureanu, Victor, Reddy, Chandan K. A., Schüldt, Christian, Chatterjee, Saikat |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Impairments are Clustered in Latents of Deep Neural Network-based Speech Quality Models
von: Cumlin, Fredrik, et al.
Veröffentlicht: (2025)
von: Cumlin, Fredrik, et al.
Veröffentlicht: (2025)
Selection of Layers from Self-supervised Learning Models for Predicting Mean-Opinion-Score of Speech
von: Liang, Xinyu, et al.
Veröffentlicht: (2025)
von: Liang, Xinyu, et al.
Veröffentlicht: (2025)
SA-SSL-MOS: Self-supervised Learning MOS Prediction with Spectral Augmentation for Generalized Multi-Rate Speech Assessment
von: Cao, Fengyuan, et al.
Veröffentlicht: (2026)
von: Cao, Fengyuan, et al.
Veröffentlicht: (2026)
Leveraging LLMs for Scalable Non-intrusive Speech Quality Assessment
von: Cumlin, Fredrik, et al.
Veröffentlicht: (2025)
von: Cumlin, Fredrik, et al.
Veröffentlicht: (2025)
Rho-Perfect: Correlation Ceiling For Subjective Evaluation Datasets
von: Cumlin, Fredrik
Veröffentlicht: (2026)
von: Cumlin, Fredrik
Veröffentlicht: (2026)
SCOREQ: Speech Quality Assessment with Contrastive Regression
von: Ragano, Alessandro, et al.
Veröffentlicht: (2024)
von: Ragano, Alessandro, et al.
Veröffentlicht: (2024)
HighRateMOS: Sampling-Rate Aware Modeling for Speech Quality Assessment
von: Ren, Wenze, et al.
Veröffentlicht: (2025)
von: Ren, Wenze, et al.
Veröffentlicht: (2025)
A Pre-training Framework that Encodes Noise Information for Speech Quality Assessment
von: Sultana, Subrina, et al.
Veröffentlicht: (2024)
von: Sultana, Subrina, et al.
Veröffentlicht: (2024)
Speech Quality-Based Localization of Low-Quality Speech and Text-to-Speech Synthesis Artefacts
von: Kuhlmann, Michael, et al.
Veröffentlicht: (2026)
von: Kuhlmann, Michael, et al.
Veröffentlicht: (2026)
Self-Supervised Speech Quality Assessment (S3QA): Leveraging Speech Foundation Models for a Scalable Speech Quality Metric
von: Ogg, Mattson, et al.
Veröffentlicht: (2025)
von: Ogg, Mattson, et al.
Veröffentlicht: (2025)
French Listening Tests for the Assessment of Intelligibility, Quality, and Identity of Body-Conducted Speech Enhancement
von: Joubaud, Thomas, et al.
Veröffentlicht: (2025)
von: Joubaud, Thomas, et al.
Veröffentlicht: (2025)
MOS-Bias: From Hidden Gender Bias to Gender-Aware Speech Quality Assessment
von: Ren, Wenze, et al.
Veröffentlicht: (2026)
von: Ren, Wenze, et al.
Veröffentlicht: (2026)
Universal Preference-Score-based Pairwise Speech Quality Assessment
von: Shi, Yu-Fei, et al.
Veröffentlicht: (2025)
von: Shi, Yu-Fei, et al.
Veröffentlicht: (2025)
MambaRate: Speech Quality Assessment Across Different Sampling Rates
von: Kakoulidis, Panos, et al.
Veröffentlicht: (2025)
von: Kakoulidis, Panos, et al.
Veröffentlicht: (2025)
Towards Explainable Spoofed Speech Attribution and Detection:a Probabilistic Approach for Characterizing Speech Synthesizer Components
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2025)
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2025)
Speech Quality Embeddings for Improved Detection and Classification of Degradations in Speech Signals
von: Kuhlmann, Michael, et al.
Veröffentlicht: (2026)
von: Kuhlmann, Michael, et al.
Veröffentlicht: (2026)
Unifying Listener Scoring Scales: Comparison Learning Framework for Speech Quality Assessment and Continuous Speech Emotion Recognition
von: Hu, Cheng-Hung, et al.
Veröffentlicht: (2025)
von: Hu, Cheng-Hung, et al.
Veröffentlicht: (2025)
Distillation and Pruning for Scalable Self-Supervised Representation-Based Speech Quality Assessment
von: Stahl, Benjamin, et al.
Veröffentlicht: (2025)
von: Stahl, Benjamin, et al.
Veröffentlicht: (2025)
Advancing Speech Quality Assessment Through Scientific Challenges and Open-source Activities
von: Huang, Wen-Chin
Veröffentlicht: (2025)
von: Huang, Wen-Chin
Veröffentlicht: (2025)
Improving Speech Enhancement with Multi-Metric Supervision from Learned Quality Assessment
von: Wang, Wei, et al.
Veröffentlicht: (2025)
von: Wang, Wei, et al.
Veröffentlicht: (2025)
MOS-Bench: Benchmarking Generalization Abilities of Subjective Speech Quality Assessment Models
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2024)
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2024)
Calibration-Reasoning Framework for Descriptive Speech Quality Assessment
von: Kostenok, Elizaveta, et al.
Veröffentlicht: (2026)
von: Kostenok, Elizaveta, et al.
Veröffentlicht: (2026)
On the Relation Between Speech Quality and Quantized Latent Representations of Neural Codecs
von: Halimeh, Mhd Modar, et al.
Veröffentlicht: (2025)
von: Halimeh, Mhd Modar, et al.
Veröffentlicht: (2025)
Towards Sub-millisecond Latency Real-Time Speech Enhancement Models on Hearables
von: Dementyev, Artem, et al.
Veröffentlicht: (2024)
von: Dementyev, Artem, et al.
Veröffentlicht: (2024)
Towards Frame-level Quality Predictions of Synthetic Speech
von: Kuhlmann, Michael, et al.
Veröffentlicht: (2025)
von: Kuhlmann, Michael, et al.
Veröffentlicht: (2025)
Rethinking Mean Opinion Scores in Speech Quality Assessment: Aggregation through Quantized Distribution Fitting
von: Kondo, Yuto, et al.
Veröffentlicht: (2025)
von: Kondo, Yuto, et al.
Veröffentlicht: (2025)
HYFuse: Aligning Heterogeneous Speech Pre-Trained Representations in Hyperbolic Space for Speech Emotion Recognition
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2025)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2025)
Quality Assessment of Noisy and Enhanced Speech with Limited Data: UWB-NTIS System for VoiceMOS 2024
von: Kunešová, Marie, et al.
Veröffentlicht: (2025)
von: Kunešová, Marie, et al.
Veröffentlicht: (2025)
NOMAD: Unsupervised Learning of Perceptual Embeddings for Speech Enhancement and Non-matching Reference Audio Quality Assessment
von: Ragano, Alessandro, et al.
Veröffentlicht: (2023)
von: Ragano, Alessandro, et al.
Veröffentlicht: (2023)
An Explainable Probabilistic Attribute Embedding Approach for Spoofed Speech Characterization
von: Chhibber, Manasi, et al.
Veröffentlicht: (2024)
von: Chhibber, Manasi, et al.
Veröffentlicht: (2024)
From Evaluation to Optimization: Neural Speech Assessment for Downstream Applications
von: Tsao, Yu
Veröffentlicht: (2025)
von: Tsao, Yu
Veröffentlicht: (2025)
ProSE: Diffusion Priors for Speech Enhancement
von: Kumar, Sonal, et al.
Veröffentlicht: (2025)
von: Kumar, Sonal, et al.
Veröffentlicht: (2025)
Low Bitrate High-Quality RVQGAN-based Discrete Speech Tokenizer
von: Shechtman, Slava, et al.
Veröffentlicht: (2024)
von: Shechtman, Slava, et al.
Veröffentlicht: (2024)
Voice Cloning for Dysarthric Speech Synthesis: Addressing Data Scarcity in Speech-Language Pathology
von: Moell, Birger, et al.
Veröffentlicht: (2025)
von: Moell, Birger, et al.
Veröffentlicht: (2025)
Learning-based A Posteriori Speech Presence Probability Estimation and Applications
von: Tao, Shuai, et al.
Veröffentlicht: (2025)
von: Tao, Shuai, et al.
Veröffentlicht: (2025)
Augmenting Open-Vocabulary Dysarthric Speech Assessment with Human Perceptual Supervision
von: Jia, Kaimeng, et al.
Veröffentlicht: (2025)
von: Jia, Kaimeng, et al.
Veröffentlicht: (2025)
Textless Streaming Speech-to-Speech Translation using Semantic Speech Tokens
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2024)
von: Zhao, Jinzheng, et al.
Veröffentlicht: (2024)
AttentiveMOS: A Lightweight Attention-Only Model for Speech Quality Prediction
von: Kibria, Imran E, et al.
Veröffentlicht: (2024)
von: Kibria, Imran E, et al.
Veröffentlicht: (2024)
StuPASE: Towards Low-Hallucination Studio-Quality Generative Speech Enhancement
von: Rong, Xiaobin, et al.
Veröffentlicht: (2026)
von: Rong, Xiaobin, et al.
Veröffentlicht: (2026)
SpeechLLM-as-Judges: Towards General and Interpretable Speech Quality Evaluation
von: Wang, Hui, et al.
Veröffentlicht: (2025)
von: Wang, Hui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Impairments are Clustered in Latents of Deep Neural Network-based Speech Quality Models
von: Cumlin, Fredrik, et al.
Veröffentlicht: (2025) -
Selection of Layers from Self-supervised Learning Models for Predicting Mean-Opinion-Score of Speech
von: Liang, Xinyu, et al.
Veröffentlicht: (2025) -
SA-SSL-MOS: Self-supervised Learning MOS Prediction with Spectral Augmentation for Generalized Multi-Rate Speech Assessment
von: Cao, Fengyuan, et al.
Veröffentlicht: (2026) -
Leveraging LLMs for Scalable Non-intrusive Speech Quality Assessment
von: Cumlin, Fredrik, et al.
Veröffentlicht: (2025) -
Rho-Perfect: Correlation Ceiling For Subjective Evaluation Datasets
von: Cumlin, Fredrik
Veröffentlicht: (2026)