Expanding and Analyzing ODAQ -- the Open Dataset of Audio Quality
Fuente:
arXiv
Guardado en:
| Autores principales: | Dick, Sascha, Thompson, Christoph, Wu, Chih-Wei, Torcoli, Matteo, Delgado, Pablo, Williams, Phillip A., Habets, Emanuel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ODAQ: Open Dataset of Audio Quality
por: Torcoli, Matteo, et al.
Publicado: (2023)
por: Torcoli, Matteo, et al.
Publicado: (2023)
Investigating the impact of stereo processing -- a study for extending the Open Dataset of Audio Quality (ODAQ)
por: Dick, Sascha, et al.
Publicado: (2025)
por: Dick, Sascha, et al.
Publicado: (2025)
Exploring Perceptual Audio Quality Measurement on Stereo Processing Using the Open Dataset of Audio Quality
por: Delgado, Pablo M., et al.
Publicado: (2025)
por: Delgado, Pablo M., et al.
Publicado: (2025)
Navigating PESQ: Up-to-Date Versions and Open Implementations
por: Torcoli, Matteo, et al.
Publicado: (2025)
por: Torcoli, Matteo, et al.
Publicado: (2025)
On the Relation Between Speech Quality and Quantized Latent Representations of Neural Codecs
por: Halimeh, Mhd Modar, et al.
Publicado: (2025)
por: Halimeh, Mhd Modar, et al.
Publicado: (2025)
ConcateNet: Dialogue Separation Using Local And Global Feature Concatenation
por: Halimeh, Mhd Modar, et al.
Publicado: (2024)
por: Halimeh, Mhd Modar, et al.
Publicado: (2024)
Acoustic Teleportation via Disentangled Neural Audio Codec Representations
por: Grundhuber, Philipp, et al.
Publicado: (2025)
por: Grundhuber, Philipp, et al.
Publicado: (2025)
Speech Loudness in Broadcasting and Streaming
por: Torcoli, Matteo, et al.
Publicado: (2024)
por: Torcoli, Matteo, et al.
Publicado: (2024)
Sample Rate Offset Compensated Acoustic Echo Cancellation For Multi-Device Scenarios
por: Korse, Srikanth, et al.
Publicado: (2025)
por: Korse, Srikanth, et al.
Publicado: (2025)
Leveraging Discriminative Latent Representations for Conditioning GAN-Based Speech Enhancement
por: Shetu, Shrishti Saha, et al.
Publicado: (2025)
por: Shetu, Shrishti Saha, et al.
Publicado: (2025)
Neural Directional Filtering with Configurable Directivity Pattern at Inference
por: Huang, Weilong, et al.
Publicado: (2025)
por: Huang, Weilong, et al.
Publicado: (2025)
GAN-Based Multi-Microphone Spatial Target Speaker Extraction
por: Shetu, Shrishti Saha, et al.
Publicado: (2025)
por: Shetu, Shrishti Saha, et al.
Publicado: (2025)
Dynamic Slimmable Networks for Efficient Speech Separation
por: Elminshawi, Mohamed, et al.
Publicado: (2025)
por: Elminshawi, Mohamed, et al.
Publicado: (2025)
Low-Complexity Neural Wind Noise Reduction for Audio Recordings
por: Eftekhari, Hesam, et al.
Publicado: (2025)
por: Eftekhari, Hesam, et al.
Publicado: (2025)
Stereo Reproduction in the Presence of Sample Rate Offsets
por: Korse, Srikanth, et al.
Publicado: (2025)
por: Korse, Srikanth, et al.
Publicado: (2025)
VoxATtack: A Multimodal Attack on Voice Anonymization Systems
por: Aloradi, Ahmad, et al.
Publicado: (2025)
por: Aloradi, Ahmad, et al.
Publicado: (2025)
Robust Speech Activity Detection in the Presence of Singing Voice
por: Grundhuber, Philipp, et al.
Publicado: (2025)
por: Grundhuber, Philipp, et al.
Publicado: (2025)
Room Impulse Response Completion Using Signal-Prediction Diffusion Models Conditioned on Simulated Early Reflections
por: Xu, Zeyu, et al.
Publicado: (2026)
por: Xu, Zeyu, et al.
Publicado: (2026)
Training Strategies for Modality Dropout Resilient Multi-Modal Target Speaker Extraction
por: Korse, Srikanth, et al.
Publicado: (2025)
por: Korse, Srikanth, et al.
Publicado: (2025)
Comparative Analysis Of Discriminative Deep Learning-Based Noise Reduction Methods In Low SNR Scenarios
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
DeePAQ: A Perceptual Audio Quality Metric Based On Foundational Models and Weakly Supervised Learning
por: Jiang, Guanxin, et al.
Publicado: (2025)
por: Jiang, Guanxin, et al.
Publicado: (2025)
Blind Acoustic Parameter Estimation Through Task-Agnostic Embeddings Using Latent Approximations
por: Götz, Philipp, et al.
Publicado: (2024)
por: Götz, Philipp, et al.
Publicado: (2024)
NDF+: Joint Neural Directional Filtering and Diffuse Sound Extraction
por: Huang, Weilong, et al.
Publicado: (2026)
por: Huang, Weilong, et al.
Publicado: (2026)
Low-Resource Text-to-Speech Synthesis Using Noise-Augmented Training of ForwardTacotron
por: Lakshminarayana, Kishor Kayyar, et al.
Publicado: (2025)
por: Lakshminarayana, Kishor Kayyar, et al.
Publicado: (2025)
Aud-Sur: An Audio Analyzer Assistant for Audio Surveillance Applications
por: Lam, Phat, et al.
Publicado: (2025)
por: Lam, Phat, et al.
Publicado: (2025)
GAN-Based Speech Enhancement for Low SNR Using Latent Feature Conditioning
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
Neural Directional Filtering Using a Compact Microphone Array
por: Huang, Weilong, et al.
Publicado: (2025)
por: Huang, Weilong, et al.
Publicado: (2025)
Matching Reverberant Speech Through Learned Acoustic Embeddings and Feedback Delay Networks
por: Götz, Philipp, et al.
Publicado: (2025)
por: Götz, Philipp, et al.
Publicado: (2025)
Data-driven Joint Detection and Localization of Acoustic Reflectors
por: Bicer, H. Nazim, et al.
Publicado: (2024)
por: Bicer, H. Nazim, et al.
Publicado: (2024)
Towards Improved Objective Perceptual Audio Quality Assessment -- Part 1: A Novel Data-Driven Cognitive Model
por: Delgado, Pablo M., et al.
Publicado: (2024)
por: Delgado, Pablo M., et al.
Publicado: (2024)
Benchmarking Neural Speech Codec Intelligibility with SITool
por: Leschanowsky, Anna, et al.
Publicado: (2025)
por: Leschanowsky, Anna, et al.
Publicado: (2025)
Align-ULCNet: Towards Low-Complexity and Robust Acoustic Echo and Noise Reduction
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
Neural Directional Filtering: Far-Field Directivity Control With a Small Microphone Array
por: Wechsler, Julian, et al.
Publicado: (2024)
por: Wechsler, Julian, et al.
Publicado: (2024)
AudioSetCaps: An Enriched Audio-Caption Dataset using Automated Generation Pipeline with Large Audio and Language Models
por: Bai, Jisheng, et al.
Publicado: (2024)
por: Bai, Jisheng, et al.
Publicado: (2024)
A Hybrid Approach for Low-Complexity Joint Acoustic Echo and Noise Reduction
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
On the Language and Gender Biases in PSTN, VoIP and Neural Audio Codecs
por: Altwlkany, Kemal, et al.
Publicado: (2025)
por: Altwlkany, Kemal, et al.
Publicado: (2025)
A Generalized Bandsplit Neural Network for Cinematic Audio Source Separation
por: Watcharasupat, Karn N., et al.
Publicado: (2023)
por: Watcharasupat, Karn N., et al.
Publicado: (2023)
PAM: Prompting Audio-Language Models for Audio Quality Assessment
por: Deshmukh, Soham, et al.
Publicado: (2024)
por: Deshmukh, Soham, et al.
Publicado: (2024)
MLAAD: The Multi-Language Audio Anti-Spoofing Dataset
por: Müller, Nicolas M., et al.
Publicado: (2024)
por: Müller, Nicolas M., et al.
Publicado: (2024)
VoxEffects: A Speech-Oriented Audio Effects Dataset and Benchmark
por: Zhang, Zhe, et al.
Publicado: (2026)
por: Zhang, Zhe, et al.
Publicado: (2026)
Ejemplares similares
-
ODAQ: Open Dataset of Audio Quality
por: Torcoli, Matteo, et al.
Publicado: (2023) -
Investigating the impact of stereo processing -- a study for extending the Open Dataset of Audio Quality (ODAQ)
por: Dick, Sascha, et al.
Publicado: (2025) -
Exploring Perceptual Audio Quality Measurement on Stereo Processing Using the Open Dataset of Audio Quality
por: Delgado, Pablo M., et al.
Publicado: (2025) -
Navigating PESQ: Up-to-Date Versions and Open Implementations
por: Torcoli, Matteo, et al.
Publicado: (2025) -
On the Relation Between Speech Quality and Quantized Latent Representations of Neural Codecs
por: Halimeh, Mhd Modar, et al.
Publicado: (2025)