Improving Machine Hearing on Limited Data Sets
Fuente:
arXiv
Guardado en:
| Autores principales: | Harar, Pavol, Bammer, Roswitha, Breger, Anna, Dörfler, Monika, Smekal, Zdenek |
|---|---|
| Formato: | Preprint |
| Publicado: |
2019
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Remixing Music for Hearing Aids Using Ensemble of Fine-Tuned Source Separators
por: Daly, Matthew
Publicado: (2024)
por: Daly, Matthew
Publicado: (2024)
The Inverse Drum Machine: Source Separation Through Joint Transcription and Analysis-by-Synthesis
por: Torres, Bernardo, et al.
Publicado: (2025)
por: Torres, Bernardo, et al.
Publicado: (2025)
GLA-Grad++: An Improved Griffin-Lim Guided Diffusion Model for Speech Synthesis
por: Baoueb, Teysir, et al.
Publicado: (2025)
por: Baoueb, Teysir, et al.
Publicado: (2025)
EmotionCaps: Enhancing Audio Captioning Through Emotion-Augmented Data Generation
por: Manivannan, Mithun, et al.
Publicado: (2024)
por: Manivannan, Mithun, et al.
Publicado: (2024)
Joint Source-Environment Adaptation of Data-Driven Underwater Acoustic Source Ranging Based on Model Uncertainty
por: Kari, Dariush, et al.
Publicado: (2025)
por: Kari, Dariush, et al.
Publicado: (2025)
A Machine Hearing System for Robust Cough Detection Based on a High-Level Representation of Band-Specific Audio Features
por: Monge-Alvarez, Jesús, et al.
Publicado: (2024)
por: Monge-Alvarez, Jesús, et al.
Publicado: (2024)
Resampling Filter Design for Multirate Neural Audio Effect Processing
por: Carson, Alistair, et al.
Publicado: (2025)
por: Carson, Alistair, et al.
Publicado: (2025)
Joint Source-Environment Adaptation for Deep Learning-Based Underwater Acoustic Source Ranging
por: Kari, Dariush, et al.
Publicado: (2025)
por: Kari, Dariush, et al.
Publicado: (2025)
Adaptive Control Attention Network for Underwater Acoustic Localization and Domain Adaptation
por: Vo, Quoc Thinh, et al.
Publicado: (2025)
por: Vo, Quoc Thinh, et al.
Publicado: (2025)
Resource-Efficient Separation Transformer
por: Della Libera, Luca, et al.
Publicado: (2022)
por: Della Libera, Luca, et al.
Publicado: (2022)
Self-Tuning Spectral Clustering for Speaker Diarization
por: Raghav, Nikhil, et al.
Publicado: (2024)
por: Raghav, Nikhil, et al.
Publicado: (2024)
Point Neuron Learning: A New Physics-Informed Neural Network Architecture
por: Bi, Hanwen, et al.
Publicado: (2024)
por: Bi, Hanwen, et al.
Publicado: (2024)
A Power-Weighted Noncentral Complex Gaussian Distribution
por: Nakashika, Toru
Publicado: (2026)
por: Nakashika, Toru
Publicado: (2026)
Reconstruction of Sound Field through Diffusion Models
por: Miotello, Federico, et al.
Publicado: (2023)
por: Miotello, Federico, et al.
Publicado: (2023)
GLA-Grad: A Griffin-Lim Extended Waveform Generation Diffusion Model
por: Liu, Haocheng, et al.
Publicado: (2024)
por: Liu, Haocheng, et al.
Publicado: (2024)
Blind Estimation of Sub-band Acoustic Parameters from Ambisonics Recordings using Spectro-Spatial Covariance Features
por: Meng, Hanyu, et al.
Publicado: (2024)
por: Meng, Hanyu, et al.
Publicado: (2024)
Speech Watermarking with Discrete Intermediate Representations
por: Ji, Shengpeng, et al.
Publicado: (2024)
por: Ji, Shengpeng, et al.
Publicado: (2024)
FlowDec: A flow-based full-band general audio codec with high perceptual quality
por: Welker, Simon, et al.
Publicado: (2025)
por: Welker, Simon, et al.
Publicado: (2025)
Diff-TONE: Timestep Optimization for iNstrument Editing in Text-to-Music Diffusion Models
por: Baoueb, Teysir, et al.
Publicado: (2025)
por: Baoueb, Teysir, et al.
Publicado: (2025)
Listenable Maps for Zero-Shot Audio Classifiers
por: Paissan, Francesco, et al.
Publicado: (2024)
por: Paissan, Francesco, et al.
Publicado: (2024)
PeriodGrad: Towards Pitch-Controllable Neural Vocoder Based on a Diffusion Probabilistic Model
por: Hono, Yukiya, et al.
Publicado: (2024)
por: Hono, Yukiya, et al.
Publicado: (2024)
A Generalized Bandsplit Neural Network for Cinematic Audio Source Separation
por: Watcharasupat, Karn N., et al.
Publicado: (2023)
por: Watcharasupat, Karn N., et al.
Publicado: (2023)
Mismatch-Robust Underwater Acoustic Localization Using A Differentiable Modular Forward Model
por: Kari, Dariush, et al.
Publicado: (2025)
por: Kari, Dariush, et al.
Publicado: (2025)
FunnelNet: An End-to-End Deep Learning Framework to Monitor Digital Heart Murmur in Real-Time
por: Jobayer, Md, et al.
Publicado: (2024)
por: Jobayer, Md, et al.
Publicado: (2024)
A DNN Based Post-Filter to Enhance the Quality of Coded Speech in MDCT Domain
por: Gupta, Kishan, et al.
Publicado: (2022)
por: Gupta, Kishan, et al.
Publicado: (2022)
Lightweight DNN for Full-Band Speech Denoising on Mobile Devices: Exploiting Long and Short Temporal Patterns
por: Drossos, Konstantinos, et al.
Publicado: (2025)
por: Drossos, Konstantinos, et al.
Publicado: (2025)
Latent Granular Resynthesis using Neural Audio Codecs
por: Tokui, Nao, et al.
Publicado: (2025)
por: Tokui, Nao, et al.
Publicado: (2025)
XAI-Driven Spectral Analysis of Cough Sounds for Respiratory Disease Characterization
por: Amado-Caballero, Patricia, et al.
Publicado: (2025)
por: Amado-Caballero, Patricia, et al.
Publicado: (2025)
Learnable Adaptive Time-Frequency Representation via Differentiable Short-Time Fourier Transform
por: Leiber, Maxime, et al.
Publicado: (2025)
por: Leiber, Maxime, et al.
Publicado: (2025)
Automated Dysphagia Screening Using Noninvasive Neck Acoustic Sensing
por: Chng, Jade, et al.
Publicado: (2026)
por: Chng, Jade, et al.
Publicado: (2026)
PD-ADSV: An Automated Diagnosing System Using Voice Signals and Hard Voting Ensemble Method for Parkinson's Disease
por: Ghaheri, Paria, et al.
Publicado: (2023)
por: Ghaheri, Paria, et al.
Publicado: (2023)
Online speaker diarization of meetings guided by speech separation
por: Gruttadauria, Elio, et al.
Publicado: (2024)
por: Gruttadauria, Elio, et al.
Publicado: (2024)
Permutation Invariant Recurrent Neural Networks for Sound Source Tracking Applications
por: Diaz-Guerra, David, et al.
Publicado: (2023)
por: Diaz-Guerra, David, et al.
Publicado: (2023)
Listenable Maps for Audio Classifiers
por: Paissan, Francesco, et al.
Publicado: (2024)
por: Paissan, Francesco, et al.
Publicado: (2024)
SLiCK: Exploiting Subsequences for Length-Constrained Keyword Spotting
por: Nishu, Kumari, et al.
Publicado: (2024)
por: Nishu, Kumari, et al.
Publicado: (2024)
Intelligent Fault Diagnosis of Type and Severity in Low-Frequency, Low Bit-Depth Signals
por: Spadini, Tito, et al.
Publicado: (2024)
por: Spadini, Tito, et al.
Publicado: (2024)
AI-Assisted Music Production: A User Study on Text-to-Music Models
por: Ronchini, Francesca, et al.
Publicado: (2025)
por: Ronchini, Francesca, et al.
Publicado: (2025)
Unsupervised Harmonic Parameter Estimation Using Differentiable DSP and Spectral Optimal Transport
por: Torres, Bernardo, et al.
Publicado: (2023)
por: Torres, Bernardo, et al.
Publicado: (2023)
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
por: Miotello, Federico, et al.
Publicado: (2024)
por: Miotello, Federico, et al.
Publicado: (2024)
A Convolutional Framework for Mapping Imagined Auditory MEG into Listened Brain Responses
por: Maghsoudi, Maryam, et al.
Publicado: (2025)
por: Maghsoudi, Maryam, et al.
Publicado: (2025)
Ejemplares similares
-
Remixing Music for Hearing Aids Using Ensemble of Fine-Tuned Source Separators
por: Daly, Matthew
Publicado: (2024) -
The Inverse Drum Machine: Source Separation Through Joint Transcription and Analysis-by-Synthesis
por: Torres, Bernardo, et al.
Publicado: (2025) -
GLA-Grad++: An Improved Griffin-Lim Guided Diffusion Model for Speech Synthesis
por: Baoueb, Teysir, et al.
Publicado: (2025) -
EmotionCaps: Enhancing Audio Captioning Through Emotion-Augmented Data Generation
por: Manivannan, Mithun, et al.
Publicado: (2024) -
Joint Source-Environment Adaptation of Data-Driven Underwater Acoustic Source Ranging Based on Model Uncertainty
por: Kari, Dariush, et al.
Publicado: (2025)