Multispecies bird sound recognition using a fully convolutional neural network
Fuente:
arXiv
Guardado en:
| Autores principales: | García-Ordás, María Teresa, Rubio-Martín, Sergio, Benítez-Andrades, José Alberto, Alaiz-Moretón, Hector, García-Rodríguez, Isaías |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Sentiment analysis in non-fixed length audios using a Fully Convolutional Neural Network
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
Determining the severity of Parkinson's disease in patients using a multi task neural network
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
Towards detecting the pathological subharmonic voicing with fully convolutional neural networks
por: Ikuma, Takeshi, et al.
Publicado: (2025)
por: Ikuma, Takeshi, et al.
Publicado: (2025)
Adaptive high-precision sound source localization at low frequencies based on convolutional neural network
por: Ma, Wenbo, et al.
Publicado: (2024)
por: Ma, Wenbo, et al.
Publicado: (2024)
Frequency-aware convolution for sound event detection
por: Song, Tao, et al.
Publicado: (2024)
por: Song, Tao, et al.
Publicado: (2024)
What do neural networks listen to? Exploring the crucial bands in Speech Enhancement using Sinc-convolution
por: Ho, Kuan-Hsun, et al.
Publicado: (2024)
por: Ho, Kuan-Hsun, et al.
Publicado: (2024)
Full-frequency dynamic convolution: a physical frequency-dependent convolution for sound event detection
por: Yue, Haobo, et al.
Publicado: (2024)
por: Yue, Haobo, et al.
Publicado: (2024)
Toward end-to-end interpretable convolutional neural networks for waveform signals
por: Vu, Linh, et al.
Publicado: (2024)
por: Vu, Linh, et al.
Publicado: (2024)
Speech privacy-preserving methods using secret key for convolutional neural network models and their robustness evaluation
por: Niwa, Shoko, et al.
Publicado: (2024)
por: Niwa, Shoko, et al.
Publicado: (2024)
PROCTER: PROnunciation-aware ConTextual adaptER for personalized speech recognition in neural transducers
por: Pandey, Rahul, et al.
Publicado: (2023)
por: Pandey, Rahul, et al.
Publicado: (2023)
Controlling the Parameterized Multi-channel Wiener Filter using a tiny neural network
por: Grinstein, Eric, et al.
Publicado: (2025)
por: Grinstein, Eric, et al.
Publicado: (2025)
DualSep: A Light-weight dual-encoder convolutional recurrent network for real-time in-car speech separation
por: Wang, Ziqian, et al.
Publicado: (2024)
por: Wang, Ziqian, et al.
Publicado: (2024)
Resnet-conformer network with shared weights and attention mechanism for sound event localization, detection, and distance estimation
por: Vo, Quoc Thinh, et al.
Publicado: (2025)
por: Vo, Quoc Thinh, et al.
Publicado: (2025)
EZhouNet:A framework based on graph neural network and anchor interval for the respiratory sound event detection
por: Chu, Yun, et al.
Publicado: (2025)
por: Chu, Yun, et al.
Publicado: (2025)
The role of direct sound spherical harmonics representation in externalization using binaural reproduction
por: Miller, Eran, et al.
Publicado: (2024)
por: Miller, Eran, et al.
Publicado: (2024)
Binaural sound source localization using a hybrid time and frequency domain model
por: Geva, Gil, et al.
Publicado: (2024)
por: Geva, Gil, et al.
Publicado: (2024)
A circular microphone array with virtual microphones based on acoustics-informed neural networks
por: Zhao, Sipei, et al.
Publicado: (2024)
por: Zhao, Sipei, et al.
Publicado: (2024)
Differentiable physics for sound field reconstruction
por: Verburg, Samuel A., et al.
Publicado: (2025)
por: Verburg, Samuel A., et al.
Publicado: (2025)
Validation of artificial neural networks to model the acoustic behaviour of induction motors
por: Jimenez-Romero, F. J., et al.
Publicado: (2024)
por: Jimenez-Romero, F. J., et al.
Publicado: (2024)
An adaptive filter bank based neural network approach for time delay estimation and speech enhancement
por: Ma, Lu
Publicado: (2025)
por: Ma, Lu
Publicado: (2025)
Physics-informed neural network for acoustic resonance analysis in a one-dimensional acoustic tube
por: Yokota, Kazuya, et al.
Publicado: (2023)
por: Yokota, Kazuya, et al.
Publicado: (2023)
Speaker anonymization using neural audio codec language models
por: Panariello, Michele, et al.
Publicado: (2023)
por: Panariello, Michele, et al.
Publicado: (2023)
Guiding the underwater acoustic target recognition with interpretable contrastive learning
por: Xie, Yuan, et al.
Publicado: (2024)
por: Xie, Yuan, et al.
Publicado: (2024)
The Neural-SRP method for positional sound source localization
por: Grinstein, Eric, et al.
Publicado: (2024)
por: Grinstein, Eric, et al.
Publicado: (2024)
E2E-AEC: Implementing an end-to-end neural network learning approach for acoustic echo cancellation
por: Jiang, Yiheng, et al.
Publicado: (2026)
por: Jiang, Yiheng, et al.
Publicado: (2026)
Graph-based multi-Feature fusion method for speech emotion recognition
por: Liu, Xueyu, et al.
Publicado: (2024)
por: Liu, Xueyu, et al.
Publicado: (2024)
Towards interpretable emotion recognition: Identifying key features with machine learning
por: Kaloga, Yacouba, et al.
Publicado: (2025)
por: Kaloga, Yacouba, et al.
Publicado: (2025)
Some clues to build a sound analysis relevant to hearing
por: Millot, Laurent
Publicado: (2024)
por: Millot, Laurent
Publicado: (2024)
Interaural time difference loss for binaural target sound extraction
por: Hernandez-Olivan, Carlos, et al.
Publicado: (2024)
por: Hernandez-Olivan, Carlos, et al.
Publicado: (2024)
Onset and offset weighted loss function for sound event detection
por: Song, Tao
Publicado: (2024)
por: Song, Tao
Publicado: (2024)
Fine-tune the pretrained ATST model for sound event detection
por: Shao, Nian, et al.
Publicado: (2023)
por: Shao, Nian, et al.
Publicado: (2023)
Advancing automatic speech recognition using feature fusion with self-supervised learning features: A case study on Fearless Steps Apollo corpus
por: Chen, Szu-Jui, et al.
Publicado: (2026)
por: Chen, Szu-Jui, et al.
Publicado: (2026)
Serial-OE: Anomalous sound detection based on serial method with outlier exposure capable of using small amounts of anomalous data for training
por: Kuroyanagi, Ibuki, et al.
Publicado: (2025)
por: Kuroyanagi, Ibuki, et al.
Publicado: (2025)
Language model integration based on memory control for sequence to sequence speech recognition
por: Cho, Jaejin, et al.
Publicado: (2018)
por: Cho, Jaejin, et al.
Publicado: (2018)
Phoneme-based speech recognition driven by large language models and sampling marginalization
por: Ma, Te, et al.
Publicado: (2025)
por: Ma, Te, et al.
Publicado: (2025)
Representational learning for an anomalous sound detection system with source separation model
por: Shin, Seunghyeon, et al.
Publicado: (2024)
por: Shin, Seunghyeon, et al.
Publicado: (2024)
Signal processing algorithm effective for sound quality of hearing loss simulators
por: Irino, Toshio, et al.
Publicado: (2024)
por: Irino, Toshio, et al.
Publicado: (2024)
Paraformer-v2: An improved non-autoregressive transformer for noise-robust speech recognition
por: An, Keyu, et al.
Publicado: (2024)
por: An, Keyu, et al.
Publicado: (2024)
InsectSet459: an open dataset of insect sounds for bioacoustic machine learning
por: Faiß, Marius, et al.
Publicado: (2025)
por: Faiß, Marius, et al.
Publicado: (2025)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
por: Ducorroy, Alexandre, et al.
Publicado: (2025)
por: Ducorroy, Alexandre, et al.
Publicado: (2025)
Ejemplares similares
-
Sentiment analysis in non-fixed length audios using a Fully Convolutional Neural Network
por: García-Ordás, María Teresa, et al.
Publicado: (2024) -
Determining the severity of Parkinson's disease in patients using a multi task neural network
por: García-Ordás, María Teresa, et al.
Publicado: (2024) -
Towards detecting the pathological subharmonic voicing with fully convolutional neural networks
por: Ikuma, Takeshi, et al.
Publicado: (2025) -
Adaptive high-precision sound source localization at low frequencies based on convolutional neural network
por: Ma, Wenbo, et al.
Publicado: (2024) -
Frequency-aware convolution for sound event detection
por: Song, Tao, et al.
Publicado: (2024)