Temporal Feature Learning in Weakly Labelled Bioacoustic Cetacean Datasets via a Variational Autoencoder and Temporal Convolutional Network: An Interdisciplinary Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Fonollosa, Laia Garrobé, Gillespie, Douglas, Stankovic, Lina, Stankovic, Vladimir, Rendell, Luke |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Foundation Models for Bioacoustics -- a Comparative Review
by: Schwinger, Raphael, et al.
Published: (2025)
by: Schwinger, Raphael, et al.
Published: (2025)
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
by: Navine, Amanda K., et al.
Published: (2024)
by: Navine, Amanda K., et al.
Published: (2024)
Review of Cetacean's click detection algorithms
by: Gracic, Mak, et al.
Published: (2024)
by: Gracic, Mak, et al.
Published: (2024)
Computational bioacoustics with deep learning: a review and roadmap
by: Stowell, Dan
Published: (2021)
by: Stowell, Dan
Published: (2021)
Adaptive Representations of Sound for Automatic Insect Recognition
by: Faiß, Marius, et al.
Published: (2023)
by: Faiß, Marius, et al.
Published: (2023)
Learning to detect an animal sound from five examples
by: Nolasco, Inês, et al.
Published: (2023)
by: Nolasco, Inês, et al.
Published: (2023)
Fish Tracking, Counting, and Behaviour Analysis in Digital Aquaculture: A Comprehensive Survey
by: Cui, Meng, et al.
Published: (2024)
by: Cui, Meng, et al.
Published: (2024)
Robust Bioacoustic Detection via Richly Labelled Synthetic Soundscape Augmentation
by: Soltero, Kaspar, et al.
Published: (2025)
by: Soltero, Kaspar, et al.
Published: (2025)
Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data
by: Nihal, Ragib Amin, et al.
Published: (2025)
by: Nihal, Ragib Amin, et al.
Published: (2025)
Temporal Attention Pooling for Frequency Dynamic Convolution in Sound Event Detection
by: Nam, Hyeonuk, et al.
Published: (2025)
by: Nam, Hyeonuk, et al.
Published: (2025)
Convolutional Variational Autoencoders for Spectrogram Compression in Automatic Speech Recognition
by: Iakovenko, Olga, et al.
Published: (2024)
by: Iakovenko, Olga, et al.
Published: (2024)
NeXt-TDNN: Modernizing Multi-Scale Temporal Convolution Backbone for Speaker Verification
by: Heo, Hyun-Jun, et al.
Published: (2023)
by: Heo, Hyun-Jun, et al.
Published: (2023)
Rene: A Pre-trained Multi-modal Architecture for Auscultation of Respiratory Diseases
by: Zhang, Pengfei, et al.
Published: (2024)
by: Zhang, Pengfei, et al.
Published: (2024)
Cochlear Wave Propagation and Dynamics in the Human Base and Apex: Model-Based Estimates from Noninvasive Measurements
by: Alkhairy, Samiya A
Published: (2024)
by: Alkhairy, Samiya A
Published: (2024)
A Classification Benchmark for Artificial Intelligence Detection of Laryngeal Cancer from Patient Voice
by: Paterson, Mary, et al.
Published: (2024)
by: Paterson, Mary, et al.
Published: (2024)
Which Augmentation Should I Use? An Empirical Investigation of Augmentations for Self-Supervised Phonocardiogram Representation Learning
by: Ballas, Aristotelis, et al.
Published: (2023)
by: Ballas, Aristotelis, et al.
Published: (2023)
Automated Measurement of Geniohyoid Muscle Thickness During Speech Using Deep Learning and Ultrasound
by: Myrgyyassov, Alisher, et al.
Published: (2026)
by: Myrgyyassov, Alisher, et al.
Published: (2026)
Automatic detection of Mild Cognitive Impairment using high-dimensional acoustic features in spontaneous speech
by: Zhang, Cong, et al.
Published: (2024)
by: Zhang, Cong, et al.
Published: (2024)
Screening method for early dementia using sound objects as voice biomarkers
by: Pluta, Adam, et al.
Published: (2024)
by: Pluta, Adam, et al.
Published: (2024)
Learning to rumble: Automated elephant call classification, detection and endpointing using deep architectures
by: Geldenhuys, Christiaan M., et al.
Published: (2024)
by: Geldenhuys, Christiaan M., et al.
Published: (2024)
From Birdsong to Rumbles: Classifying Elephant Calls with Out-of-Species Embeddings
by: Geldenhuys, Christiaan M., et al.
Published: (2026)
by: Geldenhuys, Christiaan M., et al.
Published: (2026)
Prosody of speech production in latent post-stroke aphasia
by: Zhang, Cong, et al.
Published: (2024)
by: Zhang, Cong, et al.
Published: (2024)
Multi Modal Information Fusion of Acoustic and Linguistic Data for Decoding Dairy Cow Vocalizations in Animal Welfare Assessment
by: Jobarteh, Bubacarr, et al.
Published: (2024)
by: Jobarteh, Bubacarr, et al.
Published: (2024)
Anonymising Elderly and Pathological Speech: Voice Conversion Using DDSP and Query-by-Example
by: Ghosh, Suhita, et al.
Published: (2024)
by: Ghosh, Suhita, et al.
Published: (2024)
DEMONet: Underwater Acoustic Target Recognition based on Multi-Expert Network and Cross-Temporal Variational Autoencoder
by: Xie, Yuan, et al.
Published: (2024)
by: Xie, Yuan, et al.
Published: (2024)
Variational Autoencoder for Personalized Pathological Speech Enhancement
by: Hou, Mingchi, et al.
Published: (2025)
by: Hou, Mingchi, et al.
Published: (2025)
Towards High-Fidelity and Controllable Bioacoustic Generation via Enhanced Diffusion Learning
by: Song, Tianyu, et al.
Published: (2025)
by: Song, Tianyu, et al.
Published: (2025)
Distilling Spectrograms into Tokens: Fast and Lightweight Bioacoustic Classification for BirdCLEF+ 2025
by: Miyaguchi, Anthony, et al.
Published: (2025)
by: Miyaguchi, Anthony, et al.
Published: (2025)
Few-Shot Bioacoustic Event Detection with Frame-Level Embedding Learning System
by: Zhao, PengYuan, et al.
Published: (2024)
by: Zhao, PengYuan, et al.
Published: (2024)
Large Language Models and Non-Negative Matrix Factorization for Bioacoustic Signal Decomposition
by: Torabi, Yasaman, et al.
Published: (2025)
by: Torabi, Yasaman, et al.
Published: (2025)
BioME: A Resource-Efficient Bioacoustic Foundational Model for IoT Applications
by: Guimarães, Heitor R., et al.
Published: (2026)
by: Guimarães, Heitor R., et al.
Published: (2026)
Bridging Speech Emotion Recognition and Personality: Dataset and Temporal Interaction Condition Network
by: Gao, Yuan, et al.
Published: (2025)
by: Gao, Yuan, et al.
Published: (2025)
BirdSet: A Large-Scale Dataset for Audio Classification in Avian Bioacoustics
by: Rauch, Lukas, et al.
Published: (2024)
by: Rauch, Lukas, et al.
Published: (2024)
Adaptive Learning via a Negative Selection Strategy for Few-Shot Bioacoustic Event Detection
by: Chen, Yaxiong, et al.
Published: (2024)
by: Chen, Yaxiong, et al.
Published: (2024)
Learning Domain-Robust Bioacoustic Representations for Mosquito Species Classification with Contrastive Learning and Distribution Alignment
by: Hou, Yuanbo, et al.
Published: (2025)
by: Hou, Yuanbo, et al.
Published: (2025)
SIGNL: A Label-Efficient Audio Deepfake Detection System via Spectral-Temporal Graph Non-Contrastive Learning
by: Febrinanto, Falih Gozi, et al.
Published: (2025)
by: Febrinanto, Falih Gozi, et al.
Published: (2025)
Hierarchical Pooling Structure for Weakly Labeled Sound Event Detection
by: He, Ke-Xin, et al.
Published: (2019)
by: He, Ke-Xin, et al.
Published: (2019)
animal2vec and MeerKAT: A self-supervised transformer for rare-event raw audio input and a large-scale reference dataset for bioacoustics
by: Schäfer-Zimmermann, Julian C., et al.
Published: (2024)
by: Schäfer-Zimmermann, Julian C., et al.
Published: (2024)
Towards Deep Active Learning in Avian Bioacoustics
by: Rauch, Lukas, et al.
Published: (2024)
by: Rauch, Lukas, et al.
Published: (2024)
Perch 2.0: The Bittern Lesson for Bioacoustics
by: van Merriënboer, Bart, et al.
Published: (2025)
by: van Merriënboer, Bart, et al.
Published: (2025)
Similar Items
-
Foundation Models for Bioacoustics -- a Comparative Review
by: Schwinger, Raphael, et al.
Published: (2025) -
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
by: Navine, Amanda K., et al.
Published: (2024) -
Review of Cetacean's click detection algorithms
by: Gracic, Mak, et al.
Published: (2024) -
Computational bioacoustics with deep learning: a review and roadmap
by: Stowell, Dan
Published: (2021) -
Adaptive Representations of Sound for Automatic Insect Recognition
by: Faiß, Marius, et al.
Published: (2023)