From Birdsong to Rumbles: Classifying Elephant Calls with Out-of-Species Embeddings
Fuente:
arXiv
Saved in:
| Main Authors: | Geldenhuys, Christiaan M., Niesler, Thomas R. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to rumble: Automated elephant call classification, detection and endpointing using deep architectures
by: Geldenhuys, Christiaan M., et al.
Published: (2024)
by: Geldenhuys, Christiaan M., et al.
Published: (2024)
WhaleVAD-BPN: Improving Baleen Whale Call Detection with Boundary Proposal Networks and Post-processing Optimisation
by: Geldenhuys, Christiaan M., et al.
Published: (2025)
by: Geldenhuys, Christiaan M., et al.
Published: (2025)
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
by: Navine, Amanda K., et al.
Published: (2024)
by: Navine, Amanda K., et al.
Published: (2024)
Automated Measurement of Geniohyoid Muscle Thickness During Speech Using Deep Learning and Ultrasound
by: Myrgyyassov, Alisher, et al.
Published: (2026)
by: Myrgyyassov, Alisher, et al.
Published: (2026)
A Classification Benchmark for Artificial Intelligence Detection of Laryngeal Cancer from Patient Voice
by: Paterson, Mary, et al.
Published: (2024)
by: Paterson, Mary, et al.
Published: (2024)
Which Augmentation Should I Use? An Empirical Investigation of Augmentations for Self-Supervised Phonocardiogram Representation Learning
by: Ballas, Aristotelis, et al.
Published: (2023)
by: Ballas, Aristotelis, et al.
Published: (2023)
Screening method for early dementia using sound objects as voice biomarkers
by: Pluta, Adam, et al.
Published: (2024)
by: Pluta, Adam, et al.
Published: (2024)
Foundation Models for Bioacoustics -- a Comparative Review
by: Schwinger, Raphael, et al.
Published: (2025)
by: Schwinger, Raphael, et al.
Published: (2025)
Temporal Feature Learning in Weakly Labelled Bioacoustic Cetacean Datasets via a Variational Autoencoder and Temporal Convolutional Network: An Interdisciplinary Approach
by: Fonollosa, Laia Garrobé, et al.
Published: (2024)
by: Fonollosa, Laia Garrobé, et al.
Published: (2024)
Computational bioacoustics with deep learning: a review and roadmap
by: Stowell, Dan
Published: (2021)
by: Stowell, Dan
Published: (2021)
Adaptive Representations of Sound for Automatic Insect Recognition
by: Faiß, Marius, et al.
Published: (2023)
by: Faiß, Marius, et al.
Published: (2023)
Learning to detect an animal sound from five examples
by: Nolasco, Inês, et al.
Published: (2023)
by: Nolasco, Inês, et al.
Published: (2023)
Fish Tracking, Counting, and Behaviour Analysis in Digital Aquaculture: A Comprehensive Survey
by: Cui, Meng, et al.
Published: (2024)
by: Cui, Meng, et al.
Published: (2024)
Rene: A Pre-trained Multi-modal Architecture for Auscultation of Respiratory Diseases
by: Zhang, Pengfei, et al.
Published: (2024)
by: Zhang, Pengfei, et al.
Published: (2024)
Cochlear Wave Propagation and Dynamics in the Human Base and Apex: Model-Based Estimates from Noninvasive Measurements
by: Alkhairy, Samiya A
Published: (2024)
by: Alkhairy, Samiya A
Published: (2024)
Automatic detection of Mild Cognitive Impairment using high-dimensional acoustic features in spontaneous speech
by: Zhang, Cong, et al.
Published: (2024)
by: Zhang, Cong, et al.
Published: (2024)
Prosody of speech production in latent post-stroke aphasia
by: Zhang, Cong, et al.
Published: (2024)
by: Zhang, Cong, et al.
Published: (2024)
Multi Modal Information Fusion of Acoustic and Linguistic Data for Decoding Dairy Cow Vocalizations in Animal Welfare Assessment
by: Jobarteh, Bubacarr, et al.
Published: (2024)
by: Jobarteh, Bubacarr, et al.
Published: (2024)
Anonymising Elderly and Pathological Speech: Voice Conversion Using DDSP and Query-by-Example
by: Ghosh, Suhita, et al.
Published: (2024)
by: Ghosh, Suhita, et al.
Published: (2024)
A multimodal LLM for the non-invasive decoding of spoken text from brain recordings
by: Hmamouche, Youssef, et al.
Published: (2024)
by: Hmamouche, Youssef, et al.
Published: (2024)
I Guess That's Why They Call it the Blues: Causal Analysis for Audio Classifiers
by: Kelly, David A., et al.
Published: (2026)
by: Kelly, David A., et al.
Published: (2026)
animal2vec and MeerKAT: A self-supervised transformer for rare-event raw audio input and a large-scale reference dataset for bioacoustics
by: Schäfer-Zimmermann, Julian C., et al.
Published: (2024)
by: Schäfer-Zimmermann, Julian C., et al.
Published: (2024)
Cough activity detection for automatic tuberculosis screening
by: van Vüren, Joshua Jansen, et al.
Published: (2026)
by: van Vüren, Joshua Jansen, et al.
Published: (2026)
Atrial Fibrillation Detection System via Acoustic Sensing for Mobile Phones
by: Liu, Xuanyu, et al.
Published: (2024)
by: Liu, Xuanyu, et al.
Published: (2024)
Leveraging cough sounds to optimize chest x-ray usage in low-resource settings
by: Philip, Alexander, et al.
Published: (2024)
by: Philip, Alexander, et al.
Published: (2024)
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds
by: Chang, Andrew, et al.
Published: (2025)
by: Chang, Andrew, et al.
Published: (2025)
On the Utility of Speech and Audio Foundation Models for Marmoset Call Analysis
by: Sarkar, Eklavya, et al.
Published: (2024)
by: Sarkar, Eklavya, et al.
Published: (2024)
Audio2Tool: Speak, Call, Act -- A Dataset for Benchmarking Speech Tool Use
by: Pahwa, Ramit, et al.
Published: (2026)
by: Pahwa, Ramit, et al.
Published: (2026)
Edge Intelligence for Wildlife Conservation: Real-Time Hornbill Call Classification Using TinyML
by: Hing, Kong Ka, et al.
Published: (2025)
by: Hing, Kong Ka, et al.
Published: (2025)
Challenges in Automated Processing of Speech from Child Wearables: The Case of Voice Type Classifier
by: Kunze, Tarek, et al.
Published: (2025)
by: Kunze, Tarek, et al.
Published: (2025)
Phoneme-Level Deepfake Detection Across Emotional Conditions Using Self-Supervised Embeddings
by: Nallaguntla, Vamshi, et al.
Published: (2026)
by: Nallaguntla, Vamshi, et al.
Published: (2026)
AbsoluteNet: A Deep Learning Neural Network to Classify Cerebral Hemodynamic Responses of Auditory Processing
by: Adeli, Behtom, et al.
Published: (2025)
by: Adeli, Behtom, et al.
Published: (2025)
Listenable Maps for Audio Classifiers
by: Paissan, Francesco, et al.
Published: (2024)
by: Paissan, Francesco, et al.
Published: (2024)
Learning Spatially-Aware Language and Audio Embeddings
by: Devnani, Bhavika, et al.
Published: (2024)
by: Devnani, Bhavika, et al.
Published: (2024)
Cosine Scoring with Uncertainty for Neural Speaker Embedding
by: Wang, Qiongqiong, et al.
Published: (2024)
by: Wang, Qiongqiong, et al.
Published: (2024)
Improving Perceptual Audio Aesthetic Assessment via Triplet Loss and Self-Supervised Embeddings
by: Wisnu, Dyah A. M. G., et al.
Published: (2025)
by: Wisnu, Dyah A. M. G., et al.
Published: (2025)
Self-Supervised Embeddings for Detecting Individual Symptoms of Depression
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Dynamic Recognition of Speakers for Consent Management by Contrastive Embedding Replay
by: Shahmansoori, Arash, et al.
Published: (2022)
by: Shahmansoori, Arash, et al.
Published: (2022)
Embedding-Space Diffusion for Zero-Shot Environmental Sound Classification
by: Sims, Ysobel, et al.
Published: (2024)
by: Sims, Ysobel, et al.
Published: (2024)
Semi-supervised classification of bird vocalizations
by: Hexeberg, Simen, et al.
Published: (2025)
by: Hexeberg, Simen, et al.
Published: (2025)
Similar Items
-
Learning to rumble: Automated elephant call classification, detection and endpointing using deep architectures
by: Geldenhuys, Christiaan M., et al.
Published: (2024) -
WhaleVAD-BPN: Improving Baleen Whale Call Detection with Boundary Proposal Networks and Post-processing Optimisation
by: Geldenhuys, Christiaan M., et al.
Published: (2025) -
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
by: Navine, Amanda K., et al.
Published: (2024) -
Automated Measurement of Geniohyoid Muscle Thickness During Speech Using Deep Learning and Ultrasound
by: Myrgyyassov, Alisher, et al.
Published: (2026) -
A Classification Benchmark for Artificial Intelligence Detection of Laryngeal Cancer from Patient Voice
by: Paterson, Mary, et al.
Published: (2024)