Semi-supervised classification of bird vocalizations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hexeberg, Simen, Chitre, Mandar, Hoffmann-Kuhnt, Matthias, Low, Bing Wen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
animal2vec and MeerKAT: A self-supervised transformer for rare-event raw audio input and a large-scale reference dataset for bioacoustics
von: Schäfer-Zimmermann, Julian C., et al.
Veröffentlicht: (2024)
von: Schäfer-Zimmermann, Julian C., et al.
Veröffentlicht: (2024)
Rene: A Pre-trained Multi-modal Architecture for Auscultation of Respiratory Diseases
von: Zhang, Pengfei, et al.
Veröffentlicht: (2024)
von: Zhang, Pengfei, et al.
Veröffentlicht: (2024)
Multi Modal Information Fusion of Acoustic and Linguistic Data for Decoding Dairy Cow Vocalizations in Animal Welfare Assessment
von: Jobarteh, Bubacarr, et al.
Veröffentlicht: (2024)
von: Jobarteh, Bubacarr, et al.
Veröffentlicht: (2024)
Anonymising Elderly and Pathological Speech: Voice Conversion Using DDSP and Query-by-Example
von: Ghosh, Suhita, et al.
Veröffentlicht: (2024)
von: Ghosh, Suhita, et al.
Veröffentlicht: (2024)
Temporal Feature Learning in Weakly Labelled Bioacoustic Cetacean Datasets via a Variational Autoencoder and Temporal Convolutional Network: An Interdisciplinary Approach
von: Fonollosa, Laia Garrobé, et al.
Veröffentlicht: (2024)
von: Fonollosa, Laia Garrobé, et al.
Veröffentlicht: (2024)
Computational bioacoustics with deep learning: a review and roadmap
von: Stowell, Dan
Veröffentlicht: (2021)
von: Stowell, Dan
Veröffentlicht: (2021)
Adaptive Representations of Sound for Automatic Insect Recognition
von: Faiß, Marius, et al.
Veröffentlicht: (2023)
von: Faiß, Marius, et al.
Veröffentlicht: (2023)
Learning to detect an animal sound from five examples
von: Nolasco, Inês, et al.
Veröffentlicht: (2023)
von: Nolasco, Inês, et al.
Veröffentlicht: (2023)
Fish Tracking, Counting, and Behaviour Analysis in Digital Aquaculture: A Comprehensive Survey
von: Cui, Meng, et al.
Veröffentlicht: (2024)
von: Cui, Meng, et al.
Veröffentlicht: (2024)
Learning to rumble: Automated elephant call classification, detection and endpointing using deep architectures
von: Geldenhuys, Christiaan M., et al.
Veröffentlicht: (2024)
von: Geldenhuys, Christiaan M., et al.
Veröffentlicht: (2024)
SemiPL: A Semi-supervised Method for Event Sound Source Localization
von: Li, Yue, et al.
Veröffentlicht: (2024)
von: Li, Yue, et al.
Veröffentlicht: (2024)
WhaleVAD-BPN: Improving Baleen Whale Call Detection with Boundary Proposal Networks and Post-processing Optimisation
von: Geldenhuys, Christiaan M., et al.
Veröffentlicht: (2025)
von: Geldenhuys, Christiaan M., et al.
Veröffentlicht: (2025)
Cochlear Wave Propagation and Dynamics in the Human Base and Apex: Model-Based Estimates from Noninvasive Measurements
von: Alkhairy, Samiya A
Veröffentlicht: (2024)
von: Alkhairy, Samiya A
Veröffentlicht: (2024)
A Classification Benchmark for Artificial Intelligence Detection of Laryngeal Cancer from Patient Voice
von: Paterson, Mary, et al.
Veröffentlicht: (2024)
von: Paterson, Mary, et al.
Veröffentlicht: (2024)
Which Augmentation Should I Use? An Empirical Investigation of Augmentations for Self-Supervised Phonocardiogram Representation Learning
von: Ballas, Aristotelis, et al.
Veröffentlicht: (2023)
von: Ballas, Aristotelis, et al.
Veröffentlicht: (2023)
Automated Measurement of Geniohyoid Muscle Thickness During Speech Using Deep Learning and Ultrasound
von: Myrgyyassov, Alisher, et al.
Veröffentlicht: (2026)
von: Myrgyyassov, Alisher, et al.
Veröffentlicht: (2026)
Automatic detection of Mild Cognitive Impairment using high-dimensional acoustic features in spontaneous speech
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
Screening method for early dementia using sound objects as voice biomarkers
von: Pluta, Adam, et al.
Veröffentlicht: (2024)
von: Pluta, Adam, et al.
Veröffentlicht: (2024)
Foundation Models for Bioacoustics -- a Comparative Review
von: Schwinger, Raphael, et al.
Veröffentlicht: (2025)
von: Schwinger, Raphael, et al.
Veröffentlicht: (2025)
From Birdsong to Rumbles: Classifying Elephant Calls with Out-of-Species Embeddings
von: Geldenhuys, Christiaan M., et al.
Veröffentlicht: (2026)
von: Geldenhuys, Christiaan M., et al.
Veröffentlicht: (2026)
Prosody of speech production in latent post-stroke aphasia
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
von: Zhang, Cong, et al.
Veröffentlicht: (2024)
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
von: Navine, Amanda K., et al.
Veröffentlicht: (2024)
von: Navine, Amanda K., et al.
Veröffentlicht: (2024)
UWAV: Uncertainty-weighted Weakly-supervised Audio-Visual Video Parsing
von: Lai, Yung-Hsuan, et al.
Veröffentlicht: (2025)
von: Lai, Yung-Hsuan, et al.
Veröffentlicht: (2025)
ECHO: Environmental Sound Classification with Hierarchical Ontology-guided Semi-Supervised Learning
von: Gupta, Pranav, et al.
Veröffentlicht: (2024)
von: Gupta, Pranav, et al.
Veröffentlicht: (2024)
Ovi: Twin Backbone Cross-Modal Fusion for Audio-Video Generation
von: Low, Chetwin, et al.
Veröffentlicht: (2025)
von: Low, Chetwin, et al.
Veröffentlicht: (2025)
Atrial Fibrillation Detection System via Acoustic Sensing for Mobile Phones
von: Liu, Xuanyu, et al.
Veröffentlicht: (2024)
von: Liu, Xuanyu, et al.
Veröffentlicht: (2024)
Cross Pseudo-Labeling for Semi-Supervised Audio-Visual Source Localization
von: Guo, Yuxin, et al.
Veröffentlicht: (2024)
von: Guo, Yuxin, et al.
Veröffentlicht: (2024)
Weakly-supervised Audio Temporal Forgery Localization via Progressive Audio-language Co-learning Network
von: Wu, Junyan, et al.
Veröffentlicht: (2025)
von: Wu, Junyan, et al.
Veröffentlicht: (2025)
Utilizing synthetic training data for the supervised classification of rat ultrasonic vocalizations
von: Scott, K. Jack, et al.
Veröffentlicht: (2023)
von: Scott, K. Jack, et al.
Veröffentlicht: (2023)
Visual and audio scene classification for detecting discrepancies in video: a baseline method and experimental protocol
von: Apostolidis, Konstantinos, et al.
Veröffentlicht: (2024)
von: Apostolidis, Konstantinos, et al.
Veröffentlicht: (2024)
ThinkSound: Chain-of-Thought Reasoning in Multimodal Large Language Models for Audio Generation and Editing
von: Liu, Huadai, et al.
Veröffentlicht: (2025)
von: Liu, Huadai, et al.
Veröffentlicht: (2025)
FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models
von: Aneja, Shivangi, et al.
Veröffentlicht: (2023)
von: Aneja, Shivangi, et al.
Veröffentlicht: (2023)
DIDiffGes: Decoupled Semi-Implicit Diffusion Models for Real-time Gesture Generation from Speech
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
GaussianSpeech: Audio-Driven Gaussian Avatars
von: Aneja, Shivangi, et al.
Veröffentlicht: (2024)
von: Aneja, Shivangi, et al.
Veröffentlicht: (2024)
VGGSounder: Audio-Visual Evaluations for Foundation Models
von: Zverev, Daniil, et al.
Veröffentlicht: (2025)
von: Zverev, Daniil, et al.
Veröffentlicht: (2025)
OmniAudio: Generating Spatial Audio from 360-Degree Video
von: Liu, Huadai, et al.
Veröffentlicht: (2025)
von: Liu, Huadai, et al.
Veröffentlicht: (2025)
JavisDiT: Joint Audio-Video Diffusion Transformer with Hierarchical Spatio-Temporal Prior Synchronization
von: Liu, Kai, et al.
Veröffentlicht: (2025)
von: Liu, Kai, et al.
Veröffentlicht: (2025)
Learning Separable Hidden Unit Contributions for Speaker-Adaptive Lip-Reading
von: Luo, Songtao, et al.
Veröffentlicht: (2023)
von: Luo, Songtao, et al.
Veröffentlicht: (2023)
Towards Unconstrained Audio Splicing Detection and Localization with Neural Networks
von: Moussa, Denise, et al.
Veröffentlicht: (2022)
von: Moussa, Denise, et al.
Veröffentlicht: (2022)
Enriching Multimodal Sentiment Analysis through Textual Emotional Descriptions of Visual-Audio Content
von: Wu, Sheng, et al.
Veröffentlicht: (2024)
von: Wu, Sheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
animal2vec and MeerKAT: A self-supervised transformer for rare-event raw audio input and a large-scale reference dataset for bioacoustics
von: Schäfer-Zimmermann, Julian C., et al.
Veröffentlicht: (2024) -
Rene: A Pre-trained Multi-modal Architecture for Auscultation of Respiratory Diseases
von: Zhang, Pengfei, et al.
Veröffentlicht: (2024) -
Multi Modal Information Fusion of Acoustic and Linguistic Data for Decoding Dairy Cow Vocalizations in Animal Welfare Assessment
von: Jobarteh, Bubacarr, et al.
Veröffentlicht: (2024) -
Anonymising Elderly and Pathological Speech: Voice Conversion Using DDSP and Query-by-Example
von: Ghosh, Suhita, et al.
Veröffentlicht: (2024) -
Temporal Feature Learning in Weakly Labelled Bioacoustic Cetacean Datasets via a Variational Autoencoder and Temporal Convolutional Network: An Interdisciplinary Approach
von: Fonollosa, Laia Garrobé, et al.
Veröffentlicht: (2024)