Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nihal, Ragib Amin, Yen, Benjamin, Shi, Runwu, Ashizawa, Takeshi, Nakadai, Kazuhiro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Cross-Attention with Confidence Weighting for Multi-Channel Audio Alignment
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2025)
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2025)
Single-Channel Target Speech Extraction Utilizing Distance and Room Clues
von: Shi, Runwu, et al.
Veröffentlicht: (2025)
von: Shi, Runwu, et al.
Veröffentlicht: (2025)
Ecologically-Constrained Task Arithmetic for Multi-Taxa Bioacoustic Classifiers Without Shared Data
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2026)
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2026)
Distance Based Single-Channel Target Speech Extraction
von: Shi, Runwu, et al.
Veröffentlicht: (2024)
von: Shi, Runwu, et al.
Veröffentlicht: (2024)
Unsupervised Single-Channel Audio Separation with Diffusion Source Priors
von: Shi, Runwu, et al.
Veröffentlicht: (2025)
von: Shi, Runwu, et al.
Veröffentlicht: (2025)
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
von: Wang, Jiang, et al.
Veröffentlicht: (2025)
von: Wang, Jiang, et al.
Veröffentlicht: (2025)
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
von: Shi, Runwu, et al.
Veröffentlicht: (2024)
von: Shi, Runwu, et al.
Veröffentlicht: (2024)
Unsupervised Single-Channel Speech Separation with a Diffusion Prior under Speaker-Embedding Guidance
von: Shi, Runwu, et al.
Veröffentlicht: (2025)
von: Shi, Runwu, et al.
Veröffentlicht: (2025)
Knowledge-Augmented Vision Language Models for Underwater Bioacoustic Spectrogram Analysis
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2025)
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2025)
Can all variations within the unified mask-based beamformer framework achieve identical peak extraction performance?
von: Hiroe, Atsuo, et al.
Veröffentlicht: (2024)
von: Hiroe, Atsuo, et al.
Veröffentlicht: (2024)
Robust Bioacoustic Detection via Richly Labelled Synthetic Soundscape Augmentation
von: Soltero, Kaspar, et al.
Veröffentlicht: (2025)
von: Soltero, Kaspar, et al.
Veröffentlicht: (2025)
Few-Shot Bioacoustic Event Detection with Frame-Level Embedding Learning System
von: Zhao, PengYuan, et al.
Veröffentlicht: (2024)
von: Zhao, PengYuan, et al.
Veröffentlicht: (2024)
Temporal Feature Learning in Weakly Labelled Bioacoustic Cetacean Datasets via a Variational Autoencoder and Temporal Convolutional Network: An Interdisciplinary Approach
von: Fonollosa, Laia Garrobé, et al.
Veröffentlicht: (2024)
von: Fonollosa, Laia Garrobé, et al.
Veröffentlicht: (2024)
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
von: Navine, Amanda K., et al.
Veröffentlicht: (2024)
von: Navine, Amanda K., et al.
Veröffentlicht: (2024)
Adaptive Learning via a Negative Selection Strategy for Few-Shot Bioacoustic Event Detection
von: Chen, Yaxiong, et al.
Veröffentlicht: (2024)
von: Chen, Yaxiong, et al.
Veröffentlicht: (2024)
Lightweight Hopfield Neural Networks for Bioacoustic Detection and Call Monitoring of Captive Primates
von: Lomas, Wendy, et al.
Veröffentlicht: (2025)
von: Lomas, Wendy, et al.
Veröffentlicht: (2025)
An Efficient GPU-based Implementation for Noise Robust Sound Source Localization
von: Lin, Zirui, et al.
Veröffentlicht: (2025)
von: Lin, Zirui, et al.
Veröffentlicht: (2025)
Automatic Detection and Annotation of Sperm Whale Codas
von: Gubnitsky, Guy, et al.
Veröffentlicht: (2024)
von: Gubnitsky, Guy, et al.
Veröffentlicht: (2024)
Speaker Embeddings With Weakly Supervised Voice Activity Detection For Efficient Speaker Diarization
von: Thienpondt, Jenthe, et al.
Veröffentlicht: (2024)
von: Thienpondt, Jenthe, et al.
Veröffentlicht: (2024)
Towards High-Fidelity and Controllable Bioacoustic Generation via Enhanced Diffusion Learning
von: Song, Tianyu, et al.
Veröffentlicht: (2025)
von: Song, Tianyu, et al.
Veröffentlicht: (2025)
Distilling Spectrograms into Tokens: Fast and Lightweight Bioacoustic Classification for BirdCLEF+ 2025
von: Miyaguchi, Anthony, et al.
Veröffentlicht: (2025)
von: Miyaguchi, Anthony, et al.
Veröffentlicht: (2025)
Large Language Models and Non-Negative Matrix Factorization for Bioacoustic Signal Decomposition
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
BioME: A Resource-Efficient Bioacoustic Foundational Model for IoT Applications
von: Guimarães, Heitor R., et al.
Veröffentlicht: (2026)
von: Guimarães, Heitor R., et al.
Veröffentlicht: (2026)
Advancing Marine Bioacoustics with Deep Generative Models: A Hybrid Augmentation Strategy for Southern Resident Killer Whale Detection
von: Padovese, Bruno, et al.
Veröffentlicht: (2025)
von: Padovese, Bruno, et al.
Veröffentlicht: (2025)
Towards Weakly Supervised Text-to-Audio Grounding
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
von: Xu, Xuenan, et al.
Veröffentlicht: (2024)
Learning Domain-Robust Bioacoustic Representations for Mosquito Species Classification with Contrastive Learning and Distribution Alignment
von: Hou, Yuanbo, et al.
Veröffentlicht: (2025)
von: Hou, Yuanbo, et al.
Veröffentlicht: (2025)
Weakly Supervised Data Refinement and Flexible Sequence Compression for Efficient Thai LLM-based ASR
von: Shao, Mingchen, et al.
Veröffentlicht: (2025)
von: Shao, Mingchen, et al.
Veröffentlicht: (2025)
Regularized Contrastive Pre-training for Few-shot Bioacoustic Sound Detection
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
Learning When to Trust Which Teacher for Weakly Supervised ASR
von: Agrawal, Aakriti, et al.
Veröffentlicht: (2023)
von: Agrawal, Aakriti, et al.
Veröffentlicht: (2023)
JiTTER: Jigsaw Temporal Transformer for Event Reconstruction for Self-Supervised Sound Event Detection
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
von: Nam, Hyeonuk, et al.
Veröffentlicht: (2025)
Temporal Pooling Strategies for Training-Free Anomalous Sound Detection with Self-Supervised Audio Embeddings
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2026)
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2026)
Towards Deep Active Learning in Avian Bioacoustics
von: Rauch, Lukas, et al.
Veröffentlicht: (2024)
von: Rauch, Lukas, et al.
Veröffentlicht: (2024)
Perch 2.0: The Bittern Lesson for Bioacoustics
von: van Merriënboer, Bart, et al.
Veröffentlicht: (2025)
von: van Merriënboer, Bart, et al.
Veröffentlicht: (2025)
LENS-DF: Deepfake Detection and Temporal Localization for Long-Form Noisy Speech
von: Liu, Xuechen, et al.
Veröffentlicht: (2025)
von: Liu, Xuechen, et al.
Veröffentlicht: (2025)
Data-driven Joint Detection and Localization of Acoustic Reflectors
von: Bicer, H. Nazim, et al.
Veröffentlicht: (2024)
von: Bicer, H. Nazim, et al.
Veröffentlicht: (2024)
Efficient Long-Form Speech Recognition for General Speech In-Context Learning
von: Yen, Hao, et al.
Veröffentlicht: (2024)
von: Yen, Hao, et al.
Veröffentlicht: (2024)
Phone Duration Modeling for Speaker Age Estimation in Children
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2021)
von: Shivakumar, Prashanth Gurunath, et al.
Veröffentlicht: (2021)
Automatic Sound Event Detection and Classification of Great Ape Calls Using Neural Networks
von: Jiang, Zifan, et al.
Veröffentlicht: (2023)
von: Jiang, Zifan, et al.
Veröffentlicht: (2023)
MAGENTA: Magnitude and Geometry-ENhanced Training Approach for Robust Long-Tailed Sound Event Localization and Detection
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
von: Yeow, Jun-Wei, et al.
Veröffentlicht: (2025)
w2v-SELD: A Sound Event Localization and Detection Framework for Self-Supervised Spatial Audio Pre-Training
von: Santos, Orlem Lima dos, et al.
Veröffentlicht: (2023)
von: Santos, Orlem Lima dos, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Cross-Attention with Confidence Weighting for Multi-Channel Audio Alignment
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2025) -
Single-Channel Target Speech Extraction Utilizing Distance and Room Clues
von: Shi, Runwu, et al.
Veröffentlicht: (2025) -
Ecologically-Constrained Task Arithmetic for Multi-Taxa Bioacoustic Classifiers Without Shared Data
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2026) -
Distance Based Single-Channel Target Speech Extraction
von: Shi, Runwu, et al.
Veröffentlicht: (2024) -
Unsupervised Single-Channel Audio Separation with Diffusion Source Priors
von: Shi, Runwu, et al.
Veröffentlicht: (2025)