Perch 2.0: The Bittern Lesson for Bioacoustics
Fuente:
arXiv
Guardado en:
| Autores principales: | van Merriënboer, Bart, Dumoulin, Vincent, Hamer, Jenny, Harrell, Lauren, Burns, Andrea, Denton, Tom |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Perch 2.0 transfers 'whale' to underwater tasks
por: Burns, Andrea, et al.
Publicado: (2025)
por: Burns, Andrea, et al.
Publicado: (2025)
The Search for Squawk: Agile Modeling in Bioacoustics
por: Dumoulin, Vincent, et al.
Publicado: (2025)
por: Dumoulin, Vincent, et al.
Publicado: (2025)
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
por: Navine, Amanda K., et al.
Publicado: (2024)
por: Navine, Amanda K., et al.
Publicado: (2024)
Leveraging tropical reef, bird and unrelated sounds for superior transfer learning in marine bioacoustics
por: Williams, Ben, et al.
Publicado: (2024)
por: Williams, Ben, et al.
Publicado: (2024)
Foundation Models for Bioacoustics -- a Comparative Review
por: Schwinger, Raphael, et al.
Publicado: (2025)
por: Schwinger, Raphael, et al.
Publicado: (2025)
Regularized Contrastive Pre-training for Few-shot Bioacoustic Sound Detection
por: Moummad, Ilyass, et al.
Publicado: (2023)
por: Moummad, Ilyass, et al.
Publicado: (2023)
Aggregation Strategies for Efficient Annotation of Bioacoustic Sound Events Using Active Learning
por: Lindholm, Richard, et al.
Publicado: (2025)
por: Lindholm, Richard, et al.
Publicado: (2025)
Multi-Representation Attention Framework for Underwater Bioacoustic Denoising and Recognition
por: Razig, Amine, et al.
Publicado: (2025)
por: Razig, Amine, et al.
Publicado: (2025)
Robust Bioacoustic Detection via Richly Labelled Synthetic Soundscape Augmentation
por: Soltero, Kaspar, et al.
Publicado: (2025)
por: Soltero, Kaspar, et al.
Publicado: (2025)
Towards High-Fidelity and Controllable Bioacoustic Generation via Enhanced Diffusion Learning
por: Song, Tianyu, et al.
Publicado: (2025)
por: Song, Tianyu, et al.
Publicado: (2025)
Distilling Spectrograms into Tokens: Fast and Lightweight Bioacoustic Classification for BirdCLEF+ 2025
por: Miyaguchi, Anthony, et al.
Publicado: (2025)
por: Miyaguchi, Anthony, et al.
Publicado: (2025)
Few-Shot Bioacoustic Event Detection with Frame-Level Embedding Learning System
por: Zhao, PengYuan, et al.
Publicado: (2024)
por: Zhao, PengYuan, et al.
Publicado: (2024)
Large Language Models and Non-Negative Matrix Factorization for Bioacoustic Signal Decomposition
por: Torabi, Yasaman, et al.
Publicado: (2025)
por: Torabi, Yasaman, et al.
Publicado: (2025)
BioME: A Resource-Efficient Bioacoustic Foundational Model for IoT Applications
por: Guimarães, Heitor R., et al.
Publicado: (2026)
por: Guimarães, Heitor R., et al.
Publicado: (2026)
Adaptive Learning via a Negative Selection Strategy for Few-Shot Bioacoustic Event Detection
por: Chen, Yaxiong, et al.
Publicado: (2024)
por: Chen, Yaxiong, et al.
Publicado: (2024)
Learning Domain-Robust Bioacoustic Representations for Mosquito Species Classification with Contrastive Learning and Distribution Alignment
por: Hou, Yuanbo, et al.
Publicado: (2025)
por: Hou, Yuanbo, et al.
Publicado: (2025)
NatureLM-audio: an Audio-Language Foundation Model for Bioacoustics
por: Robinson, David, et al.
Publicado: (2024)
por: Robinson, David, et al.
Publicado: (2024)
Hybrid Disagreement-Diversity Active Learning for Bioacoustic Sound Event Detection
por: Zhang, Shiqi, et al.
Publicado: (2025)
por: Zhang, Shiqi, et al.
Publicado: (2025)
Who Said What WSW 2.0? Enhanced Automated Analysis of Preschool Classroom Speech
por: Sun, Anchen, et al.
Publicado: (2025)
por: Sun, Anchen, et al.
Publicado: (2025)
Towards Deep Active Learning in Avian Bioacoustics
por: Rauch, Lukas, et al.
Publicado: (2024)
por: Rauch, Lukas, et al.
Publicado: (2024)
Lightweight Hopfield Neural Networks for Bioacoustic Detection and Call Monitoring of Captive Primates
por: Lomas, Wendy, et al.
Publicado: (2025)
por: Lomas, Wendy, et al.
Publicado: (2025)
Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data
por: Nihal, Ragib Amin, et al.
Publicado: (2025)
por: Nihal, Ragib Amin, et al.
Publicado: (2025)
Janssen 2.0: Audio Inpainting in the Time-frequency Domain
por: Mokrý, Ondřej, et al.
Publicado: (2024)
por: Mokrý, Ondřej, et al.
Publicado: (2024)
Mixture of Experts Fusion for Fake Audio Detection Using Frozen wav2vec 2.0
por: Wang, Zhiyong, et al.
Publicado: (2024)
por: Wang, Zhiyong, et al.
Publicado: (2024)
Enhancing Speaker Verification with w2v-BERT 2.0 and Knowledge Distillation guided Structured Pruning
por: Li, Ze, et al.
Publicado: (2025)
por: Li, Ze, et al.
Publicado: (2025)
Advancing Marine Bioacoustics with Deep Generative Models: A Hybrid Augmentation Strategy for Southern Resident Killer Whale Detection
por: Padovese, Bruno, et al.
Publicado: (2025)
por: Padovese, Bruno, et al.
Publicado: (2025)
BirdSet: A Large-Scale Dataset for Audio Classification in Avian Bioacoustics
por: Rauch, Lukas, et al.
Publicado: (2024)
por: Rauch, Lukas, et al.
Publicado: (2024)
LiLAC: A Lightweight Latent ControlNet for Musical Audio Generation
por: Baker, Tom, et al.
Publicado: (2025)
por: Baker, Tom, et al.
Publicado: (2025)
Optimising MFCC parameters for the automatic detection of respiratory diseases
por: Yan, Yuyang, et al.
Publicado: (2024)
por: Yan, Yuyang, et al.
Publicado: (2024)
RepCodec: A Speech Representation Codec for Speech Tokenization
por: Huang, Zhichao, et al.
Publicado: (2023)
por: Huang, Zhichao, et al.
Publicado: (2023)
SCRAPL: Scattering Transform with Random Paths for Machine Learning
por: Mitcheltree, Christopher, et al.
Publicado: (2026)
por: Mitcheltree, Christopher, et al.
Publicado: (2026)
Instabilities in Convnets for Raw Audio
por: Haider, Daniel, et al.
Publicado: (2023)
por: Haider, Daniel, et al.
Publicado: (2023)
Mixture of Mixups for Multi-label Classification of Rare Anuran Sounds
por: Moummad, Ilyass, et al.
Publicado: (2024)
por: Moummad, Ilyass, et al.
Publicado: (2024)
Hold Me Tight: Stable Encoder-Decoder Design for Speech Enhancement
por: Haider, Daniel, et al.
Publicado: (2024)
por: Haider, Daniel, et al.
Publicado: (2024)
Big Data Approaches to Bovine Bioacoustics: A FAIR-Compliant Dataset and Scalable ML Framework for Precision Livestock Welfare
por: Kate, Mayuri, et al.
Publicado: (2025)
por: Kate, Mayuri, et al.
Publicado: (2025)
Enhancement by postfiltering for speech and audio coding in ad-hoc sensor networks
por: Das, Sneha, et al.
Publicado: (2020)
por: Das, Sneha, et al.
Publicado: (2020)
Latent Granular Resynthesis using Neural Audio Codecs
por: Tokui, Nao, et al.
Publicado: (2025)
por: Tokui, Nao, et al.
Publicado: (2025)
The Effect of Batch Size on Contrastive Self-Supervised Speech Representation Learning
por: Vaessen, Nik, et al.
Publicado: (2024)
por: Vaessen, Nik, et al.
Publicado: (2024)
Musical Metamerism with Time--Frequency Scattering
por: Lostanlen, Vincent, et al.
Publicado: (2026)
por: Lostanlen, Vincent, et al.
Publicado: (2026)
NoiseBandNet: Controllable Time-Varying Neural Synthesis of Sound Effects Using Filterbanks
por: Barahona-Ríos, Adrián, et al.
Publicado: (2023)
por: Barahona-Ríos, Adrián, et al.
Publicado: (2023)
Ejemplares similares
-
Perch 2.0 transfers 'whale' to underwater tasks
por: Burns, Andrea, et al.
Publicado: (2025) -
The Search for Squawk: Agile Modeling in Bioacoustics
por: Dumoulin, Vincent, et al.
Publicado: (2025) -
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
por: Navine, Amanda K., et al.
Publicado: (2024) -
Leveraging tropical reef, bird and unrelated sounds for superior transfer learning in marine bioacoustics
por: Williams, Ben, et al.
Publicado: (2024) -
Foundation Models for Bioacoustics -- a Comparative Review
por: Schwinger, Raphael, et al.
Publicado: (2025)