Uncertainty Calibration of Multi-Label Bird Sound Classifiers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Schwinger, Raphael, McEwen, Ben, Kather, Vincent S., Heinrich, René, Rauch, Lukas, Tomforde, Sven |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Foundation Models for Bioacoustics -- a Comparative Review
von: Schwinger, Raphael, et al.
Veröffentlicht: (2025)
von: Schwinger, Raphael, et al.
Veröffentlicht: (2025)
Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio Classification
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
BirdSet: A Large-Scale Dataset for Audio Classification in Avian Bioacoustics
von: Rauch, Lukas, et al.
Veröffentlicht: (2024)
von: Rauch, Lukas, et al.
Veröffentlicht: (2024)
Can Masked Autoencoders Also Listen to Birds?
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
von: Rauch, Lukas, et al.
Veröffentlicht: (2025)
Audio-to-Image Bird Species Retrieval without Audio-Image Pairs via Text Distillation
von: Moummad, Ilyass, et al.
Veröffentlicht: (2026)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2026)
Sound Signal Synthesis with Auxiliary Classifier GAN, COVID-19 cough as an example
von: Saleh, Yahya Sherif Solayman Mohamed, et al.
Veröffentlicht: (2025)
von: Saleh, Yahya Sherif Solayman Mohamed, et al.
Veröffentlicht: (2025)
Sound and Music Biases in Deep Music Transcription Models: A Systematic Analysis
von: Marták, Lukáš Samuel, et al.
Veröffentlicht: (2025)
von: Marták, Lukáš Samuel, et al.
Veröffentlicht: (2025)
Self-Supervised Learning for Few-Shot Bird Sound Classification
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2023)
Audio Question Answering with GRPO-Based Fine-Tuning and Calibrated Segment-Level Predictions
von: Gibier, Marcel, et al.
Veröffentlicht: (2025)
von: Gibier, Marcel, et al.
Veröffentlicht: (2025)
Ecologically-Constrained Task Arithmetic for Multi-Taxa Bioacoustic Classifiers Without Shared Data
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2026)
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2026)
Mixture of Mixups for Multi-label Classification of Rare Anuran Sounds
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2024)
Spectrotemporal Modulation: Efficient and Interpretable Feature Representation for Classifying Speech, Music, and Environmental Sounds
von: Chang, Andrew, et al.
Veröffentlicht: (2025)
von: Chang, Andrew, et al.
Veröffentlicht: (2025)
Audio Flamingo Sound-CoT Technical Report: Improving Chain-of-Thought Reasoning in Sound Understanding
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
DeepForestSound: a multi-species automatic detector for passive acoustic monitoring in African tropical forests, a case study in Kibale National Park
von: Dubus, Gabriel, et al.
Veröffentlicht: (2026)
von: Dubus, Gabriel, et al.
Veröffentlicht: (2026)
Exploring Meta Information for Audio-based Zero-shot Bird Classification
von: Gebhard, Alexander, et al.
Veröffentlicht: (2023)
von: Gebhard, Alexander, et al.
Veröffentlicht: (2023)
Semantic-Aware Confidence Calibration for Automated Audio Captioning
von: Dunker, Lucas, et al.
Veröffentlicht: (2025)
von: Dunker, Lucas, et al.
Veröffentlicht: (2025)
Benchmarking LLMs on the Massive Sound Embedding Benchmark (MSEB)
von: Allauzen, Cyril, et al.
Veröffentlicht: (2026)
von: Allauzen, Cyril, et al.
Veröffentlicht: (2026)
Coverage-Guaranteed Speech Emotion Recognition via Calibrated Uncertainty-Adaptive Prediction Sets
von: Jia, Zijun, et al.
Veröffentlicht: (2025)
von: Jia, Zijun, et al.
Veröffentlicht: (2025)
MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms
von: Ristori, Eleonora, et al.
Veröffentlicht: (2025)
von: Ristori, Eleonora, et al.
Veröffentlicht: (2025)
Unleashing the Power of Natural Audio Featuring Multiple Sound Sources
von: Cheng, Xize, et al.
Veröffentlicht: (2025)
von: Cheng, Xize, et al.
Veröffentlicht: (2025)
From Weak to Strong Sound Event Labels using Adaptive Change-Point Detection and Active Learning
von: Martinsson, John, et al.
Veröffentlicht: (2024)
von: Martinsson, John, et al.
Veröffentlicht: (2024)
ESTM: An Enhanced Dual-Branch Spectral-Temporal Mamba for Anomalous Sound Detection
von: Ma, Chengyuan, et al.
Veröffentlicht: (2025)
von: Ma, Chengyuan, et al.
Veröffentlicht: (2025)
Improving Deep Learning-based Respiratory Sound Analysis with Frequency Selection and Attention Mechanism
von: Fraihi, Nouhaila, et al.
Veröffentlicht: (2025)
von: Fraihi, Nouhaila, et al.
Veröffentlicht: (2025)
CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents
von: Banerjee, Adhiraj, et al.
Veröffentlicht: (2025)
von: Banerjee, Adhiraj, et al.
Veröffentlicht: (2025)
How to Label Resynthesized Audio: The Dual Role of Neural Audio Codecs in Audio Deepfake Detection
von: Xiao, Yixuan, et al.
Veröffentlicht: (2026)
von: Xiao, Yixuan, et al.
Veröffentlicht: (2026)
A Framework for Evaluating Faithfulness in Explainable AI for Machine Anomalous Sound Detection Using Frequency-Band Perturbation
von: Buck, Alexander, et al.
Veröffentlicht: (2026)
von: Buck, Alexander, et al.
Veröffentlicht: (2026)
Feature Aggregation in Joint Sound Classification and Localization Neural Networks
von: Healy, Brendan, et al.
Veröffentlicht: (2023)
von: Healy, Brendan, et al.
Veröffentlicht: (2023)
Sound Tagging in Infant-centric Home Soundscapes
von: Khan, Mohammad Nur Hossain, et al.
Veröffentlicht: (2024)
von: Khan, Mohammad Nur Hossain, et al.
Veröffentlicht: (2024)
Improving Out-of-Domain Audio Deepfake Detection via Layer Selection and Fusion of SSL-Based Countermeasures
von: Serrano, Pierre, et al.
Veröffentlicht: (2025)
von: Serrano, Pierre, et al.
Veröffentlicht: (2025)
Segmentwise Pruning in Audio-Language Models
von: Gibier, Marcel, et al.
Veröffentlicht: (2025)
von: Gibier, Marcel, et al.
Veröffentlicht: (2025)
From Uncertainty to Precision: Enhancing Binary Classifier Performance through Calibration
von: Machado, Agathe Fernandes, et al.
Veröffentlicht: (2024)
von: Machado, Agathe Fernandes, et al.
Veröffentlicht: (2024)
SoundMorpher: Perceptually-Uniform Sound Morphing with Diffusion Model
von: Niu, Xinlei, et al.
Veröffentlicht: (2024)
von: Niu, Xinlei, et al.
Veröffentlicht: (2024)
Woosh: A Sound Effects Foundation Model
von: Hadjeres, Gaëtan, et al.
Veröffentlicht: (2026)
von: Hadjeres, Gaëtan, et al.
Veröffentlicht: (2026)
SoundSculpt: Direction and Semantics Driven Ambisonic Target Sound Extraction
von: Chen, Tuochao, et al.
Veröffentlicht: (2025)
von: Chen, Tuochao, et al.
Veröffentlicht: (2025)
LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
Multi-Task Pseudo-Label Learning for Non-Intrusive Speech Quality Assessment Model
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
von: Zezario, Ryandhimas E., et al.
Veröffentlicht: (2023)
Count The Notes: Histogram-Based Supervision for Automatic Music Transcription
von: Yaffe, Jonathan, et al.
Veröffentlicht: (2025)
von: Yaffe, Jonathan, et al.
Veröffentlicht: (2025)
Motif Mining and Unsupervised Representation Learning for BirdCLEF 2022
von: Miyaguchi, Anthony, et al.
Veröffentlicht: (2022)
von: Miyaguchi, Anthony, et al.
Veröffentlicht: (2022)
VioPTT: Violin Technique-Aware Transcription from Synthetic Data Augmentation
von: Wang, Ting-Kang, et al.
Veröffentlicht: (2025)
von: Wang, Ting-Kang, et al.
Veröffentlicht: (2025)
The iNaturalist Sounds Dataset
von: Chasmai, Mustafa, et al.
Veröffentlicht: (2025)
von: Chasmai, Mustafa, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Foundation Models for Bioacoustics -- a Comparative Review
von: Schwinger, Raphael, et al.
Veröffentlicht: (2025) -
Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio Classification
von: Rauch, Lukas, et al.
Veröffentlicht: (2025) -
BirdSet: A Large-Scale Dataset for Audio Classification in Avian Bioacoustics
von: Rauch, Lukas, et al.
Veröffentlicht: (2024) -
Can Masked Autoencoders Also Listen to Birds?
von: Rauch, Lukas, et al.
Veröffentlicht: (2025) -
Audio-to-Image Bird Species Retrieval without Audio-Image Pairs via Text Distillation
von: Moummad, Ilyass, et al.
Veröffentlicht: (2026)