Utilizing synthetic training data for the supervised classification of rat ultrasonic vocalizations
Fuente:
arXiv
Salvato in:
| Autori principali: | Scott, K. Jack, Speers, Lucinda J., Bilkey, David K. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enhancing the analysis of murine neonatal ultrasonic vocalizations: Development, evaluation, and application of different mathematical models
di: Herdt, Rudolf, et al.
Pubblicazione: (2024)
di: Herdt, Rudolf, et al.
Pubblicazione: (2024)
Tempo estimation as fully self-supervised binary classification
di: Henkel, Florian, et al.
Pubblicazione: (2024)
di: Henkel, Florian, et al.
Pubblicazione: (2024)
Acoustic-to-articulatory inversion for dysarthric speech: Are pre-trained self-supervised representations favorable?
di: Maharana, Sarthak Kumar, et al.
Pubblicazione: (2023)
di: Maharana, Sarthak Kumar, et al.
Pubblicazione: (2023)
The impact of non-target events in synthetic soundscapes for sound event detection
di: Ronchini, Francesca, et al.
Pubblicazione: (2021)
di: Ronchini, Francesca, et al.
Pubblicazione: (2021)
A benchmark of state-of-the-art sound event detection systems evaluated on synthetic soundscapes
di: Ronchini, Francesca, et al.
Pubblicazione: (2022)
di: Ronchini, Francesca, et al.
Pubblicazione: (2022)
An LSTM-Based Chord Generation System Using Chroma Histogram Representations
di: Hardwick, Jack
Pubblicazione: (2024)
di: Hardwick, Jack
Pubblicazione: (2024)
Generalization in birdsong classification: impact of transfer learning methods and dataset characteristics
di: Ghani, Burooj, et al.
Pubblicazione: (2024)
di: Ghani, Burooj, et al.
Pubblicazione: (2024)
An Experimental Comparison Of Multi-view Self-supervised Methods For Music Tagging
di: Meseguer-Brocal, Gabriel, et al.
Pubblicazione: (2024)
di: Meseguer-Brocal, Gabriel, et al.
Pubblicazione: (2024)
Clustering-based hard negative sampling for supervised contrastive speaker verification
di: Masztalski, Piotr, et al.
Pubblicazione: (2025)
di: Masztalski, Piotr, et al.
Pubblicazione: (2025)
Causal Self-supervised Pretrained Frontend with Predictive Code for Speech Separation
di: Wang, Wupeng, et al.
Pubblicazione: (2025)
di: Wang, Wupeng, et al.
Pubblicazione: (2025)
RCT: Random Consistency Training for Semi-supervised Sound Event Detection
di: Shao, Nian, et al.
Pubblicazione: (2021)
di: Shao, Nian, et al.
Pubblicazione: (2021)
Sound event localization and classification using WASN in Outdoor Environment
di: Zhang, Dongzhe, et al.
Pubblicazione: (2024)
di: Zhang, Dongzhe, et al.
Pubblicazione: (2024)
Audio Classification of Low Feature Spectrograms Utilizing Convolutional Neural Networks
di: Elias, Noel
Pubblicazione: (2024)
di: Elias, Noel
Pubblicazione: (2024)
Utilizing TTS Synthesized Data for Efficient Development of Keyword Spotting Model
di: Park, Hyun Jin, et al.
Pubblicazione: (2024)
di: Park, Hyun Jin, et al.
Pubblicazione: (2024)
Self-supervised learning of speech representations with Dutch archival data
di: Vaessen, Nik, et al.
Pubblicazione: (2025)
di: Vaessen, Nik, et al.
Pubblicazione: (2025)
Cross-Referencing Self-Training Network for Sound Event Detection in Audio Mixtures
di: Park, Sangwook, et al.
Pubblicazione: (2021)
di: Park, Sangwook, et al.
Pubblicazione: (2021)
Dirichlet process mixture model based on topologically augmented signal representation for clustering infant vocalizations
di: Bonafos, Guillem, et al.
Pubblicazione: (2024)
di: Bonafos, Guillem, et al.
Pubblicazione: (2024)
Dementia classification from spontaneous speech using wrapper-based feature selection
di: Niemelä, Marko, et al.
Pubblicazione: (2025)
di: Niemelä, Marko, et al.
Pubblicazione: (2025)
An Attention Long Short-Term Memory based system for automatic classification of speech intelligibility
di: Fernández-Díaz, Miguel, et al.
Pubblicazione: (2024)
di: Fernández-Díaz, Miguel, et al.
Pubblicazione: (2024)
Description on IEEE ICME 2024 Grand Challenge: Semi-supervised Acoustic Scene Classification under Domain Shift
di: Bai, Jisheng, et al.
Pubblicazione: (2024)
di: Bai, Jisheng, et al.
Pubblicazione: (2024)
Towards Sub-millisecond Latency Real-Time Speech Enhancement Models on Hearables
di: Dementyev, Artem, et al.
Pubblicazione: (2024)
di: Dementyev, Artem, et al.
Pubblicazione: (2024)
Towards generalisable and calibrated synthetic speech detection with self-supervised representations
di: Pascu, Octavian, et al.
Pubblicazione: (2023)
di: Pascu, Octavian, et al.
Pubblicazione: (2023)
Adversarial training of Keyword Spotting to Minimize TTS Data Overfitting
di: Park, Hyun Jin, et al.
Pubblicazione: (2024)
di: Park, Hyun Jin, et al.
Pubblicazione: (2024)
Regularized Contrastive Pre-training for Few-shot Bioacoustic Sound Detection
di: Moummad, Ilyass, et al.
Pubblicazione: (2023)
di: Moummad, Ilyass, et al.
Pubblicazione: (2023)
Efficient Adapter Tuning of Pre-trained Speech Models for Automatic Speaker Verification
di: Sang, Mufan, et al.
Pubblicazione: (2024)
di: Sang, Mufan, et al.
Pubblicazione: (2024)
An Enhanced Audio Feature Tailored for Anomalous Sound Detection Based on Pre-trained Models
di: Zhong, Guirui, et al.
Pubblicazione: (2025)
di: Zhong, Guirui, et al.
Pubblicazione: (2025)
Cross-Domain Knowledge Transfer for Underwater Acoustic Classification Using Pre-trained Models
di: Mohammadi, Amirmohammad, et al.
Pubblicazione: (2024)
di: Mohammadi, Amirmohammad, et al.
Pubblicazione: (2024)
Identifying birdsong syllables without labelled data
di: Teng, Mélisande, et al.
Pubblicazione: (2025)
di: Teng, Mélisande, et al.
Pubblicazione: (2025)
SLEEPING-DISCO 9M: A large-scale pre-training dataset for generative music modeling
di: Ahmed, Tawsif, et al.
Pubblicazione: (2025)
di: Ahmed, Tawsif, et al.
Pubblicazione: (2025)
Adaptive vector steering: A training-free, layer-wise intervention for hallucination mitigation in large audio and multimodal models
di: Lin, Tsung-En, et al.
Pubblicazione: (2025)
di: Lin, Tsung-En, et al.
Pubblicazione: (2025)
UniPET-SPK: A Unified Framework for Parameter-Efficient Tuning of Pre-trained Speech Models for Robust Speaker Verification
di: Sang, Mufan, et al.
Pubblicazione: (2025)
di: Sang, Mufan, et al.
Pubblicazione: (2025)
Towards Robust Few-shot Class Incremental Learning in Audio Classification using Contrastive Representation
di: Singh, Riyansha, et al.
Pubblicazione: (2024)
di: Singh, Riyansha, et al.
Pubblicazione: (2024)
Synthetic data enables context-aware bioacoustic sound event detection
di: Hoffman, Benjamin, et al.
Pubblicazione: (2025)
di: Hoffman, Benjamin, et al.
Pubblicazione: (2025)
Fast Timing-Conditioned Latent Audio Diffusion
di: Evans, Zach, et al.
Pubblicazione: (2024)
di: Evans, Zach, et al.
Pubblicazione: (2024)
Test-Time Training for Speech Enhancement
di: Behera, Avishkar, et al.
Pubblicazione: (2025)
di: Behera, Avishkar, et al.
Pubblicazione: (2025)
Context-aware child-directed speech detection from long-form recordings
di: Charlot, Théo, et al.
Pubblicazione: (2026)
di: Charlot, Théo, et al.
Pubblicazione: (2026)
Meta-Learning in Audio and Speech Processing: An End to End Comprehensive Review
di: Raimon, Athul, et al.
Pubblicazione: (2024)
di: Raimon, Athul, et al.
Pubblicazione: (2024)
LLM supervised Pre-training for Multimodal Emotion Recognition in Conversations
di: Dutta, Soumya, et al.
Pubblicazione: (2025)
di: Dutta, Soumya, et al.
Pubblicazione: (2025)
Who Said What WSW 2.0? Enhanced Automated Analysis of Preschool Classroom Speech
di: Sun, Anchen, et al.
Pubblicazione: (2025)
di: Sun, Anchen, et al.
Pubblicazione: (2025)
[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic
di: Choi, Kwanghee, et al.
Pubblicazione: (2026)
di: Choi, Kwanghee, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Enhancing the analysis of murine neonatal ultrasonic vocalizations: Development, evaluation, and application of different mathematical models
di: Herdt, Rudolf, et al.
Pubblicazione: (2024) -
Tempo estimation as fully self-supervised binary classification
di: Henkel, Florian, et al.
Pubblicazione: (2024) -
Acoustic-to-articulatory inversion for dysarthric speech: Are pre-trained self-supervised representations favorable?
di: Maharana, Sarthak Kumar, et al.
Pubblicazione: (2023) -
The impact of non-target events in synthetic soundscapes for sound event detection
di: Ronchini, Francesca, et al.
Pubblicazione: (2021) -
A benchmark of state-of-the-art sound event detection systems evaluated on synthetic soundscapes
di: Ronchini, Francesca, et al.
Pubblicazione: (2022)