Improving acoustic drone detection generalization through pretraining and data augmentation
Fuente:
arXiv
Salvato in:
| Autori principali: | Reuter, Paul M., Ohlenbusch, Mattes, Rollwage, Christian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Speech-dependent Data Augmentation for Own Voice Reconstruction with Hearable Microphones in Noisy Environments
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2024)
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2024)
Multi-Microphone Noise Data Augmentation for DNN-based Own Voice Reconstruction for Hearables in Noisy Environments
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2023)
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2023)
Low-Complexity Own Voice Reconstruction for Hearables with an In-Ear Microphone
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2024)
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2024)
Modeling of Speech-dependent Own Voice Transfer Characteristics for Hearables with In-ear Microphones
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2023)
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2023)
Speech-dependent Modeling of Own Voice Transfer Characteristics for In-ear Microphones in Hearables
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2023)
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2023)
Comparison of Knowledge Distillation Methods for Low-complexity Multi-microphone Speech Enhancement using the FT-JNF Architecture
di: Metzger, Robert, et al.
Pubblicazione: (2025)
di: Metzger, Robert, et al.
Pubblicazione: (2025)
Subjective quality evaluation of personalized own voice reconstruction systems
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2025)
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2025)
PAS-SE: Personalized Auxiliary-Sensor Speech Enhancement for Voice Pickup in Hearables
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2025)
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2025)
Evaluating pretrained speech embedding systems for dysarthria detection across heterogenous datasets
di: Wihlborg, Lovisa, et al.
Pubblicazione: (2025)
di: Wihlborg, Lovisa, et al.
Pubblicazione: (2025)
Fine-tune the pretrained ATST model for sound event detection
di: Shao, Nian, et al.
Pubblicazione: (2023)
di: Shao, Nian, et al.
Pubblicazione: (2023)
Stereo sound event localization and detection based on PSELDnet pretraining and BiMamba sequence modeling
di: Gao, Wenmiao, et al.
Pubblicazione: (2025)
di: Gao, Wenmiao, et al.
Pubblicazione: (2025)
On the influence of language similarity in non-target speaker verification trials
di: Reuter, Paul M., et al.
Pubblicazione: (2025)
di: Reuter, Paul M., et al.
Pubblicazione: (2025)
Online incremental learning for audio classification using a pretrained audio model
di: Mulimani, Manjunath, et al.
Pubblicazione: (2025)
di: Mulimani, Manjunath, et al.
Pubblicazione: (2025)
Automatic acoustic detection of birds through deep learning: the first Bird Audio Detection challenge
di: Stowell, Dan, et al.
Pubblicazione: (2018)
di: Stowell, Dan, et al.
Pubblicazione: (2018)
Sound event localization and detection based on crnn using rectangular filters and channel rotation data augmentation
di: Ronchini, Francesca, et al.
Pubblicazione: (2020)
di: Ronchini, Francesca, et al.
Pubblicazione: (2020)
Complete reconstruction of the tongue contour through acoustic to articulatory inversion using real-time MRI data
di: Azzouz, Sofiane, et al.
Pubblicazione: (2024)
di: Azzouz, Sofiane, et al.
Pubblicazione: (2024)
CardioPHON: Quality assessment and self-supervised pretraining for screening of cardiac function based on phonocardiogram recordings
di: Despotovic, Vladimir, et al.
Pubblicazione: (2025)
di: Despotovic, Vladimir, et al.
Pubblicazione: (2025)
Human-CLAP: Human-perception-based contrastive language-audio pretraining
di: Takano, Taisei, et al.
Pubblicazione: (2025)
di: Takano, Taisei, et al.
Pubblicazione: (2025)
Sample adaptive data augmentation with progressive scheduling
di: Lu, Hongxuan, et al.
Pubblicazione: (2024)
di: Lu, Hongxuan, et al.
Pubblicazione: (2024)
Ultralow-power standoff acoustic leak detection
di: Hasselbeck, Michael P.
Pubblicazione: (2025)
di: Hasselbeck, Michael P.
Pubblicazione: (2025)
Monaural speech enhancement on drone via Adapter based transfer learning
di: Chen, Xingyu, et al.
Pubblicazione: (2024)
di: Chen, Xingyu, et al.
Pubblicazione: (2024)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
di: Deng, Qingkun, et al.
Pubblicazione: (2024)
di: Deng, Qingkun, et al.
Pubblicazione: (2024)
Towards Improved Speech Recognition through Optimized Synthetic Data Generation
di: Perrin, Yanis, et al.
Pubblicazione: (2025)
di: Perrin, Yanis, et al.
Pubblicazione: (2025)
SelfTTS: cross-speaker style transfer through explicit embedding disentanglement and self-refinement using self-augmentation
di: Ueda, Lucas H., et al.
Pubblicazione: (2026)
di: Ueda, Lucas H., et al.
Pubblicazione: (2026)
Neural acoustic multipole splatting for room impulse response synthesis
di: Baek, Geonwoo, et al.
Pubblicazione: (2025)
di: Baek, Geonwoo, et al.
Pubblicazione: (2025)
Deep, data-driven modeling of room acoustics: literature review and research perspectives
di: van Waterschoot, Toon
Pubblicazione: (2025)
di: van Waterschoot, Toon
Pubblicazione: (2025)
Analyzing the relationships between pretraining language, phonetic, tonal, and speaker information in self-supervised speech models
di: Gubian, Michele, et al.
Pubblicazione: (2025)
di: Gubian, Michele, et al.
Pubblicazione: (2025)
Real time fault detection in 3D printers using Convolutional Neural Networks and acoustic signals
di: Waheed, Muhammad Fasih, et al.
Pubblicazione: (2026)
di: Waheed, Muhammad Fasih, et al.
Pubblicazione: (2026)
Improving Query-by-Vocal Imitation with Contrastive Learning and Audio Pretraining
di: Greif, Jonathan, et al.
Pubblicazione: (2024)
di: Greif, Jonathan, et al.
Pubblicazione: (2024)
GLAP: General contrastive audio-text pretraining across domains and languages
di: Dinkel, Heinrich, et al.
Pubblicazione: (2025)
di: Dinkel, Heinrich, et al.
Pubblicazione: (2025)
On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
di: Li, Junjie, et al.
Pubblicazione: (2024)
di: Li, Junjie, et al.
Pubblicazione: (2024)
Perceptual implications of simplifying geometrical acoustics models for Ambisonics-based binaural reverberation
di: Martin, Vincent, et al.
Pubblicazione: (2024)
di: Martin, Vincent, et al.
Pubblicazione: (2024)
Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment
di: Boeddeker, Christoph, et al.
Pubblicazione: (2024)
di: Boeddeker, Christoph, et al.
Pubblicazione: (2024)
Improving child speech recognition with augmented child-like speech
di: Zhang, Yuanyuan, et al.
Pubblicazione: (2024)
di: Zhang, Yuanyuan, et al.
Pubblicazione: (2024)
Physics-informed neural network for acoustic resonance analysis in a one-dimensional acoustic tube
di: Yokota, Kazuya, et al.
Pubblicazione: (2023)
di: Yokota, Kazuya, et al.
Pubblicazione: (2023)
Low-resource keyword spotting using contrastively trained transformer acoustic word embeddings
di: Herreilers, Julian, et al.
Pubblicazione: (2025)
di: Herreilers, Julian, et al.
Pubblicazione: (2025)
Generative AI-based data augmentation for improved bioacoustic classification in noisy environments
di: Gibbons, Anthony, et al.
Pubblicazione: (2024)
di: Gibbons, Anthony, et al.
Pubblicazione: (2024)
Directional reflection modeling via wavenumber-domain reflection coefficient for 3D acoustic field simulation
di: Hoshika, Satoshi, et al.
Pubblicazione: (2026)
di: Hoshika, Satoshi, et al.
Pubblicazione: (2026)
Discovering phoneme-specific critical articulators through a data-driven approach
di: Bandekar, Jesuraj, et al.
Pubblicazione: (2025)
di: Bandekar, Jesuraj, et al.
Pubblicazione: (2025)
Effectively obtaining acoustic, visual and textual data from videos
di: León, Jorge E., et al.
Pubblicazione: (2025)
di: León, Jorge E., et al.
Pubblicazione: (2025)
Documenti analoghi
-
Speech-dependent Data Augmentation for Own Voice Reconstruction with Hearable Microphones in Noisy Environments
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2024) -
Multi-Microphone Noise Data Augmentation for DNN-based Own Voice Reconstruction for Hearables in Noisy Environments
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2023) -
Low-Complexity Own Voice Reconstruction for Hearables with an In-Ear Microphone
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2024) -
Modeling of Speech-dependent Own Voice Transfer Characteristics for Hearables with In-ear Microphones
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2023) -
Speech-dependent Modeling of Own Voice Transfer Characteristics for In-ear Microphones in Hearables
di: Ohlenbusch, Mattes, et al.
Pubblicazione: (2023)