Decodable but not structured: linear probing enables Underwater Acoustic Target Recognition with pretrained audio embeddings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hummel, Hilde I., Bhulai, Sandjai, van der Mei, Rob D., Ghani, Burooj |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Computation of Generalized Embeddings for Underwater Acoustic Target Recognition using Contrastive Learning
von: Hummel, Hilde I., et al.
Veröffentlicht: (2025)
von: Hummel, Hilde I., et al.
Veröffentlicht: (2025)
Automated data curation for self-supervised learning in underwater acoustic analysis
von: Hummel, Hilde I, et al.
Veröffentlicht: (2025)
von: Hummel, Hilde I, et al.
Veröffentlicht: (2025)
InsectSet459: an open dataset of insect sounds for bioacoustic machine learning
von: Faiß, Marius, et al.
Veröffentlicht: (2025)
von: Faiß, Marius, et al.
Veröffentlicht: (2025)
Human-CLAP: Human-perception-based contrastive language-audio pretraining
von: Takano, Taisei, et al.
Veröffentlicht: (2025)
von: Takano, Taisei, et al.
Veröffentlicht: (2025)
Unraveling Complex Data Diversity in Underwater Acoustic Target Recognition through Convolution-based Mixture of Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
From Human Speech to Ocean Signals: Transferring Speech Large Models for Underwater Acoustic Target Recognition
von: Huang, Mengcheng, et al.
Veröffentlicht: (2026)
von: Huang, Mengcheng, et al.
Veröffentlicht: (2026)
Neural Edge Histogram Descriptors for Underwater Acoustic Target Recognition
von: Agashe, Atharva, et al.
Veröffentlicht: (2025)
von: Agashe, Atharva, et al.
Veröffentlicht: (2025)
Underwater-Art: Expanding Information Perspectives With Text Templates For Underwater Acoustic Target Recognition
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
GLAP: General contrastive audio-text pretraining across domains and languages
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2025)
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2025)
Underwater Acoustic Signal Recognition Based on Salient Feature
von: Chen, Minghao
Veröffentlicht: (2023)
von: Chen, Minghao
Veröffentlicht: (2023)
WavJEPA: Semantic learning unlocks robust audio foundation models for raw waveforms
von: Yuksel, Goksenin, et al.
Veröffentlicht: (2025)
von: Yuksel, Goksenin, et al.
Veröffentlicht: (2025)
DEMONet: Underwater Acoustic Target Recognition based on Multi-Expert Network and Cross-Temporal Variational Autoencoder
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Advancing Robust Underwater Acoustic Target Recognition through Multi-task Learning and Multi-Gate Mixture-of-Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Generalization in birdsong classification: impact of transfer learning methods and dataset characteristics
von: Ghani, Burooj, et al.
Veröffentlicht: (2024)
von: Ghani, Burooj, et al.
Veröffentlicht: (2024)
A Multi-task Learning Balanced Attention Convolutional Neural Network Model for Few-shot Underwater Acoustic Target Recognition
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Transformation of audio embeddings into interpretable, concept-based representations
von: Zhang, Alice, et al.
Veröffentlicht: (2025)
von: Zhang, Alice, et al.
Veröffentlicht: (2025)
Scaling up masked audio encoder learning for general audio classification
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2024)
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2024)
GRAM: Spatial general-purpose audio representation models for real-world applications
von: Yuksel, Goksenin, et al.
Veröffentlicht: (2025)
von: Yuksel, Goksenin, et al.
Veröffentlicht: (2025)
Fine-tune the pretrained ATST model for sound event detection
von: Shao, Nian, et al.
Veröffentlicht: (2023)
von: Shao, Nian, et al.
Veröffentlicht: (2023)
DashengTokenizer: One layer is enough for unified audio understanding and generation
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2026)
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2026)
Towards audio language modeling -- an overview
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
Evaluation of Virtual Acoustic Environments with Different Acoustic Level of Detail
von: Fichna, Stefan, et al.
Veröffentlicht: (2023)
von: Fichna, Stefan, et al.
Veröffentlicht: (2023)
Comparison Performance of Spectrogram and Scalogram as Input of Acoustic Recognition Task
von: Phan, Dang Thoai
Veröffentlicht: (2024)
von: Phan, Dang Thoai
Veröffentlicht: (2024)
Are audio DeepFake detection models polyglots?
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
von: Marek, Bartłomiej, et al.
Veröffentlicht: (2024)
Stereo sound event localization and detection based on PSELDnet pretraining and BiMamba sequence modeling
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
von: Gao, Wenmiao, et al.
Veröffentlicht: (2025)
In-Materia Speech Recognition
von: Zolfagharinejad, Mohamadreza, et al.
Veröffentlicht: (2024)
von: Zolfagharinejad, Mohamadreza, et al.
Veröffentlicht: (2024)
Tweaking autoregressive methods for inpainting of gaps in audio signals
von: Mokrý, Ondřej, et al.
Veröffentlicht: (2024)
von: Mokrý, Ondřej, et al.
Veröffentlicht: (2024)
MBCodec:Thorough disentangle for high-fidelity audio compression
von: Zhang, Ruonan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruonan, et al.
Veröffentlicht: (2025)
Real-time implementation of vibrato transfer as an audio effect
von: Hyrkas, Jeremy
Veröffentlicht: (2025)
von: Hyrkas, Jeremy
Veröffentlicht: (2025)
On Time Delay Interpolation for Improved Acoustic Reflector Localization
von: Rosseel, Hannes, et al.
Veröffentlicht: (2025)
von: Rosseel, Hannes, et al.
Veröffentlicht: (2025)
AquaSignal: An Integrated Framework for Robust Underwater Acoustic Analysis
von: Panteli, Eirini, et al.
Veröffentlicht: (2025)
von: Panteli, Eirini, et al.
Veröffentlicht: (2025)
End-to-end Acoustic-linguistic Emotion and Intent Recognition Enhanced by Semi-supervised Learning
von: Ren, Zhao, et al.
Veröffentlicht: (2025)
von: Ren, Zhao, et al.
Veröffentlicht: (2025)
Speaker anonymization using neural audio codec language models
von: Panariello, Michele, et al.
Veröffentlicht: (2023)
von: Panariello, Michele, et al.
Veröffentlicht: (2023)
FxSearcher: gradient-free text-driven audio transformation
von: Ki, Hojoon, et al.
Veröffentlicht: (2025)
von: Ki, Hojoon, et al.
Veröffentlicht: (2025)
Regularized autoregressive modeling and its application to audio signal reconstruction
von: Mokrý, Ondřej, et al.
Veröffentlicht: (2024)
von: Mokrý, Ondřej, et al.
Veröffentlicht: (2024)
EDTC: enhance depth of text comprehension in automated audio captioning
von: Tan, Liwen, et al.
Veröffentlicht: (2024)
von: Tan, Liwen, et al.
Veröffentlicht: (2024)
Adapting Whisper for Streaming Speech Recognition via Two-Pass Decoding
von: Zhou, Haoran, et al.
Veröffentlicht: (2025)
von: Zhou, Haoran, et al.
Veröffentlicht: (2025)
Discriminative-Generative Target Speaker Extraction with Decoder-Only Language Models
von: Zeng, Bang, et al.
Veröffentlicht: (2026)
von: Zeng, Bang, et al.
Veröffentlicht: (2026)
AxLSTMs: learning self-supervised audio representations with xLSTMs
von: Yadav, Sarthak, et al.
Veröffentlicht: (2024)
von: Yadav, Sarthak, et al.
Veröffentlicht: (2024)
STASE: A spatialized text-to-audio synthesis engine for music generation
von: Chi, Tutti, et al.
Veröffentlicht: (2025)
von: Chi, Tutti, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Computation of Generalized Embeddings for Underwater Acoustic Target Recognition using Contrastive Learning
von: Hummel, Hilde I., et al.
Veröffentlicht: (2025) -
Automated data curation for self-supervised learning in underwater acoustic analysis
von: Hummel, Hilde I, et al.
Veröffentlicht: (2025) -
InsectSet459: an open dataset of insect sounds for bioacoustic machine learning
von: Faiß, Marius, et al.
Veröffentlicht: (2025) -
Human-CLAP: Human-perception-based contrastive language-audio pretraining
von: Takano, Taisei, et al.
Veröffentlicht: (2025) -
Unraveling Complex Data Diversity in Underwater Acoustic Target Recognition through Convolution-based Mixture of Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024)