Adaptive ship-radiated noise recognition with learnable fine-grained wavelet transform
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xie, Yuan, Ren, Jiawei, Xu, Ji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Underwater-Art: Expanding Information Perspectives With Text Templates For Underwater Acoustic Target Recognition
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
Guiding the underwater acoustic target recognition with interpretable contrastive learning
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
DEMONet: Underwater Acoustic Target Recognition based on Multi-Expert Network and Cross-Temporal Variational Autoencoder
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Advancing Robust Underwater Acoustic Target Recognition through Multi-task Learning and Multi-Gate Mixture-of-Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Single-channel speech enhancement using learnable loss mixup
von: Chang, Oscar, et al.
Veröffentlicht: (2023)
von: Chang, Oscar, et al.
Veröffentlicht: (2023)
Boosting keyword spotting through on-device learnable user speech characteristics
von: Cioflan, Cristian, et al.
Veröffentlicht: (2024)
von: Cioflan, Cristian, et al.
Veröffentlicht: (2024)
Unraveling Complex Data Diversity in Underwater Acoustic Target Recognition through Convolution-based Mixture of Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Paraformer-v2: An improved non-autoregressive transformer for noise-robust speech recognition
von: An, Keyu, et al.
Veröffentlicht: (2024)
von: An, Keyu, et al.
Veröffentlicht: (2024)
Zipformer: A faster and better encoder for automatic speech recognition
von: Yao, Zengwei, et al.
Veröffentlicht: (2023)
von: Yao, Zengwei, et al.
Veröffentlicht: (2023)
CR-CTC: Consistency regularization on CTC for improved speech recognition
von: Yao, Zengwei, et al.
Veröffentlicht: (2024)
von: Yao, Zengwei, et al.
Veröffentlicht: (2024)
Robustifying automatic speech recognition by extracting slowly varying features
von: Pizarro, Matías, et al.
Veröffentlicht: (2021)
von: Pizarro, Matías, et al.
Veröffentlicht: (2021)
Late fusion ensembles for speech recognition on diverse input audio representations
von: Jezidžić, Marin, et al.
Veröffentlicht: (2024)
von: Jezidžić, Marin, et al.
Veröffentlicht: (2024)
Fusion approaches for emotion recognition from speech using acoustic and text-based features
von: Pepino, Leonardo, et al.
Veröffentlicht: (2024)
von: Pepino, Leonardo, et al.
Veröffentlicht: (2024)
SeMaScore : a new evaluation metric for automatic speech recognition tasks
von: Sasindran, Zitha, et al.
Veröffentlicht: (2024)
von: Sasindran, Zitha, et al.
Veröffentlicht: (2024)
Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
von: Maiti, Soumi, et al.
Veröffentlicht: (2023)
von: Maiti, Soumi, et al.
Veröffentlicht: (2023)
Fine-grained Soundscape Control for Augmented Hearing
von: Oh, Seunghyun, et al.
Veröffentlicht: (2026)
von: Oh, Seunghyun, et al.
Veröffentlicht: (2026)
A noise-robust acoustic method for recognizing foraging activities of grazing cattle
von: Martinez-Rau, Luciano S., et al.
Veröffentlicht: (2023)
von: Martinez-Rau, Luciano S., et al.
Veröffentlicht: (2023)
Versatile audio-visual learning for emotion recognition
von: Goncalves, Lucas, et al.
Veröffentlicht: (2023)
von: Goncalves, Lucas, et al.
Veröffentlicht: (2023)
EM-TTS: Efficiently Trained Low-Resource Mongolian Lightweight Text-to-Speech
von: Liang, Ziqi, et al.
Veröffentlicht: (2024)
von: Liang, Ziqi, et al.
Veröffentlicht: (2024)
A vector quantized masked autoencoder for audiovisual speech emotion recognition
von: Sadok, Samir, et al.
Veröffentlicht: (2023)
von: Sadok, Samir, et al.
Veröffentlicht: (2023)
Towards objective and interpretable speech disorder assessment: a comparative analysis of CNN and transformer-based models
von: Maisonneuve, Malo, et al.
Veröffentlicht: (2024)
von: Maisonneuve, Malo, et al.
Veröffentlicht: (2024)
Introduction to speech recognition
von: Dauphin, Gabriel
Veröffentlicht: (2024)
von: Dauphin, Gabriel
Veröffentlicht: (2024)
Adaptive Slimming for Scalable and Efficient Speech Enhancement
von: Miccini, Riccardo, et al.
Veröffentlicht: (2025)
von: Miccini, Riccardo, et al.
Veröffentlicht: (2025)
AdaPTwin: Low-Cost Adaptive Compression of Product Twins in Transformers
von: Biju, Emil, et al.
Veröffentlicht: (2024)
von: Biju, Emil, et al.
Veröffentlicht: (2024)
Adaptive Noise Resilient Keyword Spotting Using One-Shot Learning
von: Martinez-Rau, Luciano Sebastian, et al.
Veröffentlicht: (2025)
von: Martinez-Rau, Luciano Sebastian, et al.
Veröffentlicht: (2025)
Revisit Micro-batch Clipping: Adaptive Data Pruning via Gradient Manipulation
von: Wang, Lun
Veröffentlicht: (2024)
von: Wang, Lun
Veröffentlicht: (2024)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
von: Ducorroy, Alexandre, et al.
Veröffentlicht: (2025)
Coverage-Guaranteed Speech Emotion Recognition via Calibrated Uncertainty-Adaptive Prediction Sets
von: Jia, Zijun, et al.
Veröffentlicht: (2025)
von: Jia, Zijun, et al.
Veröffentlicht: (2025)
Adaptive Rotary Steering with Joint Autoregression for Robust Extraction of Closely Moving Speakers in Dynamic Scenarios
von: Kienegger, Jakob, et al.
Veröffentlicht: (2026)
von: Kienegger, Jakob, et al.
Veröffentlicht: (2026)
From Weak to Strong Sound Event Labels using Adaptive Change-Point Detection and Active Learning
von: Martinsson, John, et al.
Veröffentlicht: (2024)
von: Martinsson, John, et al.
Veröffentlicht: (2024)
Meta-Learning-Based Delayless Subband Adaptive Filter using Complex Self-Attention for Active Noise Control
von: Feng, Pengxing, et al.
Veröffentlicht: (2024)
von: Feng, Pengxing, et al.
Veröffentlicht: (2024)
MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
von: Jiang, Ziyue, et al.
Veröffentlicht: (2025)
von: Jiang, Ziyue, et al.
Veröffentlicht: (2025)
Adaptive vector steering: A training-free, layer-wise intervention for hallucination mitigation in large audio and multimodal models
von: Lin, Tsung-En, et al.
Veröffentlicht: (2025)
von: Lin, Tsung-En, et al.
Veröffentlicht: (2025)
Multi-label Zero-Shot Audio Classification with Temporal Attention
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
Bayesian Learning for Deep Neural Network Adaptation
von: Xie, Xurong, et al.
Veröffentlicht: (2020)
von: Xie, Xurong, et al.
Veröffentlicht: (2020)
Improving Generalization for AI-Synthesized Voice Detection
von: Ren, Hainan, et al.
Veröffentlicht: (2024)
von: Ren, Hainan, et al.
Veröffentlicht: (2024)
Geo-ATBench: A Benchmark for Geospatial Audio Tagging with Geospatial Semantic Context
von: Hou, Yuanbo, et al.
Veröffentlicht: (2026)
von: Hou, Yuanbo, et al.
Veröffentlicht: (2026)
Music102: An $D_{12}$-equivariant transformer for chord progression accompaniment
von: Luo, Weiliang
Veröffentlicht: (2024)
von: Luo, Weiliang
Veröffentlicht: (2024)
ACAVCaps: Enabling large-scale training for fine-grained and diverse audio understanding
von: Niu, Yadong, et al.
Veröffentlicht: (2026)
von: Niu, Yadong, et al.
Veröffentlicht: (2026)
Towards Robust Overlapping Speech Detection: A Speaker-Aware Progressive Approach Using WavLM
von: Sun, Zhaokai, et al.
Veröffentlicht: (2025)
von: Sun, Zhaokai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Underwater-Art: Expanding Information Perspectives With Text Templates For Underwater Acoustic Target Recognition
von: Xie, Yuan, et al.
Veröffentlicht: (2023) -
Guiding the underwater acoustic target recognition with interpretable contrastive learning
von: Xie, Yuan, et al.
Veröffentlicht: (2024) -
DEMONet: Underwater Acoustic Target Recognition based on Multi-Expert Network and Cross-Temporal Variational Autoencoder
von: Xie, Yuan, et al.
Veröffentlicht: (2024) -
Advancing Robust Underwater Acoustic Target Recognition through Multi-task Learning and Multi-Gate Mixture-of-Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024) -
Single-channel speech enhancement using learnable loss mixup
von: Chang, Oscar, et al.
Veröffentlicht: (2023)