Gespeichert in:
| Hauptverfasser: | Xie, Yuan, Xu, Ji, Ren, Jiawei, Li, Junfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2411.02848 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Guiding the underwater acoustic target recognition with interpretable contrastive learning
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Advancing Robust Underwater Acoustic Target Recognition through Multi-task Learning and Multi-Gate Mixture-of-Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Adaptive ship-radiated noise recognition with learnable fine-grained wavelet transform
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
Underwater-Art: Expanding Information Perspectives With Text Templates For Underwater Acoustic Target Recognition
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
von: Xie, Yuan, et al.
Veröffentlicht: (2023)
DEMONet: Underwater Acoustic Target Recognition based on Multi-Expert Network and Cross-Temporal Variational Autoencoder
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Beyond saliency: enhancing explanation of speech emotion recognition with expert-referenced acoustic cues
von: Nasr, Seham, et al.
Veröffentlicht: (2025)
von: Nasr, Seham, et al.
Veröffentlicht: (2025)
Underwater Acoustic Target Recognition based on Smoothness-inducing Regularization and Spectrogram-based Data Augmentation
von: Xu, Ji, et al.
Veröffentlicht: (2023)
von: Xu, Ji, et al.
Veröffentlicht: (2023)
IsoNet: Spatially-aware audio-visual target speech extraction in complex acoustic environments
von: Padhya, Dinanath, et al.
Veröffentlicht: (2026)
von: Padhya, Dinanath, et al.
Veröffentlicht: (2026)
SeMaScore : a new evaluation metric for automatic speech recognition tasks
von: Sasindran, Zitha, et al.
Veröffentlicht: (2024)
von: Sasindran, Zitha, et al.
Veröffentlicht: (2024)
Fusion approaches for emotion recognition from speech using acoustic and text-based features
von: Pepino, Leonardo, et al.
Veröffentlicht: (2024)
von: Pepino, Leonardo, et al.
Veröffentlicht: (2024)
Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
von: Maiti, Soumi, et al.
Veröffentlicht: (2023)
von: Maiti, Soumi, et al.
Veröffentlicht: (2023)
A noise-robust acoustic method for recognizing foraging activities of grazing cattle
von: Martinez-Rau, Luciano S., et al.
Veröffentlicht: (2023)
von: Martinez-Rau, Luciano S., et al.
Veröffentlicht: (2023)
Training chord recognition models on artificially generated audio
von: Majchrzak, Martyna, et al.
Veröffentlicht: (2025)
von: Majchrzak, Martyna, et al.
Veröffentlicht: (2025)
DeepForestSound: a multi-species automatic detector for passive acoustic monitoring in African tropical forests, a case study in Kibale National Park
von: Dubus, Gabriel, et al.
Veröffentlicht: (2026)
von: Dubus, Gabriel, et al.
Veröffentlicht: (2026)
Determining the severity of Parkinson's disease in patients using a multi task neural network
von: García-Ordás, María Teresa, et al.
Veröffentlicht: (2024)
von: García-Ordás, María Teresa, et al.
Veröffentlicht: (2024)
Virtual boundary integral neural network for three-dimensional exterior acoustic problems
von: Li, Jiahao, et al.
Veröffentlicht: (2026)
von: Li, Jiahao, et al.
Veröffentlicht: (2026)
Surface impedance inference via neural fields and sparse acoustic data obtained by a compact array
von: Xia, Yuanxin, et al.
Veröffentlicht: (2026)
von: Xia, Yuanxin, et al.
Veröffentlicht: (2026)
Unraveling Complex Data Diversity in Underwater Acoustic Target Recognition through Convolution-based Mixture of Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
CR-CTC: Consistency regularization on CTC for improved speech recognition
von: Yao, Zengwei, et al.
Veröffentlicht: (2024)
von: Yao, Zengwei, et al.
Veröffentlicht: (2024)
Better audio representations are more brain-like: linking model-brain alignment with performance in downstream auditory tasks
von: Pepino, Leonardo, et al.
Veröffentlicht: (2025)
von: Pepino, Leonardo, et al.
Veröffentlicht: (2025)
Introduction to speech recognition
von: Dauphin, Gabriel
Veröffentlicht: (2024)
von: Dauphin, Gabriel
Veröffentlicht: (2024)
Benchmarks and leaderboards for sound demixing tasks
von: Solovyev, Roman, et al.
Veröffentlicht: (2023)
von: Solovyev, Roman, et al.
Veröffentlicht: (2023)
A Systematic Evaluation of Adversarial Attacks against Speech Emotion Recognition Models
von: Facchinetti, Nicolas, et al.
Veröffentlicht: (2024)
von: Facchinetti, Nicolas, et al.
Veröffentlicht: (2024)
Time-Varying Audio Effect Modeling by End-to-End Adversarial Training
von: Bourdin, Yann, et al.
Veröffentlicht: (2025)
von: Bourdin, Yann, et al.
Veröffentlicht: (2025)
Zipformer: A faster and better encoder for automatic speech recognition
von: Yao, Zengwei, et al.
Veröffentlicht: (2023)
von: Yao, Zengwei, et al.
Veröffentlicht: (2023)
Robustifying automatic speech recognition by extracting slowly varying features
von: Pizarro, Matías, et al.
Veröffentlicht: (2021)
von: Pizarro, Matías, et al.
Veröffentlicht: (2021)
Versatile audio-visual learning for emotion recognition
von: Goncalves, Lucas, et al.
Veröffentlicht: (2023)
von: Goncalves, Lucas, et al.
Veröffentlicht: (2023)
Improving the Adversarial Robustness for Speaker Verification by Self-Supervised Learning
von: Wu, Haibin, et al.
Veröffentlicht: (2021)
von: Wu, Haibin, et al.
Veröffentlicht: (2021)
Late fusion ensembles for speech recognition on diverse input audio representations
von: Jezidžić, Marin, et al.
Veröffentlicht: (2024)
von: Jezidžić, Marin, et al.
Veröffentlicht: (2024)
CAARMA: Class Augmentation with Adversarial Mixup Regularization
von: Baali, Massa, et al.
Veröffentlicht: (2025)
von: Baali, Massa, et al.
Veröffentlicht: (2025)
Generative Adversarial Post-Training Mitigates Reward Hacking in Live Human-AI Music Interaction
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
Unraveling Adversarial Examples against Speaker Identification -- Techniques for Attack Detection and Victim Model Classification
von: Joshi, Sonal, et al.
Veröffentlicht: (2024)
von: Joshi, Sonal, et al.
Veröffentlicht: (2024)
BioSEN: A Bio-acoustic Signal Enhancement Network for Animal Vocalizations
von: Song, Tianyu, et al.
Veröffentlicht: (2026)
von: Song, Tianyu, et al.
Veröffentlicht: (2026)
MAIA: An Inpainting-Based Approach for Music Adversarial Attacks
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Liu, Yuxuan, et al.
Veröffentlicht: (2025)
DFKI-Speech System for WildSpoof Challenge: A robust framework for SASV In-the-Wild
von: Das, Arnab, et al.
Veröffentlicht: (2026)
von: Das, Arnab, et al.
Veröffentlicht: (2026)
Adversarial Data Augmentation for Robust Speaker Verification
von: Zhou, Zhenyu, et al.
Veröffentlicht: (2024)
von: Zhou, Zhenyu, et al.
Veröffentlicht: (2024)
A vector quantized masked autoencoder for audiovisual speech emotion recognition
von: Sadok, Samir, et al.
Veröffentlicht: (2023)
von: Sadok, Samir, et al.
Veröffentlicht: (2023)
ALIGN: Adversarial Learning for Generalizable Speech Neuroprosthesis
von: Zhang, Zhanqi, et al.
Veröffentlicht: (2026)
von: Zhang, Zhanqi, et al.
Veröffentlicht: (2026)
LipsAM: Lipschitz-Continuous Amplitude Modifier for Audio Signal Processing and its Application to Plug-and-Play Dereverberation
von: Matsumoto, Kazuki, et al.
Veröffentlicht: (2026)
von: Matsumoto, Kazuki, et al.
Veröffentlicht: (2026)
StrADiff: A Structured Source-Wise Adaptive Diffusion Framework for Linear and Nonlinear Blind Source Separation
von: Wei, Yuan-Hao
Veröffentlicht: (2026)
von: Wei, Yuan-Hao
Veröffentlicht: (2026)
Ähnliche Einträge
-
Guiding the underwater acoustic target recognition with interpretable contrastive learning
von: Xie, Yuan, et al.
Veröffentlicht: (2024) -
Advancing Robust Underwater Acoustic Target Recognition through Multi-task Learning and Multi-Gate Mixture-of-Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024) -
Adaptive ship-radiated noise recognition with learnable fine-grained wavelet transform
von: Xie, Yuan, et al.
Veröffentlicht: (2023) -
Underwater-Art: Expanding Information Perspectives With Text Templates For Underwater Acoustic Target Recognition
von: Xie, Yuan, et al.
Veröffentlicht: (2023) -
DEMONet: Underwater Acoustic Target Recognition based on Multi-Expert Network and Cross-Temporal Variational Autoencoder
von: Xie, Yuan, et al.
Veröffentlicht: (2024)