An Attention-Assisted Multi-Modal Data Fusion Model for Real-Time Estimation of Underwater Sound Velocity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Pengfei, Huang, Wei, Shi, Yujie, Zhang, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Multimodal Data Fusion Attention-Empowered Generative Adversarial Network for Real Time 3D Underwater Sound Speed Field Construction
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
STNet: Prediction of Underwater Sound Speed Profiles with An Advanced Semi-Transformer Neural Network
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Modulation Feature Enhancement with a Multi-Stage Attention Network for Underwater Acoustic Target Recognition
von: Yu, Jiaping, et al.
Veröffentlicht: (2026)
von: Yu, Jiaping, et al.
Veröffentlicht: (2026)
Future Full-Ocean Deep SSPs Prediction based on Hierarchical Long Short-Term Memory Neural Networks
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
FullSubNet: A Full-Band and Sub-Band Fusion Model for Real-Time Single-Channel Speech Enhancement
von: Hao, Xiang, et al.
Veröffentlicht: (2020)
von: Hao, Xiang, et al.
Veröffentlicht: (2020)
State Space and Self-Attention Collaborative Network with Feature Aggregation for DOA Estimation
von: You, Qi, et al.
Veröffentlicht: (2025)
von: You, Qi, et al.
Veröffentlicht: (2025)
Piezoelectric Electromechanical Transducers for Underwater Sound, Part I
von: Aronov, Boris S.
Veröffentlicht: (2022)
von: Aronov, Boris S.
Veröffentlicht: (2022)
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
von: Zhu, Haolin, et al.
Veröffentlicht: (2024)
von: Zhu, Haolin, et al.
Veröffentlicht: (2024)
Decomposing the Influence of Physical Acoustic Modeling on Neural Personal Sound Zone Rendering: An Ablation Study
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
SONIC: Sound Optimization for Noise In Crowds
von: N, Pranav M, et al.
Veröffentlicht: (2025)
von: N, Pranav M, et al.
Veröffentlicht: (2025)
Adaptive Control Attention Network for Underwater Acoustic Localization and Domain Adaptation
von: Vo, Quoc Thinh, et al.
Veröffentlicht: (2025)
von: Vo, Quoc Thinh, et al.
Veröffentlicht: (2025)
Differentiable Modal Synthesis for Physical Modeling of Planar String Sound and Motion Simulation
von: Lee, Jin Woo, et al.
Veröffentlicht: (2024)
von: Lee, Jin Woo, et al.
Veröffentlicht: (2024)
Reverberation-based Features for Sound Event Localization and Detection with Distance Estimation
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
von: Berghi, Davide, et al.
Veröffentlicht: (2025)
Velocity Potential Neural Field for Efficient Ambisonics Impulse Response Modeling
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2026)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2026)
Dynamic Prediction of Full-Ocean Depth SSP by Hierarchical LSTM: An Experimental Result
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
Investigation of Feature Selection and Pooling Methods for Environmental Sound Classification
von: Dehaghani, Parinaz Binandeh, et al.
Veröffentlicht: (2025)
von: Dehaghani, Parinaz Binandeh, et al.
Veröffentlicht: (2025)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
Binaural Selective Attention Model for Target Speaker Extraction
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
Comparison of Frequency-Fusion Mechanisms for Binaural Direction-of-Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2024)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2024)
Reading to Listen at the Cocktail Party: Multi-Modal Speech Separation
von: Rahimi, Akam, et al.
Veröffentlicht: (2025)
von: Rahimi, Akam, et al.
Veröffentlicht: (2025)
Stimulus Modality Matters: Impact of Perceptual Evaluations from Different Modalities on Speech Emotion Recognition System Performance
von: Chou, Huang-Cheng, et al.
Veröffentlicht: (2024)
von: Chou, Huang-Cheng, et al.
Veröffentlicht: (2024)
Continuous Speech Tokens Makes LLMs Robust Multi-Modality Learners
von: Yuan, Ze, et al.
Veröffentlicht: (2024)
von: Yuan, Ze, et al.
Veröffentlicht: (2024)
A Data-Centric Approach to Generalizable Speech Deepfake Detection
von: Huang, Wen, et al.
Veröffentlicht: (2025)
von: Huang, Wen, et al.
Veröffentlicht: (2025)
SoundSpring: Loss-Resilient Audio Transceiver with Dual-Functional Masked Language Modeling
von: Yao, Shengshi, et al.
Veröffentlicht: (2025)
von: Yao, Shengshi, et al.
Veröffentlicht: (2025)
Joint Source-Environment Adaptation of Data-Driven Underwater Acoustic Source Ranging Based on Model Uncertainty
von: Kari, Dariush, et al.
Veröffentlicht: (2025)
von: Kari, Dariush, et al.
Veröffentlicht: (2025)
DECAF: Dynamic Envelope Context-Aware Fusion for Speech-Envelope Reconstruction from EEG
von: Thakkar, Karan, et al.
Veröffentlicht: (2026)
von: Thakkar, Karan, et al.
Veröffentlicht: (2026)
Large Language Model-based Nonnegative Matrix Factorization For Cardiorespiratory Sound Separation
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
ShipEcho -- An Interactive Tool for Global Mapping of Underwater Radiated Noise from Vessels
von: Shipton, Mark, et al.
Veröffentlicht: (2026)
von: Shipton, Mark, et al.
Veröffentlicht: (2026)
Efficient Solutions for Mitigating Initialization Bias in Unsupervised Self-Adaptive Auditory Attention Decoding
von: Yao, Yuanyuan, et al.
Veröffentlicht: (2025)
von: Yao, Yuanyuan, et al.
Veröffentlicht: (2025)
AnyAccomp: Generalizable Accompaniment Generation via Quantized Melodic Bottleneck
von: Zhang, Junan, et al.
Veröffentlicht: (2025)
von: Zhang, Junan, et al.
Veröffentlicht: (2025)
Underwater Sound Speed Profile Construction: A Review
von: Huang, Wei, et al.
Veröffentlicht: (2023)
von: Huang, Wei, et al.
Veröffentlicht: (2023)
Mind the Prompt: Prompting Strategies in Audio Generations for Improving Sound Classification
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
Real-Time Streamable Generative Speech Restoration with Flow Matching
von: Welker, Simon, et al.
Veröffentlicht: (2025)
von: Welker, Simon, et al.
Veröffentlicht: (2025)
RIFT: Entropy-Optimised Fractional Wavelet Constellations for Ideal Time-Frequency Estimation
von: Cozens, James M., et al.
Veröffentlicht: (2025)
von: Cozens, James M., et al.
Veröffentlicht: (2025)
Mel-McNet: A Mel-Scale Framework for Online Multichannel Speech Enhancement
von: Yang, Yujie, et al.
Veröffentlicht: (2025)
von: Yang, Yujie, et al.
Veröffentlicht: (2025)
AADNet: An End-to-End Deep Learning Model for Auditory Attention Decoding
von: Nguyen, Nhan Duc Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Nhan Duc Thanh, et al.
Veröffentlicht: (2024)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
An Investigation of Time-Frequency Representation Discriminators for High-Fidelity Vocoder
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Multimodal Data Fusion Attention-Empowered Generative Adversarial Network for Real Time 3D Underwater Sound Speed Field Construction
von: Huang, Wei, et al.
Veröffentlicht: (2025) -
STNet: Prediction of Underwater Sound Speed Profiles with An Advanced Semi-Transformer Neural Network
von: Huang, Wei, et al.
Veröffentlicht: (2025) -
Modulation Feature Enhancement with a Multi-Stage Attention Network for Underwater Acoustic Target Recognition
von: Yu, Jiaping, et al.
Veröffentlicht: (2026) -
Future Full-Ocean Deep SSPs Prediction based on Hierarchical Long Short-Term Memory Neural Networks
von: Lu, Jiajun, et al.
Veröffentlicht: (2023) -
FullSubNet: A Full-Band and Sub-Band Fusion Model for Real-Time Single-Channel Speech Enhancement
von: Hao, Xiang, et al.
Veröffentlicht: (2020)