Guardado en:
| Autores principales: | Zhao, Aite, Liu, Yongcan, Yu, Xinglin, Xing, Xinyue |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2502.10703 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SleepGMUformer: A gated multimodal temporal neural network for sleep staging
por: Zhao, Chenjun, et al.
Publicado: (2025)
por: Zhao, Chenjun, et al.
Publicado: (2025)
Synthetic data enables context-aware bioacoustic sound event detection
por: Hoffman, Benjamin, et al.
Publicado: (2025)
por: Hoffman, Benjamin, et al.
Publicado: (2025)
Determining the severity of Parkinson's disease in patients using a multi task neural network
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
Optimising MFCC parameters for the automatic detection of respiratory diseases
por: Yan, Yuyang, et al.
Publicado: (2024)
por: Yan, Yuyang, et al.
Publicado: (2024)
Evaluating Echo State Network for Parkinson's Disease Prediction using Voice Features
por: Hosseininian, Seyedeh Zahra Seyedi, et al.
Publicado: (2024)
por: Hosseininian, Seyedeh Zahra Seyedi, et al.
Publicado: (2024)
A multimodal Bayesian Network for symptom-level depression and anxiety prediction from voice and speech data
por: Norbury, Agnes, et al.
Publicado: (2025)
por: Norbury, Agnes, et al.
Publicado: (2025)
Acoustic evaluation of a neural network dedicated to the detection of animal vocalisations
por: Rouch, Jérémy, et al.
Publicado: (2025)
por: Rouch, Jérémy, et al.
Publicado: (2025)
An AI-enabled Bias-Free Respiratory Disease Diagnosis Model using Cough Audio: A Case Study for COVID-19
por: Saeed, Tabish, et al.
Publicado: (2024)
por: Saeed, Tabish, et al.
Publicado: (2024)
A multimodal dynamical variational autoencoder for audiovisual speech representation learning
por: Sadok, Samir, et al.
Publicado: (2023)
por: Sadok, Samir, et al.
Publicado: (2023)
Robust detection of overlapping bioacoustic sound events
por: Mahon, Louis, et al.
Publicado: (2025)
por: Mahon, Louis, et al.
Publicado: (2025)
Speech foundation models on intelligibility prediction for hearing-impaired listeners
por: Cuervo, Santiago, et al.
Publicado: (2024)
por: Cuervo, Santiago, et al.
Publicado: (2024)
Adaptive vector steering: A training-free, layer-wise intervention for hallucination mitigation in large audio and multimodal models
por: Lin, Tsung-En, et al.
Publicado: (2025)
por: Lin, Tsung-En, et al.
Publicado: (2025)
Echo: Towards Advanced Audio Comprehension via Audio-Interleaved Reasoning
por: Wu, Daiqing, et al.
Publicado: (2026)
por: Wu, Daiqing, et al.
Publicado: (2026)
Sparse deepfake detection promotes better disentanglement
por: Teissier, Antoine, et al.
Publicado: (2025)
por: Teissier, Antoine, et al.
Publicado: (2025)
SpikCommander: A High-performance Spiking Transformer with Multi-view Learning for Efficient Speech Command Recognition
por: Wang, Jiaqi, et al.
Publicado: (2025)
por: Wang, Jiaqi, et al.
Publicado: (2025)
A Semi-Supervised Framework for Speech Confidence Detection using Whisper
por: Wynn, Adam, et al.
Publicado: (2026)
por: Wynn, Adam, et al.
Publicado: (2026)
An Attention Long Short-Term Memory based system for automatic classification of speech intelligibility
por: Fernández-Díaz, Miguel, et al.
Publicado: (2024)
por: Fernández-Díaz, Miguel, et al.
Publicado: (2024)
A contrastive-learning approach for auditory attention detection
por: Bajestan, Seyed Ali Alavi, et al.
Publicado: (2024)
por: Bajestan, Seyed Ali Alavi, et al.
Publicado: (2024)
Decodable but not structured: linear probing enables Underwater Acoustic Target Recognition with pretrained audio embeddings
por: Hummel, Hilde I., et al.
Publicado: (2026)
por: Hummel, Hilde I., et al.
Publicado: (2026)
BenSParX: A Robust Explainable Machine Learning Framework for Parkinson's Disease Detection from Bengali Conversational Speech
por: Hossain, Riad, et al.
Publicado: (2025)
por: Hossain, Riad, et al.
Publicado: (2025)
ADNAC: Audio Denoiser using Neural Audio Codec
por: Jimon, Daniel, et al.
Publicado: (2025)
por: Jimon, Daniel, et al.
Publicado: (2025)
Selfsupervised learning for pathological speech detection
por: Sheikh, Shakeel Ahmad
Publicado: (2024)
por: Sheikh, Shakeel Ahmad
Publicado: (2024)
High-Fidelity Music Vocoder using Neural Audio Codecs
por: Lanzendörfer, Luca A., et al.
Publicado: (2025)
por: Lanzendörfer, Luca A., et al.
Publicado: (2025)
Efficient Continual Learning in Keyword Spotting using Binary Neural Networks
por: Vu, Quynh Nguyen-Phuong, et al.
Publicado: (2025)
por: Vu, Quynh Nguyen-Phuong, et al.
Publicado: (2025)
Denoising by neural network for muzzle blast detection
por: Pujol, Hadrien, et al.
Publicado: (2025)
por: Pujol, Hadrien, et al.
Publicado: (2025)
Cough activity detection for automatic tuberculosis screening
por: van Vüren, Joshua Jansen, et al.
Publicado: (2026)
por: van Vüren, Joshua Jansen, et al.
Publicado: (2026)
SAO-Instruct: Free-form Audio Editing using Natural Language Instructions
por: Ungersböck, Michael, et al.
Publicado: (2025)
por: Ungersböck, Michael, et al.
Publicado: (2025)
Multi-Task Learning for Lung sound & Lung disease classification
por: K V, Suma, et al.
Publicado: (2024)
por: K V, Suma, et al.
Publicado: (2024)
Investigating the Effectiveness of Explainability Methods in Parkinson's Detection from Speech
por: Mancini, Eleonora, et al.
Publicado: (2024)
por: Mancini, Eleonora, et al.
Publicado: (2024)
voice2mode: Phonation Mode Classification in Singing using Self-Supervised Speech Models
por: Justus, Aju Ani, et al.
Publicado: (2026)
por: Justus, Aju Ani, et al.
Publicado: (2026)
Towards generalizing deep-audio fake detection networks
por: Gasenzer, Konstantin, et al.
Publicado: (2023)
por: Gasenzer, Konstantin, et al.
Publicado: (2023)
A benchmark of state-of-the-art sound event detection systems evaluated on synthetic soundscapes
por: Ronchini, Francesca, et al.
Publicado: (2022)
por: Ronchini, Francesca, et al.
Publicado: (2022)
Surface impedance inference via neural fields and sparse acoustic data obtained by a compact array
por: Xia, Yuanxin, et al.
Publicado: (2026)
por: Xia, Yuanxin, et al.
Publicado: (2026)
Unsupervised outlier detection to improve bird audio dataset labels
por: Collins, Bruce
Publicado: (2025)
por: Collins, Bruce
Publicado: (2025)
Learning to rumble: Automated elephant call classification, detection and endpointing using deep architectures
por: Geldenhuys, Christiaan M., et al.
Publicado: (2024)
por: Geldenhuys, Christiaan M., et al.
Publicado: (2024)
A Novel Fusion Architecture for PD Detection Using Semi-Supervised Speech Embeddings
por: Adnan, Tariq, et al.
Publicado: (2024)
por: Adnan, Tariq, et al.
Publicado: (2024)
Unleashing the Power of Natural Audio Featuring Multiple Sound Sources
por: Cheng, Xize, et al.
Publicado: (2025)
por: Cheng, Xize, et al.
Publicado: (2025)
Generalizable speech deepfake detection via meta-learned LoRA
por: Laakkonen, Janne, et al.
Publicado: (2025)
por: Laakkonen, Janne, et al.
Publicado: (2025)
The impact of non-target events in synthetic soundscapes for sound event detection
por: Ronchini, Francesca, et al.
Publicado: (2021)
por: Ronchini, Francesca, et al.
Publicado: (2021)
Voxceleb-ESP: preliminary experiments detecting Spanish celebrities from their voices
por: Labrador, Beltrán, et al.
Publicado: (2023)
por: Labrador, Beltrán, et al.
Publicado: (2023)
Ejemplares similares
-
SleepGMUformer: A gated multimodal temporal neural network for sleep staging
por: Zhao, Chenjun, et al.
Publicado: (2025) -
Synthetic data enables context-aware bioacoustic sound event detection
por: Hoffman, Benjamin, et al.
Publicado: (2025) -
Determining the severity of Parkinson's disease in patients using a multi task neural network
por: García-Ordás, María Teresa, et al.
Publicado: (2024) -
Optimising MFCC parameters for the automatic detection of respiratory diseases
por: Yan, Yuyang, et al.
Publicado: (2024) -
Evaluating Echo State Network for Parkinson's Disease Prediction using Voice Features
por: Hosseininian, Seyedeh Zahra Seyedi, et al.
Publicado: (2024)