Automatic Voice Classification Of Autistic Subjects
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vacca, Jessica, Brondino, Natascia, Dell'Acqua, Fabio, Vizziello, Anna, Savazzi, Pietro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mutual Information Analysis of Neuromorphic Coding for Distributed Wireless Spiking Neural Networks
von: Savazzi, Pietro, et al.
Veröffentlicht: (2024)
von: Savazzi, Pietro, et al.
Veröffentlicht: (2024)
Singing Voice Graph Modeling for SingFake Detection
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
Phoneme Discretized Saliency Maps for Explainable Detection of AI-Generated Voice
von: Gupta, Shubham, et al.
Veröffentlicht: (2024)
von: Gupta, Shubham, et al.
Veröffentlicht: (2024)
PromptEVC: Controllable Emotional Voice Conversion with Natural Language Prompts
von: Qi, Tianhua, et al.
Veröffentlicht: (2025)
von: Qi, Tianhua, et al.
Veröffentlicht: (2025)
AutoMashup: Automatic Music Mashups Creation
von: Delabaere, Marine, et al.
Veröffentlicht: (2025)
von: Delabaere, Marine, et al.
Veröffentlicht: (2025)
PAVITS: Exploring Prosody-aware VITS for End-to-End Emotional Voice Conversion
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
Synthetic training set generation using text-to-audio models for environmental sound classification
von: Ronchini, Francesca, et al.
Veröffentlicht: (2024)
von: Ronchini, Francesca, et al.
Veröffentlicht: (2024)
Comparison of Classification Algorithms for COVID19 Detection using Cough Acoustic Signals
von: Erdoğan, Yunus Emre, et al.
Veröffentlicht: (2022)
von: Erdoğan, Yunus Emre, et al.
Veröffentlicht: (2022)
VoicePrompter: Robust Zero-Shot Voice Conversion with Voice Prompt and Conditional Flow Matching
von: Choi, Ha-Yeong, et al.
Veröffentlicht: (2025)
von: Choi, Ha-Yeong, et al.
Veröffentlicht: (2025)
Mind the Prompt: Prompting Strategies in Audio Generations for Improving Sound Classification
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
Automatic Assessment of Oral Reading Accuracy for Reading Diagnostics
von: Molenaar, Bo, et al.
Veröffentlicht: (2023)
von: Molenaar, Bo, et al.
Veröffentlicht: (2023)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
von: Joysingh, S. Johanan, et al.
Veröffentlicht: (2024)
von: Joysingh, S. Johanan, et al.
Veröffentlicht: (2024)
Cross-Talk Reduction
von: Wang, Zhong-Qiu, et al.
Veröffentlicht: (2024)
von: Wang, Zhong-Qiu, et al.
Veröffentlicht: (2024)
Towards Improved Objective Perceptual Audio Quality Assessment -- Part 1: A Novel Data-Driven Cognitive Model
von: Delgado, Pablo M., et al.
Veröffentlicht: (2024)
von: Delgado, Pablo M., et al.
Veröffentlicht: (2024)
Binaural Selective Attention Model for Target Speaker Extraction
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
von: Wang, Kuan-Chen, et al.
Veröffentlicht: (2024)
Can all variations within the unified mask-based beamformer framework achieve identical peak extraction performance?
von: Hiroe, Atsuo, et al.
Veröffentlicht: (2024)
von: Hiroe, Atsuo, et al.
Veröffentlicht: (2024)
Constant Directivity Loudspeaker Beamforming
von: Luo, Yuancheng
Veröffentlicht: (2024)
von: Luo, Yuancheng
Veröffentlicht: (2024)
Acoustical Features as Knee Health Biomarkers: A Critical Analysis
von: Kechris, Christodoulos, et al.
Veröffentlicht: (2024)
von: Kechris, Christodoulos, et al.
Veröffentlicht: (2024)
GAN-Based Speech Enhancement for Low SNR Using Latent Feature Conditioning
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
LiSenNet: Lightweight Sub-band and Dual-Path Modeling for Real-Time Speech Enhancement
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
von: Yan, Haoyin, et al.
Veröffentlicht: (2024)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
The CARFAC v2 Cochlear Model in Matlab, NumPy, and JAX
von: Lyon, Richard F., et al.
Veröffentlicht: (2024)
von: Lyon, Richard F., et al.
Veröffentlicht: (2024)
What is Learnt by the LEArnable Front-end (LEAF)? Adapting Per-Channel Energy Normalisation (PCEN) to Noisy Conditions
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
von: Meng, Hanyu, et al.
Veröffentlicht: (2024)
HiRIS: an Airborne Sonar Sensor with a 1024 Channel Microphone Array for In-Air Acoustic Imaging
von: Laurijssen, Dennis, et al.
Veröffentlicht: (2024)
von: Laurijssen, Dennis, et al.
Veröffentlicht: (2024)
Exploiting spatial diversity for increasing the robustness of sound source localization systems against reverberation
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
von: Garcia-Barrios, Guillermo, et al.
Veröffentlicht: (2024)
SpeakerBeam-SS: Real-time Target Speaker Extraction with Lightweight Conv-TasNet and State Space Modeling
von: Sato, Hiroshi, et al.
Veröffentlicht: (2024)
von: Sato, Hiroshi, et al.
Veröffentlicht: (2024)
Self-supervised Multimodal Speech Representations for the Assessment of Schizophrenia Symptoms
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
von: Premananth, Gowtham, et al.
Veröffentlicht: (2024)
Informed FastICA: Semi-Blind Minimum Variance Distortionless Beamformer
von: Koldovský, Zbyněk, et al.
Veröffentlicht: (2024)
von: Koldovský, Zbyněk, et al.
Veröffentlicht: (2024)
Linear stimulus reconstruction works on the KU Leuven audiovisual, gaze-controlled auditory attention decoding dataset
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
von: Geirnaert, Simon, et al.
Veröffentlicht: (2024)
Computational Analysis of Yaredawi YeZema Silt in Ethiopian Orthodox Tewahedo Church Chants
von: Muluneh, Mequanent Argaw, et al.
Veröffentlicht: (2024)
von: Muluneh, Mequanent Argaw, et al.
Veröffentlicht: (2024)
Characteristics-Based Design of Generalized-Exponent Bandpass Filters
von: Alkhairy, Samiya A
Veröffentlicht: (2024)
von: Alkhairy, Samiya A
Veröffentlicht: (2024)
Interpolation Filter Design for Sample Rate Independent Audio Effect RNNs
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
On Improving Error Resilience of Neural End-to-End Speech Coders
von: Gupta, Kishan, et al.
Veröffentlicht: (2024)
von: Gupta, Kishan, et al.
Veröffentlicht: (2024)
Align-ULCNet: Towards Low-Complexity and Robust Acoustic Echo and Noise Reduction
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
von: Wang, Bo, et al.
Veröffentlicht: (2024)
von: Wang, Bo, et al.
Veröffentlicht: (2024)
Multiple Mobile Target Detection and Tracking in Active Sonar Array Using a Track-Before-Detect Approach
von: Abu, Avi, et al.
Veröffentlicht: (2024)
von: Abu, Avi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Mutual Information Analysis of Neuromorphic Coding for Distributed Wireless Spiking Neural Networks
von: Savazzi, Pietro, et al.
Veröffentlicht: (2024) -
Singing Voice Graph Modeling for SingFake Detection
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024) -
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
von: Qi, Tianhua, et al.
Veröffentlicht: (2024) -
Phoneme Discretized Saliency Maps for Explainable Detection of AI-Generated Voice
von: Gupta, Shubham, et al.
Veröffentlicht: (2024) -
PromptEVC: Controllable Emotional Voice Conversion with Natural Language Prompts
von: Qi, Tianhua, et al.
Veröffentlicht: (2025)