Clustering and Data Augmentation to Improve Accuracy of Sleep Assessment and Sleep Individuality Analysis
Fuente:
arXiv
Guardado en:
| Autores principales: | Tamai, Shintaro, Numao, Masayuki, Fukui, Ken-ichi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Deep Learning-Based Automatic Multi-Level Airway Collapse Monitoring on Obstructive Sleep Apnea Patients
por: Hsu, Ying-Chieh, et al.
Publicado: (2024)
por: Hsu, Ying-Chieh, et al.
Publicado: (2024)
Predicting Individual Depression Symptoms from Acoustic Features During Speech
por: Rodriguez, Sebastian, et al.
Publicado: (2024)
por: Rodriguez, Sebastian, et al.
Publicado: (2024)
Tuning In: Analysis of Audio Classifier Performance in Clinical Settings with Limited Data
por: Mahdi, Hamza, et al.
Publicado: (2024)
por: Mahdi, Hamza, et al.
Publicado: (2024)
Estimating Respiratory Effort from Nocturnal Breathing Sounds for Obstructive Sleep Apnoea Screening
por: Xu, Xiaolei, et al.
Publicado: (2025)
por: Xu, Xiaolei, et al.
Publicado: (2025)
Continuous Autoregressive Models with Noise Augmentation Avoid Error Accumulation
por: Pasini, Marco, et al.
Publicado: (2024)
por: Pasini, Marco, et al.
Publicado: (2024)
Evaluating Fake Music Detection Performance Under Audio Augmentations
por: Sroka, Tomasz, et al.
Publicado: (2025)
por: Sroka, Tomasz, et al.
Publicado: (2025)
Device-Robust Acoustic Scene Classification via Impulse Response Augmentation
por: Morocutti, Tobias, et al.
Publicado: (2023)
por: Morocutti, Tobias, et al.
Publicado: (2023)
CAtCh: Cognitive Assessment through Cookie Thief
por: Colonel, Joseph T, et al.
Publicado: (2025)
por: Colonel, Joseph T, et al.
Publicado: (2025)
A Recall-First CNN for Sleep Apnea Screening from Snoring Audio
por: Mallick, Anushka, et al.
Publicado: (2025)
por: Mallick, Anushka, et al.
Publicado: (2025)
Futga: Towards Fine-grained Music Understanding through Temporally-enhanced Generative Augmentation
por: Wu, Junda, et al.
Publicado: (2024)
por: Wu, Junda, et al.
Publicado: (2024)
Scaling Ambiguity: Augmenting Human Annotation in Speech Emotion Recognition with Audio-Language Models
por: Zhang, Wenda, et al.
Publicado: (2026)
por: Zhang, Wenda, et al.
Publicado: (2026)
A Conditioned UNet for Music Source Separation
por: O'Hanlon, Ken, et al.
Publicado: (2025)
por: O'Hanlon, Ken, et al.
Publicado: (2025)
Soft Clustering Anchors for Self-Supervised Speech Representation Learning in Joint Embedding Prediction Architectures
por: Ioannides, Georgios, et al.
Publicado: (2026)
por: Ioannides, Georgios, et al.
Publicado: (2026)
Unsupervised Rhythm and Voice Conversion to Improve ASR on Dysarthric Speech
por: Hajal, Karl El, et al.
Publicado: (2025)
por: Hajal, Karl El, et al.
Publicado: (2025)
Advancing Marine Bioacoustics with Deep Generative Models: A Hybrid Augmentation Strategy for Southern Resident Killer Whale Detection
por: Padovese, Bruno, et al.
Publicado: (2025)
por: Padovese, Bruno, et al.
Publicado: (2025)
Improving Generalization of Speech Separation in Real-World Scenarios: Strategies in Simulation, Optimization, and Evaluation
por: Chen, Ke, et al.
Publicado: (2024)
por: Chen, Ke, et al.
Publicado: (2024)
Pareto Data Framework: Steps Towards Resource-Efficient Decision Making Using Minimum Viable Data (MVD)
por: Ahmed, Tashfain, et al.
Publicado: (2024)
por: Ahmed, Tashfain, et al.
Publicado: (2024)
Exploring and Applying Audio-Based Sentiment Analysis in Music
por: Jhanji, Etash
Publicado: (2024)
por: Jhanji, Etash
Publicado: (2024)
Open-Amp: Synthetic Data Framework for Audio Effect Foundation Models
por: Wright, Alec, et al.
Publicado: (2024)
por: Wright, Alec, et al.
Publicado: (2024)
Discrete Speech Unit Extraction via Independent Component Analysis
por: Nakamura, Tomohiko, et al.
Publicado: (2025)
por: Nakamura, Tomohiko, et al.
Publicado: (2025)
DAISY: Data Adaptive Self-Supervised Early Exit for Speech Representation Models
por: Lin, Tzu-Quan, et al.
Publicado: (2024)
por: Lin, Tzu-Quan, et al.
Publicado: (2024)
Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data
por: Nihal, Ragib Amin, et al.
Publicado: (2025)
por: Nihal, Ragib Amin, et al.
Publicado: (2025)
Comparative Analysis of CNN and Transformer Architectures with Heart Cycle Normalization for Automated Phonocardiogram Classification
por: Sondermann, Martin, et al.
Publicado: (2025)
por: Sondermann, Martin, et al.
Publicado: (2025)
Enhancing Synthetic Training Data for Speech Commands: From ASR-Based Filtering to Domain Adaptation in SSL Latent Space
por: Quintas, Sebastião, et al.
Publicado: (2024)
por: Quintas, Sebastião, et al.
Publicado: (2024)
Towards Robust Transcription: Exploring Noise Injection Strategies for Training Data Augmentation
por: Kim, Yonghyun, et al.
Publicado: (2024)
por: Kim, Yonghyun, et al.
Publicado: (2024)
Gaussian Process Regression of Steering Vectors With Physics-Aware Deep Composite Kernels for Augmented Listening
por: Di Carlo, Diego, et al.
Publicado: (2025)
por: Di Carlo, Diego, et al.
Publicado: (2025)
From Discord to Harmony: Decomposed Consonance-based Training for Improved Audio Chord Estimation
por: Poltronieri, Andrea, et al.
Publicado: (2025)
por: Poltronieri, Andrea, et al.
Publicado: (2025)
Can Layer-wise SSL Features Improve Zero-Shot ASR Performance for Children's Speech?
por: Sinha, Abhijit, et al.
Publicado: (2025)
por: Sinha, Abhijit, et al.
Publicado: (2025)
Personalized Speech Enhancement Without a Separate Speaker Embedding Model
por: Pärnamaa, Tanel, et al.
Publicado: (2024)
por: Pärnamaa, Tanel, et al.
Publicado: (2024)
Speech Emotion Recognition Using CNN and Its Use Case in Digital Healthcare
por: Nigar, Nishargo
Publicado: (2024)
por: Nigar, Nishargo
Publicado: (2024)
DisMix: Disentangling Mixtures of Musical Instruments for Source-level Pitch and Timbre Manipulation
por: Luo, Yin-Jyun, et al.
Publicado: (2024)
por: Luo, Yin-Jyun, et al.
Publicado: (2024)
Run-Time Adaptation of Neural Beamforming for Robust Speech Dereverberation and Denoising
por: Fujita, Yoto, et al.
Publicado: (2024)
por: Fujita, Yoto, et al.
Publicado: (2024)
Towards Open Respiratory Acoustic Foundation Models: Pretraining and Benchmarking
por: Zhang, Yuwei, et al.
Publicado: (2024)
por: Zhang, Yuwei, et al.
Publicado: (2024)
aTENNuate: Optimized Real-time Speech Enhancement with Deep SSMs on Raw Audio
por: Pei, Yan Ru, et al.
Publicado: (2024)
por: Pei, Yan Ru, et al.
Publicado: (2024)
Bridging Modalities: Knowledge Distillation and Masked Training for Translating Multi-Modal Emotion Recognition to Uni-Modal, Speech-Only Emotion Recognition
por: Muaz, Muhammad, et al.
Publicado: (2024)
por: Muaz, Muhammad, et al.
Publicado: (2024)
DITTO: Diffusion Inference-Time T-Optimization for Music Generation
por: Novack, Zachary, et al.
Publicado: (2024)
por: Novack, Zachary, et al.
Publicado: (2024)
TRAMBA: A Hybrid Transformer and Mamba Architecture for Practical Audio and Bone Conduction Speech Super Resolution and Enhancement on Mobile and Wearable Platforms
por: Sui, Yueyuan, et al.
Publicado: (2024)
por: Sui, Yueyuan, et al.
Publicado: (2024)
Leveraging tropical reef, bird and unrelated sounds for superior transfer learning in marine bioacoustics
por: Williams, Ben, et al.
Publicado: (2024)
por: Williams, Ben, et al.
Publicado: (2024)
Learning Source Disentanglement in Neural Audio Codec
por: Bie, Xiaoyu, et al.
Publicado: (2024)
por: Bie, Xiaoyu, et al.
Publicado: (2024)
Detection and Forecasting of Parkinson Disease Progression from Speech Signal Features Using MultiLayer Perceptron and LSTM
por: Ali, Majid, et al.
Publicado: (2024)
por: Ali, Majid, et al.
Publicado: (2024)
Ejemplares similares
-
Deep Learning-Based Automatic Multi-Level Airway Collapse Monitoring on Obstructive Sleep Apnea Patients
por: Hsu, Ying-Chieh, et al.
Publicado: (2024) -
Predicting Individual Depression Symptoms from Acoustic Features During Speech
por: Rodriguez, Sebastian, et al.
Publicado: (2024) -
Tuning In: Analysis of Audio Classifier Performance in Clinical Settings with Limited Data
por: Mahdi, Hamza, et al.
Publicado: (2024) -
Estimating Respiratory Effort from Nocturnal Breathing Sounds for Obstructive Sleep Apnoea Screening
por: Xu, Xiaolei, et al.
Publicado: (2025) -
Continuous Autoregressive Models with Noise Augmentation Avoid Error Accumulation
por: Pasini, Marco, et al.
Publicado: (2024)