SimuSOE: A Simulated Snoring Dataset for Obstructive Sleep Apnea-Hypopnea Syndrome Evaluation during Wakefulness
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Jie, Yang, Xiuping, Xiao, Li, Li, Xinhong, Yi, Weiyan, Yang, Yuhong, Tu, Weiping, Chen, Xiong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Recall-First CNN for Sleep Apnea Screening from Snoring Audio
von: Mallick, Anushka, et al.
Veröffentlicht: (2025)
von: Mallick, Anushka, et al.
Veröffentlicht: (2025)
ICSD: An Open-source Dataset for Infant Cry and Snoring Detection
von: Liu, Qingyu, et al.
Veröffentlicht: (2024)
von: Liu, Qingyu, et al.
Veröffentlicht: (2024)
EMALG: An Enhanced Mandarin Lombard Grid Corpus with Meaningful Sentences
von: Li, Baifeng, et al.
Veröffentlicht: (2023)
von: Li, Baifeng, et al.
Veröffentlicht: (2023)
Improving Speech Enhancement by Cross- and Sub-band Processing with State Space Model
von: Li, Jizhen, et al.
Veröffentlicht: (2025)
von: Li, Jizhen, et al.
Veröffentlicht: (2025)
Improving Speech Enhancement by Integrating Inter-Channel and Band Features with Dual-branch Conformer
von: Li, Jizhen, et al.
Veröffentlicht: (2024)
von: Li, Jizhen, et al.
Veröffentlicht: (2024)
FreeCodec: A disentangled neural speech codec with fewer tokens
von: Zheng, Youqiang, et al.
Veröffentlicht: (2024)
von: Zheng, Youqiang, et al.
Veröffentlicht: (2024)
SuperCodec: A Neural Speech Codec with Selective Back-Projection Network
von: Zheng, Youqiang, et al.
Veröffentlicht: (2024)
von: Zheng, Youqiang, et al.
Veröffentlicht: (2024)
Exploring Sentence Type Effects on the Lombard Effect and Intelligibility Enhancement: A Comparative Study of Natural and Grid Sentences
von: Chen, Hongyang, et al.
Veröffentlicht: (2023)
von: Chen, Hongyang, et al.
Veröffentlicht: (2023)
Deep Learning-Based Automatic Multi-Level Airway Collapse Monitoring on Obstructive Sleep Apnea Patients
von: Hsu, Ying-Chieh, et al.
Veröffentlicht: (2024)
von: Hsu, Ying-Chieh, et al.
Veröffentlicht: (2024)
Estimating Respiratory Effort from Nocturnal Breathing Sounds for Obstructive Sleep Apnoea Screening
von: Xu, Xiaolei, et al.
Veröffentlicht: (2025)
von: Xu, Xiaolei, et al.
Veröffentlicht: (2025)
LlamaPartialSpoof: An LLM-Driven Fake Speech Dataset Simulating Disinformation Generation
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
von: Luong, Hieu-Thi, et al.
Veröffentlicht: (2024)
AdvSV: An Over-the-Air Adversarial Attack Dataset for Speaker Verification
von: Wang, Li, et al.
Veröffentlicht: (2023)
von: Wang, Li, et al.
Veröffentlicht: (2023)
Implementation and Applications of WakeWords Integrated with Speaker Recognition: A Case Study
von: Filho, Alexandre Costa Ferro, et al.
Veröffentlicht: (2024)
von: Filho, Alexandre Costa Ferro, et al.
Veröffentlicht: (2024)
RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization
von: Yang, Bing, et al.
Veröffentlicht: (2024)
von: Yang, Bing, et al.
Veröffentlicht: (2024)
Robust Wake Word Spotting With Frame-Level Cross-Modal Attention Based Audio-Visual Conformer
von: Wang, Haoxu, et al.
Veröffentlicht: (2024)
von: Wang, Haoxu, et al.
Veröffentlicht: (2024)
PB-LRDWWS System for the SLT 2024 Low-Resource Dysarthria Wake-Up Word Spotting Challenge
von: Wang, Shiyao, et al.
Veröffentlicht: (2024)
von: Wang, Shiyao, et al.
Veröffentlicht: (2024)
MSceneSpeech: A Multi-Scene Speech Dataset For Expressive Speech Synthesis
von: Yang, Qian, et al.
Veröffentlicht: (2024)
von: Yang, Qian, et al.
Veröffentlicht: (2024)
WildDESED: An LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection System
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
Self-Supervised Learning of Spatial Acoustic Representation with Cross-Channel Signal Reconstruction and Multi-Channel Conformer
von: Yang, Bing, et al.
Veröffentlicht: (2023)
von: Yang, Bing, et al.
Veröffentlicht: (2023)
PSELDNets: Pre-trained Neural Networks on a Large-scale Synthetic Dataset for Sound Event Localization and Detection
von: Hu, Jinbo, et al.
Veröffentlicht: (2024)
von: Hu, Jinbo, et al.
Veröffentlicht: (2024)
SE Territory: Monaural Speech Enhancement Meets the Fixed Virtual Perceptual Space Mapping
von: Xu, Xinmeng, et al.
Veröffentlicht: (2023)
von: Xu, Xinmeng, et al.
Veröffentlicht: (2023)
CompSpoof: A Dataset and Joint Learning Framework for Component-Level Audio Anti-spoofing Countermeasures
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
Scale This, Not That: Investigating Key Dataset Attributes for Efficient Speech Enhancement Scaling
von: Zhang, Leying, et al.
Veröffentlicht: (2024)
von: Zhang, Leying, et al.
Veröffentlicht: (2024)
ACMID: Automatic Curation of Musical Instrument Dataset for 7-Stem Music Source Separation
von: Yu, Ji, et al.
Veröffentlicht: (2025)
von: Yu, Ji, et al.
Veröffentlicht: (2025)
An Initial Investigation of Neural Replay Simulator for Over-the-Air Adversarial Perturbations to Automatic Speaker Verification
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
Text2Move: Text-to-moving sound generation via trajectory prediction and temporal alignment
von: Liu, Yunyi, et al.
Veröffentlicht: (2025)
von: Liu, Yunyi, et al.
Veröffentlicht: (2025)
Refining Self-Supervised Learnt Speech Representation using Brain Activations
von: Li, Hengyu, et al.
Veröffentlicht: (2024)
von: Li, Hengyu, et al.
Veröffentlicht: (2024)
FMSG-JLESS Submission for DCASE 2024 Task4 on Sound Event Detection with Heterogeneous Training Dataset and Potentially Missing Labels
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
von: Xiao, Yang, et al.
Veröffentlicht: (2024)
CoDiff-VC: A Codec-Assisted Diffusion Model for Zero-shot Voice Conversion
von: Li, Yuke, et al.
Veröffentlicht: (2024)
von: Li, Yuke, et al.
Veröffentlicht: (2024)
CUSIDE-T: Chunking, Simulating Future and Decoding for Transducer based Streaming ASR
von: Zhao, Wenbo, et al.
Veröffentlicht: (2024)
von: Zhao, Wenbo, et al.
Veröffentlicht: (2024)
IPDnet: A Universal Direct-Path IPD Estimation Network for Sound Source Localization
von: Wang, Yabo, et al.
Veröffentlicht: (2024)
von: Wang, Yabo, et al.
Veröffentlicht: (2024)
Complex-Cycle-Consistent Diffusion Model for Monaural Speech Enhancement
von: Li, Yi, et al.
Veröffentlicht: (2024)
von: Li, Yi, et al.
Veröffentlicht: (2024)
Pureformer-VC: Non-parallel Voice Conversion with Pure Stylized Transformer Blocks and Triplet Discriminative Training
von: Yao, Wenhan, et al.
Veröffentlicht: (2025)
von: Yao, Wenhan, et al.
Veröffentlicht: (2025)
Polyphonia: Zero-Shot Timbre Transfer in Polyphonic Music with Acoustic-Informed Attention Calibration
von: Li, Haowen, et al.
Veröffentlicht: (2026)
von: Li, Haowen, et al.
Veröffentlicht: (2026)
Toward Multimodal Industrial Fault Analysis: A Single-Speed Chain Conveyor Dataset with Audio and Vibration Signals
von: Chen, Zhang, et al.
Veröffentlicht: (2026)
von: Chen, Zhang, et al.
Veröffentlicht: (2026)
3S-TSE: Efficient Three-Stage Target Speaker Extraction for Real-Time and Low-Resource Applications
von: He, Shulin, et al.
Veröffentlicht: (2023)
von: He, Shulin, et al.
Veröffentlicht: (2023)
DOTA-ME-CS: Daily Oriented Text Audio-Mandarin English-Code Switching Dataset
von: Li, Yupei, et al.
Veröffentlicht: (2025)
von: Li, Yupei, et al.
Veröffentlicht: (2025)
DrawSpeech: Expressive Speech Synthesis Using Prosodic Sketches as Control Conditions
von: Chen, Weidong, et al.
Veröffentlicht: (2025)
von: Chen, Weidong, et al.
Veröffentlicht: (2025)
Enhancing Spectrogram Realism in Singing Voice Synthesis via Explicit Bandwidth Extension Prior to Vocoder
von: Yang, Runxuan, et al.
Veröffentlicht: (2025)
von: Yang, Runxuan, et al.
Veröffentlicht: (2025)
MT2KD: Towards A General-Purpose Encoder for Speech, Speaker, and Audio Events
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Yang, Xiaoyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Recall-First CNN for Sleep Apnea Screening from Snoring Audio
von: Mallick, Anushka, et al.
Veröffentlicht: (2025) -
ICSD: An Open-source Dataset for Infant Cry and Snoring Detection
von: Liu, Qingyu, et al.
Veröffentlicht: (2024) -
EMALG: An Enhanced Mandarin Lombard Grid Corpus with Meaningful Sentences
von: Li, Baifeng, et al.
Veröffentlicht: (2023) -
Improving Speech Enhancement by Cross- and Sub-band Processing with State Space Model
von: Li, Jizhen, et al.
Veröffentlicht: (2025) -
Improving Speech Enhancement by Integrating Inter-Channel and Band Features with Dual-branch Conformer
von: Li, Jizhen, et al.
Veröffentlicht: (2024)