Effects of Recording Condition and Number of Monitored Days on Discriminative Power of the Daily Phonotrauma Index
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ghasemzadeh, Hamzeh, Hillman, Robert E., Van Stan, Jarrad H., Mehta, Daryush D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improvements of Discriminative Feature Space Training for Anomalous Sound Detection in Unlabeled Conditions
von: Fujimura, Takuya, et al.
Veröffentlicht: (2024)
von: Fujimura, Takuya, et al.
Veröffentlicht: (2024)
Noise-Robust Voice Conversion by Conditional Denoising Training Using Latent Variables of Recording Quality and Environment
von: Igarashi, Takuto, et al.
Veröffentlicht: (2024)
von: Igarashi, Takuto, et al.
Veröffentlicht: (2024)
Stream-based Active Learning for Anomalous Sound Detection in Machine Condition Monitoring
von: Ho, Tuan Vu, et al.
Veröffentlicht: (2024)
von: Ho, Tuan Vu, et al.
Veröffentlicht: (2024)
TBDM-Net: Bidirectional Dense Networks with Gender Information for Speech Emotion Recognition
von: Striletchi, Vlad, et al.
Veröffentlicht: (2024)
von: Striletchi, Vlad, et al.
Veröffentlicht: (2024)
SaD: A Scenario-Aware Discriminator for Speech Enhancement
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
A Hybrid Discriminative and Generative System for Universal Speech Enhancement
von: Liu, Yinghao, et al.
Veröffentlicht: (2026)
von: Liu, Yinghao, et al.
Veröffentlicht: (2026)
Streaming Keyword Spotting Boosted by Cross-layer Discrimination Consistency
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
von: Nishida, Tomoya, et al.
Veröffentlicht: (2026)
von: Nishida, Tomoya, et al.
Veröffentlicht: (2026)
Description and Discussion on DCASE 2025 Challenge Task 2: First-shot Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
von: Nishida, Tomoya, et al.
Veröffentlicht: (2025)
von: Nishida, Tomoya, et al.
Veröffentlicht: (2025)
Towards generalisable and calibrated synthetic speech detection with self-supervised representations
von: Pascu, Octavian, et al.
Veröffentlicht: (2023)
von: Pascu, Octavian, et al.
Veröffentlicht: (2023)
Diffusion-based Generative Modeling with Discriminative Guidance for Streamable Speech Enhancement
von: Li, Chenda, et al.
Veröffentlicht: (2024)
von: Li, Chenda, et al.
Veröffentlicht: (2024)
Discriminative-Generative Target Speaker Extraction with Decoder-Only Language Models
von: Zeng, Bang, et al.
Veröffentlicht: (2026)
von: Zeng, Bang, et al.
Veröffentlicht: (2026)
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
Peransformer: Improving Low-informed Expressive Performance Rendering with Score-aware Discriminator
von: He, Xian, et al.
Veröffentlicht: (2025)
von: He, Xian, et al.
Veröffentlicht: (2025)
Index-ASR Technical Report
von: Song, Zheshu, et al.
Veröffentlicht: (2025)
von: Song, Zheshu, et al.
Veröffentlicht: (2025)
Hyper Recurrent Neural Network: Condition Mechanisms for Black-box Audio Effect Modeling
von: Yeh, Yen-Tung, et al.
Veröffentlicht: (2024)
von: Yeh, Yen-Tung, et al.
Veröffentlicht: (2024)
Music2Fail: Transfer Music to Failed Recorder Style
von: Leong, Chon In, et al.
Veröffentlicht: (2024)
von: Leong, Chon In, et al.
Veröffentlicht: (2024)
DIFFRENT: A Diffusion Model for Recording Environment Transfer of Speech
von: Im, Jaekwon, et al.
Veröffentlicht: (2024)
von: Im, Jaekwon, et al.
Veröffentlicht: (2024)
Comparative Analysis Of Discriminative Deep Learning-Based Noise Reduction Methods In Low SNR Scenarios
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
Automated Analysis of Naturalistic Recordings in Early Childhood: Applications, Challenges, and Opportunities
von: Li, Jialu, et al.
Veröffentlicht: (2025)
von: Li, Jialu, et al.
Veröffentlicht: (2025)
SRC4VC: Smartphone-Recorded Corpus for Voice Conversion Benchmark
von: Saito, Yuki, et al.
Veröffentlicht: (2024)
von: Saito, Yuki, et al.
Veröffentlicht: (2024)
BickGraphing: Web-Based Application for Visual Inspection of Audio Recordings
von: Seow, Kayley, et al.
Veröffentlicht: (2026)
von: Seow, Kayley, et al.
Veröffentlicht: (2026)
VoiceRestore: Flow-Matching Transformers for Speech Recording Quality Restoration
von: Kirdey, Stanislav
Veröffentlicht: (2025)
von: Kirdey, Stanislav
Veröffentlicht: (2025)
Pureformer-VC: Non-parallel Voice Conversion with Pure Stylized Transformer Blocks and Triplet Discriminative Training
von: Yao, Wenhan, et al.
Veröffentlicht: (2025)
von: Yao, Wenhan, et al.
Veröffentlicht: (2025)
Evaluating the Impact of Discriminative and Generative E2E Speech Enhancement Models on Syllable Stress Preservation
von: Bharadwaj, Rangavajjala Sankara, et al.
Veröffentlicht: (2024)
von: Bharadwaj, Rangavajjala Sankara, et al.
Veröffentlicht: (2024)
Recursive Attentive Pooling for Extracting Speaker Embeddings from Multi-Speaker Recordings
von: Horiguchi, Shota, et al.
Veröffentlicht: (2024)
von: Horiguchi, Shota, et al.
Veröffentlicht: (2024)
Continuous Target Speech Extraction: Enhancing Personalized Diarization and Extraction on Complex Recordings
von: Zhao, He, et al.
Veröffentlicht: (2024)
von: Zhao, He, et al.
Veröffentlicht: (2024)
Audio Enhancement from Multiple Crowdsourced Recordings: A Simple and Effective Baseline
von: Aziz, Shiran, et al.
Veröffentlicht: (2024)
von: Aziz, Shiran, et al.
Veröffentlicht: (2024)
Mobile Recording Device Recognition Based Cross-Scale and Multi-Level Representation Learning
von: Zeng, Chunyan, et al.
Veröffentlicht: (2024)
von: Zeng, Chunyan, et al.
Veröffentlicht: (2024)
Estimating the Number and Locations of Boundaries in Reverberant Environments with Deep Learning
von: Arikan, Toros, et al.
Veröffentlicht: (2024)
von: Arikan, Toros, et al.
Veröffentlicht: (2024)
A MATLAB toolbox for Computation of Speech Transmission Index (STI)
von: Rajmic, Pavel, et al.
Veröffentlicht: (2025)
von: Rajmic, Pavel, et al.
Veröffentlicht: (2025)
Enforcing Speech Content Privacy in Environmental Sound Recordings using Segment-wise Waveform Reversal
von: Tailleur, Modan, et al.
Veröffentlicht: (2025)
von: Tailleur, Modan, et al.
Veröffentlicht: (2025)
Attractor-Based Speech Separation of Multiple Utterances by Unknown Number of Speakers
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
Computational Extraction of Intonation and Tuning Systems from Multiple Microtonal Monophonic Vocal Recordings with Diverse Modes
von: Shafiei, Sepideh, et al.
Veröffentlicht: (2025)
von: Shafiei, Sepideh, et al.
Veröffentlicht: (2025)
RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization
von: Yang, Bing, et al.
Veröffentlicht: (2024)
von: Yang, Bing, et al.
Veröffentlicht: (2024)
Training Generative Adversarial Network-Based Vocoder with Limited Data Using Augmentation-Conditional Discriminator
von: Kaneko, Takuhiro, et al.
Veröffentlicht: (2024)
von: Kaneko, Takuhiro, et al.
Veröffentlicht: (2024)
Multichannel Keyword Spotting for Noisy Conditions
von: Saladukha, Dzmitry, et al.
Veröffentlicht: (2025)
von: Saladukha, Dzmitry, et al.
Veröffentlicht: (2025)
Crab: Multi Layer Contrastive Supervision to Improve Speech Emotion Recognition Under Both Acted and Natural Speech Condition
von: Ueda, Lucas H., et al.
Veröffentlicht: (2026)
von: Ueda, Lucas H., et al.
Veröffentlicht: (2026)
HumDial-EIBench: A Human-Recorded Multi-Turn Emotional Intelligence Benchmark for Audio Language Models
von: Wang, Shuiyuan, et al.
Veröffentlicht: (2026)
von: Wang, Shuiyuan, et al.
Veröffentlicht: (2026)
GESI: Gammachirp Envelope Similarity Index for Predicting Intelligibility of Simulated Hearing Loss Sounds
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2023)
von: Yamamoto, Ayako, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Improvements of Discriminative Feature Space Training for Anomalous Sound Detection in Unlabeled Conditions
von: Fujimura, Takuya, et al.
Veröffentlicht: (2024) -
Noise-Robust Voice Conversion by Conditional Denoising Training Using Latent Variables of Recording Quality and Environment
von: Igarashi, Takuto, et al.
Veröffentlicht: (2024) -
Stream-based Active Learning for Anomalous Sound Detection in Machine Condition Monitoring
von: Ho, Tuan Vu, et al.
Veröffentlicht: (2024) -
TBDM-Net: Bidirectional Dense Networks with Gender Information for Speech Emotion Recognition
von: Striletchi, Vlad, et al.
Veröffentlicht: (2024) -
SaD: A Scenario-Aware Discriminator for Speech Enhancement
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)