Salvato in:
| Autori principali: | Ravuri, Aditya, Muir, Jen, Lawrence, Neil D. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2407.01452 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Attention-Guided Adaptation for Code-Switching Speech Recognition
di: Aditya, Bobbi, et al.
Pubblicazione: (2023)
di: Aditya, Bobbi, et al.
Pubblicazione: (2023)
End-to-End Integration of Speech Emotion Recognition with Voice Activity Detection using Self-Supervised Learning Features
di: Yamashita, Natsuo, et al.
Pubblicazione: (2024)
di: Yamashita, Natsuo, et al.
Pubblicazione: (2024)
SCDNet: Self-supervised Learning Feature-based Speaker Change Detection
di: Li, Yue, et al.
Pubblicazione: (2024)
di: Li, Yue, et al.
Pubblicazione: (2024)
Generalizable Audio Deepfake Detection via Hierarchical Structure Learning and Feature Whitening in Poincaré sphere
di: Yang, Mingru, et al.
Pubblicazione: (2025)
di: Yang, Mingru, et al.
Pubblicazione: (2025)
Musical Chords: A Novel Java Algorithm and App Utility to Enumerate Chord-Progressions Adhering to Music Theory Guidelines
di: Lakshminarasimhan, Aditya
Pubblicazione: (2024)
di: Lakshminarasimhan, Aditya
Pubblicazione: (2024)
Profile-Error-Tolerant Target-Speaker Voice Activity Detection
di: Wang, Dongmei, et al.
Pubblicazione: (2023)
di: Wang, Dongmei, et al.
Pubblicazione: (2023)
Abusive Speech Detection in Indic Languages Using Acoustic Features
di: Spiesberger, Anika A., et al.
Pubblicazione: (2024)
di: Spiesberger, Anika A., et al.
Pubblicazione: (2024)
ComFeAT: Combination of Neural and Spectral Features for Improved Depression Detection
di: Phukan, Orchid Chetia, et al.
Pubblicazione: (2024)
di: Phukan, Orchid Chetia, et al.
Pubblicazione: (2024)
A Noval Feature via Color Quantisation for Fake Audio Detection
di: Wang, Zhiyong, et al.
Pubblicazione: (2024)
di: Wang, Zhiyong, et al.
Pubblicazione: (2024)
Disentangling Hierarchical Features for Anomalous Sound Detection Under Domain Shift
di: Guan, Jian, et al.
Pubblicazione: (2025)
di: Guan, Jian, et al.
Pubblicazione: (2025)
Unmasking Deepfakes: Leveraging Augmentations and Features Variability for Deepfake Speech Detection
di: Rimon, Inbal, et al.
Pubblicazione: (2025)
di: Rimon, Inbal, et al.
Pubblicazione: (2025)
Single-Microphone Speaker Separation and Voice Activity Detection in Noisy and Reverberant Environments
di: Opochinsky, Renana, et al.
Pubblicazione: (2024)
di: Opochinsky, Renana, et al.
Pubblicazione: (2024)
Channel-Combination Algorithms for Robust Distant Voice Activity and Overlapped Speech Detection
di: Mariotte, Théo, et al.
Pubblicazione: (2024)
di: Mariotte, Théo, et al.
Pubblicazione: (2024)
Speaker Embeddings With Weakly Supervised Voice Activity Detection For Efficient Speaker Diarization
di: Thienpondt, Jenthe, et al.
Pubblicazione: (2024)
di: Thienpondt, Jenthe, et al.
Pubblicazione: (2024)
SEABAD: A Tropical Bird Activity Detection Dataset for Passive Acoustic Monitoring
di: Zabidi, Muhammad Mun'im Ahmad, et al.
Pubblicazione: (2026)
di: Zabidi, Muhammad Mun'im Ahmad, et al.
Pubblicazione: (2026)
Attention Is Not Always the Answer: Optimizing Voice Activity Detection with Simple Feature Fusion
di: Tripathi, Kumud, et al.
Pubblicazione: (2025)
di: Tripathi, Kumud, et al.
Pubblicazione: (2025)
Improvements of Discriminative Feature Space Training for Anomalous Sound Detection in Unlabeled Conditions
di: Fujimura, Takuya, et al.
Pubblicazione: (2024)
di: Fujimura, Takuya, et al.
Pubblicazione: (2024)
Flow-TSVAD: Target-Speaker Voice Activity Detection via Latent Flow Matching
di: Chen, Zhengyang, et al.
Pubblicazione: (2024)
di: Chen, Zhengyang, et al.
Pubblicazione: (2024)
Universal Speaker Embedding Free Target Speaker Extraction and Personal Voice Activity Detection
di: Zeng, Bang, et al.
Pubblicazione: (2025)
di: Zeng, Bang, et al.
Pubblicazione: (2025)
Contrastive Loss Based Frame-wise Feature disentanglement for Polyphonic Sound Event Detection
di: Guan, Yadong, et al.
Pubblicazione: (2024)
di: Guan, Yadong, et al.
Pubblicazione: (2024)
Device Feature based on Graph Fourier Transformation with Logarithmic Processing For Detection of Replay Speech Attacks
di: He, Mingrui, et al.
Pubblicazione: (2024)
di: He, Mingrui, et al.
Pubblicazione: (2024)
BreathNet: Generalizable Audio Deepfake Detection via Breath-Cue-Guided Feature Refinement
di: Ye, Zhe, et al.
Pubblicazione: (2026)
di: Ye, Zhe, et al.
Pubblicazione: (2026)
Joint Speaker Features Learning for Audio-visual Multichannel Speech Separation and Recognition
di: Li, Guinan, et al.
Pubblicazione: (2024)
di: Li, Guinan, et al.
Pubblicazione: (2024)
Selective-Memory Meta-Learning with Environment Representations for Sound Event Localization and Detection
di: Hu, Jinbo, et al.
Pubblicazione: (2023)
di: Hu, Jinbo, et al.
Pubblicazione: (2023)
Causal Speech Enhancement with Predicting Semantics based on Quantized Self-supervised Learning Features
di: Tsunoo, Emiru, et al.
Pubblicazione: (2024)
di: Tsunoo, Emiru, et al.
Pubblicazione: (2024)
An Efficient End-to-End Approach to Noise Invariant Speech Features via Multi-Task Learning
di: Guimarães, Heitor R., et al.
Pubblicazione: (2024)
di: Guimarães, Heitor R., et al.
Pubblicazione: (2024)
Freeze and Learn: Continual Learning with Selective Freezing for Speech Deepfake Detection
di: Salvi, Davide, et al.
Pubblicazione: (2024)
di: Salvi, Davide, et al.
Pubblicazione: (2024)
Robust AI-Synthesized Speech Detection Using Feature Decomposition Learning and Synthesizer Feature Augmentation
di: Zhang, Kuiyuan, et al.
Pubblicazione: (2024)
di: Zhang, Kuiyuan, et al.
Pubblicazione: (2024)
Continuous Learning of Transformer-based Audio Deepfake Detection
di: Le, Tuan Duy Nguyen, et al.
Pubblicazione: (2024)
di: Le, Tuan Duy Nguyen, et al.
Pubblicazione: (2024)
Water Flow Detection Device Based on Sound Data Analysis and Machine Learning to Detect Water Leakage
di: Pourmehrani, Hossein, et al.
Pubblicazione: (2025)
di: Pourmehrani, Hossein, et al.
Pubblicazione: (2025)
Reverberation-based Features for Sound Event Localization and Detection with Distance Estimation
di: Berghi, Davide, et al.
Pubblicazione: (2025)
di: Berghi, Davide, et al.
Pubblicazione: (2025)
Generalized Fake Audio Detection via Deep Stable Learning
di: Wang, Zhiyong, et al.
Pubblicazione: (2024)
di: Wang, Zhiyong, et al.
Pubblicazione: (2024)
Leveraging Prompt Learning and Pause Encoding for Alzheimer's Disease Detection
di: Liu, Yin-Long, et al.
Pubblicazione: (2024)
di: Liu, Yin-Long, et al.
Pubblicazione: (2024)
Spatial-Temporal Activity-Informed Diarization and Separation
di: Hsu, Yicheng, et al.
Pubblicazione: (2024)
di: Hsu, Yicheng, et al.
Pubblicazione: (2024)
Self-Attention and Hybrid Features for Replay and Deep-Fake Audio Detection
di: Huang, Lian, et al.
Pubblicazione: (2024)
di: Huang, Lian, et al.
Pubblicazione: (2024)
Low-power SNN-based audio source localisation using a Hilbert Transform spike encoding scheme
di: Haghighatshoar, Saeid, et al.
Pubblicazione: (2024)
di: Haghighatshoar, Saeid, et al.
Pubblicazione: (2024)
FGCL: Fine-grained Contrastive Learning For Mandarin Stuttering Event Detection
di: Jiang, Han, et al.
Pubblicazione: (2024)
di: Jiang, Han, et al.
Pubblicazione: (2024)
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
di: Xiao, Yang, et al.
Pubblicazione: (2024)
di: Xiao, Yang, et al.
Pubblicazione: (2024)
Naturalness-Aware Curriculum Learning with Dynamic Temperature for Speech Deepfake Detection
di: Kim, Taewoo, et al.
Pubblicazione: (2025)
di: Kim, Taewoo, et al.
Pubblicazione: (2025)
Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection
di: Amiri, Mahdi, et al.
Pubblicazione: (2025)
di: Amiri, Mahdi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Attention-Guided Adaptation for Code-Switching Speech Recognition
di: Aditya, Bobbi, et al.
Pubblicazione: (2023) -
End-to-End Integration of Speech Emotion Recognition with Voice Activity Detection using Self-Supervised Learning Features
di: Yamashita, Natsuo, et al.
Pubblicazione: (2024) -
SCDNet: Self-supervised Learning Feature-based Speaker Change Detection
di: Li, Yue, et al.
Pubblicazione: (2024) -
Generalizable Audio Deepfake Detection via Hierarchical Structure Learning and Feature Whitening in Poincaré sphere
di: Yang, Mingru, et al.
Pubblicazione: (2025) -
Musical Chords: A Novel Java Algorithm and App Utility to Enumerate Chord-Progressions Adhering to Music Theory Guidelines
di: Lakshminarasimhan, Aditya
Pubblicazione: (2024)