Multi-Channel Replay Speech Detection using Acoustic Maps
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Neri, Michael, Virtanen, Tuomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Acoustic Simulation Framework for Multi-channel Replay Speech Detection
von: Neri, Michael, et al.
Veröffentlicht: (2025)
von: Neri, Michael, et al.
Veröffentlicht: (2025)
Multi-channel Replay Speech Detection using an Adaptive Learnable Beamformer
von: Neri, Michael, et al.
Veröffentlicht: (2025)
von: Neri, Michael, et al.
Veröffentlicht: (2025)
Knowledge Distillation for Speech Denoising by Latent Representation Alignment with Cosine Distance
von: Luong, Diep, et al.
Veröffentlicht: (2025)
von: Luong, Diep, et al.
Veröffentlicht: (2025)
From Weak to Strong Sound Event Labels using Adaptive Change-Point Detection and Active Learning
von: Martinsson, John, et al.
Veröffentlicht: (2024)
von: Martinsson, John, et al.
Veröffentlicht: (2024)
Impact of Microphone Array Mismatches to Learning-based Replay Speech Detection
von: Neri, Michael, et al.
Veröffentlicht: (2025)
von: Neri, Michael, et al.
Veröffentlicht: (2025)
Multi-label Zero-Shot Audio Classification with Temporal Attention
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
Speaker Distance Estimation in Enclosures from Single-Channel Audio
von: Neri, Michael, et al.
Veröffentlicht: (2024)
von: Neri, Michael, et al.
Veröffentlicht: (2024)
Hybrid Disagreement-Diversity Active Learning for Bioacoustic Sound Event Detection
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
Dependence on Early and Late Reverberation of Single-Channel Speaker Distance Estimation
von: Neri, Michael, et al.
Veröffentlicht: (2026)
von: Neri, Michael, et al.
Veröffentlicht: (2026)
Automatic Contextual Audio Denoising
von: Luong, Diep, et al.
Veröffentlicht: (2026)
von: Luong, Diep, et al.
Veröffentlicht: (2026)
Gen-A: Generalizing Ambisonics Neural Encoding to Unseen Microphone Arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025)
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025)
Score-informed Music Source Separation: Improving Synthetic-to-real Generalization in Classical Music
von: Tunturi, Eetu, et al.
Veröffentlicht: (2025)
von: Tunturi, Eetu, et al.
Veröffentlicht: (2025)
Multi-Utterance Speech Separation and Association Trained on Short Segments
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
Inter-Speaker Relative Cues for Text-Guided Target Speech Extraction
von: Dai, Wang, et al.
Veröffentlicht: (2025)
von: Dai, Wang, et al.
Veröffentlicht: (2025)
Evaluating the Temporal Detection Capability of Integrated Gradients Applied on Sound Classifier
von: Dumpis, Martynas, et al.
Veröffentlicht: (2026)
von: Dumpis, Martynas, et al.
Veröffentlicht: (2026)
Attractor-Based Speech Separation of Multiple Utterances by Unknown Number of Speakers
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
Automatic Live Music Song Identification Using Multi-level Deep Sequence Similarity Learning
von: Hakala, Aapo, et al.
Veröffentlicht: (2025)
von: Hakala, Aapo, et al.
Veröffentlicht: (2025)
Unsupervised Acoustic Scene Mapping Based on Acoustic Features and Dimensionality Reduction
von: Cohen, Idan, et al.
Veröffentlicht: (2023)
von: Cohen, Idan, et al.
Veröffentlicht: (2023)
Autoregressive Speech Enhancement via Acoustic Tokens
von: Della Libera, Luca, et al.
Veröffentlicht: (2025)
von: Della Libera, Luca, et al.
Veröffentlicht: (2025)
Benchmarking Representations for Speech, Music, and Acoustic Events
von: La Quatra, Moreno, et al.
Veröffentlicht: (2024)
von: La Quatra, Moreno, et al.
Veröffentlicht: (2024)
Representation Learning for Audio Privacy Preservation using Source Separation and Robust Adversarial Learning
von: Luong, Diep, et al.
Veröffentlicht: (2023)
von: Luong, Diep, et al.
Veröffentlicht: (2023)
Toward Faithful Explanations in Acoustic Anomaly Detection
von: Elrashid, Maab, et al.
Veröffentlicht: (2026)
von: Elrashid, Maab, et al.
Veröffentlicht: (2026)
SynthSOD: Developing an Heterogeneous Dataset for Orchestra Music Source Separation
von: Garcia-Martinez, Jaime, et al.
Veröffentlicht: (2024)
von: Garcia-Martinez, Jaime, et al.
Veröffentlicht: (2024)
Permutation Invariant Recurrent Neural Networks for Sound Source Tracking Applications
von: Diaz-Guerra, David, et al.
Veröffentlicht: (2023)
von: Diaz-Guerra, David, et al.
Veröffentlicht: (2023)
Scalable Speech Enhancement with Dynamic Channel Pruning
von: Miccini, Riccardo, et al.
Veröffentlicht: (2024)
von: Miccini, Riccardo, et al.
Veröffentlicht: (2024)
Neural Ambisonics encoding for compact irregular microphone arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2024)
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2024)
Noise-to-mask Ratio Loss for Deep Neural Network based Audio Watermarking
von: Moritz, Martin, et al.
Veröffentlicht: (2024)
von: Moritz, Martin, et al.
Veröffentlicht: (2024)
Adversarial Representation Learning for Robust Privacy Preservation in Audio
von: Gharib, Shayan, et al.
Veröffentlicht: (2023)
von: Gharib, Shayan, et al.
Veröffentlicht: (2023)
LMFCA-Net: A Lightweight Model for Multi-Channel Speech Enhancement with Efficient Narrow-Band and Cross-Band Attention
von: Zhang, Yaokai, et al.
Veröffentlicht: (2025)
von: Zhang, Yaokai, et al.
Veröffentlicht: (2025)
Acoustic and Machine Learning Methods for Speech-Based Suicide Risk Assessment: A Systematic Review
von: Marie, Ambre, et al.
Veröffentlicht: (2025)
von: Marie, Ambre, et al.
Veröffentlicht: (2025)
Impact of Speech Mode in Automatic Pathological Speech Detection
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
von: Sheikh, Shakeel A., et al.
Veröffentlicht: (2024)
Dynamic Recognition of Speakers for Consent Management by Contrastive Embedding Replay
von: Shahmansoori, Arash, et al.
Veröffentlicht: (2022)
von: Shahmansoori, Arash, et al.
Veröffentlicht: (2022)
The Spheres Dataset: Multitrack Orchestral Recordings for Music Source Separation and Information Retrieval
von: Garcia-Martinez, Jaime, et al.
Veröffentlicht: (2025)
von: Garcia-Martinez, Jaime, et al.
Veröffentlicht: (2025)
Acoustic Classification of Maritime Vessels using Learnable Filterbanks
von: Elsborg, Jonas, et al.
Veröffentlicht: (2025)
von: Elsborg, Jonas, et al.
Veröffentlicht: (2025)
RelUNet: Relative Channel Fusion U-Net for Multichannel Speech Enhancement
von: Aldarmaki, Ibrahim, et al.
Veröffentlicht: (2024)
von: Aldarmaki, Ibrahim, et al.
Veröffentlicht: (2024)
Acoustic Scene Classification: A Competition Review
von: Gharib, Shayan, et al.
Veröffentlicht: (2018)
von: Gharib, Shayan, et al.
Veröffentlicht: (2018)
Multi-blank Transducers for Speech Recognition
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
Detecting Throat Cancer from Speech Signals using Machine Learning: A Scoping Literature Review
von: Paterson, Mary, et al.
Veröffentlicht: (2023)
von: Paterson, Mary, et al.
Veröffentlicht: (2023)
Moving Speaker Separation via Parallel Spectral-Spatial Processing
von: Wang, Yuzhu, et al.
Veröffentlicht: (2026)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2026)
Advancing Robust Underwater Acoustic Target Recognition through Multi-task Learning and Multi-Gate Mixture-of-Experts
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Acoustic Simulation Framework for Multi-channel Replay Speech Detection
von: Neri, Michael, et al.
Veröffentlicht: (2025) -
Multi-channel Replay Speech Detection using an Adaptive Learnable Beamformer
von: Neri, Michael, et al.
Veröffentlicht: (2025) -
Knowledge Distillation for Speech Denoising by Latent Representation Alignment with Cosine Distance
von: Luong, Diep, et al.
Veröffentlicht: (2025) -
From Weak to Strong Sound Event Labels using Adaptive Change-Point Detection and Active Learning
von: Martinsson, John, et al.
Veröffentlicht: (2024) -
Impact of Microphone Array Mismatches to Learning-based Replay Speech Detection
von: Neri, Michael, et al.
Veröffentlicht: (2025)