Impact of Microphone Array Mismatches to Learning-based Replay Speech Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Neri, Michael, Virtanen, Tuomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-channel Replay Speech Detection using an Adaptive Learnable Beamformer
von: Neri, Michael, et al.
Veröffentlicht: (2025)
von: Neri, Michael, et al.
Veröffentlicht: (2025)
Acoustic Simulation Framework for Multi-channel Replay Speech Detection
von: Neri, Michael, et al.
Veröffentlicht: (2025)
von: Neri, Michael, et al.
Veröffentlicht: (2025)
Multi-Channel Replay Speech Detection using Acoustic Maps
von: Neri, Michael, et al.
Veröffentlicht: (2026)
von: Neri, Michael, et al.
Veröffentlicht: (2026)
Microphone Array Signal Processing and Deep Learning for Speech Enhancement
von: Haeb-Umbach, Reinhold, et al.
Veröffentlicht: (2025)
von: Haeb-Umbach, Reinhold, et al.
Veröffentlicht: (2025)
Evaluating the Temporal Detection Capability of Integrated Gradients Applied on Sound Classifier
von: Dumpis, Martynas, et al.
Veröffentlicht: (2026)
von: Dumpis, Martynas, et al.
Veröffentlicht: (2026)
Dependence on Early and Late Reverberation of Single-Channel Speaker Distance Estimation
von: Neri, Michael, et al.
Veröffentlicht: (2026)
von: Neri, Michael, et al.
Veröffentlicht: (2026)
HyBeam: Hybrid Microphone-Beamforming Array-Agnostic Speech Enhancement for Wearables
von: Ilan, Yuval Bar, et al.
Veröffentlicht: (2025)
von: Ilan, Yuval Bar, et al.
Veröffentlicht: (2025)
HOMULA-RIR: A Room Impulse Response Dataset for Teleconferencing and Spatial Audio Applications Acquired Through Higher-Order Microphones and Uniform Linear Microphone Arrays
von: Miotello, Federico, et al.
Veröffentlicht: (2024)
von: Miotello, Federico, et al.
Veröffentlicht: (2024)
Discriminating real and synthetic super-resolved audio samples using embedding-based classifiers
von: Silaev, Mikhail, et al.
Veröffentlicht: (2026)
von: Silaev, Mikhail, et al.
Veröffentlicht: (2026)
PLDNet: PLD-Guided Lightweight Deep Network Boosted by Efficient Attention for Handheld Dual-Microphone Speech Enhancement
von: Zhou, Nan, et al.
Veröffentlicht: (2024)
von: Zhou, Nan, et al.
Veröffentlicht: (2024)
Microphone Array Geometry Independent Multi-Talker Distant ASR: NTT System for the DASR Task of the CHiME-8 Challenge
von: Kamo, Naoyuki, et al.
Veröffentlicht: (2025)
von: Kamo, Naoyuki, et al.
Veröffentlicht: (2025)
High-Density MIMO Localization Using a 32x64 Ultrasonic Transducer-Microphone Array with Real-Time Data Streaming
von: Baeyens, Rens, et al.
Veröffentlicht: (2025)
von: Baeyens, Rens, et al.
Veröffentlicht: (2025)
HiRIS: an Airborne Sonar Sensor with a 1024 Channel Microphone Array for In-Air Acoustic Imaging
von: Laurijssen, Dennis, et al.
Veröffentlicht: (2024)
von: Laurijssen, Dennis, et al.
Veröffentlicht: (2024)
Independent Feature Enhanced Crossmodal Fusion for Match-Mismatch Classification of Speech Stimulus and EEG Response
von: Fan, Shitong, et al.
Veröffentlicht: (2024)
von: Fan, Shitong, et al.
Veröffentlicht: (2024)
Advanced Signal Analysis in Detecting Replay Attacks for Automatic Speaker Verification Systems
von: Kuang, Lee Shih
Veröffentlicht: (2024)
von: Kuang, Lee Shih
Veröffentlicht: (2024)
Gen-A: Generalizing Ambisonics Neural Encoding to Unseen Microphone Arrays
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025)
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2025)
Beyond Omnidirectional: Neural Ambisonics Encoding for Arbitrary Microphone Directivity Patterns using Cross-Attention
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2026)
von: Heikkinen, Mikko, et al.
Veröffentlicht: (2026)
Tool Wear Prediction in CNC Turning Operations using Ultrasonic Microphone Arrays and CNNs
von: Steckel, Jan, et al.
Veröffentlicht: (2024)
von: Steckel, Jan, et al.
Veröffentlicht: (2024)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
BRUDEX Database: Binaural Room Impulse Responses with Uniformly Distributed External Microphones
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
Analyzing the Impact of Accent on English Speech: Acoustic and Articulatory Perspectives
von: Premananth, Gowtham, et al.
Veröffentlicht: (2025)
von: Premananth, Gowtham, et al.
Veröffentlicht: (2025)
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
von: Miotello, Federico, et al.
Veröffentlicht: (2024)
von: Miotello, Federico, et al.
Veröffentlicht: (2024)
Speech Enhancement based on cascaded two flows
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025)
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025)
Exploiting an External Microphone for Binaural RTF-Vector-Based Direction of Arrival Estimation for Multiple Speakers
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
von: Fejgin, Daniel, et al.
Veröffentlicht: (2023)
Robust Detection of Underwater Target Against Non-Uniform Noise With Optical Fiber DAS Array
von: Cang, Siyuan, et al.
Veröffentlicht: (2025)
von: Cang, Siyuan, et al.
Veröffentlicht: (2025)
FlowSE: Flow Matching-based Speech Enhancement
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025)
von: Lee, Seonggyu, et al.
Veröffentlicht: (2025)
Advances in Microphone Array Processing and Multichannel Speech Enhancement
von: Huang, Gongping, et al.
Veröffentlicht: (2025)
von: Huang, Gongping, et al.
Veröffentlicht: (2025)
Permutation Invariant Recurrent Neural Networks for Sound Source Tracking Applications
von: Diaz-Guerra, David, et al.
Veröffentlicht: (2023)
von: Diaz-Guerra, David, et al.
Veröffentlicht: (2023)
Inter-Speaker Relative Cues for Two-Stage Text-Guided Target Speech Extraction
von: Dai, Wang, et al.
Veröffentlicht: (2026)
von: Dai, Wang, et al.
Veröffentlicht: (2026)
SpeechMLC: Speech Multi-label Classification
von: Kim, Miseul, et al.
Veröffentlicht: (2025)
von: Kim, Miseul, et al.
Veröffentlicht: (2025)
Speak in the Scene: Diffusion-based Acoustic Scene Transfer toward Immersive Speech Generation
von: Kim, Miseul, et al.
Veröffentlicht: (2024)
von: Kim, Miseul, et al.
Veröffentlicht: (2024)
Ambisonics Encoder for Wearable Array with Improved Binaural Reproduction
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
Detecting Post-Stroke Aphasia Via Brain Responses to Speech in a Deep Learning Framework
von: De Clercq, Pieter, et al.
Veröffentlicht: (2024)
von: De Clercq, Pieter, et al.
Veröffentlicht: (2024)
A Speech Production Model for Radar: Connecting Speech Acoustics with Radar-Measured Vibrations
von: Lenz, Isabella, et al.
Veröffentlicht: (2025)
von: Lenz, Isabella, et al.
Veröffentlicht: (2025)
ParaS2S: Benchmarking and Aligning Spoken Language Models for Paralinguistic-aware Speech-to-Speech Interaction
von: Yang, Shu-wen, et al.
Veröffentlicht: (2025)
von: Yang, Shu-wen, et al.
Veröffentlicht: (2025)
Speaker and Style Disentanglement of Speech Based on Contrastive Predictive Coding Supported Factorized Variational Autoencoder
von: Xie, Yuying, et al.
Veröffentlicht: (2024)
von: Xie, Yuying, et al.
Veröffentlicht: (2024)
Binaural Localization Model for Speech in Noise
von: Tokala, Vikas, et al.
Veröffentlicht: (2025)
von: Tokala, Vikas, et al.
Veröffentlicht: (2025)
Speech-Based Prioritization for Schizophrenia Intervention
von: Premananth, Gowtham, et al.
Veröffentlicht: (2025)
von: Premananth, Gowtham, et al.
Veröffentlicht: (2025)
Prompt-driven Target Speech Diarization
von: Jiang, Yidi, et al.
Veröffentlicht: (2023)
von: Jiang, Yidi, et al.
Veröffentlicht: (2023)
Multiple Mobile Target Detection and Tracking in Active Sonar Array Using a Track-Before-Detect Approach
von: Abu, Avi, et al.
Veröffentlicht: (2024)
von: Abu, Avi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Multi-channel Replay Speech Detection using an Adaptive Learnable Beamformer
von: Neri, Michael, et al.
Veröffentlicht: (2025) -
Acoustic Simulation Framework for Multi-channel Replay Speech Detection
von: Neri, Michael, et al.
Veröffentlicht: (2025) -
Multi-Channel Replay Speech Detection using Acoustic Maps
von: Neri, Michael, et al.
Veröffentlicht: (2026) -
Microphone Array Signal Processing and Deep Learning for Speech Enhancement
von: Haeb-Umbach, Reinhold, et al.
Veröffentlicht: (2025) -
Evaluating the Temporal Detection Capability of Integrated Gradients Applied on Sound Classifier
von: Dumpis, Martynas, et al.
Veröffentlicht: (2026)