Cepstral Smoothing of Binary Masks for Convolutive Blind Separation of Speech Mixtures
Fuente:
arXiv
Saved in:
| Main Authors: | Missaoui, Ibrahim, Lachiri, Zied |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Proficiency-Aware Adaptation and Data Augmentation for Robust L2 ASR
by: Sun, Ling, et al.
Published: (2025)
by: Sun, Ling, et al.
Published: (2025)
Matlab-based Epoch Extraction for Speaker Differentiation
by: Li, Kunlun, et al.
Published: (2024)
by: Li, Kunlun, et al.
Published: (2024)
Understanding the Algorithm Behind Audio Key Detection
by: Silva, Henrique Perez G.
Published: (2025)
by: Silva, Henrique Perez G.
Published: (2025)
HELIX: Scaling Raw Audio Understanding with Hybrid Mamba-Attention Beyond the Quadratic Limit
by: Khushiyant, et al.
Published: (2026)
by: Khushiyant, et al.
Published: (2026)
Machine learning based animal emotion classification using audio signals
by: Slobodian, Mariia, et al.
Published: (2025)
by: Slobodian, Mariia, et al.
Published: (2025)
IF-D: A High-Frequency, General-Purpose Inertial Foundation Dataset for Self-Supervised Learning
by: Ferreira, Patrick, et al.
Published: (2025)
by: Ferreira, Patrick, et al.
Published: (2025)
A study on audio synchronous steganography detection and distributed guide inference model based on sliding spectral features and intelligent inference drive
by: Meng, Wei
Published: (2025)
by: Meng, Wei
Published: (2025)
Person detection and re-identification in open-world settings of retail stores and public spaces
by: Brkljač, Branko, et al.
Published: (2025)
by: Brkljač, Branko, et al.
Published: (2025)
Skullptor: High Fidelity 3D Head Reconstruction in Seconds with Multi-View Normal Prediction
by: Artru, Noé, et al.
Published: (2026)
by: Artru, Noé, et al.
Published: (2026)
Contract-Driven QoE Auditing for Speech and Singing Services: From MOS Regression to Service Graphs
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
ISO/IEC-Compliant Match-on-Card Face Verification with Short Binary Templates
by: Ganmati, Abdelilah, et al.
Published: (2025)
by: Ganmati, Abdelilah, et al.
Published: (2025)
Maternal and Fetal Health Status Assessment by Using Machine Learning on Optical 3D Body Scans
by: Cheng, Ruting, et al.
Published: (2025)
by: Cheng, Ruting, et al.
Published: (2025)
Attenuation-adjusted deep learning of pore defects in 2D radiographs of additive manufacturing powders
by: Bjerregaard, Andreas, et al.
Published: (2024)
by: Bjerregaard, Andreas, et al.
Published: (2024)
Transforming faces into video stories -- VideoFace2.0
by: Brkljač, Branko, et al.
Published: (2025)
by: Brkljač, Branko, et al.
Published: (2025)
Personalised aesthetics with residual adapters
by: Rodríguez-Pardo, Carlos, et al.
Published: (2019)
by: Rodríguez-Pardo, Carlos, et al.
Published: (2019)
STOPA: A Database of Systematic VariaTion Of DeePfake Audio for Open-Set Source Tracing and Attribution
by: Firc, Anton, et al.
Published: (2025)
by: Firc, Anton, et al.
Published: (2025)
Texture Discrimination via Hilbert Curve Path Based Information Quantifiers
by: Bariviera, Aurelio F., et al.
Published: (2024)
by: Bariviera, Aurelio F., et al.
Published: (2024)
Delayed Fusion: Integrating Large Language Models into First-Pass Decoding in End-to-end Speech Recognition
by: Hori, Takaaki, et al.
Published: (2025)
by: Hori, Takaaki, et al.
Published: (2025)
On the Insecurity of Keystroke-Based AI Authorship Detection: Timing-Forgery Attacks Against Motor-Signal Verification
by: Condrey, David
Published: (2026)
by: Condrey, David
Published: (2026)
Detection of high-frequency oscillations using time-frequency analysis
by: Mohammadpour, Mostafa, et al.
Published: (2025)
by: Mohammadpour, Mostafa, et al.
Published: (2025)
Deep Learning-Based Multi-Object Tracking: A Comprehensive Survey from Foundations to State-of-the-Art
by: Adžemović, Momir
Published: (2025)
by: Adžemović, Momir
Published: (2025)
SSTAF: Spatial-Spectral-Temporal Attention Fusion Transformer for Motor Imagery Classification
by: Muna, Ummay Maria, et al.
Published: (2025)
by: Muna, Ummay Maria, et al.
Published: (2025)
From Black Box to Glass Box: Cross-Model ASR Disagreement to Prioto Review in Ambient AI Scribe Documentation
by: Karbalaie, Abdolamir, et al.
Published: (2026)
by: Karbalaie, Abdolamir, et al.
Published: (2026)
Nearest Neighbor Projection Removal Adversarial Training
by: Singh, Himanshu, et al.
Published: (2025)
by: Singh, Himanshu, et al.
Published: (2025)
VQToken: Neural Discrete Token Representation Learning for Extreme Token Reduction in Video Large Language Models
by: Zhang, Haichao, et al.
Published: (2025)
by: Zhang, Haichao, et al.
Published: (2025)
MEG-to-MEG Transfer Learning and Cross-Task Speech/Silence Detection with Limited Data
by: de Zuazo, Xabier, et al.
Published: (2026)
by: de Zuazo, Xabier, et al.
Published: (2026)
Audio-based Kinship Verification Using Age Domain Conversion
by: Sun, Qiyang, et al.
Published: (2024)
by: Sun, Qiyang, et al.
Published: (2024)
Passive Underwater Acoustic Signal Separation based on Feature Decoupling Dual-path Network
by: Liu, Yucheng, et al.
Published: (2025)
by: Liu, Yucheng, et al.
Published: (2025)
Splitformer: An improved early-exit architecture for automatic speech recognition on edge devices
by: Lasbordes, Maxence, et al.
Published: (2025)
by: Lasbordes, Maxence, et al.
Published: (2025)
Associative Syntax and Maximal Repetitions reveal context-dependent complexity in fruit bat communication
by: Assom, Luigi
Published: (2025)
by: Assom, Luigi
Published: (2025)
CS-SHRED: Enhancing SHRED for Robust Recovery of Spatiotemporal Dynamics
by: da Silva, Romulo B., et al.
Published: (2025)
by: da Silva, Romulo B., et al.
Published: (2025)
Data Augmentation and Resolution Enhancement using GANs and Diffusion Models for Tree Segmentation
by: Ferreira, Alessandro dos Santos, et al.
Published: (2025)
by: Ferreira, Alessandro dos Santos, et al.
Published: (2025)
Simultaneous source separation of unknown numbers of single-channel underwater acoustic signals based on deep neural networks with separator-decoder structure
by: Sun, Qinggang, et al.
Published: (2022)
by: Sun, Qinggang, et al.
Published: (2022)
Enhancing Speaker Verification with Whispered Speech via Post-Processing
by: Gołębiowska, Magdalena, et al.
Published: (2026)
by: Gołębiowska, Magdalena, et al.
Published: (2026)
SonicMaster: Towards Controllable All-in-One Music Restoration and Mastering
by: Melechovsky, Jan, et al.
Published: (2025)
by: Melechovsky, Jan, et al.
Published: (2025)
Crossing the Species Divide: Transfer Learning from Speech to Animal Sounds
by: Cauzinille, Jules, et al.
Published: (2025)
by: Cauzinille, Jules, et al.
Published: (2025)
RF-BayesPhysNet: A Bayesian rPPG Uncertainty Estimation Method for Complex Scenarios
by: Ma, Rufei, et al.
Published: (2025)
by: Ma, Rufei, et al.
Published: (2025)
Goal-Oriented Source Coding using LDPC Codes for Compressed-Domain Image Classification
by: Aliouat, Ahcen, et al.
Published: (2025)
by: Aliouat, Ahcen, et al.
Published: (2025)
MVTamperBench: Evaluating Robustness of Vision-Language Models
by: Agarwal, Amit, et al.
Published: (2024)
by: Agarwal, Amit, et al.
Published: (2024)
Corn Ear Detection and Orientation Estimation Using Deep Learning
by: Sprague, Nathan, et al.
Published: (2024)
by: Sprague, Nathan, et al.
Published: (2024)
Similar Items
-
Proficiency-Aware Adaptation and Data Augmentation for Robust L2 ASR
by: Sun, Ling, et al.
Published: (2025) -
Matlab-based Epoch Extraction for Speaker Differentiation
by: Li, Kunlun, et al.
Published: (2024) -
Understanding the Algorithm Behind Audio Key Detection
by: Silva, Henrique Perez G.
Published: (2025) -
HELIX: Scaling Raw Audio Understanding with Hybrid Mamba-Attention Beyond the Quadratic Limit
by: Khushiyant, et al.
Published: (2026) -
Machine learning based animal emotion classification using audio signals
by: Slobodian, Mariia, et al.
Published: (2025)