mmWave Radar Aware Dual-Conditioned GAN for Speech Reconstruction of Signals With Low SNR
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Karani, Jash, Chittem, Adithya, Roy, Deepan, Joshi, Sandeep |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GAN-Based Speech Enhancement for Low SNR Using Latent Feature Conditioning
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
Sound Signal Synthesis with Auxiliary Classifier GAN, COVID-19 cough as an example
von: Saleh, Yahya Sherif Solayman Mohamed, et al.
Veröffentlicht: (2025)
von: Saleh, Yahya Sherif Solayman Mohamed, et al.
Veröffentlicht: (2025)
We Can Hear You with mmWave Radar! An End-to-End Eavesdropping System
von: Han, Dachao, et al.
Veröffentlicht: (2025)
von: Han, Dachao, et al.
Veröffentlicht: (2025)
mmWave-Whisper: Phone Call Eavesdropping and Transcription Using Millimeter-Wave Radar
von: Basak, Suryoday, et al.
Veröffentlicht: (2024)
von: Basak, Suryoday, et al.
Veröffentlicht: (2024)
WaveSSM: Multiscale State-Space Models for Non-stationary Signal Attention
von: Solozabal, Ruben, et al.
Veröffentlicht: (2026)
von: Solozabal, Ruben, et al.
Veröffentlicht: (2026)
SpecDiff-GAN: A Spectrally-Shaped Noise Diffusion GAN for Speech and Music Synthesis
von: Baoueb, Teysir, et al.
Veröffentlicht: (2024)
von: Baoueb, Teysir, et al.
Veröffentlicht: (2024)
CMGAN: Conformer-based Metric GAN for Speech Enhancement
von: Cao, Ruizhe, et al.
Veröffentlicht: (2022)
von: Cao, Ruizhe, et al.
Veröffentlicht: (2022)
CMGAN: Conformer-Based Metric-GAN for Monaural Speech Enhancement
von: Abdulatif, Sherif, et al.
Veröffentlicht: (2022)
von: Abdulatif, Sherif, et al.
Veröffentlicht: (2022)
Brain-to-Speech: Prosody Feature Engineering and Transformer-Based Reconstruction
von: Al-Radhi, Mohammed Salah, et al.
Veröffentlicht: (2026)
von: Al-Radhi, Mohammed Salah, et al.
Veröffentlicht: (2026)
Flowing Straighter with Conditional Flow Matching for Accurate Speech Enhancement
von: Cross, Mattias, et al.
Veröffentlicht: (2025)
von: Cross, Mattias, et al.
Veröffentlicht: (2025)
SNR-Progressive Model with Harmonic Compensation for Low-SNR Speech Enhancement
von: Hou, Zhongshu, et al.
Veröffentlicht: (2024)
von: Hou, Zhongshu, et al.
Veröffentlicht: (2024)
Investigating the Effects of Diffusion-based Conditional Generative Speech Models Used for Speech Enhancement on Dysarthric Speech
von: Reszka, Joanna, et al.
Veröffentlicht: (2024)
von: Reszka, Joanna, et al.
Veröffentlicht: (2024)
U-Codec: Ultra Low Frame-rate Neural Speech Codec for Fast High-fidelity Speech Generation
von: Yang, Xusheng, et al.
Veröffentlicht: (2025)
von: Yang, Xusheng, et al.
Veröffentlicht: (2025)
CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing
von: Lu, Yen-Ju, et al.
Veröffentlicht: (2024)
von: Lu, Yen-Ju, et al.
Veröffentlicht: (2024)
Enhancing Speech Emotion Recognition with Graph-Based Multimodal Fusion and Prosodic Features for the Speech Emotion Recognition in Naturalistic Conditions Challenge at Interspeech 2025
von: Ferreira, Alef Iury Siqueira, et al.
Veröffentlicht: (2025)
von: Ferreira, Alef Iury Siqueira, et al.
Veröffentlicht: (2025)
Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder
von: Melechovsky, Jan, et al.
Veröffentlicht: (2022)
von: Melechovsky, Jan, et al.
Veröffentlicht: (2022)
Masked Autoencoders as Universal Speech Enhancer
von: Rajagopalan, Rajalaxmi, et al.
Veröffentlicht: (2026)
von: Rajagopalan, Rajalaxmi, et al.
Veröffentlicht: (2026)
Phase-Aware Deep Learning with Complex-Valued CNNs for Audio Signal Applications
von: Agrawal, Naman
Veröffentlicht: (2025)
von: Agrawal, Naman
Veröffentlicht: (2025)
Bone-conduction Guided Multimodal Speech Enhancement with Conditional Diffusion Models
von: Khanagha, Sina, et al.
Veröffentlicht: (2026)
von: Khanagha, Sina, et al.
Veröffentlicht: (2026)
Koopman Regularized Deep Speech Disentanglement for Speaker Verification
von: Chazaridis, Nikos, et al.
Veröffentlicht: (2026)
von: Chazaridis, Nikos, et al.
Veröffentlicht: (2026)
Assessing the Impact of Speaker Identity in Speech Spoofing Detection
von: Dao, Anh-Tuan, et al.
Veröffentlicht: (2026)
von: Dao, Anh-Tuan, et al.
Veröffentlicht: (2026)
Speech Emotion Recognition with Phonation Excitation Information and Articulatory Kinematics
von: Zhang, Ziqian, et al.
Veröffentlicht: (2025)
von: Zhang, Ziqian, et al.
Veröffentlicht: (2025)
A Semi-Supervised Framework for Speech Confidence Detection using Whisper
von: Wynn, Adam, et al.
Veröffentlicht: (2026)
von: Wynn, Adam, et al.
Veröffentlicht: (2026)
Investigating the Impact of Speech Enhancement on Audio Deepfake Detection in Noisy Environments
von: Anacin, et al.
Veröffentlicht: (2026)
von: Anacin, et al.
Veröffentlicht: (2026)
Optimizing Neural Architectures for Hindi Speech Separation and Enhancement in Noisy Environments
von: Ramamoorthy, Arnav
Veröffentlicht: (2025)
von: Ramamoorthy, Arnav
Veröffentlicht: (2025)
Improving Speech Emotion Recognition with Mutual Information Regularized Generative Model
von: Ahn, Chung-Soo, et al.
Veröffentlicht: (2025)
von: Ahn, Chung-Soo, et al.
Veröffentlicht: (2025)
Task Vector in TTS: Toward Emotionally Expressive Dialectal Speech Synthesis
von: Feng, Pengchao, et al.
Veröffentlicht: (2025)
von: Feng, Pengchao, et al.
Veröffentlicht: (2025)
RO-N3WS: Enhancing Generalization in Low-Resource ASR with Diverse Romanian Speech Benchmarks
von: Diaconu, Alexandra, et al.
Veröffentlicht: (2026)
von: Diaconu, Alexandra, et al.
Veröffentlicht: (2026)
Objective Evaluation of Prosody and Intelligibility in Speech Synthesis via Conditional Prediction of Discrete Tokens
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2025)
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2025)
FastWave: Optimized Diffusion Model for Audio Super-Resolution
von: Kuznetsov, Nikita, et al.
Veröffentlicht: (2026)
von: Kuznetsov, Nikita, et al.
Veröffentlicht: (2026)
PROCESS-2: A Benchmark Speech Corpus for Early Cognitive Impairment Detection
von: Pahar, Madhurananda, et al.
Veröffentlicht: (2026)
von: Pahar, Madhurananda, et al.
Veröffentlicht: (2026)
EmoHRNet: High-Resolution Neural Network Based Speech Emotion Recognition
von: Muppidi, Akshay, et al.
Veröffentlicht: (2025)
von: Muppidi, Akshay, et al.
Veröffentlicht: (2025)
EmoAugNet: A Signal-Augmented Hybrid CNN-LSTM Framework for Speech Emotion Recognition
von: Paul, Durjoy Chandra, et al.
Veröffentlicht: (2025)
von: Paul, Durjoy Chandra, et al.
Veröffentlicht: (2025)
Detecting Throat Cancer from Speech Signals using Machine Learning: A Scoping Literature Review
von: Paterson, Mary, et al.
Veröffentlicht: (2023)
von: Paterson, Mary, et al.
Veröffentlicht: (2023)
A Novel Fusion Architecture for PD Detection Using Semi-Supervised Speech Embeddings
von: Adnan, Tariq, et al.
Veröffentlicht: (2024)
von: Adnan, Tariq, et al.
Veröffentlicht: (2024)
Edge-ASR: Towards Low-Bit Quantization of Automatic Speech Recognition Models
von: Feng, Chen, et al.
Veröffentlicht: (2025)
von: Feng, Chen, et al.
Veröffentlicht: (2025)
EM-TTS: Efficiently Trained Low-Resource Mongolian Lightweight Text-to-Speech
von: Liang, Ziqi, et al.
Veröffentlicht: (2024)
von: Liang, Ziqi, et al.
Veröffentlicht: (2024)
Diffusion-Based Speech Enhancement in Matched and Mismatched Conditions Using a Heun-Based Sampler
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2023)
von: Gonzalez, Philippe, et al.
Veröffentlicht: (2023)
AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs
von: Nguyen, Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Hoang, et al.
Veröffentlicht: (2025)
voice2mode: Phonation Mode Classification in Singing using Self-Supervised Speech Models
von: Justus, Aju Ani, et al.
Veröffentlicht: (2026)
von: Justus, Aju Ani, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
GAN-Based Speech Enhancement for Low SNR Using Latent Feature Conditioning
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024) -
Sound Signal Synthesis with Auxiliary Classifier GAN, COVID-19 cough as an example
von: Saleh, Yahya Sherif Solayman Mohamed, et al.
Veröffentlicht: (2025) -
We Can Hear You with mmWave Radar! An End-to-End Eavesdropping System
von: Han, Dachao, et al.
Veröffentlicht: (2025) -
mmWave-Whisper: Phone Call Eavesdropping and Transcription Using Millimeter-Wave Radar
von: Basak, Suryoday, et al.
Veröffentlicht: (2024) -
WaveSSM: Multiscale State-Space Models for Non-stationary Signal Attention
von: Solozabal, Ruben, et al.
Veröffentlicht: (2026)