Enhancing Anti-spoofing Countermeasures Robustness through Joint Optimization and Transfer Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Yikang, Wang, Xingming, Nishizaki, Hiromitsu, Li, Ming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CompSpoof: A Dataset and Joint Learning Framework for Component-Level Audio Anti-spoofing Countermeasures
por: Zhang, Xueping, et al.
Publicado: (2025)
por: Zhang, Xueping, et al.
Publicado: (2025)
Emo-DPO: Controllable Emotional Speech Synthesis through Direct Preference Optimization
por: Gao, Xiaoxue, et al.
Publicado: (2024)
por: Gao, Xiaoxue, et al.
Publicado: (2024)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
por: Wang, Kuan-Chen, et al.
Publicado: (2024)
por: Wang, Kuan-Chen, et al.
Publicado: (2024)
Differentiable Acoustic Radiance Transfer
por: Lee, Sungho, et al.
Publicado: (2025)
por: Lee, Sungho, et al.
Publicado: (2025)
Amplifying Artifacts with Speech Enhancement in Voice Anti-spoofing
por: Trachu, Thanapat, et al.
Publicado: (2025)
por: Trachu, Thanapat, et al.
Publicado: (2025)
Joint Fullband-Subband Modeling for High-Resolution SingFake Detection
por: Chen, Xuanjun, et al.
Publicado: (2026)
por: Chen, Xuanjun, et al.
Publicado: (2026)
Adaptive Per-Channel Energy Normalization Front-end for Robust Audio Signal Processing
por: Meng, Hanyu, et al.
Publicado: (2025)
por: Meng, Hanyu, et al.
Publicado: (2025)
Generative Deep Learning and Signal Processing for Data Augmentation of Cardiac Auscultation Signals: Improving Model Robustness Using Synthetic Audio
por: Abbott, Leigh, et al.
Publicado: (2024)
por: Abbott, Leigh, et al.
Publicado: (2024)
Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
por: Yu, Chin-Yun, et al.
Publicado: (2024)
por: Yu, Chin-Yun, et al.
Publicado: (2024)
Completing Sets of Prototype Transfer Functions for Subspace-based Direction of Arrival Estimation of Multiple Speakers
por: Fejgin, Daniel, et al.
Publicado: (2025)
por: Fejgin, Daniel, et al.
Publicado: (2025)
Lessons Learned from the URGENT 2024 Speech Enhancement Challenge
por: Zhang, Wangyou, et al.
Publicado: (2025)
por: Zhang, Wangyou, et al.
Publicado: (2025)
Adaptive Diagonal Loading using Krylov Subspaces for Robust Beamforming
por: Mittal, Manan, et al.
Publicado: (2026)
por: Mittal, Manan, et al.
Publicado: (2026)
Continuous Speech Tokens Makes LLMs Robust Multi-Modality Learners
por: Yuan, Ze, et al.
Publicado: (2024)
por: Yuan, Ze, et al.
Publicado: (2024)
RawTFNet: A Lightweight CNN Architecture for Speech Anti-spoofing
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
Align-ULCNet: Towards Low-Complexity and Robust Acoustic Echo and Noise Reduction
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
por: Shetu, Shrishti Saha, et al.
Publicado: (2024)
Confidence-Based Self-Training for EMG-to-Speech: Leveraging Synthetic EMG for Robust Modeling
por: Chen, Xiaodan, et al.
Publicado: (2025)
por: Chen, Xiaodan, et al.
Publicado: (2025)
A Robust Method for Pitch Tracking in the Frequency Following Response using Harmonic Amplitude Summation Filterbank
por: Sadeghkhani, Sajad, et al.
Publicado: (2025)
por: Sadeghkhani, Sajad, et al.
Publicado: (2025)
Aliasing-Free Neural Audio Synthesis
por: Gu, Yicheng, et al.
Publicado: (2025)
por: Gu, Yicheng, et al.
Publicado: (2025)
A Machine Hearing System for Robust Cough Detection Based on a High-Level Representation of Band-Specific Audio Features
por: Monge-Alvarez, Jesús, et al.
Publicado: (2024)
por: Monge-Alvarez, Jesús, et al.
Publicado: (2024)
SoundSpring: Loss-Resilient Audio Transceiver with Dual-Functional Masked Language Modeling
por: Yao, Shengshi, et al.
Publicado: (2025)
por: Yao, Shengshi, et al.
Publicado: (2025)
Directional Selective Fixed-Filter Active Noise Control Based on a Convolutional Neural Network in Reverberant Environments
por: Wang, Boxiang, et al.
Publicado: (2026)
por: Wang, Boxiang, et al.
Publicado: (2026)
Cross-Talk Reduction
por: Wang, Zhong-Qiu, et al.
Publicado: (2024)
por: Wang, Zhong-Qiu, et al.
Publicado: (2024)
U-SAM: An audio language Model for Unified Speech, Audio, and Music Understanding
por: Wang, Ziqian, et al.
Publicado: (2025)
por: Wang, Ziqian, et al.
Publicado: (2025)
Learning Perceptually Relevant Temporal Envelope Morphing
por: Dixit, Satvik, et al.
Publicado: (2025)
por: Dixit, Satvik, et al.
Publicado: (2025)
Microphone Array Signal Processing and Deep Learning for Speech Enhancement
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
por: Haeb-Umbach, Reinhold, et al.
Publicado: (2025)
Machine Learning in Acoustics: A Review and Open-Source Repository
por: McCarthy, Ryan A., et al.
Publicado: (2025)
por: McCarthy, Ryan A., et al.
Publicado: (2025)
BR-ASR: Efficient and Scalable Bias Retrieval Framework for Contextual Biasing ASR in Speech LLM
por: Gong, Xun, et al.
Publicado: (2025)
por: Gong, Xun, et al.
Publicado: (2025)
AADNet: An End-to-End Deep Learning Model for Auditory Attention Decoding
por: Nguyen, Nhan Duc Thanh, et al.
Publicado: (2024)
por: Nguyen, Nhan Duc Thanh, et al.
Publicado: (2024)
LocaGen: Sub-Sample Time-Delay Learning for Beam Localization
por: Kunwar, Ishaan, et al.
Publicado: (2025)
por: Kunwar, Ishaan, et al.
Publicado: (2025)
Toward Universal Speech Enhancement for Diverse Input Conditions
por: Zhang, Wangyou, et al.
Publicado: (2023)
por: Zhang, Wangyou, et al.
Publicado: (2023)
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
por: Shi, Runwu, et al.
Publicado: (2024)
por: Shi, Runwu, et al.
Publicado: (2024)
Blind Source Separation of Radar Signals in Time Domain Using Deep Learning
por: Hinderer, Sven
Publicado: (2025)
por: Hinderer, Sven
Publicado: (2025)
Optimal Scalogram for Computational Complexity Reduction in Acoustic Recognition Using Deep Learning
por: Phan, Dang Thoai, et al.
Publicado: (2025)
por: Phan, Dang Thoai, et al.
Publicado: (2025)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
por: Qi, Tianhua, et al.
Publicado: (2024)
por: Qi, Tianhua, et al.
Publicado: (2024)
Detecting Post-Stroke Aphasia Via Brain Responses to Speech in a Deep Learning Framework
por: De Clercq, Pieter, et al.
Publicado: (2024)
por: De Clercq, Pieter, et al.
Publicado: (2024)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
por: Hou, Yuanbo, et al.
Publicado: (2024)
por: Hou, Yuanbo, et al.
Publicado: (2024)
30+ Years of Source Separation Research: Achievements and Future Challenges
por: Araki, Shoko, et al.
Publicado: (2025)
por: Araki, Shoko, et al.
Publicado: (2025)
PromptEVC: Controllable Emotional Voice Conversion with Natural Language Prompts
por: Qi, Tianhua, et al.
Publicado: (2025)
por: Qi, Tianhua, et al.
Publicado: (2025)
A Study on Speech Assessment with Visual Cues
por: Ahmed, Shafique, et al.
Publicado: (2025)
por: Ahmed, Shafique, et al.
Publicado: (2025)
Ultrasensitive Textile Strain Sensors Redefine Wearable Silent Speech Interfaces with High Machine Learning Efficiency
por: Tang, Chenyu, et al.
Publicado: (2023)
por: Tang, Chenyu, et al.
Publicado: (2023)
Ejemplares similares
-
CompSpoof: A Dataset and Joint Learning Framework for Component-Level Audio Anti-spoofing Countermeasures
por: Zhang, Xueping, et al.
Publicado: (2025) -
Emo-DPO: Controllable Emotional Speech Synthesis through Direct Preference Optimization
por: Gao, Xiaoxue, et al.
Publicado: (2024) -
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
por: Wang, Kuan-Chen, et al.
Publicado: (2024) -
Differentiable Acoustic Radiance Transfer
por: Lee, Sungho, et al.
Publicado: (2025) -
Amplifying Artifacts with Speech Enhancement in Voice Anti-spoofing
por: Trachu, Thanapat, et al.
Publicado: (2025)