UR Channel-Robust Synthetic Speech Detection System for ASVspoof 2021
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Xinhui, Zhang, You, Zhu, Ge, Duan, Zhiyao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2021
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An Empirical Study on Channel Effects for Synthetic Voice Spoofing Countermeasure Systems
di: Zhang, You, et al.
Pubblicazione: (2021)
di: Zhang, You, et al.
Pubblicazione: (2021)
A Probabilistic Fusion Framework for Spoofing Aware Speaker Verification
di: Zhang, You, et al.
Pubblicazione: (2022)
di: Zhang, You, et al.
Pubblicazione: (2022)
ASASVIcomtech: The Vicomtech-UGR Speech Deepfake Detection and SASV Systems for the ASVspoof5 Challenge
di: Martín-Doñas, Juan M., et al.
Pubblicazione: (2024)
di: Martín-Doñas, Juan M., et al.
Pubblicazione: (2024)
PartialEdit: Identifying Partial Deepfakes in the Era of Neural Speech Editing
di: Zhang, You, et al.
Pubblicazione: (2025)
di: Zhang, You, et al.
Pubblicazione: (2025)
BUT Systems and Analyses for the ASVspoof 5 Challenge
di: Rohdin, Johan, et al.
Pubblicazione: (2024)
di: Rohdin, Johan, et al.
Pubblicazione: (2024)
XMUspeech Systems for the ASVspoof 5 Challenge
di: Li, Wangjie, et al.
Pubblicazione: (2025)
di: Li, Wangjie, et al.
Pubblicazione: (2025)
Cacophony: An Improved Contrastive Audio-Text Model
di: Zhu, Ge, et al.
Pubblicazione: (2024)
di: Zhu, Ge, et al.
Pubblicazione: (2024)
Audio Generation Through Score-Based Generative Modeling: Design Principles and Implementation
di: Zhu, Ge, et al.
Pubblicazione: (2025)
di: Zhu, Ge, et al.
Pubblicazione: (2025)
Spoofing-Robust Speaker Verification Using Parallel Embedding Fusion: BTU Speech Group's Approach for ASVspoof5 Challenge
di: Kurnaz, Oğuzhan, et al.
Pubblicazione: (2024)
di: Kurnaz, Oğuzhan, et al.
Pubblicazione: (2024)
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
di: Tian, Jingguang, et al.
Pubblicazione: (2024)
di: Tian, Jingguang, et al.
Pubblicazione: (2024)
Predicting Global HRTFs From Scanned Head Geometry Using Deep Learning and Compact Representations
di: Wang, Yuxiang, et al.
Pubblicazione: (2022)
di: Wang, Yuxiang, et al.
Pubblicazione: (2022)
ControlVC: Zero-Shot Voice Conversion with Time-Varying Controls on Pitch and Speed
di: Chen, Meiying, et al.
Pubblicazione: (2022)
di: Chen, Meiying, et al.
Pubblicazione: (2022)
Dual-Branch Knowledge Distillation for Noise-Robust Synthetic Speech Detection
di: Fan, Cunhang, et al.
Pubblicazione: (2023)
di: Fan, Cunhang, et al.
Pubblicazione: (2023)
SingFake: Singing Voice Deepfake Detection
di: Zang, Yongyi, et al.
Pubblicazione: (2023)
di: Zang, Yongyi, et al.
Pubblicazione: (2023)
ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale
di: Wang, Xin, et al.
Pubblicazione: (2024)
di: Wang, Xin, et al.
Pubblicazione: (2024)
USTC-KXDIGIT System Description for ASVspoof5 Challenge
di: Chen, Yihao, et al.
Pubblicazione: (2024)
di: Chen, Yihao, et al.
Pubblicazione: (2024)
Channel-Combination Algorithms for Robust Distant Voice Activity and Overlapped Speech Detection
di: Mariotte, Théo, et al.
Pubblicazione: (2024)
di: Mariotte, Théo, et al.
Pubblicazione: (2024)
A Multi-Stream Fusion Approach with One-Class Learning for Audio-Visual Deepfake Detection
di: Lee, Kyungbok, et al.
Pubblicazione: (2024)
di: Lee, Kyungbok, et al.
Pubblicazione: (2024)
SZU-AFS Antispoofing System for the ASVspoof 5 Challenge
di: Xu, Yuxiong, et al.
Pubblicazione: (2024)
di: Xu, Yuxiong, et al.
Pubblicazione: (2024)
SVDD 2024: The Inaugural Singing Voice Deepfake Detection Challenge
di: Zhang, You, et al.
Pubblicazione: (2024)
di: Zhang, You, et al.
Pubblicazione: (2024)
Towards Perception-Informed Latent HRTF Representations
di: Zhang, You, et al.
Pubblicazione: (2025)
di: Zhang, You, et al.
Pubblicazione: (2025)
Generating Novel and Realistic Speakers for Voice Conversion
di: Chen, Meiying Melissa, et al.
Pubblicazione: (2025)
di: Chen, Meiying Melissa, et al.
Pubblicazione: (2025)
MusicHiFi: Fast High-Fidelity Stereo Vocoding
di: Zhu, Ge, et al.
Pubblicazione: (2024)
di: Zhu, Ge, et al.
Pubblicazione: (2024)
Generating Data with Text-to-Speech and Large-Language Models for Conversational Speech Recognition
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
di: Cornell, Samuele, et al.
Pubblicazione: (2024)
AASIST3: KAN-Enhanced AASIST Speech Deepfake Detection using SSL Features and Additional Regularization for the ASVspoof 2024 Challenge
di: Borodin, Kirill, et al.
Pubblicazione: (2024)
di: Borodin, Kirill, et al.
Pubblicazione: (2024)
SynHate: Detecting Hate Speech in Synthetic Deepfake Audio
di: Ranjan, Rishabh, et al.
Pubblicazione: (2025)
di: Ranjan, Rishabh, et al.
Pubblicazione: (2025)
A Domain Adaptation Framework for Speech Recognition Systems with Only Synthetic data
di: Tran, Minh, et al.
Pubblicazione: (2025)
di: Tran, Minh, et al.
Pubblicazione: (2025)
Attention-Based Beamformer For Multi-Channel Speech Enhancement
di: Bai, Jinglin, et al.
Pubblicazione: (2024)
di: Bai, Jinglin, et al.
Pubblicazione: (2024)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
di: Bondaruk, Łukasz, et al.
Pubblicazione: (2024)
di: Bondaruk, Łukasz, et al.
Pubblicazione: (2024)
Multi-Channel Differential ASR for Robust Wearer Speech Recognition on Smart Glasses
di: Yang, Yufeng, et al.
Pubblicazione: (2025)
di: Yang, Yufeng, et al.
Pubblicazione: (2025)
Speaker Anonymisation for Speech-based Suicide Risk Detection
di: Cui, Ziyun, et al.
Pubblicazione: (2025)
di: Cui, Ziyun, et al.
Pubblicazione: (2025)
Temporal-Channel Modeling in Multi-head Self-Attention for Synthetic Speech Detection
di: Truong, Duc-Tuan, et al.
Pubblicazione: (2024)
di: Truong, Duc-Tuan, et al.
Pubblicazione: (2024)
ShiftySpeech: A Large-Scale Synthetic Speech Dataset with Distribution Shifts
di: Garg, Ashi, et al.
Pubblicazione: (2025)
di: Garg, Ashi, et al.
Pubblicazione: (2025)
Source Tracing of Synthetic Speech Systems Through Paralinguistic Pre-Trained Representations
di: Girish, et al.
Pubblicazione: (2025)
di: Girish, et al.
Pubblicazione: (2025)
A Semantic Information-based Hierarchical Speech Enhancement Method Using Factorized Codec and Diffusion Model
di: Xiang, Yang, et al.
Pubblicazione: (2025)
di: Xiang, Yang, et al.
Pubblicazione: (2025)
Confidence-Based Self-Training for EMG-to-Speech: Leveraging Synthetic EMG for Robust Modeling
di: Chen, Xiaodan, et al.
Pubblicazione: (2025)
di: Chen, Xiaodan, et al.
Pubblicazione: (2025)
Robust Bioacoustic Detection via Richly Labelled Synthetic Soundscape Augmentation
di: Soltero, Kaspar, et al.
Pubblicazione: (2025)
di: Soltero, Kaspar, et al.
Pubblicazione: (2025)
The 1st SpeechWellness Challenge: Detecting Suicide Risk Among Adolescents
di: Wu, Wen, et al.
Pubblicazione: (2025)
di: Wu, Wen, et al.
Pubblicazione: (2025)
GMM-ResNet2: Ensemble of Group ResNet Networks for Synthetic Speech Detection
di: Lei, Zhenchun, et al.
Pubblicazione: (2024)
di: Lei, Zhenchun, et al.
Pubblicazione: (2024)
Bridging the Gap: Integrating Pre-trained Speech Enhancement and Recognition Models for Robust Speech Recognition
di: Wang, Kuan-Chen, et al.
Pubblicazione: (2024)
di: Wang, Kuan-Chen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
An Empirical Study on Channel Effects for Synthetic Voice Spoofing Countermeasure Systems
di: Zhang, You, et al.
Pubblicazione: (2021) -
A Probabilistic Fusion Framework for Spoofing Aware Speaker Verification
di: Zhang, You, et al.
Pubblicazione: (2022) -
ASASVIcomtech: The Vicomtech-UGR Speech Deepfake Detection and SASV Systems for the ASVspoof5 Challenge
di: Martín-Doñas, Juan M., et al.
Pubblicazione: (2024) -
PartialEdit: Identifying Partial Deepfakes in the Era of Neural Speech Editing
di: Zhang, You, et al.
Pubblicazione: (2025) -
BUT Systems and Analyses for the ASVspoof 5 Challenge
di: Rohdin, Johan, et al.
Pubblicazione: (2024)