BUT Systems for Environmental Sound Deepfake Detection in the ESDD 2026 Challenge
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Peng, Junyi, Zhang, Lin, Li, Jin, Plchot, Oldrich, Cernocky, Jan |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
BUT Systems for WildSpoof Challenge: SASV in the Wild
par: Peng, Junyi, et autres
Publié: (2025)
par: Peng, Junyi, et autres
Publié: (2025)
Target Speech Extraction with Pre-trained Self-supervised Learning Models
par: Peng, Junyi, et autres
Publié: (2024)
par: Peng, Junyi, et autres
Publié: (2024)
CA-MHFA: A Context-Aware Multi-Head Factorized Attentive Pooling for SSL-Based Speaker Verification
par: Peng, Junyi, et autres
Publié: (2024)
par: Peng, Junyi, et autres
Publié: (2024)
State-of-the-art Embeddings with Video-free Segmentation of the Source VoxCeleb Data
par: Barahona, Sara, et autres
Publié: (2024)
par: Barahona, Sara, et autres
Publié: (2024)
Hybrid Pruning: In-Situ Compression of Self-Supervised Speech Models for Speaker Verification and Anti-Spoofing
par: Peng, Junyi, et autres
Publié: (2025)
par: Peng, Junyi, et autres
Publié: (2025)
Probing Self-supervised Learning Models with Target Speech Extraction
par: Peng, Junyi, et autres
Publié: (2024)
par: Peng, Junyi, et autres
Publié: (2024)
TS-SUPERB: A Target Speech Processing Benchmark for Speech Self-Supervised Learning Models
par: Peng, Junyi, et autres
Publié: (2025)
par: Peng, Junyi, et autres
Publié: (2025)
BUT Systems and Analyses for the ASVspoof 5 Challenge
par: Rohdin, Johan, et autres
Publié: (2024)
par: Rohdin, Johan, et autres
Publié: (2024)
Technical Report of Nomi Team in the Environmental Sound Deepfake Detection Challenge 2026
par: Mawalim, Candy Olivia, et autres
Publié: (2025)
par: Mawalim, Candy Olivia, et autres
Publié: (2025)
Approaching Dialogue State Tracking via Aligning Speech Encoders and LLMs
par: Sedláček, Šimon, et autres
Publié: (2025)
par: Sedláček, Šimon, et autres
Publié: (2025)
Joint Speech and Text Training for LLM-Based End-to-End Spoken Dialogue State Tracking
par: Vendrame, Katia, et autres
Publié: (2025)
par: Vendrame, Katia, et autres
Publié: (2025)
EnvSSLAM-FFN: Lightweight Layer-Fused System for ESDD 2026 Challenge
par: Guo, Xiaoxuan, et autres
Publié: (2025)
par: Guo, Xiaoxuan, et autres
Publié: (2025)
Challenging margin-based speaker embedding extractors by using the variational information bottleneck
par: Stafylakis, Themos, et autres
Publié: (2024)
par: Stafylakis, Themos, et autres
Publié: (2024)
BUT System for the MLC-SLM Challenge
par: Polok, Alexander, et autres
Publié: (2025)
par: Polok, Alexander, et autres
Publié: (2025)
Audio Deepfake Verification
par: Wang, Li, et autres
Publié: (2025)
par: Wang, Li, et autres
Publié: (2025)
Detection of Deepfake Environmental Audio
par: Ouajdi, Hafsa, et autres
Publié: (2024)
par: Ouajdi, Hafsa, et autres
Publié: (2024)
EnvSDD: Benchmarking Environmental Sound Deepfake Detection
par: Yin, Han, et autres
Publié: (2025)
par: Yin, Han, et autres
Publié: (2025)
RADAR Challenge 2026: Robust Audio Deepfake Recognition under Media Transformations
par: Luong, Hieu-Thi, et autres
Publié: (2026)
par: Luong, Hieu-Thi, et autres
Publié: (2026)
Analysis of ABC Frontend Audio Systems for the NIST-SRE24
par: Barahona, Sara, et autres
Publié: (2025)
par: Barahona, Sara, et autres
Publié: (2025)
Improving Automatic Speech Recognition with Decoder-Centric Regularisation in Encoder-Decoder Models
par: Polok, Alexander, et autres
Publié: (2024)
par: Polok, Alexander, et autres
Publié: (2024)
Domain-Agnostic Incremental Learning for Sound Classification. A DCASE 2026 Challenge task
par: Casciotti, Riccardo, et autres
Publié: (2026)
par: Casciotti, Riccardo, et autres
Publié: (2026)
ASASVIcomtech: The Vicomtech-UGR Speech Deepfake Detection and SASV Systems for the ASVspoof5 Challenge
par: Martín-Doñas, Juan M., et autres
Publié: (2024)
par: Martín-Doñas, Juan M., et autres
Publié: (2024)
Streaming Endpointer for Spoken Dialogue using Neural Audio Codecs and Label-Delayed Training
par: Udupa, Sathvik, et autres
Publié: (2025)
par: Udupa, Sathvik, et autres
Publié: (2025)
Description and Discussion on DCASE 2026 Challenge Task 4: Spatial Semantic Segmentation of Sound Scenes
par: Nguyen, Binh Thien, et autres
Publié: (2026)
par: Nguyen, Binh Thien, et autres
Publié: (2026)
Audio Deepfake Detection at the First Greeting: "Hi!"
par: Shi, Haohan, et autres
Publié: (2026)
par: Shi, Haohan, et autres
Publié: (2026)
DeCRED: Decoder-Centric Regularization for Encoder-Decoder Based Speech Recognition
par: Polok, Alexander, et autres
Publié: (2025)
par: Polok, Alexander, et autres
Publié: (2025)
Mind the Gap: Impact of Synthetic Conversational Data on Multi-Talker ASR and Speaker Diarization
par: Polok, Alexander, et autres
Publié: (2026)
par: Polok, Alexander, et autres
Publié: (2026)
Pretraining End-to-End Keyword Search with Automatically Discovered Acoustic Units
par: Yusuf, Bolaji, et autres
Publié: (2024)
par: Yusuf, Bolaji, et autres
Publié: (2024)
Post-training for Deepfake Speech Detection
par: Ge, Wanying, et autres
Publié: (2025)
par: Ge, Wanying, et autres
Publié: (2025)
SVDD 2024: The Inaugural Singing Voice Deepfake Detection Challenge
par: Zhang, You, et autres
Publié: (2024)
par: Zhang, You, et autres
Publié: (2024)
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
par: Nishida, Tomoya, et autres
Publié: (2026)
par: Nishida, Tomoya, et autres
Publié: (2026)
Fine-tune Before Structured Pruning: Towards Compact and Accurate Self-Supervised Models for Speaker Diarization
par: Han, Jiangyu, et autres
Publié: (2025)
par: Han, Jiangyu, et autres
Publié: (2025)
Efficient and Generalizable Speaker Diarization via Structured Pruning of Self-Supervised Models
par: Han, Jiangyu, et autres
Publié: (2025)
par: Han, Jiangyu, et autres
Publié: (2025)
HCFD: A Benchmark for Audio Deepfake Detection in Healthcare
par: Akhtar, Mohd Mujtaba, et autres
Publié: (2026)
par: Akhtar, Mohd Mujtaba, et autres
Publié: (2026)
Multilingual Dataset Integration Strategies for Robust Audio Deepfake Detection: A SAFE Challenge System
par: Ali, Hashim, et autres
Publié: (2025)
par: Ali, Hashim, et autres
Publié: (2025)
Unmasking Deepfakes: Leveraging Augmentations and Features Variability for Deepfake Speech Detection
par: Rimon, Inbal, et autres
Publié: (2025)
par: Rimon, Inbal, et autres
Publié: (2025)
Generalizable Detection of Audio Deepfakes
par: Lopez, Jose A., et autres
Publié: (2025)
par: Lopez, Jose A., et autres
Publié: (2025)
Beyond the Labels: Unveiling Text-Dependency in Paralinguistic Speech Recognition Datasets
par: Pešán, Jan, et autres
Publié: (2024)
par: Pešán, Jan, et autres
Publié: (2024)
Language-Invariant Multilingual Speaker Verification for the TidyVoice 2026 Challenge
par: Li, Ze, et autres
Publié: (2026)
par: Li, Ze, et autres
Publié: (2026)
A Unified Framework for Modality-Agnostic Deepfakes Detection
par: Yu, Cai, et autres
Publié: (2023)
par: Yu, Cai, et autres
Publié: (2023)
Documents similaires
-
BUT Systems for WildSpoof Challenge: SASV in the Wild
par: Peng, Junyi, et autres
Publié: (2025) -
Target Speech Extraction with Pre-trained Self-supervised Learning Models
par: Peng, Junyi, et autres
Publié: (2024) -
CA-MHFA: A Context-Aware Multi-Head Factorized Attentive Pooling for SSL-Based Speaker Verification
par: Peng, Junyi, et autres
Publié: (2024) -
State-of-the-art Embeddings with Video-free Segmentation of the Source VoxCeleb Data
par: Barahona, Sara, et autres
Publié: (2024) -
Hybrid Pruning: In-Situ Compression of Self-Supervised Speech Models for Speaker Verification and Anti-Spoofing
par: Peng, Junyi, et autres
Publié: (2025)