Analysis of ABC Frontend Audio Systems for the NIST-SRE24
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Barahona, Sara, Silnova, Anna, Mošner, Ladislav, Peng, Junyi, Plchot, Oldřich, Rohdin, Johan, Zhang, Lin, Han, Jiangyu, Palka, Petr, Landini, Federico, Burget, Lukáš, Stafylakis, Themos, Cumani, Sandro, Boboš, Dominik, Hlavaček, Miroslav, Kodovsky, Martin, Pavlíček, Tomáš |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Challenging margin-based speaker embedding extractors by using the variational information bottleneck
von: Stafylakis, Themos, et al.
Veröffentlicht: (2024)
von: Stafylakis, Themos, et al.
Veröffentlicht: (2024)
State-of-the-art Embeddings with Video-free Segmentation of the Source VoxCeleb Data
von: Barahona, Sara, et al.
Veröffentlicht: (2024)
von: Barahona, Sara, et al.
Veröffentlicht: (2024)
CA-MHFA: A Context-Aware Multi-Head Factorized Attentive Pooling for SSL-Based Speaker Verification
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
Hybrid Pruning: In-Situ Compression of Self-Supervised Speech Models for Speaker Verification and Anti-Spoofing
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
BUT Systems and Analyses for the ASVspoof 5 Challenge
von: Rohdin, Johan, et al.
Veröffentlicht: (2024)
von: Rohdin, Johan, et al.
Veröffentlicht: (2024)
BUT Systems for WildSpoof Challenge: SASV in the Wild
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
Do End-to-End Neural Diarization Attractors Need to Encode Speaker Characteristic Information?
von: Zhang, Lin, et al.
Veröffentlicht: (2024)
von: Zhang, Lin, et al.
Veröffentlicht: (2024)
Alpha Divergence Losses for Biometric Verification
von: Koutsianos, Dimitrios, et al.
Veröffentlicht: (2025)
von: Koutsianos, Dimitrios, et al.
Veröffentlicht: (2025)
Leveraging Self-Supervised Learning for Speaker Diarization
von: Han, Jiangyu, et al.
Veröffentlicht: (2024)
von: Han, Jiangyu, et al.
Veröffentlicht: (2024)
Efficient and Generalizable Speaker Diarization via Structured Pruning of Self-Supervised Models
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
DiaPer: End-to-End Neural Diarization with Perceiver-Based Attractors
von: Landini, Federico, et al.
Veröffentlicht: (2023)
von: Landini, Federico, et al.
Veröffentlicht: (2023)
Fine-tune Before Structured Pruning: Towards Compact and Accurate Self-Supervised Models for Speaker Diarization
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
Joint Training of Speaker Embedding Extractor, Speech and Overlap Detection for Diarization
von: Pálka, Petr, et al.
Veröffentlicht: (2024)
von: Pálka, Petr, et al.
Veröffentlicht: (2024)
VBx for End-to-End Neural and Clustering-based Diarization
von: Pálka, Petr, et al.
Veröffentlicht: (2025)
von: Pálka, Petr, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Greek Medical Dictation
von: Georgilas, Vardis, et al.
Veröffentlicht: (2025)
von: Georgilas, Vardis, et al.
Veröffentlicht: (2025)
Comparing Data Augmentation Methods for End-to-End Task-Oriented Dialog Systems
von: Vlachos, Christos, et al.
Veröffentlicht: (2024)
von: Vlachos, Christos, et al.
Veröffentlicht: (2024)
BUT Systems for Environmental Sound Deepfake Detection in the ESDD 2026 Challenge
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
Spatially Aware Self-Supervised Models for Multi-Channel Neural Speaker Diarization
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
von: Han, Jiangyu, et al.
Veröffentlicht: (2025)
FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings
von: Kesiraju, Santosh, et al.
Veröffentlicht: (2026)
von: Kesiraju, Santosh, et al.
Veröffentlicht: (2026)
CL-UZH submission to the NIST SRE 2024 Speaker Recognition Evaluation
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2025)
Synthetic Speech Source Tracing using Metric Learning
von: Koutsianos, Dimitrios, et al.
Veröffentlicht: (2025)
von: Koutsianos, Dimitrios, et al.
Veröffentlicht: (2025)
Target Speech Extraction with Pre-trained Self-supervised Learning Models
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
LabelBuddy: An Open Source Music and Audio Language Annotation Tagging Tool Using AI Assistance
von: Prokopiou, Ioannis, et al.
Veröffentlicht: (2026)
von: Prokopiou, Ioannis, et al.
Veröffentlicht: (2026)
TS-SUPERB: A Target Speech Processing Benchmark for Speech Self-Supervised Learning Models
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
von: Peng, Junyi, et al.
Veröffentlicht: (2025)
Probing Self-supervised Learning Models with Target Speech Extraction
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
von: Peng, Junyi, et al.
Veröffentlicht: (2024)
Latent Space Disentanglement via Activation Steering for Interpretable Attribute Control in Symbolic Music Generation
von: Prokopiou, Ioannis, et al.
Veröffentlicht: (2026)
von: Prokopiou, Ioannis, et al.
Veröffentlicht: (2026)
Fine-tuning multilingual language models in Twitter/X sentiment analysis: a study on Eastern-European V4 languages
von: Filip, Tomáš, et al.
Veröffentlicht: (2024)
von: Filip, Tomáš, et al.
Veröffentlicht: (2024)
BIPOLAR: Polarization-based granular framework for LLM bias evaluation
von: Pavlíček, Martin, et al.
Veröffentlicht: (2025)
von: Pavlíček, Martin, et al.
Veröffentlicht: (2025)
Joint Speech and Text Training for LLM-Based End-to-End Spoken Dialogue State Tracking
von: Vendrame, Katia, et al.
Veröffentlicht: (2025)
von: Vendrame, Katia, et al.
Veröffentlicht: (2025)
Building Open-Retrieval Conversational Question Answering Systems by Generating Synthetic Data and Decontextualizing User Questions
von: Vlachos, Christos, et al.
Veröffentlicht: (2025)
von: Vlachos, Christos, et al.
Veröffentlicht: (2025)
DiCoW: Diarization-Conditioned Whisper for Target Speaker Automatic Speech Recognition
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
WeDefense: A Toolkit to Defend Against Fake Audio
von: Zhang, Lin, et al.
Veröffentlicht: (2026)
von: Zhang, Lin, et al.
Veröffentlicht: (2026)
Almost minimal models of log surfaces
von: Palka, Karol
Veröffentlicht: (2024)
von: Palka, Karol
Veröffentlicht: (2024)
Approaching Dialogue State Tracking via Aligning Speech Encoders and LLMs
von: Sedláček, Šimon, et al.
Veröffentlicht: (2025)
von: Sedláček, Šimon, et al.
Veröffentlicht: (2025)
BUT System for the MLC-SLM Challenge
von: Polok, Alexander, et al.
Veröffentlicht: (2025)
von: Polok, Alexander, et al.
Veröffentlicht: (2025)
FIG. 7 in Remains of the digestive system in the middle Cambrian trilobite Ptychoparioides henkli Kordule, 2006 (Barrandian area, Czech Republic)
von: Fatka, Oldřich, et al.
Veröffentlicht: (2026)
von: Fatka, Oldřich, et al.
Veröffentlicht: (2026)
1-Bit Unlimited Sampling Beyond Fourier Domain: Low-Resolution Sampling of Quantization Noise
von: Pavlicek, Vaclav, et al.
Veröffentlicht: (2025)
von: Pavlicek, Vaclav, et al.
Veröffentlicht: (2025)
Sparse Sampling in Fractional Fourier Domain: Recovery Guarantees and Cramér-Rao Bounds
von: Pavlíček, Václav, et al.
Veröffentlicht: (2024)
von: Pavlíček, Václav, et al.
Veröffentlicht: (2024)
On spectator dependence of Jacobi-Lie T-plurality
von: Petr, Ivo, et al.
Veröffentlicht: (2025)
von: Petr, Ivo, et al.
Veröffentlicht: (2025)
Plane-parallel waves as Jacobi-Lie models
von: Petr, Ivo, et al.
Veröffentlicht: (2024)
von: Petr, Ivo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Challenging margin-based speaker embedding extractors by using the variational information bottleneck
von: Stafylakis, Themos, et al.
Veröffentlicht: (2024) -
State-of-the-art Embeddings with Video-free Segmentation of the Source VoxCeleb Data
von: Barahona, Sara, et al.
Veröffentlicht: (2024) -
CA-MHFA: A Context-Aware Multi-Head Factorized Attentive Pooling for SSL-Based Speaker Verification
von: Peng, Junyi, et al.
Veröffentlicht: (2024) -
Hybrid Pruning: In-Situ Compression of Self-Supervised Speech Models for Speaker Verification and Anti-Spoofing
von: Peng, Junyi, et al.
Veröffentlicht: (2025) -
BUT Systems and Analyses for the ASVspoof 5 Challenge
von: Rohdin, Johan, et al.
Veröffentlicht: (2024)