Data Selection Effects on Self-Supervised Learning of Audio Representations for French Audiovisual Broadcasts
Fuente:
arXiv
Salvato in:
| Autori principali: | Pelloin, Valentin, Bekkali, Lina, Dehak, Reda, Doukhan, David |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Automatic Classification of News Subjects in Broadcast News: Application to a Gender Bias Representation Analysis
di: Pelloin, Valentin, et al.
Pubblicazione: (2024)
di: Pelloin, Valentin, et al.
Pubblicazione: (2024)
spINAch: A Diachronic Corpus of French Broadcast Speech Controlled for Speakers' Age and Gender
di: Devauchelle, Simon, et al.
Pubblicazione: (2026)
di: Devauchelle, Simon, et al.
Pubblicazione: (2026)
Additive Margin in Contrastive Self-Supervised Frameworks to Learn Discriminative Speaker Representations
di: Lepage, Theo, et al.
Pubblicazione: (2024)
di: Lepage, Theo, et al.
Pubblicazione: (2024)
Self-Supervised Learning for Speaker Recognition: A study and review
di: Lepage, Theo, et al.
Pubblicazione: (2026)
di: Lepage, Theo, et al.
Pubblicazione: (2026)
Experimenting with Additive Margins for Contrastive Self-Supervised Speaker Verification
di: Lepage, Theo, et al.
Pubblicazione: (2023)
di: Lepage, Theo, et al.
Pubblicazione: (2023)
InaGVAD : a Challenging French TV and Radio Corpus Annotated for Speech Activity Detection and Speaker Gender Segmentation
di: Doukhan, David, et al.
Pubblicazione: (2024)
di: Doukhan, David, et al.
Pubblicazione: (2024)
Label-Efficient Self-Supervised Speaker Verification With Information Maximization and Contrastive Learning
di: Lepage, Théo, et al.
Pubblicazione: (2022)
di: Lepage, Théo, et al.
Pubblicazione: (2022)
Self-Supervised Frameworks for Speaker Verification via Bootstrapped Positive Sampling
di: Lepage, Theo, et al.
Pubblicazione: (2025)
di: Lepage, Theo, et al.
Pubblicazione: (2025)
SSPS: Self-Supervised Positive Sampling for Robust Self-Supervised Speaker Verification
di: Lepage, Theo, et al.
Pubblicazione: (2025)
di: Lepage, Theo, et al.
Pubblicazione: (2025)
Towards Supervised Performance on Speaker Verification with Self-Supervised Learning by Leveraging Large-Scale ASR Models
di: Miara, Victor, et al.
Pubblicazione: (2024)
di: Miara, Victor, et al.
Pubblicazione: (2024)
Evolution of Voices in French Audiovisual Media Across Genders and Age in a Diachronic Perspective
di: Rilliard, Albert, et al.
Pubblicazione: (2024)
di: Rilliard, Albert, et al.
Pubblicazione: (2024)
Gender Representation in TV and Radio: Automatic Information Extraction methods versus Manual Analyses
di: Doukhan, David, et al.
Pubblicazione: (2024)
di: Doukhan, David, et al.
Pubblicazione: (2024)
Audio Mamba: Selective State Spaces for Self-Supervised Audio Representations
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
Self-Supervised Multi-View Learning for Disentangled Music Audio Representations
di: Wilkins, Julia, et al.
Pubblicazione: (2024)
di: Wilkins, Julia, et al.
Pubblicazione: (2024)
Reconstruct! Don't Encode: Self-Supervised Representation Reconstruction Loss for High-Intelligibility and Low-Latency Streaming Neural Audio Codec
di: Lee, Junhyeok, et al.
Pubblicazione: (2026)
di: Lee, Junhyeok, et al.
Pubblicazione: (2026)
SSAMBA: Self-Supervised Audio Representation Learning with Mamba State Space Model
di: Shams, Siavash, et al.
Pubblicazione: (2024)
di: Shams, Siavash, et al.
Pubblicazione: (2024)
On the Use of Self-Supervised Representation Learning for Speaker Diarization and Separation
di: Baroudi, Séverin, et al.
Pubblicazione: (2025)
di: Baroudi, Séverin, et al.
Pubblicazione: (2025)
SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer
di: Wang, Helin, et al.
Pubblicazione: (2024)
di: Wang, Helin, et al.
Pubblicazione: (2024)
Exploring WavLM Back-ends for Speech Spoofing and Deepfake Detection
di: Stourbe, Theophile, et al.
Pubblicazione: (2024)
di: Stourbe, Theophile, et al.
Pubblicazione: (2024)
Natural Language Supervision for General-Purpose Audio Representations
di: Elizalde, Benjamin, et al.
Pubblicazione: (2023)
di: Elizalde, Benjamin, et al.
Pubblicazione: (2023)
The Effect of Batch Size on Contrastive Self-Supervised Speech Representation Learning
di: Vaessen, Nik, et al.
Pubblicazione: (2024)
di: Vaessen, Nik, et al.
Pubblicazione: (2024)
Leveraging Self-supervised Audio Representations for Data-Efficient Acoustic Scene Classification
di: Cai, Yiqiang, et al.
Pubblicazione: (2024)
di: Cai, Yiqiang, et al.
Pubblicazione: (2024)
CoughViT: A Self-Supervised Vision Transformer for Cough Audio Representation Learning
di: Luong, Justin, et al.
Pubblicazione: (2025)
di: Luong, Justin, et al.
Pubblicazione: (2025)
Audio Deepfake Detection with Self-Supervised WavLM and Multi-Fusion Attentive Classifier
di: Guo, Yinlin, et al.
Pubblicazione: (2023)
di: Guo, Yinlin, et al.
Pubblicazione: (2023)
Self-Supervised Convolutional Audio Models are Flexible Acoustic Feature Learners: A Domain Specificity and Transfer-Learning Study
di: Ogg, Mattson
Pubblicazione: (2025)
di: Ogg, Mattson
Pubblicazione: (2025)
Adaptive Federated Fine-Tuning of Self-Supervised Speech Representations
di: Guo, Xin, et al.
Pubblicazione: (2026)
di: Guo, Xin, et al.
Pubblicazione: (2026)
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning
di: Yang, Yifan, et al.
Pubblicazione: (2024)
di: Yang, Yifan, et al.
Pubblicazione: (2024)
CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing
di: Lu, Yen-Ju, et al.
Pubblicazione: (2024)
di: Lu, Yen-Ju, et al.
Pubblicazione: (2024)
Compositional Audio Representation Learning
di: Sridhar, Sripathi, et al.
Pubblicazione: (2024)
di: Sridhar, Sripathi, et al.
Pubblicazione: (2024)
AudioNet: Supervised Deep Hashing for Retrieval of Similar Audio Events
di: Dutta, Sagar, et al.
Pubblicazione: (2025)
di: Dutta, Sagar, et al.
Pubblicazione: (2025)
Semi-Supervised Contrastive Learning of Musical Representations
di: Guinot, Julien, et al.
Pubblicazione: (2024)
di: Guinot, Julien, et al.
Pubblicazione: (2024)
Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning
di: Wilkins, Julia, et al.
Pubblicazione: (2025)
di: Wilkins, Julia, et al.
Pubblicazione: (2025)
Emotion-Aware Speech Self-Supervised Representation Learning with Intensity Knowledge
di: Liu, Rui, et al.
Pubblicazione: (2024)
di: Liu, Rui, et al.
Pubblicazione: (2024)
ULTRAS -- Unified Learning of Transformer Representations for Audio and Speech Signals
di: E, Ameenudeen P, et al.
Pubblicazione: (2026)
di: E, Ameenudeen P, et al.
Pubblicazione: (2026)
Exploring Self-Supervised Audio Models for Generalized Anomalous Sound Detection
di: Han, Bing, et al.
Pubblicazione: (2025)
di: Han, Bing, et al.
Pubblicazione: (2025)
Cyclostationarity Analysis as a Complement to Self-Supervised Representations for Speech Deepfake Detection
di: Hanilçi, Cemal, et al.
Pubblicazione: (2026)
di: Hanilçi, Cemal, et al.
Pubblicazione: (2026)
Speech Loudness in Broadcasting and Streaming
di: Torcoli, Matteo, et al.
Pubblicazione: (2024)
di: Torcoli, Matteo, et al.
Pubblicazione: (2024)
Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervision
di: Chen, Yafeng, et al.
Pubblicazione: (2024)
di: Chen, Yafeng, et al.
Pubblicazione: (2024)
Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervision
di: Chen, Yafeng, et al.
Pubblicazione: (2023)
di: Chen, Yafeng, et al.
Pubblicazione: (2023)
SPEAR: A Unified SSL Framework for Learning Speech and Audio Representations
di: Yang, Xiaoyu, et al.
Pubblicazione: (2025)
di: Yang, Xiaoyu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Automatic Classification of News Subjects in Broadcast News: Application to a Gender Bias Representation Analysis
di: Pelloin, Valentin, et al.
Pubblicazione: (2024) -
spINAch: A Diachronic Corpus of French Broadcast Speech Controlled for Speakers' Age and Gender
di: Devauchelle, Simon, et al.
Pubblicazione: (2026) -
Additive Margin in Contrastive Self-Supervised Frameworks to Learn Discriminative Speaker Representations
di: Lepage, Theo, et al.
Pubblicazione: (2024) -
Self-Supervised Learning for Speaker Recognition: A study and review
di: Lepage, Theo, et al.
Pubblicazione: (2026) -
Experimenting with Additive Margins for Contrastive Self-Supervised Speaker Verification
di: Lepage, Theo, et al.
Pubblicazione: (2023)