A Semi-Automatic Approach to Create Large Gender- and Age-Balanced Speaker Corpora: Usefulness of Speaker Diarization & Identification
Fuente:
arXiv
Saved in:
| Main Authors: | Uro, Rémi, Doukhan, David, Rilliard, Albert, Larcher, Laëtitia, Adgharouamane, Anissa-Claire, Tahon, Marie, Laurent, Antoine |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Detecting the terminality of speech-turn boundary for spoken interactions in French TV and Radio content
by: Uro, Rémi, et al.
Published: (2024)
by: Uro, Rémi, et al.
Published: (2024)
Evolution of Voices in French Audiovisual Media Across Genders and Age in a Diachronic Perspective
by: Rilliard, Albert, et al.
Published: (2024)
by: Rilliard, Albert, et al.
Published: (2024)
spINAch: A Diachronic Corpus of French Broadcast Speech Controlled for Speakers' Age and Gender
by: Devauchelle, Simon, et al.
Published: (2026)
by: Devauchelle, Simon, et al.
Published: (2026)
Articulatory Configurations across Genders and Periods in French Radio and TV archives
by: Elie, Benjamin, et al.
Published: (2024)
by: Elie, Benjamin, et al.
Published: (2024)
InaGVAD : a Challenging French TV and Radio Corpus Annotated for Speech Activity Detection and Speaker Gender Segmentation
by: Doukhan, David, et al.
Published: (2024)
by: Doukhan, David, et al.
Published: (2024)
Automatic Voice Identification after Speech Resynthesis using PPG
by: Gaudier, Thibault, et al.
Published: (2024)
by: Gaudier, Thibault, et al.
Published: (2024)
Pretraining Multi-Speaker Identification for Neural Speaker Diarization
by: Horiguchi, Shota, et al.
Published: (2025)
by: Horiguchi, Shota, et al.
Published: (2025)
ASoBO: Attentive Beamformer Selection for Distant Speaker Diarization in Meetings
by: Mariotte, Theo, et al.
Published: (2024)
by: Mariotte, Theo, et al.
Published: (2024)
How to build an Open Science Monitor based on publications? A French perspective
by: Bracco, Laetitia, et al.
Published: (2025)
by: Bracco, Laetitia, et al.
Published: (2025)
ExtracTable: Human-in-the-Loop Transformation of Scientific Corpora into Structured Knowledge
by: John, Lena, et al.
Published: (2025)
by: John, Lena, et al.
Published: (2025)
A Toolkit for Joint Speaker Diarization and Identification with Application to Speaker-Attributed ASR
by: Morrone, Giovanni, et al.
Published: (2024)
by: Morrone, Giovanni, et al.
Published: (2024)
On the Limitations of Speaker Diarization
by: Joana Amorim, et al.
Published: (2026)
by: Joana Amorim, et al.
Published: (2026)
Open Political Corpora: Structuring, Searching, and Analyzing Political Text Collections with PoliCorp
by: Smirnova, Nina, et al.
Published: (2025)
by: Smirnova, Nina, et al.
Published: (2025)
Mens Sana In Corpore Sano: Sound Firmware Corpora for Vulnerability Research
by: Helmke, René, et al.
Published: (2024)
by: Helmke, René, et al.
Published: (2024)
Sequence-to-Sequence Neural Diarization with Automatic Speaker Detection and Representation
by: Cheng, Ming, et al.
Published: (2024)
by: Cheng, Ming, et al.
Published: (2024)
Uncertainty Quantification in Machine Learning for Joint Speaker Diarization and Identification
by: McKnight, Simon W., et al.
Published: (2023)
by: McKnight, Simon W., et al.
Published: (2023)
Balancing the Byline: Exploring Gender and Authorship Patterns in Canadian Science Publishing Journals
by: Hennessey, Eden J., et al.
Published: (2025)
by: Hennessey, Eden J., et al.
Published: (2025)
Leveraging LLMs for Semi-Automatic Corpus Filtration in Systematic Literature Reviews
by: Joos, Lucas, et al.
Published: (2025)
by: Joos, Lucas, et al.
Published: (2025)
ميزان الحقوق والحريات الرقمية في التشريعات الشرق أوسطية Balancing Digital Rights and Freedoms in Middle Eastern Legislations
by: Fawz Ahmed awad, Dr.Amal
Published: (2019)
by: Fawz Ahmed awad, Dr.Amal
Published: (2019)
GLiSE: A Prompt-Driven and ML-Powered Tool for Automated Grey Literature Extraction in Software Engineering
by: Cherief, Houcine Abdelkader, et al.
Published: (2025)
by: Cherief, Houcine Abdelkader, et al.
Published: (2025)
Exploring Speaker Diarization with Mixture of Experts
by: Yang, Gaobin, et al.
Published: (2025)
by: Yang, Gaobin, et al.
Published: (2025)
USED: Universal Speaker Extraction and Diarization
by: Ao, Junyi, et al.
Published: (2023)
by: Ao, Junyi, et al.
Published: (2023)
ASR-Synchronized Speaker-Role Diarization
by: Ghosh, Arindam, et al.
Published: (2025)
by: Ghosh, Arindam, et al.
Published: (2025)
Creating a Digital Twin of the Early Helladic Cemetery of Asteria Glyfadas: Challenges and Innovations in Field Documentation
by: Konstantakis, Markos, et al.
Published: (2025)
by: Konstantakis, Markos, et al.
Published: (2025)
Creating a Digital Twin of the Early Helladic Cemetery of Asteria Glyfadas: Challenges and Innovations in Field Documentation
by: Konstantakis, Markos, et al.
Published: (2025)
by: Konstantakis, Markos, et al.
Published: (2025)
Analysis of the Usability of Automatically Enriched Cultural Heritage Data
by: Raemy, Julien Antoine, et al.
Published: (2023)
by: Raemy, Julien Antoine, et al.
Published: (2023)
A Hybrid AI Methodology for Generating Ontologies of Research Topics from Scientific Paper Corpora
by: Pisu, Alessia, et al.
Published: (2025)
by: Pisu, Alessia, et al.
Published: (2025)
Can We Really Repurpose Multi-Speaker ASR Corpus for Speaker Diarization?
by: Horiguchi, Shota, et al.
Published: (2025)
by: Horiguchi, Shota, et al.
Published: (2025)
Leveraging Speaker Embeddings in End-to-End Neural Diarization for Two-Speaker Scenarios
by: Alvarez-Trejos, Juan Ignacio, et al.
Published: (2024)
by: Alvarez-Trejos, Juan Ignacio, et al.
Published: (2024)
Speaker Embeddings With Weakly Supervised Voice Activity Detection For Efficient Speaker Diarization
by: Thienpondt, Jenthe, et al.
Published: (2024)
by: Thienpondt, Jenthe, et al.
Published: (2024)
Robust Target Speaker Diarization and Separation via Augmented Speaker Embedding Sampling
by: Jalal, Md Asif, et al.
Published: (2025)
by: Jalal, Md Asif, et al.
Published: (2025)
Ocean Decade Vision 2030 White Papers – Challenge 8: Create a digital representation of the ocean.
by: Calewaert, J.-B., et al.
Published: (2024)
by: Calewaert, J.-B., et al.
Published: (2024)
Towards Industrial Convergence : Understanding the evolution of scientific norms and practices in the field of AI
by: Houssard, Antoine
Published: (2025)
by: Houssard, Antoine
Published: (2025)
International Research Collaboration Among Top Performers: A Gender Gap Persists
by: Kwiek, Marek, et al.
Published: (2025)
by: Kwiek, Marek, et al.
Published: (2025)
DiarizationLM: Speaker Diarization Post-Processing with Large Language Models
by: Wang, Quan, et al.
Published: (2024)
by: Wang, Quan, et al.
Published: (2024)
Self-Tuning Spectral Clustering for Speaker Diarization
by: Raghav, Nikhil, et al.
Published: (2024)
by: Raghav, Nikhil, et al.
Published: (2024)
Investigating Confidence Estimation Measures for Speaker Diarization
by: Chowdhury, Anurag, et al.
Published: (2024)
by: Chowdhury, Anurag, et al.
Published: (2024)
Multi-Stage Speaker Diarization for Noisy Classrooms
by: Khan, Ali Sartaz, et al.
Published: (2025)
by: Khan, Ali Sartaz, et al.
Published: (2025)
Leveraging Self-Supervised Learning for Speaker Diarization
by: Han, Jiangyu, et al.
Published: (2024)
by: Han, Jiangyu, et al.
Published: (2024)
Language Modelling for Speaker Diarization in Telephonic Interviews
by: India, Miquel, et al.
Published: (2025)
by: India, Miquel, et al.
Published: (2025)
Similar Items
-
Detecting the terminality of speech-turn boundary for spoken interactions in French TV and Radio content
by: Uro, Rémi, et al.
Published: (2024) -
Evolution of Voices in French Audiovisual Media Across Genders and Age in a Diachronic Perspective
by: Rilliard, Albert, et al.
Published: (2024) -
spINAch: A Diachronic Corpus of French Broadcast Speech Controlled for Speakers' Age and Gender
by: Devauchelle, Simon, et al.
Published: (2026) -
Articulatory Configurations across Genders and Periods in French Radio and TV archives
by: Elie, Benjamin, et al.
Published: (2024) -
InaGVAD : a Challenging French TV and Radio Corpus Annotated for Speech Activity Detection and Speaker Gender Segmentation
by: Doukhan, David, et al.
Published: (2024)