Saar-Voice: A Multi-Speaker Saarbrücken Dialect Speech Corpus
Fuente:
arXiv
Saved in:
| Main Authors: | Oberkircher, Lena S., Alabi, Jesujoba O., Klakow, Dietrich, Trouvain, Jürgen |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AAdaM at SemEval-2024 Task 1: Augmentation and Adaptation for Multilingual Semantic Textual Relatedness
by: Zhang, Miaoran, et al.
Published: (2024)
by: Zhang, Miaoran, et al.
Published: (2024)
AfriHuBERT: A self-supervised speech representation model for African languages
by: Alabi, Jesujoba O., et al.
Published: (2024)
by: Alabi, Jesujoba O., et al.
Published: (2024)
Charting the Landscape of African NLP: Mapping Progress and Shaping the Road Ahead
by: Alabi, Jesujoba O., et al.
Published: (2025)
by: Alabi, Jesujoba O., et al.
Published: (2025)
The Hidden Space of Transformer Language Adapters
by: Alabi, Jesujoba O., et al.
Published: (2024)
by: Alabi, Jesujoba O., et al.
Published: (2024)
The Impact of Demonstrations on Multilingual In-Context Learning: A Multidimensional Analysis
by: Zhang, Miaoran, et al.
Published: (2024)
by: Zhang, Miaoran, et al.
Published: (2024)
TRepLiNa: Layer-wise CKA+REPINA Alignment Improves Low-Resource Machine Translation in Aya-23 8B
by: Nakai, Toshiki, et al.
Published: (2025)
by: Nakai, Toshiki, et al.
Published: (2025)
Voice Conversion Improves Cross-Domain Robustness for Spoken Arabic Dialect Identification
by: Abdullah, Badr M., et al.
Published: (2025)
by: Abdullah, Badr M., et al.
Published: (2025)
Joint vs Sequential Speaker-Role Detection and Automatic Speech Recognition for Air-traffic Control
by: Blatt, Alexander, et al.
Published: (2024)
by: Blatt, Alexander, et al.
Published: (2024)
YoNER: A New Yorùbá Multi-domain Named Entity Recognition Dataset
by: Falola, Peace Busola, et al.
Published: (2026)
by: Falola, Peace Busola, et al.
Published: (2026)
AFRIDOC-MT: Document-level MT Corpus for African Languages
by: Alabi, Jesujoba O., et al.
Published: (2025)
by: Alabi, Jesujoba O., et al.
Published: (2025)
SIB-200: A Simple, Inclusive, and Big Evaluation Dataset for Topic Classification in 200+ Languages and Dialects
by: Adelani, David Ifeoluwa, et al.
Published: (2023)
by: Adelani, David Ifeoluwa, et al.
Published: (2023)
Improving Semantic Understanding in Speech Language Models via Brain-tuning
by: Moussa, Omer, et al.
Published: (2024)
by: Moussa, Omer, et al.
Published: (2024)
Utilizing Multimodal Data for Edge Case Robust Call-sign Recognition and Understanding
by: Blatt, Alexander, et al.
Published: (2024)
by: Blatt, Alexander, et al.
Published: (2024)
RegSpeech12: A Regional Corpus of Bengali Spontaneous Speech Across Dialects
by: Hassan, Md. Rezuwan, et al.
Published: (2025)
by: Hassan, Md. Rezuwan, et al.
Published: (2025)
Human Speech Perception in Noise: Can Large Language Models Paraphrase to Improve It?
by: Chingacham, Anupama, et al.
Published: (2024)
by: Chingacham, Anupama, et al.
Published: (2024)
NaijaRC: A Multi-choice Reading Comprehension Dataset for Nigerian Languages
by: Aremu, Anuoluwapo, et al.
Published: (2023)
by: Aremu, Anuoluwapo, et al.
Published: (2023)
Arab Voices: Mapping Standard and Dialectal Arabic Speech Technology
by: Sullivan, Peter, et al.
Published: (2026)
by: Sullivan, Peter, et al.
Published: (2026)
YAD: Leveraging T5 for Improved Automatic Diacritization of Yorùbá Text
by: Olawole, Akindele Michael, et al.
Published: (2024)
by: Olawole, Akindele Michael, et al.
Published: (2024)
Pashto Common Voice: Building the First Open Speech Corpus for a 60-Million-Speaker Low-Resource Language
by: Rahman, Hanif, et al.
Published: (2026)
by: Rahman, Hanif, et al.
Published: (2026)
IGC: Integrating a Gated Calculator into an LLM to Solve Arithmetic Tasks Reliably and Efficiently
by: Dietz, Florian, et al.
Published: (2025)
by: Dietz, Florian, et al.
Published: (2025)
EkoHate: Abusive Language and Hate Speech Detection for Code-switched Political Discussions on Nigerian Twitter
by: Ilevbare, Comfort Eseohen, et al.
Published: (2024)
by: Ilevbare, Comfort Eseohen, et al.
Published: (2024)
Tarab: A Multi-Dialect Corpus of Arabic Lyrics and Poetry
by: El-Haj, Mo
Published: (2026)
by: El-Haj, Mo
Published: (2026)
WenetSpeech-Chuan: A Large-Scale Sichuanese Corpus with Rich Annotation for Dialectal Speech Processing
by: Dai, Yuhang, et al.
Published: (2025)
by: Dai, Yuhang, et al.
Published: (2025)
UNIQORN: Unified Question Answering over RDF Knowledge Graphs and Natural Language Text
by: Pramanik, Soumajit, et al.
Published: (2021)
by: Pramanik, Soumajit, et al.
Published: (2021)
A Multi-Dialectal Dataset for German Dialect ASR and Dialect-to-Standard Speech Translation
by: Blaschke, Verena, et al.
Published: (2025)
by: Blaschke, Verena, et al.
Published: (2025)
ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis
by: Toyin, Hawau Olamide, et al.
Published: (2025)
by: Toyin, Hawau Olamide, et al.
Published: (2025)
Connecting Voices: LoReSpeech as a Low-Resource Speech Parallel Corpus
by: Ouzerrout, Samy
Published: (2025)
by: Ouzerrout, Samy
Published: (2025)
Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages
by: Bayes, Edward, et al.
Published: (2024)
by: Bayes, Edward, et al.
Published: (2024)
It's Not a Walk in the Park! Challenges of Idiom Translation in Speech-to-text Systems
by: Zaitova, Iuliia, et al.
Published: (2025)
by: Zaitova, Iuliia, et al.
Published: (2025)
On the Encoding of Gender in Transformer-based ASR Representations
by: Krishnan, Aravind, et al.
Published: (2024)
by: Krishnan, Aravind, et al.
Published: (2024)
Evaluating Intermediate Reasoning of Code-Assisted Large Language Models for Mathematics
by: Al-Khalili, Zena, et al.
Published: (2025)
by: Al-Khalili, Zena, et al.
Published: (2025)
Ethio-ASR: Joint Multilingual Speech Recognition and Language Identification for Ethiopian Languages
by: Abdullah, Badr M., et al.
Published: (2026)
by: Abdullah, Badr M., et al.
Published: (2026)
MERLIN: Multi-Stage Curriculum Alignment for Multilingual Encoder-LLM Integration in Cross-Lingual Reasoning
by: Uemura, Kosei, et al.
Published: (2025)
by: Uemura, Kosei, et al.
Published: (2025)
What Do Dialect Speakers Want? A Survey of Attitudes Towards Language Technology for German Dialects
by: Blaschke, Verena, et al.
Published: (2024)
by: Blaschke, Verena, et al.
Published: (2024)
mSTEB: Massively Multilingual Evaluation of LLMs on Speech and Text Tasks
by: Beyene, Luel Hagos, et al.
Published: (2025)
by: Beyene, Luel Hagos, et al.
Published: (2025)
Understanding "Democratization" in NLP and ML Research
by: Subramonian, Arjun, et al.
Published: (2024)
by: Subramonian, Arjun, et al.
Published: (2024)
Finding A Voice: Exploring the Potential of African American Dialect and Voice Generation for Chatbots
by: Finch, Sarah E., et al.
Published: (2025)
by: Finch, Sarah E., et al.
Published: (2025)
Less Stress, More Privacy: Stress Detection on Anonymized Speech of Air Traffic Controllers
by: Viswanathan, Janaki, et al.
Published: (2025)
by: Viswanathan, Janaki, et al.
Published: (2025)
Large Language Models Discriminate Against Speakers of German Dialects
by: Bui, Minh Duc, et al.
Published: (2025)
by: Bui, Minh Duc, et al.
Published: (2025)
FMSD-TTS: Few-shot Multi-Speaker Multi-Dialect Text-to-Speech Synthesis for Ü-Tsang, Amdo and Kham Speech Dataset Generation
by: Liu, Yutong, et al.
Published: (2025)
by: Liu, Yutong, et al.
Published: (2025)
Similar Items
-
AAdaM at SemEval-2024 Task 1: Augmentation and Adaptation for Multilingual Semantic Textual Relatedness
by: Zhang, Miaoran, et al.
Published: (2024) -
AfriHuBERT: A self-supervised speech representation model for African languages
by: Alabi, Jesujoba O., et al.
Published: (2024) -
Charting the Landscape of African NLP: Mapping Progress and Shaping the Road Ahead
by: Alabi, Jesujoba O., et al.
Published: (2025) -
The Hidden Space of Transformer Language Adapters
by: Alabi, Jesujoba O., et al.
Published: (2024) -
The Impact of Demonstrations on Multilingual In-Context Learning: A Multidimensional Analysis
by: Zhang, Miaoran, et al.
Published: (2024)