Connecting Voices: LoReSpeech as a Low-Resource Speech Parallel Corpus
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Ouzerrout, Samy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pashto Common Voice: Building the First Open Speech Corpus for a 60-Million-Speaker Low-Resource Language
von: Rahman, Hanif, et al.
Veröffentlicht: (2026)
von: Rahman, Hanif, et al.
Veröffentlicht: (2026)
Giving Voice to the Constitution: Low-Resource Text-to-Speech for Quechua and Spanish Using a Bilingual Legal Corpus
von: Ortega, John E., et al.
Veröffentlicht: (2026)
von: Ortega, John E., et al.
Veröffentlicht: (2026)
DrVoice: Parallel Speech-Text Voice Conversation Model via Dual-Resolution Speech Representations
von: Tan, Chao-Hong, et al.
Veröffentlicht: (2025)
von: Tan, Chao-Hong, et al.
Veröffentlicht: (2025)
Saar-Voice: A Multi-Speaker Saarbrücken Dialect Speech Corpus
von: Oberkircher, Lena S., et al.
Veröffentlicht: (2026)
von: Oberkircher, Lena S., et al.
Veröffentlicht: (2026)
Speech-to-Speech Translation Pipelines for Conversations in Low-Resource Languages
von: Popescu-Belis, Andrei, et al.
Veröffentlicht: (2025)
von: Popescu-Belis, Andrei, et al.
Veröffentlicht: (2025)
Mixture of LoRA Experts for Low-Resourced Multi-Accent Automatic Speech Recognition
von: Bagat, Raphaël, et al.
Veröffentlicht: (2025)
von: Bagat, Raphaël, et al.
Veröffentlicht: (2025)
EuroSpeech: A Multilingual Speech Corpus
von: Pfisterer, Samuel, et al.
Veröffentlicht: (2025)
von: Pfisterer, Samuel, et al.
Veröffentlicht: (2025)
LoASR-Bench: Evaluating Large Speech Language Models on Low-Resource Automatic Speech Recognition Across Language Families
von: Chen, Jianan, et al.
Veröffentlicht: (2026)
von: Chen, Jianan, et al.
Veröffentlicht: (2026)
Enhancing Voice Wake-Up for Dysarthria: Mandarin Dysarthria Speech Corpus Release and Customized System Design
von: Gao, Ming, et al.
Veröffentlicht: (2024)
von: Gao, Ming, et al.
Veröffentlicht: (2024)
FFSTC: Fongbe to French Speech Translation Corpus
von: Kponou, D. Fortune, et al.
Veröffentlicht: (2024)
von: Kponou, D. Fortune, et al.
Veröffentlicht: (2024)
Towards Inclusive ASR: Investigating Voice Conversion for Dysarthric Speech Recognition in Low-Resource Languages
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
RegSpeech12: A Regional Corpus of Bengali Spontaneous Speech Across Dialects
von: Hassan, Md. Rezuwan, et al.
Veröffentlicht: (2025)
von: Hassan, Md. Rezuwan, et al.
Veröffentlicht: (2025)
DATASHI: A Parallel English-Tashlhiyt Corpus for Orthography Normalization and Low-Resource Language Processing
von: Monir, Nasser-Eddine, et al.
Veröffentlicht: (2026)
von: Monir, Nasser-Eddine, et al.
Veröffentlicht: (2026)
ÌròyìnSpeech: A multi-purpose Yorùbá Speech Corpus
von: Ogunremi, Tolulope, et al.
Veröffentlicht: (2023)
von: Ogunremi, Tolulope, et al.
Veröffentlicht: (2023)
Speech Vecalign: an Embedding-based Method for Aligning Parallel Speech Documents
von: Meng, Chutong, et al.
Veröffentlicht: (2025)
von: Meng, Chutong, et al.
Veröffentlicht: (2025)
Speechless: Speech Instruction Training Without Speech for Low Resource Languages
von: Dao, Alan, et al.
Veröffentlicht: (2025)
von: Dao, Alan, et al.
Veröffentlicht: (2025)
Advancing STT for Low-Resource Real-World Speech
von: D'Intino, Flavio, et al.
Veröffentlicht: (2025)
von: D'Intino, Flavio, et al.
Veröffentlicht: (2025)
Leveraging the Cross-Domain & Cross-Linguistic Corpus for Low Resource NMT: A Case Study On Bhili-Hindi-English Parallel Corpus
von: Singh, Pooja, et al.
Veröffentlicht: (2025)
von: Singh, Pooja, et al.
Veröffentlicht: (2025)
An Annotated Corpus of Arabic Tweets for Hate Speech Analysis
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2025)
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2025)
Developing an Open Conversational Speech Corpus for the Isan Language
von: Na-Thalang, Adisai, et al.
Veröffentlicht: (2025)
von: Na-Thalang, Adisai, et al.
Veröffentlicht: (2025)
Gradient-Informed Training for Low-Resource Multilingual Speech Translation
von: Sun, Ruiyan, et al.
Veröffentlicht: (2026)
von: Sun, Ruiyan, et al.
Veröffentlicht: (2026)
MimicLM: Zero-Shot Voice Imitation through Autoregressive Modeling of Pseudo-Parallel Speech Corpora
von: Feng, Tao, et al.
Veröffentlicht: (2026)
von: Feng, Tao, et al.
Veröffentlicht: (2026)
VoiceCraft-X: Unifying Multilingual, Voice-Cloning Speech Synthesis and Speech Editing
von: Zheng, Zhisheng, et al.
Veröffentlicht: (2025)
von: Zheng, Zhisheng, et al.
Veröffentlicht: (2025)
Swedish Whispers; Leveraging a Massive Speech Corpus for Swedish Speech Recognition
von: Vesterbacka, Leonora, et al.
Veröffentlicht: (2025)
von: Vesterbacka, Leonora, et al.
Veröffentlicht: (2025)
Textless Speech-to-Speech Translation With Limited Parallel Data
von: Diwan, Anuj, et al.
Veröffentlicht: (2023)
von: Diwan, Anuj, et al.
Veröffentlicht: (2023)
RosettaSpeech: Zero-Shot Speech-to-Speech Translation without Parallel Speech
von: Zheng, Zhisheng, et al.
Veröffentlicht: (2025)
von: Zheng, Zhisheng, et al.
Veröffentlicht: (2025)
ParCzech4Speech: A New Speech Corpus Derived from Czech Parliamentary Data
von: Stankov, Vladislav, et al.
Veröffentlicht: (2025)
von: Stankov, Vladislav, et al.
Veröffentlicht: (2025)
WorldSpeech: A Multilingual Speech Corpus from Around the World
von: Asonitis, Antonis, et al.
Veröffentlicht: (2026)
von: Asonitis, Antonis, et al.
Veröffentlicht: (2026)
Speak & Improve Corpus 2025: an L2 English Speech Corpus for Language Assessment and Feedback
von: Knill, Kate, et al.
Veröffentlicht: (2024)
von: Knill, Kate, et al.
Veröffentlicht: (2024)
Voice of a Continent: Mapping Africa's Speech Technology Frontier
von: Elmadany, AbdelRahim, et al.
Veröffentlicht: (2025)
von: Elmadany, AbdelRahim, et al.
Veröffentlicht: (2025)
OpenWHO: A Document-Level Parallel Corpus for Health Translation in Low-Resource Languages
von: Merx, Raphaël, et al.
Veröffentlicht: (2025)
von: Merx, Raphaël, et al.
Veröffentlicht: (2025)
IndicVoices-R: Unlocking a Massive Multilingual Multi-speaker Speech Corpus for Scaling Indian TTS
von: Sankar, Ashwin, et al.
Veröffentlicht: (2024)
von: Sankar, Ashwin, et al.
Veröffentlicht: (2024)
ViDia2Std: A Parallel Corpus and Methods for Low-Resource Vietnamese Dialect-to-Standard Translation
von: Ta, Khoa Anh, et al.
Veröffentlicht: (2026)
von: Ta, Khoa Anh, et al.
Veröffentlicht: (2026)
Advancing Speech Translation: A Corpus of Mandarin-English Conversational Telephone Speech
von: Wotherspoon, Shannon, et al.
Veröffentlicht: (2024)
von: Wotherspoon, Shannon, et al.
Veröffentlicht: (2024)
WenetSpeech-Chuan: A Large-Scale Sichuanese Corpus with Rich Annotation for Dialectal Speech Processing
von: Dai, Yuhang, et al.
Veröffentlicht: (2025)
von: Dai, Yuhang, et al.
Veröffentlicht: (2025)
SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech Recognition
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2024)
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2024)
EthioMT: Parallel Corpus for Low-resource Ethiopian Languages
von: Tonja, Atnafu Lambebo, et al.
Veröffentlicht: (2024)
von: Tonja, Atnafu Lambebo, et al.
Veröffentlicht: (2024)
SpeechLLM: Unified Speech and Language Model for Enhanced Multi-Task Understanding in Low Resource Settings
von: Yoo, Jaekwon, et al.
Veröffentlicht: (2025)
von: Yoo, Jaekwon, et al.
Veröffentlicht: (2025)
ZAEBUC-Spoken: A Multilingual Multidialectal Arabic-English Speech Corpus
von: Hamed, Injy, et al.
Veröffentlicht: (2024)
von: Hamed, Injy, et al.
Veröffentlicht: (2024)
GMU Systems for the IWSLT 2025 Low-Resource Speech Translation Shared Task
von: Meng, Chutong, et al.
Veröffentlicht: (2025)
von: Meng, Chutong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Pashto Common Voice: Building the First Open Speech Corpus for a 60-Million-Speaker Low-Resource Language
von: Rahman, Hanif, et al.
Veröffentlicht: (2026) -
Giving Voice to the Constitution: Low-Resource Text-to-Speech for Quechua and Spanish Using a Bilingual Legal Corpus
von: Ortega, John E., et al.
Veröffentlicht: (2026) -
DrVoice: Parallel Speech-Text Voice Conversation Model via Dual-Resolution Speech Representations
von: Tan, Chao-Hong, et al.
Veröffentlicht: (2025) -
Saar-Voice: A Multi-Speaker Saarbrücken Dialect Speech Corpus
von: Oberkircher, Lena S., et al.
Veröffentlicht: (2026) -
Speech-to-Speech Translation Pipelines for Conversations in Low-Resource Languages
von: Popescu-Belis, Andrei, et al.
Veröffentlicht: (2025)