OLaPh: Optimal Language Phonemizer
Fuente:
arXiv
Saved in:
| Main Author: | Wirth, Johannes |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Using Phonemes in cascaded S2S translation pipeline
by: Pilz, Rene, et al.
Published: (2025)
by: Pilz, Rene, et al.
Published: (2025)
From Babble to Words: Pre-Training Language Models on Continuous Streams of Phonemes
by: Goriely, Zébulon, et al.
Published: (2024)
by: Goriely, Zébulon, et al.
Published: (2024)
From Phonemes to Meaning: Evaluating Large Language Models on Tamil
by: Varsha, Jeyarajalingam, et al.
Published: (2025)
by: Varsha, Jeyarajalingam, et al.
Published: (2025)
Modelling the Diachronic Emergence of Phoneme Frequency Distributions
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026)
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026)
The Distribution of Phoneme Frequencies across the World's Languages: Macroscopic and Microscopic Information-Theoretic Models
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026)
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026)
AraS2P: Arabic Speech-to-Phonemes System
by: Matar, Bassam, et al.
Published: (2025)
by: Matar, Bassam, et al.
Published: (2025)
Handling Korean Out-of-Vocabulary Words with Phoneme Representation Learning
by: Kim, Nayeon, et al.
Published: (2025)
by: Kim, Nayeon, et al.
Published: (2025)
LLM-Powered Grapheme-to-Phoneme Conversion: Benchmark and Case Study
by: Qharabagh, Mahta Fetrat, et al.
Published: (2024)
by: Qharabagh, Mahta Fetrat, et al.
Published: (2024)
Zero-Shot Cross-Lingual NER Using Phonemic Representations for Low-Resource Languages
by: Sohn, Jimin, et al.
Published: (2024)
by: Sohn, Jimin, et al.
Published: (2024)
Prompting with Phonemes: Enhancing LLMs' Multilinguality for Non-Latin Script Languages
by: Nguyen, Hoang H, et al.
Published: (2024)
by: Nguyen, Hoang H, et al.
Published: (2024)
IPA-CHILDES & G2P+: Feature-Rich Resources for Cross-Lingual Phonology and Phonemic Language Modeling
by: Goriely, Zébulon, et al.
Published: (2025)
by: Goriely, Zébulon, et al.
Published: (2025)
The exception of humour: Iconicity, Phonemic Surprisal, Memory Recall, and Emotional Associations
by: Kilpatrick, Alexander, et al.
Published: (2025)
by: Kilpatrick, Alexander, et al.
Published: (2025)
Phoneme-Level Visual Speech Recognition via Point-Visual Fusion and Language Model Reconstruction
by: Teng, Matthew Kit Khinn, et al.
Published: (2025)
by: Teng, Matthew Kit Khinn, et al.
Published: (2025)
Small Language Models Also Work With Small Vocabularies: Probing the Linguistic Abilities of Grapheme- and Phoneme-Based Baby Llamas
by: Bunzeck, Bastian, et al.
Published: (2024)
by: Bunzeck, Bastian, et al.
Published: (2024)
PolyIPA -- Multilingual Phoneme-to-Grapheme Conversion Model
by: Lauc, Davor
Published: (2024)
by: Lauc, Davor
Published: (2024)
ViSpeechFormer: A Phonemic Approach for Vietnamese Automatic Speech Recognition
by: Nguyen, Khoa Anh, et al.
Published: (2026)
by: Nguyen, Khoa Anh, et al.
Published: (2026)
Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages
by: Akavarapu, V. S. D. S. Mahesh, et al.
Published: (2026)
by: Akavarapu, V. S. D. S. Mahesh, et al.
Published: (2026)
Emergence of Phonemic, Syntactic, and Semantic Representations in Artificial Neural Networks
by: Orhan, Pierre, et al.
Published: (2026)
by: Orhan, Pierre, et al.
Published: (2026)
Improving Spoken Language Modeling with Phoneme Classification: A Simple Fine-tuning Approach
by: Poli, Maxime, et al.
Published: (2024)
by: Poli, Maxime, et al.
Published: (2024)
PhonemeFake: Redefining Deepfake Realism with Language-Driven Segmental Manipulation and Adaptive Bilevel Detection
by: Baser, Oguzhan, et al.
Published: (2025)
by: Baser, Oguzhan, et al.
Published: (2025)
Speech Rhythm-Based Speaker Embeddings Extraction from Phonemes and Phoneme Duration for Multi-Speaker Speech Synthesis
by: Fujita, Kenichi, et al.
Published: (2024)
by: Fujita, Kenichi, et al.
Published: (2024)
Whisper based Cross-Lingual Phoneme Recognition between Vietnamese and English
by: Minh, Nguyen Huu Nhat, et al.
Published: (2025)
by: Minh, Nguyen Huu Nhat, et al.
Published: (2025)
Mitigating the Linguistic Gap with Phonemic Representations for Robust Cross-lingual Transfer
by: Jung, Haeji, et al.
Published: (2024)
by: Jung, Haeji, et al.
Published: (2024)
Can LLMs Simulate Human Behavioral Variability? A Case Study in the Phonemic Fluency Task
by: Qiu, Mengyang, et al.
Published: (2025)
by: Qiu, Mengyang, et al.
Published: (2025)
CUPE: Contextless Universal Phoneme Encoder for Language-Agnostic Speech Processing
by: Rehman, Abdul, et al.
Published: (2025)
by: Rehman, Abdul, et al.
Published: (2025)
Multilingual Dysarthric Speech Assessment Using Universal Phone Recognition and Language-Specific Phonemic Contrast Modeling
by: Yeo, Eunjung, et al.
Published: (2026)
by: Yeo, Eunjung, et al.
Published: (2026)
Cross-Dialect Text-To-Speech in Pitch-Accent Language Incorporating Multi-Dialect Phoneme-Level BERT
by: Yamauchi, Kazuki, et al.
Published: (2024)
by: Yamauchi, Kazuki, et al.
Published: (2024)
CIPHER: Conformer-based Inference of Phonemes from High-density EEG
by: Madishetty, Varshith
Published: (2026)
by: Madishetty, Varshith
Published: (2026)
Comparison of sEMG Encoding Accuracy Across Speech Modes Using Articulatory and Phoneme Features
by: Le, Chenqian, et al.
Published: (2026)
by: Le, Chenqian, et al.
Published: (2026)
Speaker Style-Aware Phoneme Anchoring for Improved Cross-Lingual Speech Emotion Recognition
by: Upadhyay, Shreya G., et al.
Published: (2025)
by: Upadhyay, Shreya G., et al.
Published: (2025)
iPhoneme: Brain-to-Text Communication for ALS Using ConformerXL Decoding
by: Cha, Yoonmin, et al.
Published: (2026)
by: Cha, Yoonmin, et al.
Published: (2026)
Phonikud: Hebrew Grapheme-to-Phoneme Conversion for Real-Time Text-to-Speech
by: Kolani, Yakov, et al.
Published: (2025)
by: Kolani, Yakov, et al.
Published: (2025)
Multitask Learning for Grapheme-to-Phoneme Conversion of Anglicisms in German Speech Recognition
by: Pritzen, Julia, et al.
Published: (2021)
by: Pritzen, Julia, et al.
Published: (2021)
Low-Resourced Speech Recognition for Iu Mien Language via Weakly-Supervised Phoneme-based Multilingual Pre-training
by: Dong, Lukuan, et al.
Published: (2024)
by: Dong, Lukuan, et al.
Published: (2024)
DyPCL: Dynamic Phoneme-level Contrastive Learning for Dysarthric Speech Recognition
by: Lee, Wonjun, et al.
Published: (2025)
by: Lee, Wonjun, et al.
Published: (2025)
Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning
by: Ohnaka, Hien, et al.
Published: (2025)
by: Ohnaka, Hien, et al.
Published: (2025)
DiscoPhon: Benchmarking the Unsupervised Discovery of Phoneme Inventories With Discrete Speech Units
by: Poli, Maxime, et al.
Published: (2026)
by: Poli, Maxime, et al.
Published: (2026)
Bridging the Gap: An Intermediate Language for Enhanced and Cost-Effective Grapheme-to-Phoneme Conversion with Homographs with Multiple Pronunciations Disambiguation
by: Bertina, Abbas, et al.
Published: (2025)
by: Bertina, Abbas, et al.
Published: (2025)
RoboPhD: Self-Improving Text-to-SQL Through Autonomous Agent Evolution
by: Borthwick, Andrew, et al.
Published: (2026)
by: Borthwick, Andrew, et al.
Published: (2026)
PARCO: Phoneme-Augmented Robust Contextual ASR via Contrastive Entity Disambiguation
by: He, Jiajun, et al.
Published: (2025)
by: He, Jiajun, et al.
Published: (2025)
Similar Items
-
Using Phonemes in cascaded S2S translation pipeline
by: Pilz, Rene, et al.
Published: (2025) -
From Babble to Words: Pre-Training Language Models on Continuous Streams of Phonemes
by: Goriely, Zébulon, et al.
Published: (2024) -
From Phonemes to Meaning: Evaluating Large Language Models on Tamil
by: Varsha, Jeyarajalingam, et al.
Published: (2025) -
Modelling the Diachronic Emergence of Phoneme Frequency Distributions
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026) -
The Distribution of Phoneme Frequencies across the World's Languages: Macroscopic and Microscopic Information-Theoretic Models
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026)