LLM-Powered Grapheme-to-Phoneme Conversion: Benchmark and Case Study
Fuente:
arXiv
Saved in:
| Main Authors: | Qharabagh, Mahta Fetrat, Dehghanian, Zahra, Rabiee, Hamid R. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast, Not Fancy: Rethinking G2P with Rich Data and Rule-Based Models
by: Qharabagh, Mahta Fetrat, et al.
Published: (2025)
by: Qharabagh, Mahta Fetrat, et al.
Published: (2025)
ManaTTS Persian: a recipe for creating TTS datasets for lower resource languages
by: Qharabagh, Mahta Fetrat, et al.
Published: (2024)
by: Qharabagh, Mahta Fetrat, et al.
Published: (2024)
Beyond Unified Models: A Service-Oriented Approach to Low Latency, Context Aware Phonemization for Real Time TTS
by: Fetrat, Mahta, et al.
Published: (2025)
by: Fetrat, Mahta, et al.
Published: (2025)
PolyIPA -- Multilingual Phoneme-to-Grapheme Conversion Model
by: Lauc, Davor
Published: (2024)
by: Lauc, Davor
Published: (2024)
Multitask Learning for Grapheme-to-Phoneme Conversion of Anglicisms in German Speech Recognition
by: Pritzen, Julia, et al.
Published: (2021)
by: Pritzen, Julia, et al.
Published: (2021)
Phonikud: Hebrew Grapheme-to-Phoneme Conversion for Real-Time Text-to-Speech
by: Kolani, Yakov, et al.
Published: (2025)
by: Kolani, Yakov, et al.
Published: (2025)
Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning
by: Ohnaka, Hien, et al.
Published: (2025)
by: Ohnaka, Hien, et al.
Published: (2025)
Bridging the Gap: An Intermediate Language for Enhanced and Cost-Effective Grapheme-to-Phoneme Conversion with Homographs with Multiple Pronunciations Disambiguation
by: Bertina, Abbas, et al.
Published: (2025)
by: Bertina, Abbas, et al.
Published: (2025)
CineLOG: A Training Free Approach for Cinematic Long Video Generation
by: Dehghanian, Zahra, et al.
Published: (2025)
by: Dehghanian, Zahra, et al.
Published: (2025)
Redefining Generalization in Visual Domains: A Two-Axis Framework for Fake Image Detection with FusionDetect
by: Amanzadi, Amirtaha, et al.
Published: (2025)
by: Amanzadi, Amirtaha, et al.
Published: (2025)
Small Language Models Also Work With Small Vocabularies: Probing the Linguistic Abilities of Grapheme- and Phoneme-Based Baby Llamas
by: Bunzeck, Bastian, et al.
Published: (2024)
by: Bunzeck, Bastian, et al.
Published: (2024)
LVLM-COUNT: Enhancing the Counting Ability of Large Vision-Language Models
by: Qharabagh, Muhammad Fetrat, et al.
Published: (2024)
by: Qharabagh, Muhammad Fetrat, et al.
Published: (2024)
Camera Trajectory Generation: A Comprehensive Survey of Methods, Metrics, and Future Directions
by: Dehghanian, Zahra, et al.
Published: (2025)
by: Dehghanian, Zahra, et al.
Published: (2025)
RealDrag: The First Dragging Benchmark with Real Target Image
by: Zafarani, Ahmad, et al.
Published: (2025)
by: Zafarani, Ahmad, et al.
Published: (2025)
ReDiF: Reinforced Distillation for Few Step Diffusion
by: Tighkhorshid, Amirhossein, et al.
Published: (2025)
by: Tighkhorshid, Amirhossein, et al.
Published: (2025)
Unicode Normalization and Grapheme Parsing of Indic Languages
by: Ansary, Nazmuddoha, et al.
Published: (2023)
by: Ansary, Nazmuddoha, et al.
Published: (2023)
LensCraft: Your Professional Virtual Cinematographer
by: Dehghanian, Zahra, et al.
Published: (2025)
by: Dehghanian, Zahra, et al.
Published: (2025)
UPL: Uncertainty-aware Pseudo-labeling for Imbalance Transductive Node Classification
by: Teimuri, Mohammad T., et al.
Published: (2025)
by: Teimuri, Mohammad T., et al.
Published: (2025)
Cueless EEG imagined speech for subject identification: dataset and benchmarks
by: Derakhshesh, Ali, et al.
Published: (2025)
by: Derakhshesh, Ali, et al.
Published: (2025)
Knowledge-Infused LLM-Powered Conversational Health Agent: A Case Study for Diabetes Patients
by: Abbasian, Mahyar, et al.
Published: (2024)
by: Abbasian, Mahyar, et al.
Published: (2024)
Can LLMs Simulate Human Behavioral Variability? A Case Study in the Phonemic Fluency Task
by: Qiu, Mengyang, et al.
Published: (2025)
by: Qiu, Mengyang, et al.
Published: (2025)
Classifying Graphemes in English Words Through the Application of a Fuzzy Inference System
by: Rose, Samuel, et al.
Published: (2024)
by: Rose, Samuel, et al.
Published: (2024)
Graphemic Normalization of the Perso-Arabic Script
by: Doctor, Raiomond, et al.
Published: (2022)
by: Doctor, Raiomond, et al.
Published: (2022)
Query Understanding in LLM-based Conversational Information Seeking
by: Yuan, Yifei, et al.
Published: (2025)
by: Yuan, Yifei, et al.
Published: (2025)
Generative LLM Powered Conversational AI Application for Personalized Risk Assessment: A Case Study in COVID-19
by: Roshani, Mohammad Amin, et al.
Published: (2024)
by: Roshani, Mohammad Amin, et al.
Published: (2024)
Improving Grapheme-to-Phoneme Conversion through In-Context Knowledge Retrieval with Large Language Models
by: Han, Dongrui, et al.
Published: (2024)
by: Han, Dongrui, et al.
Published: (2024)
OLaPh: Optimal Language Phonemizer
by: Wirth, Johannes
Published: (2025)
by: Wirth, Johannes
Published: (2025)
Who Benchmarks the Benchmarks? A Case Study of LLM Evaluation in Icelandic
by: Ingimundarson, Finnur Ágúst, et al.
Published: (2026)
by: Ingimundarson, Finnur Ágúst, et al.
Published: (2026)
Modelling the Diachronic Emergence of Phoneme Frequency Distributions
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026)
by: Martín, Fermín Moscoso del Prado, et al.
Published: (2026)
DiscoPhon: Benchmarking the Unsupervised Discovery of Phoneme Inventories With Discrete Speech Units
by: Poli, Maxime, et al.
Published: (2026)
by: Poli, Maxime, et al.
Published: (2026)
GraphemeAug: A Systematic Approach to Synthesized Hard Negative Keyword Spotting Examples
by: Zhang, Harry, et al.
Published: (2025)
by: Zhang, Harry, et al.
Published: (2025)
3DLAND: 3D Lesion Abdominal Anomaly Localization Dataset
by: Advand, Mehran, et al.
Published: (2026)
by: Advand, Mehran, et al.
Published: (2026)
AraS2P: Arabic Speech-to-Phonemes System
by: Matar, Bassam, et al.
Published: (2025)
by: Matar, Bassam, et al.
Published: (2025)
Conversational Health Agents: A Personalized LLM-Powered Agent Framework
by: Abbasian, Mahyar, et al.
Published: (2023)
by: Abbasian, Mahyar, et al.
Published: (2023)
NC-Bench: An LLM Benchmark for Evaluating Conversational Competence
by: Moore, Robert J., et al.
Published: (2026)
by: Moore, Robert J., et al.
Published: (2026)
Trust Modeling in Counseling Conversations: A Benchmark Study
by: Srivastava, Aseem, et al.
Published: (2025)
by: Srivastava, Aseem, et al.
Published: (2025)
IRLab@iKAT24: Learned Sparse Retrieval with Multi-aspect LLM Query Generation for Conversational Search
by: Lupart, Simon, et al.
Published: (2024)
by: Lupart, Simon, et al.
Published: (2024)
Handling Korean Out-of-Vocabulary Words with Phoneme Representation Learning
by: Kim, Nayeon, et al.
Published: (2025)
by: Kim, Nayeon, et al.
Published: (2025)
GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations
by: Yang, Jingbo, et al.
Published: (2026)
by: Yang, Jingbo, et al.
Published: (2026)
VayuChat: An LLM-Powered Conversational Interface for Air Quality Data Analytics
by: Acharya, Vedant, et al.
Published: (2025)
by: Acharya, Vedant, et al.
Published: (2025)
Similar Items
-
Fast, Not Fancy: Rethinking G2P with Rich Data and Rule-Based Models
by: Qharabagh, Mahta Fetrat, et al.
Published: (2025) -
ManaTTS Persian: a recipe for creating TTS datasets for lower resource languages
by: Qharabagh, Mahta Fetrat, et al.
Published: (2024) -
Beyond Unified Models: A Service-Oriented Approach to Low Latency, Context Aware Phonemization for Real Time TTS
by: Fetrat, Mahta, et al.
Published: (2025) -
PolyIPA -- Multilingual Phoneme-to-Grapheme Conversion Model
by: Lauc, Davor
Published: (2024) -
Multitask Learning for Grapheme-to-Phoneme Conversion of Anglicisms in German Speech Recognition
by: Pritzen, Julia, et al.
Published: (2021)