Linear Semantic Segmentation for Low-Resource Spoken Dialects
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chirkunov, Kirill, Samih, Younes, Freihat, Abed Alhakim, Aldarmaki, Hanan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DialectalArabicMMLU: Benchmarking Dialectal Capabilities in Arabic and Multilingual Language Models
von: Altakrori, Malik H., et al.
Veröffentlicht: (2025)
von: Altakrori, Malik H., et al.
Veröffentlicht: (2025)
Toward a Better Localization of Princeton WordNet
von: Freihat, Abed Alhakim
Veröffentlicht: (2025)
von: Freihat, Abed Alhakim
Veröffentlicht: (2025)
Instruction-Guided Poetry Generation in Arabic and Its Dialects
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2026)
Spoken Word2Vec: Learning Skipgram Embeddings from Speech
von: Sayeed, Mohammad Amaan, et al.
Veröffentlicht: (2023)
von: Sayeed, Mohammad Amaan, et al.
Veröffentlicht: (2023)
Cultural Benchmarking of LLMs in Standard and Dialectal Arabic Dialogues
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2026)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2026)
SparQLe: Speech Queries to Text Translation Through LLMs
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
JEEM: Vision-Language Understanding in Four Arabic Dialects
von: Kadaoui, Karima, et al.
Veröffentlicht: (2025)
von: Kadaoui, Karima, et al.
Veröffentlicht: (2025)
Advancing the Arabic WordNet: Elevating Content Quality
von: Freihat, Abed Alhakim, et al.
Veröffentlicht: (2024)
von: Freihat, Abed Alhakim, et al.
Veröffentlicht: (2024)
Morphemes Without Borders: Evaluating Root-Pattern Morphology in Arabic Tokenizers and LLMs
von: Alakeel, Yara, et al.
Veröffentlicht: (2026)
von: Alakeel, Yara, et al.
Veröffentlicht: (2026)
Are LLMs Good Text Diacritizers? An Arabic and Yoruba Case Study
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025)
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025)
From Multiple-Choice to Extractive QA: A Case Study for English and Arabic
von: Lynn, Teresa, et al.
Veröffentlicht: (2024)
von: Lynn, Teresa, et al.
Veröffentlicht: (2024)
Low-Resource Dialect Adaptation of Large Language Models: A French Dialect Case-Study
von: Khan, Eeham, et al.
Veröffentlicht: (2025)
von: Khan, Eeham, et al.
Veröffentlicht: (2025)
Multi-BERT: Leveraging Adapters and Prompt Tuning for Low-Resource Multi-Domain Adaptation
von: Azad, Parham Abed, et al.
Veröffentlicht: (2024)
von: Azad, Parham Abed, et al.
Veröffentlicht: (2024)
Fine-Tuning LLMs for Low-Resource Dialect Translation: The Case of Lebanese
von: Yakhni, Silvana, et al.
Veröffentlicht: (2025)
von: Yakhni, Silvana, et al.
Veröffentlicht: (2025)
STTATTS: Unified Speech-To-Text And Text-To-Speech Model
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2024)
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2024)
Personal Attribute Leakage in Federated Speech Models
von: Al-Ali, Hamdan, et al.
Veröffentlicht: (2025)
von: Al-Ali, Hamdan, et al.
Veröffentlicht: (2025)
Large Language Models as Code Executors: An Exploratory Study
von: Lyu, Chenyang, et al.
Veröffentlicht: (2024)
von: Lyu, Chenyang, et al.
Veröffentlicht: (2024)
Data Augmentation Integrating Dialogue Flow and Style to Adapt Spoken Dialogue Systems to Low-Resource User Groups
von: Qi, Zhiyang, et al.
Veröffentlicht: (2024)
von: Qi, Zhiyang, et al.
Veröffentlicht: (2024)
ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025)
von: Toyin, Hawau Olamide, et al.
Veröffentlicht: (2025)
New Semantic Task for the French Spoken Language Understanding MEDIA Benchmark
von: Alavoine, Nadège, et al.
Veröffentlicht: (2024)
von: Alavoine, Nadège, et al.
Veröffentlicht: (2024)
GLoRIA: Gated Low-Rank Interpretable Adaptation for Dialectal ASR
von: Mehralian, Pouya, et al.
Veröffentlicht: (2026)
von: Mehralian, Pouya, et al.
Veröffentlicht: (2026)
Improving Low-Resource Dialect Classification Using Retrieval-based Voice Conversion
von: Fischbach, Lea, et al.
Veröffentlicht: (2025)
von: Fischbach, Lea, et al.
Veröffentlicht: (2025)
Extracting Lexical Features from Dialects via Interpretable Dialect Classifiers
von: Xie, Roy, et al.
Veröffentlicht: (2024)
von: Xie, Roy, et al.
Veröffentlicht: (2024)
Clinical Annotations for Automatic Stuttering Severity Assessment
von: Valente, Ana Rita, et al.
Veröffentlicht: (2025)
von: Valente, Ana Rita, et al.
Veröffentlicht: (2025)
The Effectiveness of Morphology-aware Segmentation in Low-Resource Neural Machine Translation
von: Sälevä, Jonne, et al.
Veröffentlicht: (2021)
von: Sälevä, Jonne, et al.
Veröffentlicht: (2021)
Efficient Dialect-Aware Modeling and Conditioning for Low-Resource Taiwanese Hakka Speech Processing
von: Peng, An-Ci, et al.
Veröffentlicht: (2026)
von: Peng, An-Ci, et al.
Veröffentlicht: (2026)
Dialectal Coverage And Generalization in Arabic Speech Recognition
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024)
Source-Grounded Semantic Reinforcement Learning for Low-Resource Target-Language Generation
von: Su, Zeli, et al.
Veröffentlicht: (2026)
von: Su, Zeli, et al.
Veröffentlicht: (2026)
SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents
von: Si, Shuzheng, et al.
Veröffentlicht: (2023)
von: Si, Shuzheng, et al.
Veröffentlicht: (2023)
DialectLLM: A Dialect-Aware Dialog[ue] Generation Framework Beyond Standard American English
von: Oh, Jio, et al.
Veröffentlicht: (2026)
von: Oh, Jio, et al.
Veröffentlicht: (2026)
Spoken Grammar Assessment Using LLM
von: Kopparapu, Sunil Kumar, et al.
Veröffentlicht: (2024)
von: Kopparapu, Sunil Kumar, et al.
Veröffentlicht: (2024)
Mixat: A Data Set of Bilingual Emirati-English Speech
von: Ali, Maryam Al, et al.
Veröffentlicht: (2024)
von: Ali, Maryam Al, et al.
Veröffentlicht: (2024)
Predicting the Target Word of Game-playing Conversations using a Low-Rank Dialect Adapter for Decoder Models
von: Srirag, Dipankar, et al.
Veröffentlicht: (2024)
von: Srirag, Dipankar, et al.
Veröffentlicht: (2024)
AraSpot: Arabic Spoken Command Spotting
von: Salhab, Mahmoud, et al.
Veröffentlicht: (2023)
von: Salhab, Mahmoud, et al.
Veröffentlicht: (2023)
Data-Augmentation-Based Dialectal Adaptation for LLMs
von: Faisal, Fahim, et al.
Veröffentlicht: (2024)
von: Faisal, Fahim, et al.
Veröffentlicht: (2024)
"Hunt Takes Hare": Theming Games Through Game-Word Vector Translation
von: Younès, Rabii, et al.
Veröffentlicht: (2024)
von: Younès, Rabii, et al.
Veröffentlicht: (2024)
ChatGPT v.s. Media Bias: A Comparative Study of GPT-3.5 and Fine-tuned Language Models
von: Wen, Zehao, et al.
Veröffentlicht: (2024)
von: Wen, Zehao, et al.
Veröffentlicht: (2024)
Attention-guided Evidence Grounding for Spoken Question Answering
von: Yang, Ke, et al.
Veröffentlicht: (2026)
von: Yang, Ke, et al.
Veröffentlicht: (2026)
Multi-Agent Dialectical Refinement for Enhanced Argument Classification
von: Bąba, Jakub, et al.
Veröffentlicht: (2026)
von: Bąba, Jakub, et al.
Veröffentlicht: (2026)
When Every Token Counts: Optimal Segmentation for Low-Resource Language Models
von: Raj, Bharath, et al.
Veröffentlicht: (2024)
von: Raj, Bharath, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DialectalArabicMMLU: Benchmarking Dialectal Capabilities in Arabic and Multilingual Language Models
von: Altakrori, Malik H., et al.
Veröffentlicht: (2025) -
Toward a Better Localization of Princeton WordNet
von: Freihat, Abed Alhakim
Veröffentlicht: (2025) -
Instruction-Guided Poetry Generation in Arabic and Its Dialects
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2026) -
Spoken Word2Vec: Learning Skipgram Embeddings from Speech
von: Sayeed, Mohammad Amaan, et al.
Veröffentlicht: (2023) -
Cultural Benchmarking of LLMs in Standard and Dialectal Arabic Dialogues
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2026)