LLMs for Low-Resource Dialect Translation Using Context-Aware Prompting: A Case Study on Sylheti
Fuente:
arXiv
Saved in:
| Main Authors: | Prama, Tabia Tanzin, Danforth, Christopher M., Dodds, Peter Sheridan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Misalignment of LLM-Generated Personas with Human Perceptions in Low-Resource Settings
by: Prama, Tabia Tanzin, et al.
Published: (2025)
by: Prama, Tabia Tanzin, et al.
Published: (2025)
Transformer-Based Low-Resource Language Translation: A Study on Standard Bengali to Sylheti
by: Oni, Mangsura Kabir, et al.
Published: (2025)
by: Oni, Mangsura Kabir, et al.
Published: (2025)
BanglaMATH : A Bangla benchmark dataset for testing LLM mathematical reasoning at grades 6, 7, and 8
by: Prama, Tabia Tanzin, et al.
Published: (2025)
by: Prama, Tabia Tanzin, et al.
Published: (2025)
Story and essential meaning dynamics in Bangladesh's July 2024 Student-People's Uprising
by: Prama, Tabia Tanzin, et al.
Published: (2025)
by: Prama, Tabia Tanzin, et al.
Published: (2025)
Us-vs-Them bias in Large Language Models
by: Prama, Tabia Tanzin, et al.
Published: (2025)
by: Prama, Tabia Tanzin, et al.
Published: (2025)
Global brain drain and gain in high-potential student mobility
by: Pramaa, Tabia Tanzin, et al.
Published: (2026)
by: Pramaa, Tabia Tanzin, et al.
Published: (2026)
Curating corpora with classifiers: A case study of clean energy sentiment online
by: Arnold, Michael V., et al.
Published: (2023)
by: Arnold, Michael V., et al.
Published: (2023)
AI Enabled User-Specific Cyberbullying Severity Detection with Explainability
by: Prama, Tabia Tanzin, et al.
Published: (2025)
by: Prama, Tabia Tanzin, et al.
Published: (2025)
Detecting sub-populations in online health communities: A mixed-methods exploration of breastfeeding messages in BabyCenter Birth Clubs
by: Beauregard, Calla, et al.
Published: (2025)
by: Beauregard, Calla, et al.
Published: (2025)
Hollywood's misrepresentation of death: A comparison of overall and by-gender mortality causes in film and the real world
by: Beauregard, Calla, et al.
Published: (2024)
by: Beauregard, Calla, et al.
Published: (2024)
Optimized Custom CNN for Real-Time Tomato Leaf Disease Detection
by: Oni, Mangsura Kabir, et al.
Published: (2025)
by: Oni, Mangsura Kabir, et al.
Published: (2025)
Statistical laws and linguistics inform meaning in naturalistic and fictional conversation
by: Fehr, Ashley M. A., et al.
Published: (2025)
by: Fehr, Ashley M. A., et al.
Published: (2025)
Political Biases on X before the 2025 German Federal Election
by: Prama, Tabia Tanzin, et al.
Published: (2025)
by: Prama, Tabia Tanzin, et al.
Published: (2025)
Toxicity-Aware Few-Shot Prompting for Low-Resource Singlish Translation
by: Ge, Ziyu, et al.
Published: (2025)
by: Ge, Ziyu, et al.
Published: (2025)
LLM-Based Evaluation of Low-Resource Machine Translation: A Reference-less Dialect Guided Approach with a Refined Sylheti-English Benchmark
by: Rahman, Md. Atiqur, et al.
Published: (2025)
by: Rahman, Md. Atiqur, et al.
Published: (2025)
Complete asymptotic type-token relationship for growing complex systems with inverse power-law count rankings
by: Rosillo-Rodes, Pablo, et al.
Published: (2025)
by: Rosillo-Rodes, Pablo, et al.
Published: (2025)
Fine-Tuning LLMs for Low-Resource Dialect Translation: The Case of Lebanese
by: Yakhni, Silvana, et al.
Published: (2025)
by: Yakhni, Silvana, et al.
Published: (2025)
Vuyko Mistral: Adapting LLMs for Low-Resource Dialectal Translation
by: Kyslyi, Roman, et al.
Published: (2025)
by: Kyslyi, Roman, et al.
Published: (2025)
Archetypes and gender in fiction: A data-driven mapping of gender stereotypes in stories
by: Beauregard, Calla Glavin, et al.
Published: (2026)
by: Beauregard, Calla Glavin, et al.
Published: (2026)
Dialectal and Low-Resource Machine Translation for Aromanian
by: Jerpelea, Alexandru-Iulius, et al.
Published: (2024)
by: Jerpelea, Alexandru-Iulius, et al.
Published: (2024)
Prompt, Translate, Fine-Tune, Re-Initialize, or Instruction-Tune? Adapting LLMs for In-Context Learning in Low-Resource Languages
by: Toukmaji, Christopher, et al.
Published: (2025)
by: Toukmaji, Christopher, et al.
Published: (2025)
Assessing the Pedagogical Readiness of Large Language Models as AI Tutors in Low-Resource Contexts: A Case Study of Nepal's K-10 Curriculum
by: Acharya, Pratyush, et al.
Published: (2026)
by: Acharya, Pratyush, et al.
Published: (2026)
Prompting ChatGPT for Translation: A Comparative Analysis of Translation Brief and Persona Prompts
by: He, Sui
Published: (2024)
by: He, Sui
Published: (2024)
Aim High, Stay Private: Differentially Private Synthetic Data Enables Public Release of Behavioral Health Information with High Utility
by: Ghasemizade, Mohsen, et al.
Published: (2025)
by: Ghasemizade, Mohsen, et al.
Published: (2025)
A Formative Study of Brief Affective Text as a Complement to Wearable Sensing for Longitudinal Student Health Monitoring
by: Harry, Tamunotonye, et al.
Published: (2026)
by: Harry, Tamunotonye, et al.
Published: (2026)
Taste for Privacy: How Context, Identity, and Lived-Experience Shape Information Sharing Preferences
by: Lovato, Juniper, et al.
Published: (2026)
by: Lovato, Juniper, et al.
Published: (2026)
Low-Resource Dialect Adaptation of Large Language Models: A French Dialect Case-Study
by: Khan, Eeham, et al.
Published: (2025)
by: Khan, Eeham, et al.
Published: (2025)
Understanding In-Context Machine Translation for Low-Resource Languages: A Case Study on Manchu
by: Pei, Renhao, et al.
Published: (2025)
by: Pei, Renhao, et al.
Published: (2025)
Cultural Value Differences of LLMs: Prompt, Language, and Model Size
by: Zhong, Qishuai, et al.
Published: (2024)
by: Zhong, Qishuai, et al.
Published: (2024)
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination
by: Fleisig, Eve, et al.
Published: (2024)
by: Fleisig, Eve, et al.
Published: (2024)
Medical Malice: A Dataset for Context-Aware Safety in Healthcare LLMs
by: D'addario, Andrew Maranhão Ventura
Published: (2025)
by: D'addario, Andrew Maranhão Ventura
Published: (2025)
Learning from Convenience Samples: A Case Study on Fine-Tuning LLMs for Survey Non-response in the German Longitudinal Election Study
by: Holtdirk, Tobias, et al.
Published: (2025)
by: Holtdirk, Tobias, et al.
Published: (2025)
Towards New Benchmark for AI Alignment & Sentiment Analysis in Socially Important Issues: A Comparative Study of Human and LLMs in the Context of AGI
by: Bojic, Ljubisa, et al.
Published: (2025)
by: Bojic, Ljubisa, et al.
Published: (2025)
Harmful Speech Detection by Language Models Exhibits Gender-Queer Dialect Bias
by: Dorn, Rebecca, et al.
Published: (2024)
by: Dorn, Rebecca, et al.
Published: (2024)
Evaluating Machine Translation Datasets for Low-Web Data Languages: A Gendered Lens
by: Nigatu, Hellina Hailu, et al.
Published: (2025)
by: Nigatu, Hellina Hailu, et al.
Published: (2025)
Crisis-induced differences in attention towards Ukraine in Twitter 2008-2023
by: Mets, Mark, et al.
Published: (2026)
by: Mets, Mark, et al.
Published: (2026)
Can Prompting LLMs Unlock Hate Speech Detection across Languages? A Zero-shot and Few-shot Study
by: Ghorbanpour, Faeze, et al.
Published: (2025)
by: Ghorbanpour, Faeze, et al.
Published: (2025)
Prompt Refinement or Fine-tuning? Best Practices for using LLMs in Computational Social Science Tasks
by: Møller, Anders Giovanni, et al.
Published: (2024)
by: Møller, Anders Giovanni, et al.
Published: (2024)
Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selection
by: Nowshin, Afroza, et al.
Published: (2026)
by: Nowshin, Afroza, et al.
Published: (2026)
Artificial Intelligence Driven Course Generation: A Case Study Using ChatGPT
by: Rouabhia, Djaber
Published: (2024)
by: Rouabhia, Djaber
Published: (2024)
Similar Items
-
Misalignment of LLM-Generated Personas with Human Perceptions in Low-Resource Settings
by: Prama, Tabia Tanzin, et al.
Published: (2025) -
Transformer-Based Low-Resource Language Translation: A Study on Standard Bengali to Sylheti
by: Oni, Mangsura Kabir, et al.
Published: (2025) -
BanglaMATH : A Bangla benchmark dataset for testing LLM mathematical reasoning at grades 6, 7, and 8
by: Prama, Tabia Tanzin, et al.
Published: (2025) -
Story and essential meaning dynamics in Bangladesh's July 2024 Student-People's Uprising
by: Prama, Tabia Tanzin, et al.
Published: (2025) -
Us-vs-Them bias in Large Language Models
by: Prama, Tabia Tanzin, et al.
Published: (2025)