Artificially Fluent: Swahili AI Performance Benchmarks Between English-Trained and Natively-Trained Datasets
Fuente:
arXiv
Salvato in:
| Autori principali: | Jaffer, Sophie, Sayer, Simeon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Benchmarking Sociolinguistic Diversity in Swahili NLP: A Taxonomy-Guided Approach
di: Oketch, Kezia, et al.
Pubblicazione: (2025)
di: Oketch, Kezia, et al.
Pubblicazione: (2025)
Hidden Signals in Language: Inferring Sensitive Attributes from Reddit Comments Using Machine Learning
di: Agarwalla, Anay, et al.
Pubblicazione: (2026)
di: Agarwalla, Anay, et al.
Pubblicazione: (2026)
Fluent but Foreign: Even Regional LLMs Lack Cultural Alignment
di: Agarwal, Dhruv, et al.
Pubblicazione: (2025)
di: Agarwal, Dhruv, et al.
Pubblicazione: (2025)
Exploring Equality: An Investigation into Custom Loss Functions for Fairness Definitions
di: Lee, Gordon, et al.
Pubblicazione: (2025)
di: Lee, Gordon, et al.
Pubblicazione: (2025)
News Sentiment as a Predictor for American Domestic Migration
di: Lane, Benjamin, et al.
Pubblicazione: (2025)
di: Lane, Benjamin, et al.
Pubblicazione: (2025)
Machine Learning for Public Good: Predicting Urban Crime Patterns to Enhance Community Safety
di: Gupta, Sia, et al.
Pubblicazione: (2024)
di: Gupta, Sia, et al.
Pubblicazione: (2024)
SwaQuAD-24: QA Benchmark Dataset in Swahili
di: Kondoro, Alfred Malengo
Pubblicazione: (2024)
di: Kondoro, Alfred Malengo
Pubblicazione: (2024)
Towards Best Practices for Open Datasets for LLM Training
di: Baack, Stefan, et al.
Pubblicazione: (2025)
di: Baack, Stefan, et al.
Pubblicazione: (2025)
Are Models Trained on Indian Legal Data Fair?
di: Girhepuje, Sahil, et al.
Pubblicazione: (2023)
di: Girhepuje, Sahil, et al.
Pubblicazione: (2023)
Artificial Intelligence Bias on English Language Learners in Automatic Scoring
di: Guo, Shuchen, et al.
Pubblicazione: (2025)
di: Guo, Shuchen, et al.
Pubblicazione: (2025)
How Large Language Models are Designed to Hallucinate
di: Ackermann, Richard, et al.
Pubblicazione: (2025)
di: Ackermann, Richard, et al.
Pubblicazione: (2025)
Readers Prefer Outputs of AI Trained on Copyrighted Books over Expert Human Writers
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2025)
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2025)
PANORAMA: A Dataset and Benchmarks Capturing Decision Trails and Rationales in Patent Examination
di: Lim, Hyunseung, et al.
Pubblicazione: (2025)
di: Lim, Hyunseung, et al.
Pubblicazione: (2025)
Epistemological Fault Lines Between Human and Artificial Intelligence
di: Quattrociocchi, Walter, et al.
Pubblicazione: (2025)
di: Quattrociocchi, Walter, et al.
Pubblicazione: (2025)
Training in translation tools and technologies: Findings of the EMT survey 2023
di: Rothwell, Andrew, et al.
Pubblicazione: (2025)
di: Rothwell, Andrew, et al.
Pubblicazione: (2025)
Benchmark on Peer Review Toxic Detection: A Challenging Task with a New Dataset
di: Luo, Man, et al.
Pubblicazione: (2025)
di: Luo, Man, et al.
Pubblicazione: (2025)
Does Scientific Writing Converge to U.S. English? Evidence from Generative AI-Assisted Publications
di: Filimonovic, Dragan, et al.
Pubblicazione: (2025)
di: Filimonovic, Dragan, et al.
Pubblicazione: (2025)
Training LLM-based Tutors to Improve Student Learning Outcomes in Dialogues
di: Scarlatos, Alexander, et al.
Pubblicazione: (2025)
di: Scarlatos, Alexander, et al.
Pubblicazione: (2025)
Impacts of Racial Bias in Historical Training Data for News AI
di: Bhargava, Rahul, et al.
Pubblicazione: (2025)
di: Bhargava, Rahul, et al.
Pubblicazione: (2025)
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
di: Jiao, Junfeng, et al.
Pubblicazione: (2025)
di: Jiao, Junfeng, et al.
Pubblicazione: (2025)
The Cost of Perfect English: Pragmatic Flattening and the Erasure of Authorial Voice in L2 Writing Supported by GenAI
di: Liu, Ao, et al.
Pubblicazione: (2026)
di: Liu, Ao, et al.
Pubblicazione: (2026)
Unmasking and Improving Data Credibility: A Study with Datasets for Training Harmless Language Models
di: Zhu, Zhaowei, et al.
Pubblicazione: (2023)
di: Zhu, Zhaowei, et al.
Pubblicazione: (2023)
MAGPIE: Multi-Task Media-Bias Analysis Generalization for Pre-Trained Identification of Expressions
di: Horych, Tomáš, et al.
Pubblicazione: (2024)
di: Horych, Tomáš, et al.
Pubblicazione: (2024)
WorldValuesBench: A Large-Scale Benchmark Dataset for Multi-Cultural Value Awareness of Language Models
di: Zhao, Wenlong, et al.
Pubblicazione: (2024)
di: Zhao, Wenlong, et al.
Pubblicazione: (2024)
Can Grammarly and ChatGPT accelerate language change? AI-powered technologies and their impact on the English language: wordiness vs. conciseness
di: Rudnicka, Karolina
Pubblicazione: (2025)
di: Rudnicka, Karolina
Pubblicazione: (2025)
The Cambridge Law Corpus: A Dataset for Legal AI Research
di: Östling, Andreas, et al.
Pubblicazione: (2023)
di: Östling, Andreas, et al.
Pubblicazione: (2023)
From Measurement Instruments to Data: Leveraging Theory-Driven Synthetic Training Data for Classifying Social Constructs
di: Birkenmaier, Lukas, et al.
Pubblicazione: (2024)
di: Birkenmaier, Lukas, et al.
Pubblicazione: (2024)
Reinforcing Stereotypes of Anger: Emotion AI on African American Vernacular English
di: Dorn, Rebecca, et al.
Pubblicazione: (2025)
di: Dorn, Rebecca, et al.
Pubblicazione: (2025)
COMPL-AI Framework: A Technical Interpretation and LLM Benchmarking Suite for the EU Artificial Intelligence Act
di: Guldimann, Philipp, et al.
Pubblicazione: (2024)
di: Guldimann, Philipp, et al.
Pubblicazione: (2024)
Meet Your New Client: Writing Reports for AI -- Benchmarking Information Loss in Market Research Deliverables
di: Simmering, Paul F., et al.
Pubblicazione: (2025)
di: Simmering, Paul F., et al.
Pubblicazione: (2025)
How Far Are LLMs from Believable AI? A Benchmark for Evaluating the Believability of Human Behavior Simulation
di: Xiao, Yang, et al.
Pubblicazione: (2023)
di: Xiao, Yang, et al.
Pubblicazione: (2023)
A Few Good Clauses: Comparing LLMs vs Domain-Trained Small Language Models on Structured Contract Extraction
di: Lincoln, Nicole, et al.
Pubblicazione: (2026)
di: Lincoln, Nicole, et al.
Pubblicazione: (2026)
AI Slop or AI-enhancement? Student perceptions of AI-generated media for an English for Academic Purposes course
di: Woo, David James, et al.
Pubblicazione: (2026)
di: Woo, David James, et al.
Pubblicazione: (2026)
Generative Artificial Intelligence in Qualitative Research Methods: Between Hype and Risks?
di: Teixeira, Maria Couto, et al.
Pubblicazione: (2025)
di: Teixeira, Maria Couto, et al.
Pubblicazione: (2025)
What Is The Political Content in LLMs' Pre- and Post-Training Data?
di: Ceron, Tanise, et al.
Pubblicazione: (2025)
di: Ceron, Tanise, et al.
Pubblicazione: (2025)
The Last Fingerprint: How Markdown Training Shapes LLM Prose
di: Freeburg, E. M.
Pubblicazione: (2026)
di: Freeburg, E. M.
Pubblicazione: (2026)
Responsible AI for Test Equity and Quality: The Duolingo English Test as a Case Study
di: Burstein, Jill, et al.
Pubblicazione: (2024)
di: Burstein, Jill, et al.
Pubblicazione: (2024)
Conversational Alignment with Artificial Intelligence in Context
di: Sterken, Rachel Katharine, et al.
Pubblicazione: (2025)
di: Sterken, Rachel Katharine, et al.
Pubblicazione: (2025)
Value Drifts: Tracing Value Alignment During LLM Post-Training
di: Bhatia, Mehar, et al.
Pubblicazione: (2025)
di: Bhatia, Mehar, et al.
Pubblicazione: (2025)
Beyond English: Unveiling Multilingual Bias in LLM Copyright Compliance
di: Chen, Yupeng, et al.
Pubblicazione: (2025)
di: Chen, Yupeng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Benchmarking Sociolinguistic Diversity in Swahili NLP: A Taxonomy-Guided Approach
di: Oketch, Kezia, et al.
Pubblicazione: (2025) -
Hidden Signals in Language: Inferring Sensitive Attributes from Reddit Comments Using Machine Learning
di: Agarwalla, Anay, et al.
Pubblicazione: (2026) -
Fluent but Foreign: Even Regional LLMs Lack Cultural Alignment
di: Agarwal, Dhruv, et al.
Pubblicazione: (2025) -
Exploring Equality: An Investigation into Custom Loss Functions for Fairness Definitions
di: Lee, Gordon, et al.
Pubblicazione: (2025) -
News Sentiment as a Predictor for American Domestic Migration
di: Lane, Benjamin, et al.
Pubblicazione: (2025)