Neural Recovery of Historical Lexical Structure in Bantu Languages from Modern Data
Fuente:
arXiv
Saved in:
| Main Authors: | Mutisya, Hillary, Mugane, John |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Zero-Shot Morphological Discovery in Low-Resource Bantu Languages via Cross-Lingual Transfer and Unsupervised Clustering
by: Mutisya, Hillary, et al.
Published: (2026)
by: Mutisya, Hillary, et al.
Published: (2026)
Attention Sinks in Massively Multilingual Neural Machine Translation:Discovery, Analysis, and Mitigation
by: Mutisya, Hillary, et al.
Published: (2026)
by: Mutisya, Hillary, et al.
Published: (2026)
The Thiomi Dataset: A Large-Scale Multimodal Corpus for Low-Resource African Languages
by: Mutisya, Hillary, et al.
Published: (2026)
by: Mutisya, Hillary, et al.
Published: (2026)
Continued Pretraining for Low-Resource Swahili ASR: Achieving State-of-the-Art Performance with Minimal Labeled Data
by: Mutisya, Hillary, et al.
Published: (2026)
by: Mutisya, Hillary, et al.
Published: (2026)
Model Internal Sleuthing: Finding Lexical Identity and Inflectional Features in Modern Language Models
by: Li, Michael, et al.
Published: (2025)
by: Li, Michael, et al.
Published: (2025)
On Lexical Invariance on Multisets and Graphs
by: Zhang, Muhan
Published: (2024)
by: Zhang, Muhan
Published: (2024)
Meta-Tuning LLMs to Leverage Lexical Knowledge for Generalizable Language Style Understanding
by: Guo, Ruohao, et al.
Published: (2023)
by: Guo, Ruohao, et al.
Published: (2023)
Why is "Chicago" Predictive of Deceptive Reviews? Using LLMs to Discover Language Phenomena from Lexical Cues
by: Qu, Jiaming, et al.
Published: (2025)
by: Qu, Jiaming, et al.
Published: (2025)
Multilingual Lexical Feature Analysis of Spoken Language for Predicting Major Depression Symptom Severity
by: Tokareva, Anastasiia, et al.
Published: (2025)
by: Tokareva, Anastasiia, et al.
Published: (2025)
Lexical Hints of Accuracy in LLM Reasoning Chains
by: Vanhoyweghen, Arne, et al.
Published: (2025)
by: Vanhoyweghen, Arne, et al.
Published: (2025)
Luxical: High-Speed Lexical-Dense Text Embeddings
by: DatologyAI, et al.
Published: (2025)
by: DatologyAI, et al.
Published: (2025)
Lexical Tone is Hard to Quantize: Probing Discrete Speech Units in Mandarin and Yorùbá
by: Osakuade, Opeyemi, et al.
Published: (2026)
by: Osakuade, Opeyemi, et al.
Published: (2026)
Unified Lexical Representation for Interpretable Visual-Language Alignment
by: Li, Yifan, et al.
Published: (2024)
by: Li, Yifan, et al.
Published: (2024)
Transfer Learning via Lexical Relatedness: A Sarcasm and Hate Speech Case Study
by: Cabrera, Angelly, et al.
Published: (2025)
by: Cabrera, Angelly, et al.
Published: (2025)
Structural Pruning of Pre-trained Language Models via Neural Architecture Search
by: Klein, Aaron, et al.
Published: (2024)
by: Klein, Aaron, et al.
Published: (2024)
Zipf Distributions from Two-Stage Symbolic Processes: Stability Under Stochastic Lexical Filtering
by: Berman, Vladimir
Published: (2025)
by: Berman, Vladimir
Published: (2025)
Phonetic and Lexical Discovery of a Canine Language using HuBERT
by: Li, Xingyuan, et al.
Published: (2024)
by: Li, Xingyuan, et al.
Published: (2024)
How Does a Deep Neural Network Look at Lexical Stress in English Words?
by: Allouche, Itai, et al.
Published: (2025)
by: Allouche, Itai, et al.
Published: (2025)
Why Does ChatGPT "Delve" So Much? Exploring the Sources of Lexical Overrepresentation in Large Language Models
by: Juzek, Tom S., et al.
Published: (2024)
by: Juzek, Tom S., et al.
Published: (2024)
Shared Lexical Task Representations Explain Behavioral Variability In LLMs
by: Yang, Zhuonan, et al.
Published: (2026)
by: Yang, Zhuonan, et al.
Published: (2026)
Sensitivity of Generative VLMs to Semantically and Lexically Altered Prompts
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Confidence-Aware Sub-Structure Beam Search (CABS): Mitigating Hallucination in Structured Data Generation with Large Language Models
by: Wei, Chengwei, et al.
Published: (2024)
by: Wei, Chengwei, et al.
Published: (2024)
Do Lexical and Contextual Coreference Resolution Systems Degrade Differently under Mention Noise? An Empirical Study on Scientific Software Mentions
by: Alkan, Atilla Kaan, et al.
Published: (2026)
by: Alkan, Atilla Kaan, et al.
Published: (2026)
Patent Language Model Pretraining with ModernBERT
by: Yousefiramandi, Amirhossein, et al.
Published: (2025)
by: Yousefiramandi, Amirhossein, et al.
Published: (2025)
GNN: Graph Neural Network and Large Language Model for Data Discovery
by: Hoang, Thomas
Published: (2024)
by: Hoang, Thomas
Published: (2024)
Transfer of Structural Knowledge from Synthetic Languages
by: Budnikov, Mikhail, et al.
Published: (2025)
by: Budnikov, Mikhail, et al.
Published: (2025)
Automatic Classification of Arabic Literature into Historical Eras
by: Alhathloul, Zainab, et al.
Published: (2026)
by: Alhathloul, Zainab, et al.
Published: (2026)
Neural Proto-Language Reconstruction
by: Cui, Chenxuan, et al.
Published: (2024)
by: Cui, Chenxuan, et al.
Published: (2024)
Anatomy of Neural Language Models
by: Saleh, Majd, et al.
Published: (2024)
by: Saleh, Majd, et al.
Published: (2024)
Recover-LoRA: Data-Free Accuracy Recovery of Degraded Language Models via Low-Rank Adaptation
by: Das, Devleena, et al.
Published: (2025)
by: Das, Devleena, et al.
Published: (2025)
LMD3: Language Model Data Density Dependence
by: Kirchenbauer, John, et al.
Published: (2024)
by: Kirchenbauer, John, et al.
Published: (2024)
Structure-Aware RAG: Structured Retrieval Augmented Generation from Noisy Data for Conversational Agents
by: Han, Kaiqiao, et al.
Published: (2026)
by: Han, Kaiqiao, et al.
Published: (2026)
Impacts of Racial Bias in Historical Training Data for News AI
by: Bhargava, Rahul, et al.
Published: (2025)
by: Bhargava, Rahul, et al.
Published: (2025)
LLaSA: Large Language and Structured Data Assistant
by: Xu, Yao, et al.
Published: (2024)
by: Xu, Yao, et al.
Published: (2024)
On the Representational Capacity of Recurrent Neural Language Models
by: Nowak, Franz, et al.
Published: (2023)
by: Nowak, Franz, et al.
Published: (2023)
Training Neural Networks as Recognizers of Formal Languages
by: Butoi, Alexandra, et al.
Published: (2024)
by: Butoi, Alexandra, et al.
Published: (2024)
mmBERT: A Modern Multilingual Encoder with Annealed Language Learning
by: Marone, Marc, et al.
Published: (2025)
by: Marone, Marc, et al.
Published: (2025)
A Comprehensive Evaluation of Neural SPARQL Query Generation from Natural Language Questions
by: Diallo, Papa Abdou Karim Karou, et al.
Published: (2023)
by: Diallo, Papa Abdou Karim Karou, et al.
Published: (2023)
LANCET: Neural Intervention via Structural Entropy for Mitigating Faithfulness Hallucinations in LLMs
by: Wang, Chenxu, et al.
Published: (2026)
by: Wang, Chenxu, et al.
Published: (2026)
Advancing Regular Language Reasoning in Linear Recurrent Neural Networks
by: Fan, Ting-Han, et al.
Published: (2023)
by: Fan, Ting-Han, et al.
Published: (2023)
Similar Items
-
Zero-Shot Morphological Discovery in Low-Resource Bantu Languages via Cross-Lingual Transfer and Unsupervised Clustering
by: Mutisya, Hillary, et al.
Published: (2026) -
Attention Sinks in Massively Multilingual Neural Machine Translation:Discovery, Analysis, and Mitigation
by: Mutisya, Hillary, et al.
Published: (2026) -
The Thiomi Dataset: A Large-Scale Multimodal Corpus for Low-Resource African Languages
by: Mutisya, Hillary, et al.
Published: (2026) -
Continued Pretraining for Low-Resource Swahili ASR: Achieving State-of-the-Art Performance with Minimal Labeled Data
by: Mutisya, Hillary, et al.
Published: (2026) -
Model Internal Sleuthing: Finding Lexical Identity and Inflectional Features in Modern Language Models
by: Li, Michael, et al.
Published: (2025)