CWoMP: Morpheme Representation Learning for Interlinear Glossing
Fuente:
arXiv
Saved in:
| Main Authors: | Alper, Morris, Rice, Enora, Shandilya, Bhargav, Palmer, Alexis, Levin, Lori |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GlossLM: A Massively Multilingual Corpus and Pretrained Model for Interlinear Glossed Text
by: Ginn, Michael, et al.
Published: (2024)
by: Ginn, Michael, et al.
Published: (2024)
Wav2Gloss: Generating Interlinear Glossed Text from Speech
by: He, Taiqi, et al.
Published: (2024)
by: He, Taiqi, et al.
Published: (2024)
Robust Generalization Strategies for Morpheme Glossing in an Endangered Language Documentation Context
by: Ginn, Michael, et al.
Published: (2023)
by: Ginn, Michael, et al.
Published: (2023)
Boosting the Capabilities of Compact Models in Low-Data Contexts with Large Language Models and Retrieval-Augmented Generation
by: Shandilya, Bhargav, et al.
Published: (2024)
by: Shandilya, Bhargav, et al.
Published: (2024)
Massively Multilingual Joint Segmentation and Glossing
by: Ginn, Michael, et al.
Published: (2026)
by: Ginn, Michael, et al.
Published: (2026)
Interdisciplinary Research in Conversation: A Case Study in Computational Morphology for Language Documentation
by: Rice, Enora, et al.
Published: (2025)
by: Rice, Enora, et al.
Published: (2025)
Untangling the Influence of Typology, Data and Model Architecture on Ranking Transfer Languages for Cross-Lingual POS Tagging
by: Rice, Enora, et al.
Published: (2025)
by: Rice, Enora, et al.
Published: (2025)
TAMS: Translation-Assisted Morphological Segmentation
by: Rice, Enora, et al.
Published: (2024)
by: Rice, Enora, et al.
Published: (2024)
From Priest to Doctor: Domain Adaptation for Low-Resource Neural Machine Translation
by: Marashian, Ali, et al.
Published: (2024)
by: Marashian, Ali, et al.
Published: (2024)
Emergent Visual-Semantic Hierarchies in Image-Text Representations
by: Alper, Morris, et al.
Published: (2024)
by: Alper, Morris, et al.
Published: (2024)
Using Contextual Information for Sentence-level Morpheme Segmentation
by: Bhandari, Prabin, et al.
Published: (2024)
by: Bhandari, Prabin, et al.
Published: (2024)
Learning Beyond Limits: Multitask Learning and Synthetic Data for Low-Resource Canonical Morpheme Segmentation
by: Yang, Changbing, et al.
Published: (2025)
by: Yang, Changbing, et al.
Published: (2025)
A Spatio-Temporal Representation Learning as an Alternative to Traditional Glosses in Sign Language Translation and Production
by: Hwang, Eui Jun, et al.
Published: (2024)
by: Hwang, Eui Jun, et al.
Published: (2024)
Morpheme Induction for Emergent Language
by: Boldt, Brendon, et al.
Published: (2025)
by: Boldt, Brendon, et al.
Published: (2025)
Selective Contrastive Learning For Gloss Free Sign Language Translation
by: Lai, Changhao, et al.
Published: (2026)
by: Lai, Changhao, et al.
Published: (2026)
Improving Gloss-free Sign Language Translation by Reducing Representation Density
by: Ye, Jinhui, et al.
Published: (2024)
by: Ye, Jinhui, et al.
Published: (2024)
C${^2}$RL: Content and Context Representation Learning for Gloss-free Sign Language Translation and Retrieval
by: Chen, Zhigang, et al.
Published: (2024)
by: Chen, Zhigang, et al.
Published: (2024)
Embedded Translations for Low-resource Automated Glossing
by: Yang, Changbing, et al.
Published: (2024)
by: Yang, Changbing, et al.
Published: (2024)
ViConBERT: Context-Gloss Aligned Vietnamese Word Embedding for Polysemous and Sense-Aware Representations
by: Huynh, Khang T., et al.
Published: (2025)
by: Huynh, Khang T., et al.
Published: (2025)
Kiki or Bouba? Sound Symbolism in Vision-and-Language Models
by: Alper, Morris, et al.
Published: (2023)
by: Alper, Morris, et al.
Published: (2023)
Gloss2Text: Sign Language Gloss translation using LLMs and Semantically Aware Label Smoothing
by: Fayyazsanavi, Pooya, et al.
Published: (2024)
by: Fayyazsanavi, Pooya, et al.
Published: (2024)
Factorized Learning Assisted with Large Language Model for Gloss-free Sign Language Translation
by: Chen, Zhigang, et al.
Published: (2024)
by: Chen, Zhigang, et al.
Published: (2024)
The Morphemic Origin of Zipf's Law: A Factorized Combinatorial Framework
by: Berman, Vladimir
Published: (2025)
by: Berman, Vladimir
Published: (2025)
From Smør-re-brød to Subwords: Training LLMs on Danish, One Morpheme at a Time
by: Kildeberg, Mikkel Wildner, et al.
Published: (2025)
by: Kildeberg, Mikkel Wildner, et al.
Published: (2025)
Morphemes Without Borders: Evaluating Root-Pattern Morphology in Arabic Tokenizers and LLMs
by: Alakeel, Yara, et al.
Published: (2026)
by: Alakeel, Yara, et al.
Published: (2026)
ConlangCrafter: Constructing Languages with a Multi-Hop LLM Pipeline
by: Alper, Morris, et al.
Published: (2025)
by: Alper, Morris, et al.
Published: (2025)
Introduction and Analysis of an Interlinear Qur'an Translation with Unknown Old Anatolian Turkish
by: Süleyman Aksu, et al.
Published: (2023)
by: Süleyman Aksu, et al.
Published: (2023)
Leveraging the Power of MLLMs for Gloss-Free Sign Language Translation
by: Kim, Jungeun, et al.
Published: (2024)
by: Kim, Jungeun, et al.
Published: (2024)
Morpheme Boundary Detection & Grammatical Feature Prediction for Gujarati : Dataset & Model
by: Baxi, Jatayu, et al.
Published: (2021)
by: Baxi, Jatayu, et al.
Published: (2021)
Neural Induction of Finite-State Transducers
by: Ginn, Michael, et al.
Published: (2026)
by: Ginn, Michael, et al.
Published: (2026)
Is linguistically-motivated data augmentation worth it?
by: Groshan, Ray, et al.
Published: (2025)
by: Groshan, Ray, et al.
Published: (2025)
Can we teach language models to gloss endangered languages?
by: Ginn, Michael, et al.
Published: (2024)
by: Ginn, Michael, et al.
Published: (2024)
Multiple Sources are Better Than One: Incorporating External Knowledge in Low-Resource Glossing
by: Yang, Changbing, et al.
Published: (2024)
by: Yang, Changbing, et al.
Published: (2024)
Gloss-Free Sign Language Translation: An Unbiased Evaluation of Progress in the Field
by: Sincan, Ozge Mercanoglu, et al.
Published: (2026)
by: Sincan, Ozge Mercanoglu, et al.
Published: (2026)
GLOS: Sign Language Generation with Temporally Aligned Gloss-Level Conditioning
by: Lee, Taeryung, et al.
Published: (2025)
by: Lee, Taeryung, et al.
Published: (2025)
Breaking the Silence: A Dataset and Benchmark for Bangla Text-to-Gloss Translation
by: Abdullah, Sharif Mohammad, et al.
Published: (2025)
by: Abdullah, Sharif Mohammad, et al.
Published: (2025)
Multilingual Gloss-free Sign Language Translation: Towards Building a Sign Language Foundation Model
by: Tan, Sihan, et al.
Published: (2025)
by: Tan, Sihan, et al.
Published: (2025)
Phonikud: Hebrew Grapheme-to-Phoneme Conversion for Real-Time Text-to-Speech
by: Kolani, Yakov, et al.
Published: (2025)
by: Kolani, Yakov, et al.
Published: (2025)
Enhancing Structured Meaning Representations with Aspect Classification
by: Post, Claire Benét, et al.
Published: (2026)
by: Post, Claire Benét, et al.
Published: (2026)
Leveraging Long-Context Large Language Models for Multi-Document Understanding and Summarization in Enterprise Applications
by: Godbole, Aditi, et al.
Published: (2024)
by: Godbole, Aditi, et al.
Published: (2024)
Similar Items
-
GlossLM: A Massively Multilingual Corpus and Pretrained Model for Interlinear Glossed Text
by: Ginn, Michael, et al.
Published: (2024) -
Wav2Gloss: Generating Interlinear Glossed Text from Speech
by: He, Taiqi, et al.
Published: (2024) -
Robust Generalization Strategies for Morpheme Glossing in an Endangered Language Documentation Context
by: Ginn, Michael, et al.
Published: (2023) -
Boosting the Capabilities of Compact Models in Low-Data Contexts with Large Language Models and Retrieval-Augmented Generation
by: Shandilya, Bhargav, et al.
Published: (2024) -
Massively Multilingual Joint Segmentation and Glossing
by: Ginn, Michael, et al.
Published: (2026)