Automated Quality Control for Language Documentation: Detecting Phonotactic Inconsistencies in a Kokborok Wordlist
Fuente:
arXiv
Saved in:
| Main Authors: | van Dam, Kellen Parker, Stephen, Abishek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unstable Grounds for Beautiful Trees? Testing the Robustness of Concept Translations in the Compilation of Multilingual Wordlists
by: Snee, David, et al.
Published: (2025)
by: Snee, David, et al.
Published: (2025)
Annotating and Inferring Compositional Structures in Numeral Systems Across Languages
by: Rubehn, Arne, et al.
Published: (2025)
by: Rubehn, Arne, et al.
Published: (2025)
Phonotactic Complexity across Dialects
by: Shim, Ryan Soh-Eun, et al.
Published: (2024)
by: Shim, Ryan Soh-Eun, et al.
Published: (2024)
Towards High-Quality Machine Translation for Kokborok: A Low-Resource Tibeto-Burman Language of Northeast India
by: Nyalang, Badal, et al.
Published: (2026)
by: Nyalang, Badal, et al.
Published: (2026)
Learning Phonotactics from Linguistic Informants
by: Breiss, Canaan, et al.
Published: (2024)
by: Breiss, Canaan, et al.
Published: (2024)
Self-Supervised Borrowing Detection on Multilingual Wordlists
by: Wientzek, Tim
Published: (2025)
by: Wientzek, Tim
Published: (2025)
Evaluating Morphological Plausibility of Subword Tokenization via Statistical Alignment with Morpho-Syntactic Features
by: Stephen, Abishek, et al.
Published: (2026)
by: Stephen, Abishek, et al.
Published: (2026)
Trimming Phonetic Alignments Improves the Inference of Sound Correspondence Patterns from Multilingual Wordlists
by: Blum, Frederic, et al.
Published: (2023)
by: Blum, Frederic, et al.
Published: (2023)
On Finding Inconsistencies in Documents
by: Lovering, Charles J., et al.
Published: (2025)
by: Lovering, Charles J., et al.
Published: (2025)
FIZZ: Factual Inconsistency Detection by Zoom-in Summary and Zoom-out Document
by: Yang, Joonho, et al.
Published: (2024)
by: Yang, Joonho, et al.
Published: (2024)
Misleading through Inconsistency: A Benchmark for Political Inconsistencies Detection
by: Sagimbayeva, Nursulu, et al.
Published: (2025)
by: Sagimbayeva, Nursulu, et al.
Published: (2025)
Improved Evidence Extraction and Metrics for Document Inconsistency Detection with LLMs
by: Tan, Nelvin, et al.
Published: (2026)
by: Tan, Nelvin, et al.
Published: (2026)
Fast and Accurate Factual Inconsistency Detection Over Long Documents
by: Lattimer, Barrett Martin, et al.
Published: (2023)
by: Lattimer, Barrett Martin, et al.
Published: (2023)
Multitask Fine-Tuning and Generative Adversarial Learning for Improved Auxiliary Classification
by: Sun, Christopher, et al.
Published: (2024)
by: Sun, Christopher, et al.
Published: (2024)
Detecting Corpus-Level Knowledge Inconsistencies in Wikipedia with Large Language Models
by: Semnani, Sina J., et al.
Published: (2025)
by: Semnani, Sina J., et al.
Published: (2025)
How communicatively optimal are exact numeral systems? Once more on lexicon size and morphosyntactic complexity
by: Cathcart, Chundra, et al.
Published: (2026)
by: Cathcart, Chundra, et al.
Published: (2026)
DRIFT: Detecting Representational Inconsistencies for Factual Truthfulness
by: Bhatnagar, Rohan, et al.
Published: (2026)
by: Bhatnagar, Rohan, et al.
Published: (2026)
Inconsistencies in Masked Language Models
by: Young, Tom, et al.
Published: (2022)
by: Young, Tom, et al.
Published: (2022)
Human-like fleeting memory improves language learning but impairs reading time prediction in transformer language models
by: Thamma, Abishek, et al.
Published: (2025)
by: Thamma, Abishek, et al.
Published: (2025)
Measuring the Inconsistency of Large Language Models in Preferential Ranking
by: Zhao, Xiutian, et al.
Published: (2024)
by: Zhao, Xiutian, et al.
Published: (2024)
Measuring Moral Inconsistencies in Large Language Models
by: Bonagiri, Vamshi Krishna, et al.
Published: (2024)
by: Bonagiri, Vamshi Krishna, et al.
Published: (2024)
Investigating the Performance of Language Models for Completing Code in Functional Programming Languages: a Haskell Case Study
by: van Dam, Tim, et al.
Published: (2024)
by: van Dam, Tim, et al.
Published: (2024)
PrefixNLI: Detecting Factual Inconsistencies as Soon as They Arise
by: Harary, Sapir, et al.
Published: (2025)
by: Harary, Sapir, et al.
Published: (2025)
Towards Data Contamination Detection for Modern Large Language Models: Limitations, Inconsistencies, and Oracle Challenges
by: Samuel, Vinay, et al.
Published: (2024)
by: Samuel, Vinay, et al.
Published: (2024)
Discourse-Driven Evaluation: Unveiling Factual Inconsistency in Long Document Summarization
by: Zhong, Yang, et al.
Published: (2025)
by: Zhong, Yang, et al.
Published: (2025)
SynLexLM: Scaling Legal LLMs with Synthetic Data and Curriculum Learning
by: Upadhyay, Ojasw, et al.
Published: (2025)
by: Upadhyay, Ojasw, et al.
Published: (2025)
Addressing Tokenization Inconsistency in Steganography and Watermarking Based on Large Language Models
by: Yan, Ruiyi, et al.
Published: (2025)
by: Yan, Ruiyi, et al.
Published: (2025)
Evaluating Knowledge-based Cross-lingual Inconsistency in Large Language Models
by: Xing, Xiaolin, et al.
Published: (2024)
by: Xing, Xiaolin, et al.
Published: (2024)
Human-like Linguistic Biases in Neural Speech Models: Phonetic Categorization and Phonotactic Constraints in Wav2Vec2.0
by: Kloots, Marianne de Heer, et al.
Published: (2024)
by: Kloots, Marianne de Heer, et al.
Published: (2024)
Prompt-Reverse Inconsistency: LLM Self-Inconsistency Beyond Generative Randomness and Prompt Paraphrasing
by: Ahn, Jihyun Janice, et al.
Published: (2025)
by: Ahn, Jihyun Janice, et al.
Published: (2025)
SIFiD: Reassess Summary Factual Inconsistency Detection with LLM
by: Yang, Jiuding, et al.
Published: (2024)
by: Yang, Jiuding, et al.
Published: (2024)
Legal Document Summarization: Enhancing Judicial Efficiency through Automation Detection
by: Li, Yongjie, et al.
Published: (2025)
by: Li, Yongjie, et al.
Published: (2025)
Lost in Multilinguality: Dissecting Cross-lingual Factual Inconsistency in Transformer Language Models
by: Wang, Mingyang, et al.
Published: (2025)
by: Wang, Mingyang, et al.
Published: (2025)
Multimodal Language Models Cannot Spot Spatial Inconsistencies
by: Khangaonkar, Om, et al.
Published: (2026)
by: Khangaonkar, Om, et al.
Published: (2026)
Unraveling and Mitigating Retriever Inconsistencies in Retrieval-Augmented Large Language Models
by: Li, Mingda, et al.
Published: (2024)
by: Li, Mingda, et al.
Published: (2024)
Localizing Factual Inconsistencies in Attributable Text Generation
by: Cattan, Arie, et al.
Published: (2024)
by: Cattan, Arie, et al.
Published: (2024)
Losing Phonotactic Distinctions in Context
by: John R. Starr, et al.
Published: (2025)
by: John R. Starr, et al.
Published: (2025)
Large Language Models for Document-Level Event-Argument Data Augmentation for Challenging Role Types
by: Gatto, Joseph, et al.
Published: (2024)
by: Gatto, Joseph, et al.
Published: (2024)
Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency
by: Liu, Yiran, et al.
Published: (2024)
by: Liu, Yiran, et al.
Published: (2024)
Uncovering Misattributed Suicide Causes through Annotation Inconsistency Detection in Death Investigation Notes
by: Wang, Song, et al.
Published: (2024)
by: Wang, Song, et al.
Published: (2024)
Similar Items
-
Unstable Grounds for Beautiful Trees? Testing the Robustness of Concept Translations in the Compilation of Multilingual Wordlists
by: Snee, David, et al.
Published: (2025) -
Annotating and Inferring Compositional Structures in Numeral Systems Across Languages
by: Rubehn, Arne, et al.
Published: (2025) -
Phonotactic Complexity across Dialects
by: Shim, Ryan Soh-Eun, et al.
Published: (2024) -
Towards High-Quality Machine Translation for Kokborok: A Low-Resource Tibeto-Burman Language of Northeast India
by: Nyalang, Badal, et al.
Published: (2026) -
Learning Phonotactics from Linguistic Informants
by: Breiss, Canaan, et al.
Published: (2024)