Evidence for a Structured Token-Generation System in the Voynich Manuscript

Fuente: Zenodo
Guardado en:
Detalles Bibliográficos
Autor principal: Chang, Youngsan
Formato: Recurso digital
Lenguaje:inglés
Publicado: Zenodo 2026
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866901476927864832
author Chang, Youngsan
author_facet Chang, Youngsan
contents <p>This record provides the preprint and reproduction package for the study "Evidence for a Structured Token-Generation System in the Voynich Manuscript."</p> <p>The study tests whether Voynich Manuscript tokens can be modeled by a structured prefix–core–suffix (PCS) token-generation system rather than by random processes. Using the Zandbergen–Landini EVA transcription (ZL3b), the analysis evaluates token matching, coverage, Zipf distribution alignment, fixed-length validation, bootstrap significance, holdout generalization, inter-token transition entropy, morphological family structure, positional dependency, and compression-based structural regularity.</p> <p>Key results include:</p> <p>- Matching Rate: PCS 97.02% vs Random 5.07%<br>- Token Coverage: PCS 85.25% vs Random 18.50%<br>- Zipf Slope Difference: PCS 0.0795 vs Random 0.7138<br>- H(suffix | core): 0.8746<br>- Transition entropy reduction: 0.2504 bits<br>- Compression ratios: real 0.3089, shuffled 0.3240, PCS-generated 0.3163<br>- Morphological families: 767 core-sharing families and 1,017 suffix-alternating clusters<br>- Positional chi-square: prefix = 4554.31, core = 17993.46, suffix = 3329.07</p> <p>The results support a structured token-generation mechanism but do not constitute semantic decipherment.</p>
format Recurso digital
id zenodo_https___doi_org_10_5281_zenodo_19858095
institution Zenodo
language eng
publishDate 2026
publisher Zenodo
record_format zenodo
spellingShingle Evidence for a Structured Token-Generation System in the Voynich Manuscript
Chang, Youngsan
Voynich Manuscript
prefix-core-suffix
generative model
morphological decomposition
token structure
Zipf distribution
compression analysis
computational linguistics
unknown scripts
<p>This record provides the preprint and reproduction package for the study "Evidence for a Structured Token-Generation System in the Voynich Manuscript."</p> <p>The study tests whether Voynich Manuscript tokens can be modeled by a structured prefix–core–suffix (PCS) token-generation system rather than by random processes. Using the Zandbergen–Landini EVA transcription (ZL3b), the analysis evaluates token matching, coverage, Zipf distribution alignment, fixed-length validation, bootstrap significance, holdout generalization, inter-token transition entropy, morphological family structure, positional dependency, and compression-based structural regularity.</p> <p>Key results include:</p> <p>- Matching Rate: PCS 97.02% vs Random 5.07%<br>- Token Coverage: PCS 85.25% vs Random 18.50%<br>- Zipf Slope Difference: PCS 0.0795 vs Random 0.7138<br>- H(suffix | core): 0.8746<br>- Transition entropy reduction: 0.2504 bits<br>- Compression ratios: real 0.3089, shuffled 0.3240, PCS-generated 0.3163<br>- Morphological families: 767 core-sharing families and 1,017 suffix-alternating clusters<br>- Positional chi-square: prefix = 4554.31, core = 17993.46, suffix = 3329.07</p> <p>The results support a structured token-generation mechanism but do not constitute semantic decipherment.</p>
title Evidence for a Structured Token-Generation System in the Voynich Manuscript
topic Voynich Manuscript
prefix-core-suffix
generative model
morphological decomposition
token structure
Zipf distribution
compression analysis
computational linguistics
unknown scripts
url https://doi.org/10.5281/zenodo.19858095