Document indexing with a concept hierarchy

Fuente: Redalyc
Guardado en:
Detalles Bibliográficos
Autor principal: Alexander Gelbukh
Formato: Artículo científico
Lenguaje:en
Publicado: Instituto Politécnico Nacional 2005
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1876425025420001280
author Alexander Gelbukh
author_facet Alexander Gelbukh
contents Document indexing with a concept hierarchy Alexander Gelbukh Grigori Sidorov Adolfo Guzmán-Arenas Computación Ontology Document Comparison Statistical Methods Document Characterization Given a large hierarchical concept dictionary (thesaurus, or ontology), the task of selection of the concepts that describe the contents of a given document is considered. A statistical method of document indexing driven by such a dictionary is proposed. The method is insensible to inaccuracies in the dictionary, which allow for semi-automatic translation of the hierarchy into different languages. The problem of handling non-terminal and especially top-level nodes in the hierarchy is discussed. Common sense-complaint methods of automatically assigning the weights to the nodes and links in the hierarchy are presented. The application of the method in the Classifier system is discussed. 2005 artículo científico 1405-5546 https://www.redalyc.org/articulo.oa?id=61580403 en http://www.redalyc.org/revista.oa?id=615 Computación y Sistemas application/pdf Instituto Politécnico Nacional Computación y Sistemas (México) Num.4 Vol.8
format Artículo científico
id redalyc_61580403
institution Redalyc
language en
publishDate 2005
publisher Instituto Politécnico Nacional
spellingShingle Document indexing with a concept hierarchy
Alexander Gelbukh
Computación
Ontology
Document Comparison
Statistical Methods
Document Characterization
Document indexing with a concept hierarchy Alexander Gelbukh Grigori Sidorov Adolfo Guzmán-Arenas Computación Ontology Document Comparison Statistical Methods Document Characterization Given a large hierarchical concept dictionary (thesaurus, or ontology), the task of selection of the concepts that describe the contents of a given document is considered. A statistical method of document indexing driven by such a dictionary is proposed. The method is insensible to inaccuracies in the dictionary, which allow for semi-automatic translation of the hierarchy into different languages. The problem of handling non-terminal and especially top-level nodes in the hierarchy is discussed. Common sense-complaint methods of automatically assigning the weights to the nodes and links in the hierarchy are presented. The application of the method in the Classifier system is discussed. 2005 artículo científico 1405-5546 https://www.redalyc.org/articulo.oa?id=61580403 en http://www.redalyc.org/revista.oa?id=615 Computación y Sistemas application/pdf Instituto Politécnico Nacional Computación y Sistemas (México) Num.4 Vol.8
title Document indexing with a concept hierarchy
topic Computación
Ontology
Document Comparison
Statistical Methods
Document Characterization
url https://www.redalyc.org/articulo.oa?id=61580403