Document indexing with a concept hierarchy

Fuente: Redalyc
Saved in:
Bibliographic Details
Main Author: Alexander Gelbukh
Format: Artículo científico
Language:en
Published: Instituto Politécnico Nacional 2005
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1876425025420001280
author Alexander Gelbukh
author_facet Alexander Gelbukh
contents Document indexing with a concept hierarchy Alexander Gelbukh Grigori Sidorov Adolfo Guzmán-Arenas Computación Ontology Document Comparison Statistical Methods Document Characterization Given a large hierarchical concept dictionary (thesaurus, or ontology), the task of selection of the concepts that describe the contents of a given document is considered. A statistical method of document indexing driven by such a dictionary is proposed. The method is insensible to inaccuracies in the dictionary, which allow for semi-automatic translation of the hierarchy into different languages. The problem of handling non-terminal and especially top-level nodes in the hierarchy is discussed. Common sense-complaint methods of automatically assigning the weights to the nodes and links in the hierarchy are presented. The application of the method in the Classifier system is discussed. 2005 artículo científico 1405-5546 https://www.redalyc.org/articulo.oa?id=61580403 en http://www.redalyc.org/revista.oa?id=615 Computación y Sistemas application/pdf Instituto Politécnico Nacional Computación y Sistemas (México) Num.4 Vol.8
format Artículo científico
id redalyc_61580403
institution Redalyc
language en
publishDate 2005
publisher Instituto Politécnico Nacional
spellingShingle Document indexing with a concept hierarchy
Alexander Gelbukh
Computación
Ontology
Document Comparison
Statistical Methods
Document Characterization
Document indexing with a concept hierarchy Alexander Gelbukh Grigori Sidorov Adolfo Guzmán-Arenas Computación Ontology Document Comparison Statistical Methods Document Characterization Given a large hierarchical concept dictionary (thesaurus, or ontology), the task of selection of the concepts that describe the contents of a given document is considered. A statistical method of document indexing driven by such a dictionary is proposed. The method is insensible to inaccuracies in the dictionary, which allow for semi-automatic translation of the hierarchy into different languages. The problem of handling non-terminal and especially top-level nodes in the hierarchy is discussed. Common sense-complaint methods of automatically assigning the weights to the nodes and links in the hierarchy are presented. The application of the method in the Classifier system is discussed. 2005 artículo científico 1405-5546 https://www.redalyc.org/articulo.oa?id=61580403 en http://www.redalyc.org/revista.oa?id=615 Computación y Sistemas application/pdf Instituto Politécnico Nacional Computación y Sistemas (México) Num.4 Vol.8
title Document indexing with a concept hierarchy
topic Computación
Ontology
Document Comparison
Statistical Methods
Document Characterization
url https://www.redalyc.org/articulo.oa?id=61580403