Document indexing with a concept hierarchy
Fuente:
Redalyc
Guardado en:
| Autor principal: | |
|---|---|
| Formato: | Artículo científico |
| Lenguaje: | en |
| Publicado: |
Instituto Politécnico Nacional
2005
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
| _version_ | 1876425025420001280 |
|---|---|
| author | Alexander Gelbukh |
| author_facet | Alexander Gelbukh |
| contents | Document indexing with a concept hierarchy Alexander Gelbukh Grigori Sidorov Adolfo Guzmán-Arenas Computación Ontology Document Comparison Statistical Methods Document Characterization Given a large hierarchical concept dictionary (thesaurus, or ontology), the task of selection of the concepts that describe the contents of a given document is considered. A statistical method of document indexing driven by such a dictionary is proposed. The method is insensible to inaccuracies in the dictionary, which allow for semi-automatic translation of the hierarchy into different languages. The problem of handling non-terminal and especially top-level nodes in the hierarchy is discussed. Common sense-complaint methods of automatically assigning the weights to the nodes and links in the hierarchy are presented. The application of the method in the Classifier system is discussed. 2005 artículo científico 1405-5546 https://www.redalyc.org/articulo.oa?id=61580403 en http://www.redalyc.org/revista.oa?id=615 Computación y Sistemas application/pdf Instituto Politécnico Nacional Computación y Sistemas (México) Num.4 Vol.8 |
| format | Artículo científico |
| id | redalyc_61580403 |
| institution | Redalyc |
| language | en |
| publishDate | 2005 |
| publisher | Instituto Politécnico Nacional |
| spellingShingle | Document indexing with a concept hierarchy Alexander Gelbukh Computación Ontology Document Comparison Statistical Methods Document Characterization Document indexing with a concept hierarchy Alexander Gelbukh Grigori Sidorov Adolfo Guzmán-Arenas Computación Ontology Document Comparison Statistical Methods Document Characterization Given a large hierarchical concept dictionary (thesaurus, or ontology), the task of selection of the concepts that describe the contents of a given document is considered. A statistical method of document indexing driven by such a dictionary is proposed. The method is insensible to inaccuracies in the dictionary, which allow for semi-automatic translation of the hierarchy into different languages. The problem of handling non-terminal and especially top-level nodes in the hierarchy is discussed. Common sense-complaint methods of automatically assigning the weights to the nodes and links in the hierarchy are presented. The application of the method in the Classifier system is discussed. 2005 artículo científico 1405-5546 https://www.redalyc.org/articulo.oa?id=61580403 en http://www.redalyc.org/revista.oa?id=615 Computación y Sistemas application/pdf Instituto Politécnico Nacional Computación y Sistemas (México) Num.4 Vol.8 |
| title | Document indexing with a concept hierarchy |
| topic | Computación Ontology Document Comparison Statistical Methods Document Characterization |
| url | https://www.redalyc.org/articulo.oa?id=61580403 |