Exploring diseases based biomedical document clustering and visualization using self-organizing maps

Date

2017-10
Language
English

Embargo Lift Date

Committee Members

Degree

Degree Year

Department

Grantor

Journal Title

Journal ISSN

Volume Title

Found At

IEEE

Abstract

Document clustering is a text mining technique used to provide better document search and browsing in digital libraries or online corpora. In this research, a vector representation of concepts of diseases and similarity measurement between concepts are proposed. They identify the closest concepts of diseases in the context of a corpus. Each document is represented by using the vector space model. A weight scheme is proposed to consider both local content and associations between concepts. Self-Organizing Maps (SOM) are often used as document clustering algorithm. The vector projection and visualization features of SOM enable visualization and analysis of the cluster distribution and relationships on the two dimensional space. The Davies-Bouldin index is used to validate the clusters based on the visualized cluster distributions. The results show that the proposed document clustering framework generates meaningful clusters and can facilitate clustering visualization and information retrieval based on the concepts of diseases.

Description

item.page.description.tableofcontents

item.page.relation.haspart

Cite As

Shah, S., & Luo, X. (2017). Exploring diseases based biomedical document clustering and visualization using self-organizing maps. In 2017 IEEE 19th International Conference on e-Health Networking, Applications and Services (Healthcom) (pp. 1–6). https://doi.org/10.1109/HealthCom.2017.8210791

ISSN

Publisher

Series/Report

Sponsorship

Major

Extent

Identifier

Relation

Journal

2017 IEEE 19th International Conference on e-Health Networking, Applications and Services

Rights

Publisher Policy

Source

Author

Alternative Title

Type

Article

Number

Volume

Conference Dates

Conference Host

Conference Location

Conference Name

Conference Panel

Conference Secretariat Location

Version

Author's manuscript

Full Text Available at

This item is under embargo {{howLong}}