Your browser doesn't support javascript.
loading
Classification of Human Papillomavirus (HPV) Risk Type via Text Mining
Genomics & Informatics ; : 80-86, 2003.
Article in English | WPRIM | ID: wpr-197482
ABSTRACT
Human Papillomavirus (HPV) infection is known as the main factor for cervical cancer which is a leading cause of cancer deaths in women worldwide. Because there are more than 100 types in HPV, it is critical to discriminate the HPVs related with cervical cancer from those not related with it. In this paper, the risk type of HPVs using their textual explanation. The important issue in this problem is to distinguish false negatives from false positives. That is, we must find high-risk HPVs as many as possible though we may miss some low-risk HPVs. For this purpose, the AdaCost, a cost-sensitive learner is adopted to consider different costs between training examples. The experimental results on the HPV sequence database show that the consideration of costs gives higher performance. The improvement in F-score is higher than that of the accuracy, which implies that the number of high-risk HPVs found is increased.
Subject(s)

Full text: Available Index: WPRIM (Western Pacific) Main subject: Uterine Cervical Neoplasms / Classification / Data Mining Type of study: Etiology study Limits: Female / Humans Language: English Journal: Genomics & Informatics Year: 2003 Type: Article

Similar

MEDLINE

...
LILACS

LIS

Full text: Available Index: WPRIM (Western Pacific) Main subject: Uterine Cervical Neoplasms / Classification / Data Mining Type of study: Etiology study Limits: Female / Humans Language: English Journal: Genomics & Informatics Year: 2003 Type: Article