Tehran University of Medical Sciences

Science Communicator Platform

Stay connected! Follow us on X network (Twitter):
Share this content! On (X network) By
Discovery of Hidden Patterns in Breast Cancer Patients, Using Data Mining on a Real Data Set Publisher Pubmed



Atashi A1 ; Tohidinezhad F2 ; Dorri S3 ; Nazeri N3 ; Ghousi R4 ; Marashi S1 ; Hajialiasgari F1
Authors
Show Affiliations
Authors Affiliations
  1. 1. E-Health Department, Virtual School, Tehran University of Medical Sciences, Tehran, Iran
  2. 2. Department of Medical Informatics, Faculty of Medicine, Mashhad University of Medical Sciences, Mashhad, Iran
  3. 3. Clinical Research Department, Breast Cancer Research Center, Motamed Cancer Institute, ACECR, Tehran, Iran
  4. 4. School of Industrial Engineering, Iran University of Science and Technology, Tehran, Iran

Source: Studies in Health Technology and Informatics Published:2019


Abstract

The aim is to recognize the unknown atterns in a real breast cancer dataset using data mining algorithms as a new method in medicine. Due to excessive missing data in the collection only data on 665 of 809 patients were available. The other missing values were estimated using the EM algorithm in SPSS21 software. Fields have been converted into discrete fields and finally the APRIORI algorithm has been used to analyze and explore the unknown patterns. After the rule extraction, experts in the field of breast cancer eliminated redundant and meaningless relations. 100 association rules with a confidence value of more than 0.9 explored by the APRIORI algorithm and after the clinical expert feedback, 10 clinically meaningful relations have been detected and reported. Due to the high number of risk factors, the use of data mining is effective for cancer data. These patterns provide the future study hypotheses of specific clinical studies. © 2019 The authors and IOS Press. All rights reserved.