Detection of colon cancer based on microarray dataset using machine learning as a feature selection and classification techniques

Microarray data is an increasingly important tool for providing information on gene expression for analysis and interpretation. Researchers attempt to utilize the smallest possible set of relevant gene expression profiles in most gene expression studies to enhance tumor identification accuracy. This...

全面介绍

书目详细资料
Main Authors: A. S. M, Shafi, M. M., Imran Molla, Jui, Julakha Jahan, Mohammad, Motiur Rahman
格式: 文件
语言:English
出版: Springer Link 2020
主题:
在线阅读:http://umpir.ump.edu.my/id/eprint/28914/1/Shafi2020_Article_DetectionOfColonCancerBasedOnM.pdf
实物特征
总结:Microarray data is an increasingly important tool for providing information on gene expression for analysis and interpretation. Researchers attempt to utilize the smallest possible set of relevant gene expression profiles in most gene expression studies to enhance tumor identification accuracy. This research aims to analyze and predicts colon cancer data employing a machine learning approach and feature selection technique based on a random forest classifier. More particularly, our proposed method can reduce the burden of high dimensional data and allow faster calculations by combining the “Mean Decrease Accuracy” and “Mean Decrease Gini” as feature selection methods into a renowned classifier namely Random Forest, with the aim of increasing the prediction model's accuracy level. In addition, we have also shown a comparative model analysis with selection of features and model without selection of features. The extensive experimental results have demonstrated that the proposed model with feature selection is favorable and effective which triumphs the best performance of accuracy.