A bi-objective hybrid optimization algorithm to reduce noise and data dimension in diabetes diagnosis using support vector machines. (1st August 2019)
- Record Type:
- Journal Article
- Title:
- A bi-objective hybrid optimization algorithm to reduce noise and data dimension in diabetes diagnosis using support vector machines. (1st August 2019)
- Main Title:
- A bi-objective hybrid optimization algorithm to reduce noise and data dimension in diabetes diagnosis using support vector machines
- Authors:
- Alirezaei, Mahsa
Niaki, Seyed Taghi Akhavan
Niaki, Seyed Armin Akhavan - Abstract:
- Highlights: The PIMA Indian Type-2 diabetes dataset is used. Pre-processing techniques are combined together to access high-quality data. Significant features are found using SVM. Four bi-objective meta-heuristics are employed to maximize the accuracy and to minimize the number of selected features. The 10-fold cross validation method is used to validate the constructed model. Abstract: Diabetes mellitus is a medical condition examined by data miners for reasons such as significant health complications in affected people, the economic impact on healthcare networks, and so on. In order to find the main causes of this disease, researchers look into the patient's lifestyle, hereditary information, etc. The goal of data mining in this context is to find patterns that make early detection of the disease and proper treatment easier. Due to the high volume of data involved in therapeutic contexts and disease diagnosis, provision of the intended treatment method become almost impossible over a short period of time. This justifies the use of pre-processing techniques and data reduction methods in such contexts. In this regard, clustering and meta-heuristic algorithms maintain important roles. In this paper, a method based on the k-means clustering algorithm is first utilized to detect and delete outliers. Then, in order to select significant and effective features, four bi-objective meta-heuristic algorithms are employed to choose the least number of significant features with theHighlights: The PIMA Indian Type-2 diabetes dataset is used. Pre-processing techniques are combined together to access high-quality data. Significant features are found using SVM. Four bi-objective meta-heuristics are employed to maximize the accuracy and to minimize the number of selected features. The 10-fold cross validation method is used to validate the constructed model. Abstract: Diabetes mellitus is a medical condition examined by data miners for reasons such as significant health complications in affected people, the economic impact on healthcare networks, and so on. In order to find the main causes of this disease, researchers look into the patient's lifestyle, hereditary information, etc. The goal of data mining in this context is to find patterns that make early detection of the disease and proper treatment easier. Due to the high volume of data involved in therapeutic contexts and disease diagnosis, provision of the intended treatment method become almost impossible over a short period of time. This justifies the use of pre-processing techniques and data reduction methods in such contexts. In this regard, clustering and meta-heuristic algorithms maintain important roles. In this paper, a method based on the k-means clustering algorithm is first utilized to detect and delete outliers. Then, in order to select significant and effective features, four bi-objective meta-heuristic algorithms are employed to choose the least number of significant features with the highest classification accuracy using support vector machines (SVM). In addition, the 10-fold cross validation (CV) method is used to validate the constructed model. Using real case data, it is concluded that the multi-objective firefly (MOFA) and multi-objective imperialist competitive algorithm (MOICA) with a 100% classification accuracy outperform the non-dominated sorting genetic algorithm (NSGA-II) and multi-objective particle swarm optimization (MOPSO) with the accuracies of 98.2% and 94.6%, respectively. … (more)
- Is Part Of:
- Expert systems with applications. Volume 127(2019)
- Journal:
- Expert systems with applications
- Issue:
- Volume 127(2019)
- Issue Display:
- Volume 127, Issue 2019 (2019)
- Year:
- 2019
- Volume:
- 127
- Issue:
- 2019
- Issue Sort Value:
- 2019-0127-2019-0000
- Page Start:
- 47
- Page End:
- 57
- Publication Date:
- 2019-08-01
- Subjects:
- Diabetes diagnosis -- Feature selection -- Meta-heuristic algorithms -- K-means algorithms -- Support vector machine
Expert systems (Computer science) -- Periodicals
Systèmes experts (Informatique) -- Périodiques
Electronic journals
006.33 - Journal URLs:
- http://www.sciencedirect.com/science/journal/09574174 ↗
http://www.elsevier.com/journals ↗ - DOI:
- 10.1016/j.eswa.2019.02.037 ↗
- Languages:
- English
- ISSNs:
- 0957-4174
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - 3842.004220
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 9736.xml