Meida Cahyo Untoro, Mugi Praseptiawan, Mastuti Widianingsih, Ilham Firman Ashari, Aidil Afriansyah, Oktafianto
Imbalanced data causes misclassification because the majority of the dominant data is in the minority data, which results in a decrease in the value of accuracy. UCI dataset is a public dataset that can be used as a dataset in machine learning. This study aims to evaluate the Decision Tree, K-NN, Naive Bayes, and Support Vector Machine classification methods on data imbalances in MWMOTE. MWMOTE is used in resolving Imbalanced cases through weighting and grouping. This goal is achieved by evaluating the Decision Tree, K-NN, Naive Bayes, and Support Vector Machine classification methods in MWMOTE to produce more representative synthetic data and increase the accuracy value. The results obtained from this study indicate that the Decision Tree has higher evaluations of recall, precision, F-measure, and accuracy compared to K-NN, Naive Bayes, and Support Vector Machine for data that are balanced with MWMOTE. © 2020 IOP Publishing Ltd. All rights reserved.
Department of Informatics Engineering, Institut Teknologi Sumatera, Indonesia; Department of Biology, Universitas Gadjah Mada, Indonesia
Research at a Glance
Register to unlockTopics & SDG Alignment
Register to unlockCollaboration
Register to unlockAuthor Profile (Selected)
Register to unlockReferences Overview
Register to unlockJournal & Source
Register to unlockMetadata & Integrity
Register to unlock