Improving classification performance for an imbalanced educational dataset example using SMOTE
Yazarlar (3)
Dr. Öğr. Üyesi Yavuz ÜNAL Amasya Üniversitesi, Türkiye
Öğr. Gör. Ahmet Sağlam Amasya Üniversitesi, Türkiye
Dr. Öğr. Üyesi Osman Kayhan T.C. Milli Eğitim Bakanlığı
Makale Türü Açık Erişim Özgün Makale (Diğer hakemli uluslarası dergilerde yayınlanan tam makale)
Dergi Adı European Journal of Science and Technology
Dergi ISSN 2148-2683
Dergi Tarandığı Indeksler TR DİZİN
Makale Dili İngilizce Basım Tarihi 10-2019
Kabul Tarihi Yayınlanma Tarihi 31-10-2019
Cilt / Sayı / Sayfa – / 0 / 485–489 DOI 10.31590/ejosat.638608
Makale Linki https://dergipark.org.tr/tr/doi/10.31590/ejosat.638608
UAK Araştırma Alanları
Görüntü İşleme
Özet
With technology, a lot of data is formed in digital environments. One of the areas with intensive data is educational data sets. By analyzing educational data sets, students' situatiokjgjjööÖns can be predicted by foreseeing. In this way, students can be assisted by anticipating situations such as drop-out due to failure. Educational institutions can take measures to prevent such dropouts and reduce student drop-out. Thus, financial losses of students and educational institutions can be prevented. In this study, the data of five separate associate degree students who were enrolled in Amasya University Distance Education Center in 2016-2017 were used. These are associate degree programs in child development, medical documentation and secretarial, electricity, mechatronics, and internet and network technologies. It was estimated whether the students could graduate or not at the end of the IV. Semester with looking at their I. and II. semester course notes. These data were analyzed by k nearest neighbor (K-NN) and KStar algorithms. Some of the data were obtained from the distance education center as imbalanced data due to the low number of students. In Educational Data Mining, researchers usually overlook the balance of the distribution on a dataset. Unbalanced data can seriously affect the success of classification. Synthetic minority oversampling technique (SMOTE) method was applied to these unbalanced data and how it affected the success of classification was examined. First, the raw data were analyzed with K-nearest neighbors classifier and KStar classifier. In this study, the analysis results of these five chapters are given in tables and …
Anahtar Kelimeler
BM Sürdürülebilir Kalkınma Amaçları
Atıf Sayıları
Google Scholar 3
Improving classification performance for an imbalanced educational dataset example using SMOTE

Paylaş