Predicting science achievement scores with machine learning algorithms: a case study of OECD PISA 2015–2018 data

dc.contributor.authorAçışlı Çelik, Sibel
dc.contributor.authorYeşilkanat, Cafer Mert
dc.date.accessioned2025-07-17T08:01:27Z
dc.date.available2025-07-17T08:01:27Z
dc.date.issued2023
dc.departmentAÇÜ, Eğitim Fakültesi, Matematik ve Fen Bilimleri Eğitimi Bölümü
dc.description.abstractIn this study, the performance of machine learning methods was examined in terms of predicting the science education achievement scores of the students who took the exam for the next term, PISA 2018, and the science average scores of the countries, using PISA 2015 data. The research sample consists of a total of 67,329 students who took the PISA 2015 exam from 13 randomly selected countries (Brazil, Chinese Taipei, Dominican Republic, Estonia, Finland, Hungary, Italy, Japan, Lithuania, Luxembourg, Peru, Singapore, Türkiye). In this study, multiple linear regression, support vector regression, random forest, and extreme gradient boosting (XGBoost) machine learning algorithms were used. For the machine learning process, a randomly determined part from the PISA-2015 data of each country researched was divided as training data and the remaining part as testing data to evaluate model performance. As a result of the research, it was determined that the XGBoost algorithm showed the best performance in estimating both PISA-2015 test data and PISA-2018 science academic achievement scores in all researched countries. Furthermore, it was determined that the highest PISA-2018 science achievement scores of the students who participated in the exam, estimated by this algorithm, were in Luxembourg (r = 0.600, RMSE = 75.06, MAE = 59.97), while the lowest were in Finland (r = 0.467, RMSE = 79.38, MAE = 63.24). In addition, the average PISA-2018 science scores of the countries were estimated with the XGBoost algorithm, and the average science scores calculated for all the countries studied were estimated with very high accuracy.
dc.identifier.doi10.1007/s00521-023-08901-6
dc.identifier.endpage21228
dc.identifier.issn09410643
dc.identifier.issue28
dc.identifier.scopuss2.0-85166636097
dc.identifier.scopusqualityQ1
dc.identifier.startpage21201
dc.identifier.urihttps://hdl.handle.net/11494/5796
dc.identifier.volume35
dc.indekslendigikaynakScopus
dc.institutionauthorAçışlı Çelik, Sibel
dc.institutionauthorYeşilkanat, Cafer Mert
dc.language.isoen
dc.publisherSpringer Science and Business Media Deutschland GmbH
dc.relation.ispartofNeural Computing and Applications
dc.relation.publicationcategoryMakale - Uluslararası Hakemli Dergi - Kurum Öğretim Elemanı
dc.rightsinfo:eu-repo/semantics/openAccess
dc.subjectArtificial intelligence
dc.subjectInterdisciplinary/transdisciplinary
dc.subjectRandom forest
dc.subjectXGBoost
dc.titlePredicting science achievement scores with machine learning algorithms: a case study of OECD PISA 2015–2018 data
dc.typeArticle

Dosyalar

Orijinal paket
Listeleniyor 1 - 1 / 1
Yükleniyor...
Küçük Resim
İsim:
5796.pdf
Boyut:
3.91 MB
Biçim:
Adobe Portable Document Format
Lisans paketi
Listeleniyor 1 - 1 / 1
[ X ]
İsim:
license.txt
Boyut:
1.17 KB
Biçim:
Item-specific license agreed upon to submission
Açıklama: