IDEAS home Printed from https://ideas.repec.org/a/eee/teinso/v76y2024ics0160791x24000228.html
   My bibliography  Save this article

Predicting student dropouts with machine learning: An empirical study in Finnish higher education

Author

Listed:
  • Vaarma, Matti
  • Li, Hongxiu

Abstract

This study uses three machine learning models to predict student dropouts based on students' transcript, demographic, and learning management system (LMS) data from a Finnish university. The contribution of this research lies in 1) comparing the relative importance of LMS (Moodle) data with transcript and demographic data in degree program dropout prediction, 2) examining the predictive importance of different data features monthly as a function of time from enrollment, hence extending the prior end-of-semester research to a midsemester analysis, and 3) measuring the prediction performance of the models monthly. The results identify “accumulated credits” (transcript) the “number of failed courses” (transcript), and “Moodle activity count” (LMS) as the most important features, suggesting LMS has significant predictive power and should be considered alongside transcript and demographic data when predicting degree program dropouts. Moreover, we visualize how these factors' importance and prediction performance vary over time, revealing general longitudinal trends and fluctuations within semesters. Finally, we elaborate upon this study's contributions before highlighting its limitations.

Suggested Citation

  • Vaarma, Matti & Li, Hongxiu, 2024. "Predicting student dropouts with machine learning: An empirical study in Finnish higher education," Technology in Society, Elsevier, vol. 76(C).
  • Handle: RePEc:eee:teinso:v:76:y:2024:i:c:s0160791x24000228
    DOI: 10.1016/j.techsoc.2024.102474
    as

    Download full text from publisher

    File URL: http://www.sciencedirect.com/science/article/pii/S0160791X24000228
    Download Restriction: Full text for ScienceDirect subscribers only

    File URL: https://libkey.io/10.1016/j.techsoc.2024.102474?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. Terry T. Ishitani, 2006. "Studying Attrition and Degree Completion Behavior among First-Generation College Students in the United States," The Journal of Higher Education, Taylor & Francis Journals, vol. 77(5), pages 861-885, September.
    2. Stephen L. DesJardins & Dennis A. Ahlburg & Brian P. McCall, 2002. "A Temporal Investigation of Factors Related to Timely Degree Completion," The Journal of Higher Education, Taylor & Francis Journals, vol. 73(5), pages 555-581, September.
    3. Brent J. Evans & Rachel B. Baker & Thomas S. Dee, 2016. "Persistence Patterns in Massive Open Online Courses (MOOCs)," The Journal of Higher Education, Taylor & Francis Journals, vol. 87(2), pages 206-242, March.
    4. Rong Chen & Stephen L. DesJardins, 2010. "Investigating the Impact of Financial Aid on Student Dropout Risks: Racial and Ethnic Differences," The Journal of Higher Education, Taylor & Francis Journals, vol. 81(2), pages 179-208, March.
    5. Takaya Saito & Marc Rehmsmeier, 2015. "The Precision-Recall Plot Is More Informative than the ROC Plot When Evaluating Binary Classifiers on Imbalanced Datasets," PLOS ONE, Public Library of Science, vol. 10(3), pages 1-21, March.
    6. Silvia Gilardi & Chiara Guglielmetti, 2011. "University Life of Non-Traditional Students: Engagement Styles and Impact on Attrition," The Journal of Higher Education, Taylor & Francis Journals, vol. 82(1), pages 33-53, January.
    7. Sameano F. Porchea & Jeff Allen & Steve Robbins & Richard P. Phelps, 2010. "Predictors of Long-Term Enrollment and Degree Outcomes for Community College Students: Integrating Academic, Psychosocial, Socio-demographic, and Situational Factors," The Journal of Higher Education, Taylor & Francis Journals, vol. 81(6), pages 680-708, November.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Terry T. Ishitani & Lee D. Flood, 2018. "Student Transfer-Out Behavior at Four-Year Institutions," Research in Higher Education, Springer;Association for Institutional Research, vol. 59(7), pages 825-846, November.
    2. John Bound & Michael F. Lovenheim & Sarah Turner, 2012. "Increasing Time to Baccalaureate Degree in the United States," Education Finance and Policy, MIT Press, vol. 7(4), pages 375-424, September.
    3. Gloria Crisp & Charlie Potter & Amanda Taggart, 2022. "Characteristics and Predictors of Transfer and Withdrawal Among Students Who Begin College at Bachelor’s Granting Institutions," Research in Higher Education, Springer;Association for Institutional Research, vol. 63(3), pages 481-513, May.
    4. Jameson D. Lopez, 2018. "Factors Influencing American Indian and Alaska Native Postsecondary Persistence: AI/AN Millennium Falcon Persistence Model," Research in Higher Education, Springer;Association for Institutional Research, vol. 59(6), pages 792-811, September.
    5. Contini, Dalit & Salza, Guido, 2020. "Too few university graduates. Inclusiveness and effectiveness of the Italian higher education system," Socio-Economic Planning Sciences, Elsevier, vol. 71(C).
    6. Morazes, Jennifer Lynne, 2016. "Educational background, high school stress, and academic success," Children and Youth Services Review, Elsevier, vol. 69(C), pages 201-209.
    7. Damgaard, Mette Trier & Nielsen, Helena Skyt, 2018. "Nudging in education," Economics of Education Review, Elsevier, vol. 64(C), pages 313-342.
    8. Christopher J Greenwood & George J Youssef & Primrose Letcher & Jacqui A Macdonald & Lauryn J Hagg & Ann Sanson & Jenn Mcintosh & Delyse M Hutchinson & John W Toumbourou & Matthew Fuller-Tyszkiewicz &, 2020. "A comparison of penalised regression methods for informing the selection of predictive markers," PLOS ONE, Public Library of Science, vol. 15(11), pages 1-14, November.
    9. Jinhee Kim & Swarn Chatterjee, 2019. "Student Loans, Health, and Life Satisfaction of US Households: Evidence from a Panel Study," Journal of Family and Economic Issues, Springer, vol. 40(1), pages 36-50, March.
    10. Jie-Huei Wang & Cheng-Yu Liu & You-Ruei Min & Zih-Han Wu & Po-Lin Hou, 2024. "Cancer Diagnosis by Gene-Environment Interactions via Combination of SMOTE-Tomek and Overlapped Group Screening Approaches with Application to Imbalanced TCGA Clinical and Genomic Data," Mathematics, MDPI, vol. 12(14), pages 1-24, July.
    11. Le, Hong Hanh & Viviani, Jean-Laurent, 2018. "Predicting bank failure: An improvement by implementing a machine-learning approach to classical financial ratios," Research in International Business and Finance, Elsevier, vol. 44(C), pages 16-25.
    12. João Chang Junior & Fábio Binuesa & Luiz Fernando Caneo & Aida Luiza Ribeiro Turquetto & Elisandra Cristina Trevisan Calvo Arita & Aline Cristina Barbosa & Alfredo Manoel da Silva Fernandes & Evelinda, 2020. "Improving preoperative risk-of-death prediction in surgery congenital heart defects using artificial intelligence model: A pilot study," PLOS ONE, Public Library of Science, vol. 15(9), pages 1-21, September.
    13. Angela Boatman & Bridget Terry Long, 2016. "Does Financial Aid Impact College Student Engagement?," Research in Higher Education, Springer;Association for Institutional Research, vol. 57(6), pages 653-681, September.
    14. Elena Arias & Catherine Dehon, 2011. "The Roads to Success: Analyzing Dropout and Degree Completion at University," Working Papers ECARES ECARES 2011-025, ULB -- Universite Libre de Bruxelles.
    15. Arthur De Sá Ferreira & Ney Meziat-Filho & Ana Paula Antunes Ferreira, 2021. "Double threshold receiver operating characteristic plot for three-modal continuous predictors," Computational Statistics, Springer, vol. 36(3), pages 2231-2245, September.
    16. Fan, Xudong & Wang, Xiaowei & Zhang, Xijin & ASCE Xiong (Bill) Yu, P.E.F., 2022. "Machine learning based water pipe failure prediction: The effects of engineering, geology, climate and socio-economic factors," Reliability Engineering and System Safety, Elsevier, vol. 219(C).
    17. Aina, Carmen & Baici, Eliana & Casalone, Giorgia & Pastore, Francesco, 2018. "The Economics of University Dropouts and Delayed Graduation: A Survey," IZA Discussion Papers 11421, Institute of Labor Economics (IZA).
    18. Zhang, Han, 2021. "How Using Machine Learning Classification as a Variable in Regression Leads to Attenuation Bias and What to Do About It," SocArXiv 453jk, Center for Open Science.
    19. Diaeldin Osman & Conor O’Leary & Mark Brimble & Dave Thompson, 2019. "Factor That Impact Attrition And Retention Rates Among Accountancy Diploma Students: Evidence From Saudi Arabia," Business Education and Accreditation, The Institute for Business and Finance Research, vol. 11(1), pages 89-110.
    20. Masabho P Milali & Samson S Kiware & Nicodem J Govella & Fredros Okumu & Naveen Bansal & Serdar Bozdag & Jacques D Charlwood & Marta F Maia & Sheila B Ogoma & Floyd E Dowell & George F Corliss & Maggy, 2020. "An autoencoder and artificial neural network-based method to estimate parity status of wild mosquitoes from near-infrared spectra," PLOS ONE, Public Library of Science, vol. 15(6), pages 1-16, June.

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:teinso:v:76:y:2024:i:c:s0160791x24000228. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: https://www.journals.elsevier.com/technology-in-society .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.