Course-Disjoint Evaluation and Capacity-Aware Triage for Student Dropout Risk Prediction

Wijiyanto Wijiyanto, Aris Marjuni, Ahmad Zainul Fanani, Ruri Suko Basuki

Abstract


Early-warning systems for student dropout prevention require evaluation protocols and outputs that remain reliable when applied across heterogeneous academic contexts. This study quantifies how conventional random splits can overestimate performance when models are expected to generalize across different courses and proposes a decision-support layer that translates predicted risk into capacity-aware intervention policies. Using a benchmark higher-education dataset (N=4,424; 34 predictors; three classes: Dropout, Enrolled, Graduate) with 17 Course groups, phased prediction is implemented to reflect increasing evidence availability: S0 (pre-enrollment), S1 (plus semester-1 academic evidence), and S2 (plus semester-2 academic evidence). Baseline results are replicated with leakage-safe preprocessing (imputation, one-hot encoding, scaling) and Synthetic Minority Over-sampling Technique (SMOTE) applied strictly within training folds, comparing multinomial logistic regression, random forest, and tree-based boosting models. Deployment-oriented performance is assessed using StratifiedGroupKFold by Course to enforce course-disjoint testing. Discrimination is reported with Macro-F1 and Balanced Accuracy, while probability quality is evaluated using LogLoss, Brier score, expected calibration error, maximum calibration error, and reliability diagrams. Calibrated probabilities are translated into capacity-aware risk bands (Top-k% triage), selective prediction is evaluated via abstention to defer low-confidence cases, and split conformal prediction sets are optionally reported for multiclass uncertainty communication. Results show consistent performance drops under course-disjoint validation, confirming a non-trivial generalization gap. Error decomposition indicates that Enrolled is the most ambiguous class and exhibits phase-dependent confusion toward both terminal outcomes. Calibration shows phase-specific trade-offs between likelihood-based and worst-case calibration metrics, while risk bands yield high-precision triage under limited capacity, and abstention improves decision quality at reduced coverage. Overall, the study provides a deployment-oriented evaluation and decision-support workflow for translating dropout risk models into actionable capacity planning.


Keywords


Student Dropout, Course-Disjoint Validation, Probability Calibration, Risk Bands, Selective Prediction, Conformal Prediction

Full Text:

PDF

References


R. Boudjehem and Y. Lafifi, “An early warning system to predict dropouts inside e-learning environments,” Educ Inf Technol, vol. 29, no. 13, pp. 16365–16385, Sep. 2024, doi: 10.1007/s10639-024-12498-1.

R. M. Santos and R. Henriques, “Accurate, timely, and portable: Course-agnostic early prediction of student performance from LMS logs,” Computers and Education: Artificial Intelligence, vol. 5, p. 100175, 2023, doi: 10.1016/j.caeai.2023.100175.

D. Bañeres, M. E. Rodríguez-González, A.-E. Guerrero-Roldán, and P. Cortadas, “An early warning system to identify and intervene online dropout learners,” Int J Educ Technol High Educ, vol. 20, no. 1, p. 3, Jan. 2023, doi: 10.1186/s41239-022-00371-5.

M. Sailer, M. Ninaus, S. E. Huber, E. Bauer, and S. Greiff, “The End is the Beginning is the End: The closed-loop learning analytics framework,” Computers in Human Behavior, vol. 158, p. 108305, Sep. 2024, doi: 10.1016/j.chb.2024.108305.

J. E. Raffaghelli, M. E. Rodríguez, A.-E. Guerrero-Roldán, and D. Bañeres, “Applying the UTAUT model to explain the students’ acceptance of an early warning system in Higher Education,” Computers & Education, vol. 182, p. 104468, Jun. 2022, doi: 10.1016/j.compedu.2022.104468.

J. Mandi et al., “Decision-Focused Learning: Foundations, State of the Art, Benchmark and Future Opportunities,” jair, vol. 80, pp. 1623–1701, Aug. 2024, doi: 10.1613/jair.1.15320.

K. Zhou, Z. Liu, Y. Qiao, T. Xiang, and C. C. Loy, “Domain Generalization: A Survey,” IEEE Trans. Pattern Anal. Mach. Intell., pp. 1–20, 2022, doi: 10.1109/TPAMI.2022.3195549.

C. Kumar, G. Walton, P. Santi, and C. Luza, “Random Cross-Validation Produces Biased Assessment of Machine Learning Performance in Regional Landslide Susceptibility Prediction,” Remote Sensing, vol. 17, no. 2, p. 213, Jan. 2025, doi: 10.3390/rs17020213.

A. Palanci, R. M. Yılmaz, and Z. Turan, “Learning analytics in distance education: A systematic review study,” Educ Inf Technol, vol. 29, no. 17, pp. 22629–22650, Dec. 2024, doi: 10.1007/s10639-024-12737-5.

T. Silva Filho, H. Song, M. Perello-Nieto, R. Santos-Rodriguez, M. Kull, and P. Flach, “Classifier calibration: a survey on how to assess and improve predicted class probabilities,” Mach Learn, vol. 112, no. 9, pp. 3211–3260, Sep. 2023, doi: 10.1007/s10994-023-06336-7.

U. Sadana, A. Chenreddy, E. Delage, A. Forel, E. Frejinger, and T. Vidal, “A survey of contextual optimization methods for decision-making under uncertainty,” European Journal of Operational Research, vol. 320, no. 2, pp. 271–289, Jan. 2025, doi: 10.1016/j.ejor.2024.03.020.

V. Christou et al., “Performance and early drop prediction for higher education students using machine learning,” Expert Systems with Applications, vol. 225, p. 120079, Sep. 2023, doi: 10.1016/j.eswa.2023.120079.

A. Zanellati, S. P. Zingaro, and M. Gabbrielli, “Balancing Performance and Explainability in Academic Dropout Prediction,” IEEE Trans. Learning Technol., vol. 17, pp. 2086–2099, 2024, doi: 10.1109/TLT.2024.3425959.

J. Dong et al., “A Survey on Confidence Calibration of Deep Learning-Based Classification Models Under Class Imbalance Data,” IEEE Trans. Neural Netw. Learning Syst., vol. 36, no. 9, pp. 15664–15684, Sep. 2025, doi: 10.1109/TNNLS.2025.3565159.

K. Hendrickx, L. Perini, D. Van Der Plas, W. Meert, and J. Davis, “Machine learning with a reject option: a survey,” Mach Learn, vol. 113, no. 5, pp. 3073–3110, May 2024, doi: 10.1007/s10994-024-06534-x.

X. Zhou, B. Chen, Y. Gui, and L. Cheng, “Conformal Prediction: A Data Perspective,” ACM Comput. Surv., vol. 58, no. 2, pp. 1–37, Jan. 2026, doi: 10.1145/3736575.

J. G. C. Krüger, A. D. S. Britto, and J. P. Barddal, “An explainable machine learning approach for student dropout prediction,” Expert Systems with Applications, vol. 233, p. 120933, Dec. 2023, doi: 10.1016/j.eswa.2023.120933.

Z. Azizah, T. Ohyama, X. Zhao, Y. Ohkawa, and T. Mitsuishi, “Predicting at-risk students in the early stage of a blended learning course via machine learning using limited data,” Computers and Education: Artificial Intelligence, vol. 7, p. 100261, Dec. 2024, doi: 10.1016/j.caeai.2024.100261.

L. Herrmann and J. Weigert, “AI-based prediction of academic success: Support for many, disadvantage for some?,” Computers and Education: Artificial Intelligence, vol. 7, p. 100303, Dec. 2024, doi: 10.1016/j.caeai.2024.100303.

M. V. M. Valentim Realinho, “Predict Students’ Dropout and Academic Success.” UCI Machine Learning Repository, 2021. doi: 10.24432/C5MC89.

H. Waheed, S.-U. Hassan, R. Nawaz, N. R. Aljohani, G. Chen, and D. Gasevic, “Early prediction of learners at risk in self-paced education: A neural network approach,” Expert Systems with Applications, vol. 213, p. 118868, Mar. 2023, doi: 10.1016/j.eswa.2022.118868.

H. El Aouifi, M. El Hajji, and Y. Es-Saady, “A hybrid approach for early-identification of at-risk dropout students using LSTM-DNN networks,” Educ Inf Technol, vol. 29, no. 14, pp. 18839–18857, Oct. 2024, doi: 10.1007/s10639-024-12588-0.

G. I. Austin, I. Pe’er, and T. Korem, “Distributional bias compromises leave-one-out cross-validation,” Sci. Adv., vol. 11, no. 48, p. eadx6976, Nov. 2025, doi: 10.1126/sciadv.adx6976.

A. Apicella, F. Isgrò, and R. Prevete, “Don’t push the button! Exploring data leakage risks in machine learning and transfer learning,” Artif Intell Rev, vol. 58, no. 11, p. 339, Aug. 2025, doi: 10.1007/s10462-025-11326-3.

J. Allgaier and R. Pryss, “Practical approaches in evaluating validation and biases of machine learning applied to mobile health studies,” Commun Med, vol. 4, no. 1, p. 76, Apr. 2024, doi: 10.1038/s43856-024-00468-0.

M. Phan, A. De Caigny, and K. Coussement, “A decision support framework to incorporate textual data for early student dropout prediction in higher education,” Decision Support Systems, vol. 168, p. 113940, May 2023, doi: 10.1016/j.dss.2023.113940.

T. Dawood et al., “Uncertainty aware training to improve deep learning model calibration for classification of cardiac MR images,” Medical Image Analysis, vol. 88, p. 102861, Aug. 2023, doi: 10.1016/j.media.2023.102861.

W. Huang, G. Cao, J. Xia, J. Chen, H. Wang, and J. Zhang, “H-Calibration: Rethinking Classifier Recalibration With Probabilistic Error-Bounded Objective,” IEEE Trans. Pattern Anal. Mach. Intell., vol. 47, no. 10, pp. 9023–9042, Oct. 2025, doi: 10.1109/TPAMI.2025.3582796.

T. Dimitriadis, T. Gneiting, A. I. Jordan, and P. Vogel, “Evaluating probabilistic classifiers: The triptych,” International Journal of Forecasting, vol. 40, no. 3, pp. 1101–1122, Jul. 2024, doi: 10.1016/j.ijforecast.2023.09.007.

M. Kängsepp, K. Valk, and M. Kull, “On the usefulness of the fit-on-test view on evaluating calibration of classifiers,” Mach Learn, vol. 114, no. 4, p. 105, Apr. 2025, doi: 10.1007/s10994-024-06652-6.

B. Sonnleitner, T. Madou, M. Deceuninck, F. Theodosiou, and Y. R. Sagaert, “Evaluation of early student performance prediction given concept drift,” Computers and Education: Artificial Intelligence, vol. 8, p. 100369, Jun. 2025, doi: 10.1016/j.caeai.2025.100369.

D. Hooshyar et al., “Towards responsible AI for education: Hybrid human-AI to confront the elephant in the room,” Computers and Education: Artificial Intelligence, vol. 9, p. 100524, Dec. 2025, doi: 10.1016/j.caeai.2025.100524.

A. Demircioğlu, “Applying oversampling before cross-validation will lead to high bias in radiomics,” Sci Rep, vol. 14, no. 1, p. 11563, May 2024, doi: 10.1038/s41598-024-62585-z.

W. Chen, K. Yang, Z. Yu, Y. Shi, and C. L. P. Chen, “A survey on imbalanced learning: latest research, applications and future directions,” Artif Intell Rev, vol. 57, no. 6, p. 137, May 2024, doi: 10.1007/s10462-024-10759-6.

R. F. Barber, E. J. Candès, A. Ramdas, and R. J. Tibshirani, “Conformal prediction beyond exchangeability,” Ann. Statist., vol. 51, no. 2, Apr. 2023, doi: 10.1214/23-AOS2276.




DOI: https://doi.org/10.47738/jads.v7i3.1419

Refbacks

  • There are currently no refbacks.



Barcode

Journal of Applied Data Sciences

ISSN : 2723-6471 (Online)
Publisher : Bright Publisher
Website : http://bright-journal.org/JADS
Email : taqwa@amikompurwokerto.ac.id (principal contact)
    support@bright-journal.org (technical issues)

This work is licensed under a Creative Commons Attribution-ShareAlike 4.0