Student Performance Classification Using Academic
Student Performance Classification Using Academic
1, February 2026
223
Journal of Information Systems and Informatics | ISSN: 2656-5935 | e-ISSN: 2656-4882 | pp. 223-239
Published by Asosiasi Doktor Sistem Informasi Indonesia
Vol. 8, No. 1, February 2026
1. INTRODUCTION
offer new perspectives on how academic and non-academic variables jointly shape
student success [24]–[26].
Accordingly, this study aims to develop and evaluate a comprehensive machine learning
based classification framework for predicting student academic performance. The
proposed framework integrates academic variables derived from Learning Management
System (LMS) activity, socio-economic and demographic indicators, and digital behavioral
features that reflect broader patterns of student engagement.
The study addresses the following research questions (RQs): RQ1: How effectively can
machine learning classification models predict student academic performance when
integrating academic, socio-economic, and digital behavioral features?, RQ2: Which
machine learning algorithm provides the most accurate and reliable classification of
students into good (GPA ≥ 3.00) and poor (GPA < 3.00) performance categories?. Seven
machine learning algorithms Naïve Bayes, Generalized Linear Model, Logistic Regression,
Deep Learning, Decision Tree, Random Forest, and Gradient Boosted Trees are evaluated
using a dataset of 2,423 students from multiple study programs within a single university.
This comparative design enables a consistent and robust assessment of model
performance.
The scope of this study is limited to one Indonesian university; therefore, the findings
may not fully generalize to other institutional contexts. Nevertheless, the results provide
valuable empirical insights into student performance prediction in Indonesian higher
education and contribute to the broader learning analytics literature.
2. METHODS
225 | Student Performance Classification Using Academic, Socioeconomic, and Digital …..
Vol. 8, No. 1, February 2026
class distribution, with the good-performance category slightly more prevalent than the
poor-performance category.
The dataset consists of 2,423 student records categorized into two performance classes
based on GPA. A total of 1,512 students (62.4%) were classified as good performers (GPA
≥ 3.00), while 911 students (37.6%) were categorized as poor performers (GPA < 3.00). This
distribution indicates a moderate class imbalance, which was addressed during model
training using SMOTE applied exclusively to the training data. A summary of feature
descriptions, and data types is provided in Table 2.
227 | Student Performance Classification Using Academic, Socioeconomic, and Digital …..
Vol. 8, No. 1, February 2026
two hidden layers with 64 and 32 neurons, and a single-node output layer for binary
classification. ReLU activation functions were used in the hidden layers, while a sigmoid
activation function was applied in the output layer. The model was trained using the
Adam optimizer for a maximum of 100 epochs, with early stopping based on validation
loss to reduce the risk of overfitting.
This section presents the results of the classification models developed in this study and
discusses their effectiveness in predicting student academic performance. The models
were evaluated using multiple performance metrics to identify the most suitable
approach for early academic risk identification. In line with the study objective, the
classification framework focuses on identifying students at risk of poor academic
performance rather than predicting exact GPA values.
229 | Student Performance Classification Using Academic, Socioeconomic, and Digital …..
Vol. 8, No. 1, February 2026
between ±0.01 and ±0.03 across models, indicating relatively stable performance across
folds.
The results show that Gradient Boosted Trees (GBT) achieved the highest overall accuracy
(0.75), followed closely by Random Forest and Logistic Regression (0.74). Although Naïve
Bayes produced a slightly lower accuracy (0.73), its high recall value (0.89) indicates strong
sensitivity in identifying students at risk of poor academic performance, which is
particularly desirable in early warning contexts.
The Decision Tree model achieved a perfect recall score (1.00), indicating that it
successfully identified all poor-performing students. However, this result was
accompanied by lower AUC and precision values, suggesting a tendency toward
overfitting and reduced generalization capability. In contrast, ensemble-based methods
such as Random Forest and GBT exhibited more balanced performance across metrics,
highlighting their robustness in handling heterogeneous academic, socio-economic, and
behavioral features. Although AUC values across models are modest (0.61–0.65), this
outcome is expected in educational datasets characterized by overlapping feature
distributions and moderate class imbalance. In early warning applications, recall is often
prioritized over AUC, as failing to identify at-risk students (false negatives) carries greater
institutional consequences than issuing additional alerts. To further examine error
characteristics, a confusion matrix was generated for the best-performing model (GBT).
The results indicate that false negatives were relatively limited compared to false
positives, supporting the suitability of the classification framework for early intervention
scenarios where sensitivity to academic risk is essential.
231 | Student Performance Classification Using Academic, Socioeconomic, and Digital …..
Vol. 8, No. 1, February 2026
3.3. Discussion
The findings indicate that reframing student performance prediction from regression-
based GPA estimation to a binary classification task produces outputs that are more
interpretable and operationally useful for academic decision-making. Instead of reporting
an exact GPA value—which may be difficult to translate into policy actions—the
classification framework directly identifies whether a student is likely to fall into the
poor (GPA < 3.00) or good (GPA ≥ 3.00) performance category. Importantly, defining poor
performance as the positive class aligns model evaluation with the practical goal of early
risk detection. In this context, recall becomes a priority metric because false negatives
(at-risk students incorrectly classified as not at risk) can delay intervention and
potentially worsen academic outcomes.
the ensemble models provided a better balance between sensitivity and overall predictive
stability under the same experimental design (10-fold cross-validation with low variability
across folds). By contrast, the single Decision Tree achieved perfect recall (1.00), meaning
it identified all poor-performing students; however, this came with lower precision (0.74)
and modest AUC (0.62), suggesting reduced generalization and potential overfitting. For
early warning systems, such a model may generate more false alarms, which can burden
academic support units and reduce stakeholder trust in the system.
Although the AUC values across models are relatively modest (0.61–0.65), this pattern is
common in educational prediction tasks where feature distributions overlap and
outcomes are influenced by many unobserved factors. In such settings, AUC alone may
understate practical utility, especially when the institutional objective is to maximize
detection of at-risk students. The high recall achieved by most models—including Naïve
Bayes (0.89), Deep Learning (0.90), and particularly the ensemble methods (0.95)—
demonstrates strong sensitivity to academic risk. This supports the suitability of the
proposed framework for early intervention scenarios, where a manageable increase in
false positives is often preferable to missing students who genuinely require support.
Moreover, the use of SMOTE exclusively within training folds helps mitigate the moderate
class imbalance (62.4% good vs 37.6% poor) while reducing the risk of information
leakage, strengthening confidence that the reported performance reflects genuine
predictive signal rather than inflated results.
The feature importance analysis using GBT provides additional insight into which
variables most strongly differentiate performance categories in this dataset. Social media
usage emerged as the most influential predictor (weight = 0.2931), followed by total LMS
logins (0.1629) and gender (0.1218). Together with the correlation results (social media
usage showing a weak negative association with performance, r = –0.293; LMS logins
showing a weak positive association, r = 0.163), these findings suggest that digital
engagement signals can meaningfully complement traditional academic indicators. A
plausible interpretation is that high social media exposure may reflect time displacement,
reduced concentration, or fragmented study routines, whereas sustained LMS activity
may signal consistent academic engagement. However, these patterns should be
interpreted as predictive associations rather than causal relationships. Social media use,
for example, may also be a proxy for other unmeasured factors (stress, motivation, or
233 | Student Performance Classification Using Academic, Socioeconomic, and Digital …..
Vol. 8, No. 1, February 2026
learning habits). Likewise, gender’s relatively high importance warrants careful handling:
it may capture structural or behavioral differences in learning engagement, but it should
not be used to justify biased decision-making. Institutional implementation should
prioritize supportive interventions and avoid stigmatization or differential treatment
based on demographic attributes.
Among the academic variables, assignment submissions (0.0857) and forum participation
(0.0073) contributed less than login frequency, suggesting that broad engagement
intensity (regular access and presence in LMS) may be more informative than specific
activity counts within this dataset. Non-academic factors showed mixed influence:
domicile distance (0.0511) had moderate importance—possibly reflecting commuting
constraints or time availability—while organizational involvement (0.0214) contributed
modestly. Parental income (0.0027) and study group participation (0.0047) appeared least
influential, which may reflect limited measurement granularity (e.g., income recorded in
broad ranges), context-specific characteristics of the sampled institution, or the
possibility that academic behaviors captured through LMS overshadow these variables in
predictive power. Notably, the Pearson correlation matrix indicates generally weak inter-
variable correlations, suggesting minimal multicollinearity and supporting the
appropriateness of combining these predictors within machine learning models without
severe redundancy.
Ethical and privacy considerations are critical when incorporating digital behavior data
into educational analytics. Institutions should ensure transparency about what data are
collected and why, minimize the use of personally sensitive attributes, and rely on
With respect to the research questions, RQ1 is addressed by the observed improvement
in predictive capability when combining academic (LMS), socio-economic, and digital
behavioral features, as reflected in consistently high recall values across models and
competitive accuracy levels, particularly for ensemble methods. RQ2 is answered by the
comparative evaluation showing that Gradient Boosted Trees provides the most accurate
and reliable performance overall (accuracy = 0.75; recall = 0.95), closely followed by
Random Forest (accuracy = 0.74; recall = 0.95). Finally, the study’s single-institution scope
remains a limitation for generalizability; future work should validate these findings across
multiple Indonesian universities, consider hyperparameter tuning, and incorporate
interpretability analyses (e.g., local explanations) to better support responsible
deployment in real academic settings.
4. CONCLUSION
This study compared seven machine learning algorithms to classify student academic
performance into good and poor categories, adopting a classification framework to
enhance interpretability beyond regression-based GPA prediction. The results indicate
that ensemble-based models, particularly Gradient Boosted Trees and Random Forest,
achieve the most reliable performance, with the highest accuracy reaching 0.75. RQ1 is
addressed by showing that integrating academic, socio-economic, and digital behavioral
features enables effective prediction of student academic performance, while RQ2 is
answered by identifying Gradient Boosted Trees as the best-performing classifier.
235 | Student Performance Classification Using Academic, Socioeconomic, and Digital …..
Vol. 8, No. 1, February 2026
This study is limited by the use of data from a single institution and by moderate AUC
values, which reflect overlapping student performance characteristics. Future research
should explore multi-institutional or longitudinal datasets and richer behavioral
indicators to improve generalizability and predictive robustness. Overall, the findings
demonstrate that classification-oriented modeling provides actionable insights that
effectively support data-driven decision-making in higher education.
ACKNOWLEDGMENT
The authors would like to thank Universitas Muria Kudus for funding this research
through the Internal Research Scheme (PFR). We also extend our gratitude to all parties
who have contributed to the completion of this study.
REFERENCES
[1] A. Alshanqiti and A. Namoun, “Predicting Student Performance and Its Influential
Factors Using Hybrid Regression and Multi-Label Classification,” vol. 8, pp. 203827–
203844, 2020, doi: 10.1109/ACCESS.2020.3036572.
[2] S. M. Dol and P. M. Jawandhiya, “Systematic Review and Analysis of EDM for
Predicting the Academic Performance of Students,” J. Inst. Eng. Ser. B, vol. 105, no.
4, pp. 1021–1071, 2024, doi: 10.1007/s40031-024-00998-0.
[3] Khan, A., and S. K. Ghosh, “Student performance analysis and prediction in classroom
learning: A review of educational data mining studies,” Educ. Inf. Technol., vol. 26,
no. 1, pp. 205–240, 2021.
[4] I. Papadogiannis and M. Wallace, “Educational Data Mining : A Foundational
Overview,” Encyclopedia, vol. 4, no. 4, pp. 1644–1664, 2024.
[5] P. E. O. David Iyanuoluwa Ajiga, Oladimeji Hamza, Adeoluwa Eweje, Eseoghene
Kokogho, “Data-Driven Strategies for Enhancing Student Success in,” vol. 11, no. 1, pp.
411–424, 2025, doi: 10.56201/[Link].411.424.
[6] N. P.-M. Esomonu, “Utilizing AI and Big Data for Predictive Insights on Institutional
Performance and Student Success: A Data-Driven Approach to Quality Assurance,”
AI Ethics, Acad. Integr. Futur. Qual. Assur. High. Educ., vol. 29, 2025.
[7] Gonugunta, K. C., and K. Leo, “Role of data-driven decision making in enhancing
higher education performance: A comprehensive analysis of analytics in
institutional management,” Int. J. Acta Informat., vol. 3, no. 1, pp. 149–159, 2024.
[8] Wakeel, S., D. Sher, A. Kauser, and K. B. A. H. K. Niazi, “Investigating how predictive
analytics and student data modeling influence interventions, curriculum design, and
educational policy,” Crit. Rev. Soc. Sci. Stud., vol. 3, no. 2, pp. 1506–1521, 2025.
[9] H. Luo, “Prediction of Student Decision-Making Behaviour based on Machine
Learning Algorithms,” Pakistan J. Life Soc. Sci., vol. 22, no. 2, pp. 16382–16390, 2024.
[10] A. Mahmoud et al., “Policy Reviews in Higher Education Understanding the drivers
of student loan decision-making and its impact on graduation rates in Ghanaian
public universities,” Policy Rev. High. Educ., vol. 8, no. 1, pp. 85–101, 2024, doi:
10.1080/23322969.2024.2358008.
[11] Thelma, C. C., “Student retention in higher learning institutions of Zambia,” Int. J. Res.
Publ. Rev., vol. 5, no. 6, pp. 433–441, 2024.
[12] Gul, M. N., W. Abbasi, M. Z. Babar, A. Aljohani, and M. Arif, “Data driven decisions in
education using a comprehensive machine learning framework for student
performance prediction,” Discover Comput., vol. 28, no. 1, Art. no. 153, 2025.
[13] K. Patil, K. Yesugade, and K. B. Naikwadi, “A Study on Regression Based Machine
Learning Models to Predict the Student Performance,” J. Eng. Educ. Transform., pp.
177–186, 2024, doi: 10.16920/jeet/2024/v38i2/24200.
[14] H. Almaghrabi, B. Soh, and A. Li, “SoK : The Impact of Educational Data Mining on
Organisational Administration,” Information, vol. 15, no. 11, p. 738, 2024.
[15] H. Sahlaoui, E. L. Arbi, A. Alaoui, and M. M. Jaber, “Predicting and Interpreting Student
Performance Using Ensemble Models and Shapley Additive Explanations,” IEEE
Access, vol. 9, pp. 152688–152703, 2021, doi: 10.1109/ACCESS.2021.3124270.
[16] P. Boozary, S. Sheykhan, H. Ghorbantanhaei, and C. Magazzino, “International Journal
of Information Enhancing customer retention with machine learning : A
comparative analysis of ensemble models for accurate churn prediction,” Int. J. Inf.
Manag. Data Insights, vol. 5, no. 1, p. 100331, 2025, doi: 10.1016/[Link].2025.100331.
237 | Student Performance Classification Using Academic, Socioeconomic, and Digital …..
Vol. 8, No. 1, February 2026
[17] D. J. Lemay, C. Baek, and T. Doleck, “Computers and Education : Arti fi cial Intelligence
Comparison of learning analytics and educational data mining : A topic modeling
approach,” Comput. Educ. Artif. Intell., vol. 2, no. March, p. 100016, 2021, doi:
10.1016/[Link].2021.100016.
[18] M. Motevalli, “Comparative analysis of systematic, scoping, umbrella, and narrative
reviews in clinical research: critical considerations and future directions,” Int. J. Clin.
Pract., vol. 2025, no. 1, Art. no. 9929300, 2025.
[19] R. Awashreh, “Bridging Cultural Gaps : Enhancing Student Motivation and Academic
Integrity in Oman ’ s Universities,” Forum Linguist. Stud., vol. 07, no. 02, pp. 265–279,
2025.
[20] P. Banerjee, “Connecting the dots: A systematic review of explanatory factors
linking contextual indicators, institutional culture and degree awarding gaps,” High.
Educ. Eval. Dev., vol. 18, no. 1, pp. 31–52, 2024, doi: 10.1108/HEED-07-2023-0020.
[21] T. Getaneh, T. Zaw, and K. Jozsa, “Bridging theoretical gaps to improve students’
academic success in higher education in the digital era: A systematic literature
review,” Int. J. Educ. Res. Open, vol. 9, no. April, p. 100510, 2025, doi:
10.1016/[Link].2025.100510.
[22] E. N. Syamiya, S. Lestari, D. Wulandari, and D. Ekawati, “Digital literacy analysis on
online learning outcomes for macroeconomics with gender-mediated and family
socio-economics as moderating variables,” J. Kependidikan, vol. 8, no. 2, pp. 450–459,
2022.
[23] B. Maunah, “Social and cultural capital and learners’ cognitive ability: Issues and
prospects for educational relevance, access and equity towards digital
communication in Indonesia,” J. Soc. Stud. Educ. Res., vol. 11, no. 1, pp. 163–191, 2020.
[24] B. A. Al-sheeb, A. M. Hamouda, and G. M. Abdella, “Modeling of student academic
achievement in engineering education using cognitive and non-cognitive factors,”
vol. 11, no. 2, pp. 178–198, 2025, doi: 10.1108/JARHE-10-2017-0120.
[25] M. Mohzana, “The impact of the new student orientation program on the adaptation
process and academic performance,” Int. J. Educ. Narrat., vol. 2, no. 2, pp. 169–178,
2024.
[26] G. J. Palardy, “School peer non-academic skills and academic performance in high
school,” Front. Educ., vol. 4, p. 57, 2019, doi: 10.3389/feduc.2019.00057.
[27] M. Vaarma and H. Li, “Technology in Society Predicting student dropouts with
machine learning : An empirical study in Finnish higher education,” Technol. Soc., vol.
76, no. February, p. 102474, 2024, doi: 10.1016/[Link].2024.102474.
[28] S. Zhao, D. Zhou, H. Wang, and D. Chen, “Enhancing Student Academic Success
Prediction Through Ensemble Learning and Image-Based Behavioral Data
Transformation,” Appl. Sci., vol. 15, no. 3, p. 1231, 2025.
[29] T. K. Shoukath and M. Chakkaravarthy, “Predictive analytics in education: machine
learning approaches and performance metrics for student success—a systematic
literature review,” Data Metadata, vol. 4, Art. no. 730, 2025, doi: 10.56294/dm2025730.
[30] M. J. Shayegan and R. Akhtari, “A Stacking Machine Learning Model for Student
Performance Prediction Based on Class Activities in E-Learning,” Comput. Syst. Sci.
Eng., vol. 48, no. 5, 2024, doi: 10.32604/csse.2024.052587.
[31] T. T. Id and Z. Li, “Predicting learning achievement using ensemble learning with
result explanation,” PLoS One, vol. 20, no. 1, p. e0312124, 2025, doi:
10.1371/[Link].0312124.
239 | Student Performance Classification Using Academic, Socioeconomic, and Digital …..