Abstract
This study benchmarks multiple machine learning models to predict student academic performance. The research analyzes data from students in mathematics and Portuguese language courses, examining the relationship between various factors and academic performance. The benchmark implementation includes data preprocessing, exploratory data analysis, feature engineering, model training, and hyperparameter tuning for both regression (predicting final grades) and classification (predicting pass/ fail outcomes) tasks. The findings demonstrate that ensemble methods, particularly gradient boosting models, outperform other algorithms with root mean square error of 3.34 for regression and F1 score of 0.88 for classification after hyperparameter tuning. Feature importance analysis reveals that past failures, alcohol consumption, study time, and parent education level are among the most influential predictors of academic performance. The results provide valuable insights for educational stakeholders to implement targeted interventions for at-risk students and improve overall academic outcomes.
| Original language | English |
|---|---|
| Journal | International Journal of Information and Communication Technology Education |
| Volume | 22 |
| Issue number | 1 |
| DOIs | |
| State | Published - Jan 2026 |
Keywords
- Classification
- Clustering
- Data Mining
- Educational Data Mining (EDM)
- Feature Extraction
- Knowledge Discovery
- Personalized Learning
- Prediction
- Student’s Academic Performance
- Student’s Behavior
Fingerprint
Dive into the research topics of 'Forecasting Students' Academic Performance in Educational Data Using Machine Learning Techniques'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver