Feasibility Analysis of Education in Indonesian Provinces: A Machine Learning Clustering Approach for Regional Classification
Downloads
Education system feasibility varies significantly across Indonesian provinces due to geographic, socioeconomic, and infrastructural disparities. This study applies unsupervised machine learning clustering algorithms (K-Means, Hierarchical Clustering, Gaussian Mixture Models, and Self-Organizing Maps) to classify Indonesian provinces into homogeneous groups based on education feasibility indicators. Using 39 provinces and special territories, we evaluated clustering quality through internal validation metrics (Silhouette coefficient and Davies-Bouldin index) and inter-algorithm agreement measures (Adjusted Rand Index and Normalized Mutual Information). Results demonstrate that Hierarchical Clustering achieves the best Davies-Bouldin index (0.782), while K-Means and GMM produce identical partitions (ARI = 1.0, NMI = 1.0). Self-Organizing Maps identified nine distinct regional education feasibility profiles, with provincial distributions revealing significant heterogeneity in education conditions. These findings provide a quantitative framework for targeted policy interventions and resource allocation to improve education feasibility across Indonesia’s diverse regions.
UNESCO. (2020). Global monitoring report: Education for all. United Nations Educational, Scientific and Cultural Organization.
UNESCO. (2017). Indonesia education sector: Regional challenges and the global sustainable development agenda. UNESCO Regional Office for South Asia.
World Bank. (2021). Indonesia economic prospects: Regional disparities in education and development. World Bank Group, East Asia and Pacific Region.
Jain, A. K. (2010). Data clustering: 50 years beyond K-means. Pattern Recognition Letters, 31(8), 651-666.
Lanjouw, P., Pradhan, M., Saadah, F., Sayed, H., & Sparrow, R. (2002). Poverty, education and health in Indonesia: Who benefits from public spending?. Education and Health Expenditures, and Development: The Cases of Indonesia and Peru. Development Centre Studies, OECD Development Centre, Paris, 17-78.
Chen, D. (2009). The economics of teacher supply in Indonesia. World Bank Policy Research Working Paper, (4975).
Nizar, N. I., Nuryartono, N., Juanda, B., & Fauzi, A. (2024). Can knowledge and culture eradicate poverty and reduce income inequality? The evidence from Indonesia. Journal of the Knowledge Economy, 15(2), 6425-6450.
Usman, S., Akhmadi, & Suryadarma, D. (2007). Patterns of teacher absence in public primary schools in Indonesia. Asia Pacific Journal of Education, 27(2), 207-219.
Setyadi, S. (2022). Inequality of Education in Indonesia by Gender, Socioeconomic Background and Government Expenditure. Eko-Regional, 17(1), 383462.
Mufic, J. (2022). Measurable but not quantifiable’: The Swedish Schools Inspectorate on construing “quality” as “auditable. International Journal of Lifelong Education, 41(2), 199-211.
Rahma, F., & Ulfah, S. Z. (2025). Clustering students based on academic performance and social factors: an unsupervised learning approach to identify student patterns. International Journal for Applied Information Management, 5(3), 139-154.
Hooshyar, D., Yang, Y., Pedaste, M., & Huang, Y. M. (2020). Clustering algorithms in an educational context: An automatic comparative approach. IEEE Access, 8, 146994-147014.
MacQueen, J. (1967). Some methods for classification and analysis of multivariate observations. Proceedings of the 5th Berkeley Symposium on Mathematical Statistics and Probability, 1, 281-297.
Ward, J. H. (1963). Hierarchical grouping to optimize an objective function. Journal of the American Statistical Association, 58(301), 236-244.
Murtagh, F., & Contreras, P. (2012). Algorithms for hierarchical clustering: An overview. Wiley Interdisciplinary Reviews: Data Mining and Knowledge Discovery, 2(1), 86-97.
Reynolds, D. A. (2015). Gaussian mixture models. Encyclopedia of Biometrics, 827-832. Springer.
Fraley, C., & Raftery, A. E. (2002). Model-based clustering, discriminant analysis, and density estimation. Journal of the American Statistical Association, 97(458), 611-631.
Kohonen, T. (2001). Self-organizing maps (3rd ed.). Springer Series in Information Sciences.
Ultsch, A., & Mörchen, F. (2005). ESOM maps: Tools for clustering, visualization, and classification of high-dimensional data. Technical Report No. 36, Department of Mathematics and Computer Science, University of Marburg.
Rousseeuw, P. J. (1987). Silhouettes: A graphical aid to the interpretation and validation of cluster analysis. Journal of Computational and Applied Mathematics, 20, 53-65.
Hubert, L., & Arabie, P. (1985). Comparing partitions. Journal of Classification, 2(1), 193-218.
Strehl, A., Ghosh, J., & Mooney, R. (2003). Impact of similarity measures on web-page clustering. AAAI Workshop on Semantic Web, 58-64.
