Cluster Analysis of Alumni Job Waiting Times Using the K-Means Algorithm with Elbow and Silhouette Methods

Authors

  • Abdur Ro’uf Institut Teknologi dan Bisnis Widya Gama Lumajang
  • Hasyim Asy’ari Institut Teknologi dan Bisnis Widya Gama Lumajang
  • Maysas Yafi Urrochman Institut Teknologi dan Bisnis Widya Gama Lumajang

DOI:

https://doi.org/10.30741/jid.v5i1.2080

Keywords:

K-Means, Clustering, Elbow, Silhouette, Tracer Study, Employment Waiting Period

Abstract

Evaluating graduate employability is a key indicator for assessing the quality of higher education institutions. One effective approach involves analyzing alumni tracer study data using data mining techniques. This study aims to categorize alumni data based on the waiting period for employment after graduation using the K-Means algorithm. To determine the optimal number of clusters, two methods were employed: the Elbow method (Sum of Squared Errors/SSE) and the Silhouette Score. The research process included data preprocessing—comprising data selection, transformation, and normalization—to enhance data quality prior to clustering. Test results indicated that the Elbow method yielded an optimal cluster count of k=4, whereas the Silhouette method showed the best result at k=3, with a score of 0.62. Consequently, k=3 was selected as the optimal number of clusters due to superior cluster quality. The clustering results revealed that the majority of alumni (67.6%) experienced a short waiting period (averaging ±2 months), while 16.5% faced a long waiting period (±8 months), and 15.9% fell into the medium category (±5 months). These findings are expected to serve as a basis for higher education institutions to evaluate and improve graduate quality, ensuring better alignment with the demands of the job market.

References

Alam, A., & Muqeem, M. (2023). Hybridization of K-means with improved firefly algorithm for automatic clustering in high dimension. ArXiv.

Çetin, V., & Yıldız, O. (2022). A comprehensive review on data preprocessing techniques in data analysis. Pamukkale University Journal of Engineering Sciences, 28(2), 299–312. https://doi.org/10.5505/pajes.2021.62687

Dewi, D. A. I. C., & Pramita, D. A. K. (2019). Comparative Analysis of Elbow and Silhouette Methods On K-Medoids Clustering Algorithm in Balinese Craft Production Grouping. Matrix Journal, 9(3), 102–109.

Dinh, D. T., Fujinami, T., & Huynh, V. N. (2019). Estimating the Optimal Number of Clusters in Categorical Data Clustering by Silhouette Coefficient. Communications in Computer and Information Science, 1103 CCIS, 1–17. https://doi.org/10.1007/978-981-15-1209-4_1

Eliyanto, J., & Surono, S. (2022). The Suitable Distance Function for Fuzzy C-Means Clustering. AIP Conference Proceedings, 2578(November). https://doi.org/10.1063/5.0106185

Muhammad Raqib Syahkur, Hartama, D., & Solikhun, S. (2024). Evaluasi Jumlah Cluster pada Algoritma K-Means++ Menggunakan Silhouette dan Elbow dengan Validasi Nilai DBI dalam Mengelompokkan Gizi Balita. JST (Jurnal Sains Dan Teknologi), 13(3), 487–496. https://doi.org/10.23887/jstundiksha.v13i3.86419

Piao Tan, M., & A. Floudas, C. (2008). Determining the Optimal Number of Clusters. Encyclopedia of Optimization, 1, 687–694. https://doi.org/10.1007/978-0-387-74759-0_123

Prayogi, A., & , Irfandi, M. A. K. (2024). Jurnal Multidisiplin Ilmu Nasional Pendekatan Kualitatif dan Kuantitatif. Complex : Jurnal Multidisiplin Ilmu Nasional, 1(2), 31. https://ejurnal.faaslibsmedia.com/index.php/complex/article/view/7/28

Roring, M. A. C., Ginoga, P. A., Ibrahim, R., Rompas, R. C., Yusupa, A., & Paturusi, S. D. E. (2025). E-issn : 2988-1986. Multidisiplin Saintek, 8(5), 1–16. https://cibangsa.com/index.php/kohesi/article/view/6074/5284

Rouf, A., Qoritunnadyah, M., Asyari, H., & Urrohman, M. Y. (2024). Clustering of Lecturer Performance Using K-Means. Journal of Informatics Development, 3(1), 27–33. https://doi.org/10.30741/jid.v3i1.1430

Schubert, E. (2023). Stop using the elbow criterion for k-means and how to choose the number of clusters instead. ACM SIGKDD Explorations Newsletter, 25(1), 36–42. https://doi.org/10.1145/3606274.3606278

Singh, M. P., Gayathri, V., & Chaudhuri, D. (2022). A Simple Data Preprocessing and Postprocessing Techniques for SVM Classifier of Remote Sensing Multispectral Image Classification. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 15, 7248–7262. https://doi.org/10.1109/JSTARS.2022.3201273

Suraya, S., Sholeh, M., & Lestari, U. (2023). Evaluation of Data Clustering Accuracy using K-Means Algorithm. International Journal of Multidisciplinary Approach Research and Science, 2(01), 385–396. https://doi.org/10.59653/ijmars.v2i01.504

Syahfitri, N., Budianita, E., Nazir, A., & Afrianty, I. (2023). Pengelompokan Produk Berdasarkan Data Persediaan Barang Menggunakan Metode Elbow dan K-Medoid. KLIK: Kajian Ilmiah Informatika Dan Komputer, 4(3), 1668–1675. https://doi.org/10.30865/klik.v4i3.1525

Undari Sulung, M. M. (2024). Jurnal Edu Research Indonesian Institute For Corporate Learning And Studies (IICLS) Page 25. Jurnal Edu Research : Indonesian Institute For Corporate Learning And Studies (IICLS), 5(2), 28–33.

Zhao, L., Fang, J., Ji, Y., Zhang, Y., Zhou, X., Yin, J., Zhang, M., & Bao, W. (2023). K-means cluster analysis of characteristic patterns of allergen in different ages: Real life study. Clinical and Translational Allergy, 13(7). https://doi.org/10.1002/clt2.12281

Downloads

Published

2026-10-01

How to Cite

Ro’uf, A., Asy’ari, H., & Urrochman, M. Y. (2026). Cluster Analysis of Alumni Job Waiting Times Using the K-Means Algorithm with Elbow and Silhouette Methods. Journal of Informatics Development, 5(1), 7–17. https://doi.org/10.30741/jid.v5i1.2080