Studi Komparatif Algoritma K-Means dan K-Medoids untuk Segmentasi Informasi Kesehatan

Authors

  • Muhammad Dwi Ananda Universitas Pembangunan Nasional "Veteran" Jakarta
  • Mardiah Mardiah Universitas Pembangunan Nasional "Veteran" Jakarta
  • Anis Fitri Nur Masruriyah Universitas Pembangunan Nasional "Veteran" Jakarta
  • Karenina Nurmelita Malik Universitas Pembangunan Nasional "Veteran" Jakarta

DOI:

https://doi.org/10.31294/coscience.v5i2.9207

Keywords:

Information Segmentation, Clustering, K-Means, K-Medoids, Health Data

Abstract

In analyzing medical data to support clinical decisions, segmentation of health information plays a crucial role. This study presents a comparative analysis of K-Means and K-Medoids algorithms in clustering Medical Examination data. This evaluation is conducted using two main internal approaches, namely Silhouette Score and Davies-Bouldin Index in measuring the quality of separation as well as cohesion between clusters. The experiment involved varying the number of clusters to determine the optimal configuration of each algorithm. The results show that K-Means provides representative performance and is more stable against data complexity, compared to the K-Medoids algorithm which is only optimal in a small number of clusters. Statistical analysis using one-way ANOVA was applied to test the significance of performance differences between algorithms based on the average Silhouette Score value, yielding an F-value of 4.8594 with a P-value of 0.0447. This indicates that the performance difference between the two algorithms is statistically significant at 5% significance rate. This research confirms the K-Means algorithm for segmenting health data with diverse distributions and is expected to serve as a foundation for the development of more efficient health data classification systems in the future.

Downloads

Download data is not yet available.

References

Arif, Alfarez, D. A., & Ramadhan, M. R. (2023). Anova dan Tukey HSD Perbandingan Produksi Padi Antara Tiga Kabupaten di Provinsi Jambi. Multi Proximity: Jurnal Statistika, 2(1), Article 1. https://doi.org/10.22437/multiproximity.v2i1.25908

Christnatalis -, Claudyo, E., Lucky -, Manullang, H. K., & Zebua, A. I. (2023). ANALISIS PELAYANAN RUMAH SAKIT UMUM DENGAN PERBANDINGAN ANTARA METODE ALGORITMA KMEANS, DAN K-MEDOIDS CLUSTERING. JURNAL TEKNOLOGI DAN ILMU KOMPUTER PRIMA (JUTIKOMP), 6(2), Article 2. https://doi.org/10.34012/jutikomp.v6i2.4145

Fadil, A., & Fatah, Z. (2025). ALGORITMA K-MEANS CLUSTERING UNTUK MENENTUKAN SISWA UNGGULAN BERDASARKAN HASIL UJIAN DI SEKOLAH. Jurnal Riset Sistem Informasi, 2(1), Article 1. https://doi.org/10.69714/p26gcf27

Fira, A., Rozikin, C., & Garno, G. (2021). Komparasi Algoritma K-Means dan K-Medoids Untuk Pengelompokkan Penyebaran Covid-19 di Indonesia. Journal of Applied Informatics and Computing, 5(2), Article 2. https://doi.org/10.30871/jaic.v5i2.3286

Hasan, S. (2025). Medical Examination Dataset. https://www.kaggle.com/datasets/jazidesigns/medical-examination-dataset

Hendrastuty, N. (2024). Penerapan Data Mining Menggunakan Algoritma K-Means Clustering Dalam Evaluasi Hasil Pembelajaran Siswa. Jurnal Ilmiah Informatika Dan Ilmu Komputer (JIMA-ILKOM), 3(1), Article 1. https://doi.org/10.58602/jima-ilkom.v3i1.26

Joshi, M. A. P., & Patel, B. V. (2021). Data Preprocessing: The Techniques for Preparing Clean and Quality Data for Data Analytics Process. Oriental Journal of Computer Science and Technology, 13(2,3), 78–81. https://doi.org/10.13005/ojcst13.0203.03

Leis, A. M., McSpadden, E., Segaloff, H. E., Lauring, A. S., Cheng, C., Petrie, J. G., Lamerato, L. E., Patel, M., Flannery, B., Ferdinands, J., Karvonen-Gutierrez, C. A., Monto, A., & Martin, E. T. (2023). K-medoids clustering of hospital admission characteristics to classify severity of influenza virus infection. Influenza and Other Respiratory Viruses, 17(3), e13120. https://doi.org/10.1111/irv.13120

Meiriza, A., Ali, E., Rahmiati, & Agustin. (2023). Perbandingan Algoritma K-Means dan K-Medoids untuk Pengelompokan Program BPJS Ketenagakerjaan. The Indonesian Journal of Computer Science, 12(2). https://doi.org/10.33022/ijcs.v12i2.3184

Momahhed, S. S., Emamgholipour Sefiddashti, S., Minaei, B., & Shahali, Z. (2023). K-means clustering of outpatient prescription claims for health insureds in Iran. BMC Public Health, 23(1), 788. https://doi.org/10.1186/s12889-023-15753-1

Ningrum, H., Irawan, E., & Lubis, M. R. (2021). Implementasi Metode K-Medoids Clustering Dalam Pengelompokan Data Penyakit Alergi Pada Anak. Jurasik (Jurnal Riset Sistem Informasi Dan Teknik Informatika), 6(1), Article 1. https://doi.org/10.30645/jurasik.v6i1.277

Nirwana, S. D., Jambak, M. I., & Bardadi, A. (2022). PERBANDINGAN ALGORITMA K-MEANS DAN K-MEDOIDS DALAM CLUSTERING RATA-RATA PENAMBAHAN KASUS COVID-19 BERDASARKAN KOTA/KABUPATEN DI PROVINSI SUMATERA SELATAN. JSiI (Jurnal Sistem Informasi), 9(2), Article 2. https://doi.org/10.30656/jsii.v9i2.5127

Nurhalizah, R. S., Ardianto, R., & Purwono, P. (2024). Analisis Supervised dan Unsupervised Learning pada Machine Learning: Systematic Literature Review. Jurnal Ilmu Komputer dan Informatika, 4(1), Article 1. https://doi.org/10.54082/jiki.168

Permana, I., & Salisah, F. N. S. (2022). Pengaruh Normalisasi Data Terhadap Performa Hasil Klasifikasi Algoritma Backpropagation: The Effect of Data Normalization on the Performance of the Classification Results of the Backpropagation Algorithm. Indonesian Journal of Informatic Research and Software Engineering (IJIRSE), 2(1), Article 1. https://doi.org/10.57152/ijirse.v2i1.311

Permatasari, R., Sukirman, & Fazza, F. E. (2024). Perbandingan Optimasi Penyuluhan Penyakit Stunting Pada Balita Integrasi Algoritma K- Means Dan Partitioning Around Medoids (PAM). Jurnal Ilmu Komputer Dan Teknologi Informasi, 1(2), Article 2. https://doi.org/10.71466/jiktif.v1i2.43

Purba, W. N., Sembiring, G. A., Turnip, M. T., Saputra, A., & Manihuruk, B. J. I. (2023). PENERAPAN DATA MINING UNTUK PENGELOLAAN DATA REKAM MEDIS MENGGUNAKAN METODE K-MEANS CLUSTERING PADA RUMAH SAKIT ROYAL PRIMA MEDAN. Jurnal Teknik Informasi Dan Komputer (Tekinkom), 6(1), Article 1. https://doi.org/10.37600/tekinkom.v6i1.857

Razaki, A., Chrisnanto, Y. H., & Melina, M. (2024). Penanganan Outlier Pada Metode Algoritma K- Nearest Neighbors (KNN) Dengan Metode Kernel Density Estimation Pada Kasus Penyakit Diabetes. INTECOMS: Journal of Information Technology and Computer Science, 7(4), 1177–1188. https://doi.org/10.31539/intecoms.v7i4.10866

Romli, I. (2021). PENERAPAN DATA MINING MENGGUNAKAN ALGORITMA K-MEANS UNTUK KLASIFIKASI PENYAKIT ISPA. Indonesian Journal of Business Intelligence (IJUBI), 4(1), Article 1. https://doi.org/10.21927/ijubi.v4i1.1727

Safitri, E. M. (2024). Clustering Study Of Hospitals In Bojonegoro Based On Health Workers With K-Means And K-Medoids Methods. Jurnal Statistika Dan Komputasi, 3(2), Article 2. https://doi.org/10.32665/statkom.v3i2.3592

Samosir, F. V. P., Mustamu, L. P., Anggara, E. D., Wiyogo, A. I., & Widjaja, A. (2021). Exploratory Data Analysis terhadap Kepadatan Penumpang Kereta Rel Listrik. Jurnal Teknik Informatika Dan Sistem Informasi, 7(2), Article 2. https://doi.org/10.28932/jutisi.v7i2.3700

Sari, Z. D. R., Arvita, Y., & Jasmir, J. (2024). Penerapan Data Mining Untuk Prediksi Penyakit Diabetes Menggunakan Algoritma C4.5 Zudyanti Dwi Rahma Sari1, Ja. Jurnal Informatika Dan Rekayasa Komputer(JAKAKOM), 4(1), 827–834. https://doi.org/10.33998/jakakom.2024.4.1.1624

Setiawan, M. A. M., Kusrini, K., & Hartono, A. D. (2025). Menggunakan Metode Machine Learning Untuk Memprediksi Nilai Mahasiswa Dengan Model Prediksi Multiclass. Jurnal Informatika: Jurnal Pengembangan IT, 10(1), Article 1. https://doi.org/10.30591/jpit.v10i1.8334

Syamfithriani, T. S., Mirantika, N., & Trisudarmo, R. (2023). Perbandingan Algoritma K-Means dan K-Medoids Untuk Pemetaan Daerah Penanganan Diare Pada Balita di Kabupaten Kuningan. Jurnal Sistem Informasi Bisnis, 12(2), 132–139. https://doi.org/10.21456/vol12iss2pp132-139

Utomo, W. (2021). The comparison of k-means and k-medoids algorithms for clustering the spread of the covid-19 outbreak in Indonesia. ILKOM Jurnal Ilmiah, 13(1), Article 1. https://doi.org/10.33096/ilkom.v13i1.763.31-35

Wahyudi, E. E., Auzan, M., Dharmawan, A., Nuryanto, D. E., Susyanto, N., Samodra, G., & Hadmoko, D. S. (2022). Akuisisi Data Prediksi Curah Hujan Secara Periodik Menggunakan Apache Airflow. Journal of Informatics Information System Software Engineering and Applications (INISTA), 4(2), Article 2. https://doi.org/10.20895/inista.v4i2.574

Downloads

Published

2025-07-22

Issue

Section

Articles

How to Cite

Studi Komparatif Algoritma K-Means dan K-Medoids untuk Segmentasi Informasi Kesehatan. (2025). Computer Science (CO-SCIENCE), 5(2), 103-112. https://doi.org/10.31294/coscience.v5i2.9207