Penerapan Integrasi Algoritma K-Means Dan Naïve Bayes Untuk Klasifikasi Wilayah Rawan Banjir Di Jakarta
DOI:
https://doi.org/10.31294/coscience.v5i2.6900Keywords:
Flood, Data Mining, Naïve Bayes, K-MeansAbstract
Jakarta, as a metropolitan city in Indonesia, often experiences flooding caused by high rainfall, poor drainage systems, and rapid urbanization. This research aims to classify flood-prone areas in Jakarta using a combination of K-Means Clustering and Naïve Bayes Classifier algorithms. The research phase begins with data collection from the Satu Data Jakarta website, including attributes such as region, sub-district, village, average water level, number of affected RWs, number of affected families, number of affected people, and number of flood events. The collected data is then processed through cleaning and normalization stages before being analyzed using the K-Means algorithm to group areas based on their flooding characteristics. Furthermore, the Naïve Bayes algorithm was used to build a classification model that predicts flood-prone areas. The results showed that the combination of these two algorithms resulted in higher average accuracy compared to the use of conventional Naïve Bayes, having an accuracy of 98.18%% at training and testing data split ratios of 70:30, 80;20 and 90:10. The findings provide valuable insights for flood risk mitigation in Jakarta, assisting the government in taking more effective preventive measures.
Downloads
References
Anggraini, N., Pangaribuan, B., Siregar, A. P., Sintampalam, G., Muhammad, A., Ridha, M., Damanik, S., & Rahmadi, T. (2021). ANALISIS PEMETAAN DAERAH RAWAN BANJIR DI KOTA MEDAN TAHUN 2020.
Angreini, S., & Supratman, E. (2021). Visualisasi Data Lokasi Rawan Bencana Di Provinsi Sumatera Selatan Menggunakan Tableau. Jurnal Nasional Ilmu Komputer , 2.
Amril Mutoi Siregar, S. Kom. , M. Kom., & Adam Puspabhuana, S. Kom. , M. Kom. (2017). DATA MINING: Pengolahan Data Menjadi Informasi dengan RapidMiner (Cetakan pertama). CV Kekata Group
Badan Nasional Penanggulangan Bencana. (2023). BUKU IRBI 2023.
Bui, M. A., & Bahtiar, A. (2024). IMPLEMENTASI METODE ALGORITMA K-MEANS CLUSTERING UNTUK MENGELOMPOKKAN TRANSAKSI PENJUALAN BARANG DI TOKO ARINO. In Jurnal Mahasiswa Teknik Informatika (Vol. 8, Issue 2).
Chikalkar, S. N. (2020). Knowledge Discovery and Data Mining. International Journal for Research in Applied Science and Engineering Technology, 8(10), 874–876. https://doi.org/10.22214/ijraset.2020.32045
Erick, B., Akhmad, F., & Ibnu Prasetyo, W. (2024). Application of Naive Bayes Algorithm for Physical Fitness Level Classification. International Journal of Disabilities Sports and Health Sciences, 7(1), 178–187. https://doi.org/10.33438/ijdshs.1330745
Fatonah, N. S., Buana, M., Selatan, J. M., Kembangan, K., Barat, J., Khusus, D., Jakarta, I., & Com, N. (2021). Penerapan Deteksi Bencana Banjir Menggunakan Metode Machine Learning.
Khomsiyah, J., Ramdhani, A., Damayanti, A. F., & Rohman, D. (2021). PENERAPAN ALGORITMA K-MEANS CLUSTERING UNTUK PENGELOMPOKAN WILAYAH RAWAN BANJIR.
Lestari, P. I., Ratnawati, D. E., & Muflikhah, L. (2019). Implementasi Algoritme K-Means Clustering Dan Naive Bayes Classifier Untuk Klasifikasi Diagnosa Penyakit Pada Kucing (Vol. 3, Issue 1). http://j-ptiik.ub.ac.id
Makmun Effendi, M., & Siswandi, A. (2024). Analysis Prediksi Wilayah Rawan Banjir dengan Algoritma K-Means. Journal of Information System Research (JOSH), 5(2), 697–703. https://doi.org/10.47065/josh.v5i2.4770
Nandang Iriadi, Priatno, & Ahmad Ishaq. (2020). Penerapan Data Mining dengan Rapid Miner; Konsep Data Mining, Data Warehouse, Metode, Model, Teknik. Graha Ilmu.
Nigam, N., & Rajavat, A. (2020). A Systematic Literature Review of Data Classification Techniques. In International Journal of Computer Applications (Vol. 177, Issue 44). https://www.sas.com/en_us/insights/analytics/data-
Ridwan, A. (2020). Penerapan Algoritma Naïve Bayes Untuk Klasifikasi Penyakit Diabetes Mellitus.
Rikhi, N. (2015). Data Mining and Knowledge Discovery in Database. International Journal of Engineering Trends and Technology, 2. http://www.ijettjournal.org
Riyanto, S., Sitanggang, I. S., Djatna, T., & Atikah, T. D. (2023). Comparative Analysis using Various Performance Metrics in Imbalanced Data for Multi-class Text Classification. In IJACSA) International Journal of Advanced Computer Science and Applications (Vol. 14, Issue 6). http://gcancer.org/pdr
Sirichanya, C., & Kraisak, K. (2021). Semantic data mining in the information age: A systematic review. International Journal of Intelligent Systems, 36(8), 3880–3916. https://doi.org/https://doi.org/10.1002/int.22443
Wahab, A., Samarinda, S., Lishania, I., Goejantoro, R., & Nasution, Y. N. (2019). Perbandingan Klasifikasi Metode Naive Bayes dan Metode Decision Tree Algoritma (J48) pada Pasien Penderita Penyakit Stroke di RSUD Comparison of the Classification for Naive Bayes Method and the Decision Tree Algorithm (J48) for Stroke Patients in Abdul Wahab
Sjahranie Samarinda Hospital. Jurnal EKSPONENSIAL, 10(2).
Zhang, X. (2020). Research on Data Mining Algorithm Based on Pattern Recognition. International Journal of Pattern Recognition and Artificial Intelligence, 34(6). https://doi.org/10.1142/S0218001420590156
Zai, C. (2022). IMPLEMENTASI DATA MINING SEBAGAI PENGOLAHAN DATA. In Portaldata.org (Vol. 2, Issue 3).
Downloads
Published
Issue
Section
License
Copyright (c) 2025 Irfan Maulana Sinatrya Sinatrya, Achmad Baroqah Pohan, Yunita Yunita, Hilda Amalia, Ade Fitria Lestari

This work is licensed under a Creative Commons Attribution-ShareAlike 4.0 International License.


















