HYBRID DATA MINING METHODS TO SUPPORT MSME SUSTAINABILITY IN RURAL AND URBAN AREAS

Authors

  • Krisantus Jumarto Tey Seran Universitas Timor
  • Debora Chrisinta Universitas Timor
  • Yasinta Oktaviana Legu Rema Universitas Timor
  • Hevi Herlina Ullu Universitas Timor
  • Budiman Baso Universitas Timor

DOI:

https://doi.org/10.33480/jitk.v12i1.8271

Keywords:

Data Mining, K-Mode, MSME Sustainability, Naive Bayes, Categorical

Abstract

This study aims to analyze the factors influencing the sustainability of MSMEs in North Central Timor Regency by utilizing the K-Mode clustering method and Naive Bayes classification. The data used includes 550 MSMEs in North Central Timor Regency, East Nusa Tenggara Province, classified based on attributes such as location, product price, financial condition, innovation, technology utilization, and sustainability. The K-Mode method was employed to group MSMEs based on categorical similarities after the data was segmented by location attributes, while the Naive Bayes method was applied to classify MSME sustainability following clustering. The results indicate that in rural areas, MSMEs tend to dominate with high product prices and good financial conditions but show low levels of innovation and technology utilization. In contrast, MSMEs in urban areas are generally more innovative and technology-driven despite facing infra-structure challenges. The application of Naive Bayes demonstrated that a data training ratio of 70:30 yielded the best accuracy. Accordingly, the resulting model can be utilized to monitor the sustainability conditions of MSMEs. This study provides insights into sustainability patterns of MSMEs in both rural and urban areas and opens opportunities for further research on external factors affecting sustainability and barriers to technology adoption in rural areas.

Downloads

Download data is not yet available.

Author Biographies

  • Debora Chrisinta, Universitas Timor

    Teknologi Informasi

  • Yasinta Oktaviana Legu Rema, Universitas Timor

    Teknologi Informasi

  • Hevi Herlina Ullu, Universitas Timor

    Teknologi Informasi

  • Budiman Baso, Universitas Timor

    Teknologi Informasi

References

[1] R. R. Bakrie, S. Atikah Suri, S. A. Nabila, V. H Pratama, and Firmansyah., “Pengaruh Kreativitas UMKM Serta Kontribusinya Di Era Digitalisasi Terhadap Perekonomian Indonesia,” Jurnal Ekonomi Dan Bisnis, vol. 16, no. 2, pp. 82–88, Jul. 2024, doi: 10.55049/JEB.V16I2.308.

[2] S. Widaningsih, W. Muhamad, R. Hendriyanto, and H. Nugroho, “An ID3 Decision Tree Algorithm-Based Model for Predicting Student Performance Using Comprehensive Student Selection Data at Telkom University,” Ingénierie des Systèmes d’Information, vol. 28, no. 5, pp. 1205–1212, 2023, doi: 10.18280/isi.280508.

[3] M. D. Pangastuti and F. W. Nalle, “Determinan Kinerja Pelaku Usaha Kecil Menengah (UMKM) Pasar Perbatasan Kabupaten Timor Tengah Utara–Timor Leste,” Ekonomi dan Bisnis, vol. 11, no. 2, pp. 109–134, 2024, doi: 10.35590/jeb.v11i2.7334.

[4] E. Aminullah et al., “Interactive Components of Digital MSMEs Ecosystem for Inclusive Digital Economy in Indonesia,” Journal of the Knowledge Economy, vol. 15, pp. 487–517, 2024, doi: 10.1007/s13132-022-01086-8.

[5] L. Anatan and Nur, “Micro, Small, and Medium Enterprises’ Readiness for Digital Transformation in Indonesia,” Economies, vol. 11, no. 6, p. 156, 2023, doi: 10.3390/economies11060156.

[6] F. Faiz, V. Le, and E. K. Masli, “Determinants of Digital Technology Adoption in Innovative SMEs,” Journal of Innovation & Knowledge, vol. 9, no. 4, p. 100610, 2024, doi: 10.1016/j.jik.2024.100610.

[7] T. Yuwono, A. Suroso, and W. Novandari, “Information and Communication Technology in SMEs: A Systematic Literature Review,” Journal of Innovation and Entrepreneurship, vol. 13, art. 31, 2024, doi: 10.1186/s13731-024-00392-6.

[8] S. Chanmee and K. Kesorn, “Semantic Decision Trees: A New Learning System for The ID3-Based Algorithm using a Knowledge Base,” Advanced engineering informatics, vol. 58, p. 102156, 2023, doi: 10.1016/j.aei.2023.102156.

[9] K. J. T. Seran, D. Chrisinta, and Y. O. L. Rema, “Sistem Pendukung Keputusan Keberlanjutan UMKM Menggunakan Algoritma ID3 di Wilayah Perbatasan RI-RDTL,” Decode: Jurnal Pendidikan Teknologi Informasi, vol. 5, no. 1, pp. 41–53, 2025, doi: 10.51454/decode.v5i1.946.

[10] Terttiaavini, “A Hybrid Approach Using K-Means Clustering and the SAW Method for Evaluating and Determining the Priority of SMEs in Palembang City,” Journal of Intelligent System and Computation, vol. 6, no. 1, pp. 46–53, 2024, doi: 10.52985/insyst.v6i1.392.

[11] Y. Kustiyahningsih, E. Rahmanita, A. Khozaimi, Y. D. P. Negara, B. K. Khotimah, and J. Purnama, “SCM-SCOR and K-Means Approaches for Clustering MSME Batik Bangkalan Madura Indonesia,” AIP Conference Proceedings, vol. 3250, p. 050012, 2025, doi: 10.1063/5.0240736.

[12] D. E. Putri and E. P. W. Mandala, “Hybrid Data Mining berdasarkan Klasterisasi Produk untuk Klasifikasi Penjualan,” Jurnal KomtekInfo, vol. 9, no. 2, pp. 68–73, Jun. 2022, doi: 10.35134/KOMTEKINFO.V9I2.279.

[13] G. Dwilestari and T. A. Afifah, “Perbandingan Kinerja Algoritma Naive Bayes Dan Decision Tree Dalam Klasifikasi Kanker Paru-Paru,” JATI (Jurnal Mahasiswa Teknik Informatika), vol. 9, no. 1, pp. 801–807, 2025, doi: 10.36040/jati.v9i1.12463.

[14] M. Z. Haq, C. S. Octiva, A. Ayuliana, U. W. Nuryanto, and D. Suryadi, “Algoritma Naïve Bayes untuk Mengidentifikasi Hoaks di Media Sosial,” Jurnal Minfo Polgan, vol. 13, no. 1, pp. 1079–1084, Jul. 2024, doi: 10.33395/JMP.V13I1.13937.

[15] A. Holl and R. Rama, “Spatial Patterns and Drivers of SME Digitalisation,” Journal of the Knowledge Economy, vol. 15, pp. 5625–5649, 2024, doi: 10.1007/s13132-023-01257-1.

[16] V. Çetin and O. Yıldız, “A Comprehensive Review on Data Preprocessing Techniques in Data Analysis,” Pamukkale Üniversitesi Mühendislik Bilimleri Dergisi, vol. 28, no. 2, pp. 299–312, 2022, doi: 10.5505/pajes.2021.62687.

[17] P. Koukaras and C. Tjortjis, “Data Preprocessing and Feature Engineering for Data Mining: Techniques, Tools, and Best Practices,” AI, vol. 6, no. 10, p. 257, 2025, doi: 10.3390/ai6100257.

[18] D. Chrisinta and J. Eduardo Simarmata, “Eksplorasi Teknik Web Scraping pada Data Mining: Pendekatan Pencarian Data Berbasis Python,” Faktor Exacta, vol. 17, no. 1, pp. 1979–276, 2024, doi: 10.30998/faktorexacta.v17i1.22393.

[19] O. Peretz, M. Koren, and O. Koren, “Naive Bayes Classifier—An Ensemble Procedure for Recall and Precision Enrichment,” Engineering Applications of Artificial Intelligence, vol. 136, p. 108972, 2024, doi: 10.1016/j.engappai.2024.108972.

[20] X. Ye, Y. Zhu, S. Zhang, H. Deng, and P. Yu, “Non-Interactive K-Mode Clustering of High-Dimensional Categorical Data under Local Differential Privacy,” Inf. Sci. (N. Y)., vol. 718, p. 122417, 2025, doi: 10.1016/j.ins.2025.122417.

[21] B. Supri, B. Mawadah, and H. Ali, “Asian Stock Index Price Prediction Analysis Using Comparison of Split Data Training and Data Testing,” JEMSI (Jurnal Ekonomi, Manajemen, Dan Akuntansi), vol. 9, no. 4, pp. 1403–1408, Aug. 2023, doi: 10.35870/JEMSI.V9I4.1339.

[22] I. Muraina, “Ideal Dataset Splitting Ratios in Machine Learning Algorithms: General Concerns for Data Scientists and Data Analysts,” in 7th International Mardin Artuklu Scientific Researches Conference, 2022, pp. 496–504.

[23] S. S. Prasetiyowati and Y. Sibaroni, “Unlocking the Potential of Naive Bayes for Spatio Temporal Classification: A Novel Approach to Feature Expansion,” Journal of Big Data, vol. 11, art. 106, 2024, doi: 10.1186/s40537-024-00958-x.

[24] D. Chrisinta, J. S.-K. J. S. Komputer, and undefined 2023, “Analisis Sentimen Penilaian Masyarakat Terhadap Pejabat Publik Menggunakan Algoritma Naïve Bayes Classifier,” ojs.unikom.ac.idD Chrisinta, JE SimarmataKomputika: Jurnal Sistem Komputer, 2023•ojs.unikom.ac.id, vol. 12, no. 1, p. 2020, 2023, doi: 10.34010/komputika.v12i1.9638.

[25] J. E. Simarmata, D. Chrisinta, and M. Purnomo, “Implementation of K-Means Clustering to Human Development Indicators in East Nusa Tenggara,” Journal of Research in Mathematics Trends and Technology, vol. 6, no. 2, pp. 46–56, Sep. 2024, doi: 10.32734/jormtt.v6i2.17066.

[26] A. Lal, A. Sharan, K. Sharma, A. Ram, D. K. Roy, and B. Datta, “Scrutinizing Different Predictive Modeling Validation Methodologies and Data-Partitioning Strategies: New Insights Using Groundwater Modeling Case Study,” Environmental Monitoring and Assessment, vol. 196, art. 623, 2024, doi: 10.1007/s10661-024-12794-w.

[27] B. Supri, Rudianto, Abdurohim, Badriatul Mawadah, and Helmi Ali, “Asian Stock Index Price Prediction Analysis Using Comparison of Split Data Training and Data Testing,” JEMSI (Jurnal Ekonomi, Manajemen, dan Akuntansi), vol. 9, no. 4, pp. 1403–1408, 2023, doi: 10.35870/jemsi.v9i4.1339.

[28] D. El-Shahat et al., “Machine Learning and Deep Learning Models Based Grid Search Cross Validation for Short-Term Solar Irradiance Forecasting,” Journal of Big Data, vol. 11, art. 134, 2024, doi: 10.1186/s40537-024-00991-w.

[29] T. Q. A’yuni, B. N. Febriati, L. I. Effendie, M. Muhajir, and R. Yotenka, “MSME Sales Clustering Based on Business Aid Distribution Priority Using K-Affinity Propagation,” Enthusiastic: International Journal of Applied Statistics and Data Science, vol. 3, no. 1, pp. 111–124, 2023, doi: 10.20885/enthusiastic.vol3.iss1.art10.

[30] M. Conciatori, A. Valletta, and A. Segalini, “Improving the Quality Evaluation Process of Machine Learning Algorithms Applied to Landslide Time Series Analysis,” Computers & Geosciences, vol. 184, p. 105531, 2024, doi: 10.1016/j.cageo.2024.105531.

Downloads

Published

2026-08-11

How to Cite

[1]
“HYBRID DATA MINING METHODS TO SUPPORT MSME SUSTAINABILITY IN RURAL AND URBAN AREAS”, jitk, vol. 12, no. 1, pp. 152–162, Aug. 2026, doi: 10.33480/jitk.v12i1.8271.

Most read articles by the same author(s)