Beta Distribution Weighted Fuzzy C-Ordered-Means Clustering
DOI:
https://doi.org/10.32890/jict2024.23.3.6Keywords:
Fuzzy clustering, beta distribution, feature weighted, ordered mechanismAbstract
The fuzzy C-ordered-means clustering (FCOM) is a fuzzy clustering algorithm that enhances robustness and clustering accuracy through the ordered mechanism based on fuzzy C-means (FCM). However, despite these improvements, the FCOM algorithm’s effectiveness remains unsatisfactory due to the significant time cost incurred by its ordered operation. To address this problem, an investigation was conducted on the ordered weighted model of the FCOM algorithm leading to proposed enhancements by introducing the beta distribution weighted fuzzy C-ordered-means clustering (BDFCOM). The BDFCOM algorithm utilises the properties of the Beta distribution to weight sample features, thus not only circumventing the time cost problem of the traditional ordered mechanism but also reducing the influence of noise. Experiments were conducted on six UCI datasets to validate the effectiveness of the BDFCOM, comparing its performance against seven other clustering algorithms using six evaluation indices. The results show that compared to the average of the other seven algorithms, BDFCOM improves about 15 percent on F1-score, 11 percent on Rand Index, 13 percent on Adjusted Rand Index, 3 percent on Fowlkes-Mallows Index and 16 percent on Jaccard Index. For the other two ordered mechanism FCM algorithms, the time consumption was also reduced by 90.15 percent on average. The proposed algorithm, which designs a new way of feature weighting for ordered mechanisms, advances the field of ordered mechanisms.
And, this paper provides a new method in the application field where there is a lot of noise in the dataset.
References
Abu-Shareha, A. A. (2022). TOPSIS-based regression algorithms evaluation. Journal of Information and Communication Technology, 21(4), 513–547. https://doi.org/10.32890/ jict2022.21.4.3
Amorim, L., Cavalcanti, G. D. C., & Cruz, R. M. O. (2023). The choice of scaling technique matters for classification performance. Applied Soft Computing, 133, 109924. https://doi.org/10.1016/j.asoc.2022.109924
Campello, R. J. G. B. (2007). A fuzzy extension of the Rand index and other related indexes for clustering and classification assessment. Pattern Recognition Letters, 28(7), 833–841. https://doi.org/10.1016/j.patrec.2006.11.010
Chen, W., Hu, Y., Peng, M., & Zhu, B. (2024). Positional normalisation-based mixed-image data augmentation and ensemble self-distillation algorithm. Expert Systems With Applications, 124140. https://doi.org/10.1016/j.eswa.2024.124140
Christen, P., Hand, D. J., & Kirielle, N. (2023). A review of the F-Measure: its history, properties, criticism, and alternatives. ACM Computing Surveys, 56(3), 1–24. https://doi.org/10.1145/3606367
Chunhao, Z., Bin, X., Ximei, Z., & Tongtong, X. (2023). Weighted fuzzy clustering algorithm combining adaptive nearest Journal of ICT, 23, No. 3 (July) 2024, pp: 523-neighbors and density peaks. Journal of Chinese Computer Systems, 44(9), 1974-1982. https://doi.org/10.20009/j.cnki.21-1106/TP.2021-0988
Davis, R. H., & Economou, C. (1984). A review of fuzzy clustering methods. Advances in Engineering Software, 6(4), 189–191. https://doi.org/10.1016/0141-1195(84)90002-0
Fauzi, N. S. M., Kasim, M. M., & Desa, N. H. M. (2022). Construction of air pollution index with the inclusion of aggregated weights of the pollutants. Journal of Information and Communication Technology, 21(4), 495-512. https://doi.org/10.32890/ jict2022.21.4.2
Forbes, C., Evans, M., Hastings, N., & Peacock, B. (2010). Statistical distributions. John Wiley & Sons.
Garg, H., & Arora, R. (2018). Generalised intuitionistic fuzzy soft power aggregation operator based ont-norm and their application in multicriteria decision-making. International Journal of Intelligent Systems, 34(2), 215–246. https://doi.org/10.1002/int.22048
Habib, G., Malik, I. A., Ahmad, J., Ahmed, I., & Qureshi, S. (2024). Exploring the efficacy of Group-Normalization in deep learning models for Alzheimer’s disease classification. arXiv (Cornell University). https://doi.org/10.48550/arxiv.2404.00946
Harumeka, A., & Purwa, T. (2023). Does the physical type of house still affect household poverty in Indonesia? An entropy-based fuzzy weighted logistic regression approach. Journal of Information and Communication Technology, 22(3), 337–361. https://doi.org/10.32890/jict2023.22.3.2
Herbreteau, S., & Kervrann, C. (2024, February 23). On normalisation-equivariance properties of supervised and unsupervised denoising methods: a survey. arXiv.org. https://arxiv.org/ abs/2402.15352
Ho, P. L., Lee, C., Le, C., Nguyen, P. H., & Yee, J. (2024). A computational homogenisation for yield design of asymmetric microstructures using adaptive bES-FEM. Computers & Structures, 294, 107271. https://doi.org/10.1016/j. compstruc.2023.107271
Hu, J., Wu, M., Chen, L., & Pedrycz, W. (2022). A novel modeling framework based on customised Kernel-Based fuzzy C-Means clustering in iron ore sintering process. IEEE/ASME Transactions on Mechatronics, 27(2), 950–961. https://doi.org/10.1109/tmech.2021.3076208 Journal of ICT, 23, No. 3 (July) 2024, pp: 523-
Keshkeh, K., Jantan, A.., & Alieyan, K.. (2022). A machine learning classification approach to detect tls-based malware using entropy-based flow set features. Journal of Information and Communication Technology, 21(3), 279–313. https://doi.org/10.32890/jict2022.21.3.1
Khairuddin, A. R., Alwee, R., & Haron, H. (2023). Hybrid neighbourhood component analysis with Gradient tree boosting for feature selection in forecasting crime rate. Journal of Information and Communication Technology, 22(2), 207–229. https://doi.org/10.32890/jict2023.22.2.3
Krishnapuram, R., & Keller, J. M. (1993). A possibilistic approach to clustering. IEEE Transactions on Fuzzy Systems, 1(2), 98–110. https://doi.org/10.1109/91.227387
Leski, J. M. (2016). Fuzzy c-ordered-means clustering. Fuzzy Sets and Systems, 286, 114–133. https://doi.org/10.1016/j. fss.2014.12.007
Liao, H., Xiao, Y., Wu, X., & Baušys, R. (2024). Z-DNMASort: A double normalisation-based multiple aggregation sorting method with Z-numbers for multi-criterion sorting problems. Information Sciences, 653, 119782. https://doi.org/10.1016/j. ins.2023.119782
Lu, J., Ma, G., & Zhang, G. (2024). Fuzzy machine learning: A comprehensive framework and systematic review. IEEE Transactions on Fuzzy Systems, 1–18. https://doi.org/10.1109/ tfuzz.2024.3387429
Markelle, K., Rachel, L., & Kolby, N. (n.d.). Datasets. The UCI Machine Learning Repository. https://archive.ics.uci.edu/
Materum, L., & Teologo Jr., A. T. (2021). An improved K-power means technique using Minkowski distance metric and dimension weights for clustering wireless multipaths in indoor channel scenarios. Journal of Information and Communication Technology, 20(4), 541–563. https://doi.org/10.32890/ jict2021.20.4.4
Park, C. W., & Eom, I. K. (2024). Underwater image enhancement using adaptive standardisation and normalisation networks. Engineering Applications of Artificial Intelligence, 127, 107445. https://doi.org/10.1016/j.engappai.2023.107445
Radzi, A. R., Farouk, A. M., Romali, N. S., Farouk, M., Elgamal, M., & Rahman, R. A. (2024). Assessing environmental management plan implementation in water supply construction projects: Key performance indicators. Sustainability, 16(2), 600. https://doi.org/10.3390/su16020600 Journal of ICT, 23, No. 3 (July) 2024, pp: 523-
Roszkowska, E., & Wachowicz, T. (2024). Impact of normalisation on entropy-based weights in Hellwig’s method: A case study on evaluating sustainable development in the education area. Entropy, 26(5), 365. https://doi.org/10.3390/e26050365
Rovatti, R., & Fantuzzi, C. (1996). S-norm aggregation of infinite collections. Fuzzy Sets and Systems, 84(3), 255–269. https://doi.org/10.1016/0165-0114(95)00320-7
Sainin, M. S., Alfred, R., & Ahmad, F. (2021). Ensemble meta classifier with sampling and feature selection for data with imbalance multiclass problem. Journal of Information and Communication Technology, 20(2), 103–133. https://doi.org/10.32890/jict2021.20.2.1
Seman, A., & Mohd Sapawi, A. (2018). Extensions to the K-Amh algorithm for numerical clustering. Journal of Information and Communication Technology, 17(4), 587–599. https://doi.org/10.32890/jict2018.17.4.8272
Sharif, N. A. M., Harun, N. H., & Yusof, Y. (2024). Colour image enhancement model of retinal fundus image for diabetic retinopathy recognition. Journal of Information and Communication Technology, 23(2), 293-334.
Singh, A., & Kumar, M. (2023). Bayesian fuzzy clustering and deep CNN-based automatic video summarisation. Multimedia Tools and Applications, 83(1), 963–1000. https://doi.org/10.1007/ s11042-023-15431-9
Song, J., Lee, H., & Kwon, O.-Y. (2023). Investigating job mismatch in software industry through news big data. Journal of Information and Communication Technology, 22(1), 31–48. https://doi.org/10.32890/jict2023.22.1.2
Sutranggono, A. N., Riyanarto Sarno, & Imam Ghozali. (2024). Multi-class multi-level classification of mental health disorders based on textual data from social media. Journal of Information and Communication Technology, 23(1), 77–104. https://doi.org/10.32890/jict2024.23.1.4
Takamiya, K., Iwamoto, Y., Nonaka, M., & Chen, Y. (2023). CT brain image synthesization from MRI brain images using CycleGAN. 2023 IEEE International Conference on Consumer Electronics (ICCE). https://doi.org/10.1109/icce56470.2023.10043572
Tyler, D. E. (2008). Robust statistics: Theory and methods. Journal of the American Statistical Association, 103(482), 888–889. https://doi.org/10.1198/jasa.2008.s239 Journal of ICT, 23, No. 3 (July) 2024, pp: 523-
Wang, Y., & Xue, Q. (2024). Fault identification of product design using fuzzy clustering generative adversarial network (FCGAN) model. Soft Computing, 28(4), 3725–3742. https://doi.org/10.1007/s00500-024-09636-9
Wu, C. W. (2022). On rearrangement inequalities for T-norm logics. arXiv (Cornell University). https://doi.org/10.48550/ arxiv.2204.06051
Yang, A., Peng, B., Lu, C., He, Z., Chen, E. T., & Sheng, F. (2024b). A novel method for multiple targets localisation based on normalised cross-correlation adaptive variable step-size dynamic template matching. AIP Advances, 14(4). https://doi.org/10.1063/5.0194376
Yin, H., Aryani, A., Petrie, S., Nambissan, A., Astudillo, A., & Cao, S. (2024). A rapid review of clustering algorithms. arXiv (Cornell University). https://doi.org/10.48550/arxiv.2401.07389
Yongli, L., Hengda, W., Jing, L., & Lishen, Y. (2019). Feature weighted fuzzy C-ordered-means clustering algorithm. Journal of Henan Polytechnic University (Natural Science), 38(3), 123-130. https://doi.org/10.16186/j.cnki.1673-9787.2019.3.17
Published
Issue
Section
License
Copyright (c) 2024 Journal of Information and Communication Technology

This work is licensed under a Creative Commons Attribution 4.0 International License.
How to Cite
Research impact
Harvested 2026-09-06Counts differ between services because each indexes a different body of literature. None of them is the whole picture.
2002 - 2020






















