Beta Distribution Weighted Fuzzy C-Ordered-Means Clustering

Authors

  • Wang Hengda School of Computing, Universiti Utara Malaysia, Malaysia and Artificial Intelligence Department, Chongqing Institute of Engineering, China
  • Mohamad Farhan Mohamad Mohsin School of Computing, Universiti Utara Malaysia, Malaysia
  • Muhammad Syafiq Mohd Pozi School of Computing, Universiti Utara Malaysia, Malaysia

DOI:

https://doi.org/10.32890/jict2024.23.3.6

Keywords:

Fuzzy clustering, beta distribution, feature weighted, ordered mechanism

Abstract

The fuzzy C-ordered-means clustering (FCOM) is a fuzzy clustering algorithm that enhances robustness and clustering accuracy through the ordered mechanism based on fuzzy C-means (FCM). However, despite these improvements, the FCOM algorithm’s effectiveness remains unsatisfactory due to the significant time cost incurred by its ordered operation. To address this problem, an investigation was conducted on the ordered weighted model of the FCOM algorithm leading to proposed enhancements by introducing the beta distribution weighted fuzzy C-ordered-means clustering (BDFCOM). The BDFCOM algorithm utilises the properties of the Beta distribution to weight sample features, thus not only circumventing the time cost problem of the traditional ordered mechanism but also reducing the influence of noise. Experiments were conducted on six UCI datasets to validate the effectiveness of the BDFCOM, comparing its performance against seven other clustering algorithms using six evaluation indices. The results show that compared to the average of the other seven algorithms, BDFCOM improves about 15 percent on F1-score, 11 percent on Rand Index, 13 percent on Adjusted Rand Index, 3 percent on Fowlkes-Mallows Index and 16 percent on Jaccard Index. For the other two ordered mechanism FCM algorithms, the time consumption was also reduced by 90.15 percent on average. The proposed algorithm, which designs a new way of feature weighting for ordered mechanisms, advances the field of ordered mechanisms.
And, this paper provides a new method in the application field where there is a lot of noise in the dataset. 

References

Abu-Shareha, A. A. (2022). TOPSIS-based regression algorithms evaluation. Journal of Information and Communication Technology, 21(4), 513–547. https://doi.org/10.32890/ jict2022.21.4.3

Amorim, L., Cavalcanti, G. D. C., & Cruz, R. M. O. (2023). The choice of scaling technique matters for classification performance. Applied Soft Computing, 133, 109924. https://doi.org/10.1016/j.asoc.2022.109924

Campello, R. J. G. B. (2007). A fuzzy extension of the Rand index and other related indexes for clustering and classification assessment. Pattern Recognition Letters, 28(7), 833–841. https://doi.org/10.1016/j.patrec.2006.11.010

Chen, W., Hu, Y., Peng, M., & Zhu, B. (2024). Positional normalisation-based mixed-image data augmentation and ensemble self-distillation algorithm. Expert Systems With Applications, 124140. https://doi.org/10.1016/j.eswa.2024.124140

Christen, P., Hand, D. J., & Kirielle, N. (2023). A review of the F-Measure: its history, properties, criticism, and alternatives. ACM Computing Surveys, 56(3), 1–24. https://doi.org/10.1145/3606367

Chunhao, Z., Bin, X., Ximei, Z., & Tongtong, X. (2023). Weighted fuzzy clustering algorithm combining adaptive nearest Journal of ICT, 23, No. 3 (July) 2024, pp: 523-neighbors and density peaks. Journal of Chinese Computer Systems, 44(9), 1974-1982. https://doi.org/10.20009/j.cnki.21-1106/TP.2021-0988

Davis, R. H., & Economou, C. (1984). A review of fuzzy clustering methods. Advances in Engineering Software, 6(4), 189–191. https://doi.org/10.1016/0141-1195(84)90002-0

Fauzi, N. S. M., Kasim, M. M., & Desa, N. H. M. (2022). Construction of air pollution index with the inclusion of aggregated weights of the pollutants. Journal of Information and Communication Technology, 21(4), 495-512. https://doi.org/10.32890/ jict2022.21.4.2

Forbes, C., Evans, M., Hastings, N., & Peacock, B. (2010). Statistical distributions. John Wiley & Sons.

Garg, H., & Arora, R. (2018). Generalised intuitionistic fuzzy soft power aggregation operator based ont-norm and their application in multicriteria decision-making. International Journal of Intelligent Systems, 34(2), 215–246. https://doi.org/10.1002/int.22048

Habib, G., Malik, I. A., Ahmad, J., Ahmed, I., & Qureshi, S. (2024). Exploring the efficacy of Group-Normalization in deep learning models for Alzheimer’s disease classification. arXiv (Cornell University). https://doi.org/10.48550/arxiv.2404.00946

Harumeka, A., & Purwa, T. (2023). Does the physical type of house still affect household poverty in Indonesia? An entropy-based fuzzy weighted logistic regression approach. Journal of Information and Communication Technology, 22(3), 337–361. https://doi.org/10.32890/jict2023.22.3.2

Herbreteau, S., & Kervrann, C. (2024, February 23). On normalisation-equivariance properties of supervised and unsupervised denoising methods: a survey. arXiv.org. https://arxiv.org/ abs/2402.15352

Ho, P. L., Lee, C., Le, C., Nguyen, P. H., & Yee, J. (2024). A computational homogenisation for yield design of asymmetric microstructures using adaptive bES-FEM. Computers & Structures, 294, 107271. https://doi.org/10.1016/j. compstruc.2023.107271

Hu, J., Wu, M., Chen, L., & Pedrycz, W. (2022). A novel modeling framework based on customised Kernel-Based fuzzy C-Means clustering in iron ore sintering process. IEEE/ASME Transactions on Mechatronics, 27(2), 950–961. https://doi.org/10.1109/tmech.2021.3076208 Journal of ICT, 23, No. 3 (July) 2024, pp: 523-

Keshkeh, K., Jantan, A.., & Alieyan, K.. (2022). A machine learning classification approach to detect tls-based malware using entropy-based flow set features. Journal of Information and Communication Technology, 21(3), 279–313. https://doi.org/10.32890/jict2022.21.3.1

Khairuddin, A. R., Alwee, R., & Haron, H. (2023). Hybrid neighbourhood component analysis with Gradient tree boosting for feature selection in forecasting crime rate. Journal of Information and Communication Technology, 22(2), 207–229. https://doi.org/10.32890/jict2023.22.2.3

Krishnapuram, R., & Keller, J. M. (1993). A possibilistic approach to clustering. IEEE Transactions on Fuzzy Systems, 1(2), 98–110. https://doi.org/10.1109/91.227387

Leski, J. M. (2016). Fuzzy c-ordered-means clustering. Fuzzy Sets and Systems, 286, 114–133. https://doi.org/10.1016/j. fss.2014.12.007

Liao, H., Xiao, Y., Wu, X., & Baušys, R. (2024). Z-DNMASort: A double normalisation-based multiple aggregation sorting method with Z-numbers for multi-criterion sorting problems. Information Sciences, 653, 119782. https://doi.org/10.1016/j. ins.2023.119782

Lu, J., Ma, G., & Zhang, G. (2024). Fuzzy machine learning: A comprehensive framework and systematic review. IEEE Transactions on Fuzzy Systems, 1–18. https://doi.org/10.1109/ tfuzz.2024.3387429

Markelle, K., Rachel, L., & Kolby, N. (n.d.). Datasets. The UCI Machine Learning Repository. https://archive.ics.uci.edu/

Materum, L., & Teologo Jr., A. T. (2021). An improved K-power means technique using Minkowski distance metric and dimension weights for clustering wireless multipaths in indoor channel scenarios. Journal of Information and Communication Technology, 20(4), 541–563. https://doi.org/10.32890/ jict2021.20.4.4

Park, C. W., & Eom, I. K. (2024). Underwater image enhancement using adaptive standardisation and normalisation networks. Engineering Applications of Artificial Intelligence, 127, 107445. https://doi.org/10.1016/j.engappai.2023.107445

Radzi, A. R., Farouk, A. M., Romali, N. S., Farouk, M., Elgamal, M., & Rahman, R. A. (2024). Assessing environmental management plan implementation in water supply construction projects: Key performance indicators. Sustainability, 16(2), 600. https://doi.org/10.3390/su16020600 Journal of ICT, 23, No. 3 (July) 2024, pp: 523-

Roszkowska, E., & Wachowicz, T. (2024). Impact of normalisation on entropy-based weights in Hellwig’s method: A case study on evaluating sustainable development in the education area. Entropy, 26(5), 365. https://doi.org/10.3390/e26050365

Rovatti, R., & Fantuzzi, C. (1996). S-norm aggregation of infinite collections. Fuzzy Sets and Systems, 84(3), 255–269. https://doi.org/10.1016/0165-0114(95)00320-7

Sainin, M. S., Alfred, R., & Ahmad, F. (2021). Ensemble meta classifier with sampling and feature selection for data with imbalance multiclass problem. Journal of Information and Communication Technology, 20(2), 103–133. https://doi.org/10.32890/jict2021.20.2.1

Seman, A., & Mohd Sapawi, A. (2018). Extensions to the K-Amh algorithm for numerical clustering. Journal of Information and Communication Technology, 17(4), 587–599. https://doi.org/10.32890/jict2018.17.4.8272

Sharif, N. A. M., Harun, N. H., & Yusof, Y. (2024). Colour image enhancement model of retinal fundus image for diabetic retinopathy recognition. Journal of Information and Communication Technology, 23(2), 293-334.

Singh, A., & Kumar, M. (2023). Bayesian fuzzy clustering and deep CNN-based automatic video summarisation. Multimedia Tools and Applications, 83(1), 963–1000. https://doi.org/10.1007/ s11042-023-15431-9

Song, J., Lee, H., & Kwon, O.-Y. (2023). Investigating job mismatch in software industry through news big data. Journal of Information and Communication Technology, 22(1), 31–48. https://doi.org/10.32890/jict2023.22.1.2

Sutranggono, A. N., Riyanarto Sarno, & Imam Ghozali. (2024). Multi-class multi-level classification of mental health disorders based on textual data from social media. Journal of Information and Communication Technology, 23(1), 77–104. https://doi.org/10.32890/jict2024.23.1.4

Takamiya, K., Iwamoto, Y., Nonaka, M., & Chen, Y. (2023). CT brain image synthesization from MRI brain images using CycleGAN. 2023 IEEE International Conference on Consumer Electronics (ICCE). https://doi.org/10.1109/icce56470.2023.10043572

Tyler, D. E. (2008). Robust statistics: Theory and methods. Journal of the American Statistical Association, 103(482), 888–889. https://doi.org/10.1198/jasa.2008.s239 Journal of ICT, 23, No. 3 (July) 2024, pp: 523-

Wang, Y., & Xue, Q. (2024). Fault identification of product design using fuzzy clustering generative adversarial network (FCGAN) model. Soft Computing, 28(4), 3725–3742. https://doi.org/10.1007/s00500-024-09636-9

Wu, C. W. (2022). On rearrangement inequalities for T-norm logics. arXiv (Cornell University). https://doi.org/10.48550/ arxiv.2204.06051

Yang, A., Peng, B., Lu, C., He, Z., Chen, E. T., & Sheng, F. (2024b). A novel method for multiple targets localisation based on normalised cross-correlation adaptive variable step-size dynamic template matching. AIP Advances, 14(4). https://doi.org/10.1063/5.0194376

Yin, H., Aryani, A., Petrie, S., Nambissan, A., Astudillo, A., & Cao, S. (2024). A rapid review of clustering algorithms. arXiv (Cornell University). https://doi.org/10.48550/arxiv.2401.07389

Yongli, L., Hengda, W., Jing, L., & Lishen, Y. (2019). Feature weighted fuzzy C-ordered-means clustering algorithm. Journal of Henan Polytechnic University (Natural Science), 38(3), 123-130. https://doi.org/10.16186/j.cnki.1673-9787.2019.3.17

Downloads

Published

28-07-2024

How to Cite

Wang, H., Mohamad Mohsin, M. F., & Mohd Pozi, M. S. (2024). Beta Distribution Weighted Fuzzy C-Ordered-Means Clustering. Journal of Information and Communication Technology, 23(3), 523-559. https://doi.org/10.32890/jict2024.23.3.6

Research impact

Harvested 2026-09-06
1 citations, from OpenAlex — the highest of the sources checked

Counts differ between services because each indexes a different body of literature. None of them is the whole picture.

Identifiers DOI 10.32890/jict2024.23.3.6 OpenAlex W4401102313 Scopus 85200132203

Most read articles by the same author(s)