Transforming Alzheimer’s Disease Diagnosis: Implementing Vision Transformer for MRI Images Classification
DOI:
https://doi.org/10.32890/jict2025.24.1.6Keywords:
Alzheimer’s, deep learning, image classification, MRI, vision transformerAbstract
Alzheimer’s disease (AD) is a progressive neurological disorder and the leading cause of dementia, accounting for 60-80% of cases. Early detection of AD is crucial for timely intervention, as the disease significantly impacts cognitive functions and daily activities. Diagnosing AD in clinical settings remains challenging due to subtle early symptoms, diverse presentations, lengthy diagnostic processes, and inconsistent criteria that heavily rely on medical expertise. Accurate and timely diagnosis during the early stages is essential for effective treatment and intervention. Modern imaging techniques, such as Magnetic Resonance Imaging (MRI), have become essential in diagnosing Alzheimer’s by providing detailed insights into structural brain changes. This study explores the application of the Vision Transformer (ViT) model for classifying MRI images of Alzheimer’s patients, focusing on enhancing accuracy and reliability through data augmentation during pre-processing. A dataset of 8,000 MRI images, categorised into four groups—non-demented, very mild demented, mild demented, and moderate demented—was used to evaluate the ViT model. The experiment achieved promising results, with an accuracy of 98.19%, sensitivity of 96.34%, specificity of 98.80%, and an F1-score of 96.37%. These findings underscore the model’s effectiveness in distinguishing between affected and unaffected individuals, minimising misdiagnosis and enabling timely clinical interventions. However, some challenges remain, particularly in the classification between “Non-Demented” and “Very Mild Demented” cases. Future research should focus on enhancing data augmentation techniques and increasing data diversity in these categories to improve the model’s performance further. The ViT model holds great potential for advancing Alzheimer’s diagnosis, offering a valuable tool for early detection and intervention in clinical settings.
References
Abubakar, M. B., Sanusi, K. O., Ugusman, A., Mohamed, W., Kamal, H., Ibrahim, N. H., Khoo, C. S., & Kumar, J. (2022). Alzheimer’s disease: An update and insights into pathophysiology. Frontiers in Aging Neuroscience, 14. https://doi.org/10.3389/fnagi.2022.742408
Al Rahbani, R. G., Ioannou, A., & Wang, T. (2024). Alzheimer’s disease multi-class detection through deep learning models and post-processing heuristics. Computer Methods in Biomechanics and Biomedical Engineering: Imaging and Visualisation, 12(1). https://doi.org/10.1080/21681163. 2024.2383219
Almufareh, M. F., Tehsin, S., Humayun, M., & Kausar, S. (2023). Artificial cognition for detection of mental disability: A vision transformer approach for Alzheimer’s disease. Healthcare (Switzerland), 11(20). https://doi.org/10.3390/healthcare11202763
Alomar, K., Aysel, H. I., & Cai, X. (2023). Data augmentation in classification and segmentation: A survey and new strategies. Journal of Imaging, 9(2), 46. https://doi.org/10.3390/ jimaging9020046
Alzheimer’s Association. (2021). 2021 Alzheimer’s disease facts and figures. Alzheimer’s and Dementia, 17(3), 327–406. https://doi.org/10.1002/alz.12328
Alzheimer’s Association. (2023). 2023 Alzheimer’s disease facts and figures. Alzheimer’s and Dementia, 19(4), 1598–1695. https://doi.org/10.1002/alz.13016
Arevalo, J., González, F. A., Ramos-Pollán, R., Oliveira, J. L., & Guevara Lopez, M. A. (2016). Representation learning for mammography mass lesion classification with convolutional neural networks. Computer Methods and Programs in Biomedicine, 127, 248–257. https://doi.org/10.1016/j.cmpb.2015.12.014
Arnab, A., Dehghani, M., Heigold, G., Sun, C., Lučić, M., & Schmid, C. (2021). ViViT: A Video vision transformer. Proceedings of the IEEE International Conference on Computer Vision, 6816–6826. https://doi.org/10.1109/ICCV48922.2021.00676
Azad, R., Asadi-Aghbolaghi, M., Fathy, M., & Escalera, S. (2019). Bi-directional ConvLSTM U-net with densley connected convolutions. Proceedings - 2019 International Conference on Computer Vision Workshop, ICCVW 2019, 406–415. https://doi.org/10.1109/ICCVW. 2019.00052 Azad, R., Kazerouni, A., Heidari, M., Aghdam, E. K., Molaei, A., Jia, Y., Jose, A., Roy, R., & Merhof,
D. (2024). Advances in medical image analysis with vision transformers: A comprehensive review. Medical Image Analysis, 91. https://doi.org/10.1016/j.media.2023.103000
Azad, R., Khosravi, N., & Merhof, D. (2022). SMU-Net: Style matching U-Net for brain tumor segmentation with missing modalities. Proceedings of Machine Learning Research, 172, 48–62.
Bazi, Y., Bashmal, L., Al Rahhal, M. M., Dayil, R. Al, & Ajlan, N. Al. (2021). Vision transformers for remote sensing image classification. Remote Sensing, 13(3), 1–20. https://doi.org/10.3390/rs13030516
Bello, I., Zoph, B., Le, Q., Vaswani, A., & Shlens, J. (2019). Attention augmented convolutional networks. Proceedings of the IEEE International Conference on Computer Vision, 2019-Octob, 3285–3294. https://doi.org/10.1109/ICCV.2019.00338
Bengio, Y., Lecun, Y., & Hinton, G. (2021). Deep learning for AI. Communications of the ACM, 64(7), 58–65. https://doi.org/10.1145/3448250
Breijyeh, Z., & Karaman, R. (2020). Comprehensive review on Alzheimer’s disease: Causes and treatment. In Molecules (Vol. 25, Issue 24). https://doi.org/10.3390/MOLECULES25245789
Chen, Y., Wang, L., Ding, B., Shi, J., Wen, T., Huang, J., & Ye, Y. (2024). Automated Alzheimer’s disease classification using deep learning models with Soft-NMS and improved ResNet50 integration. Journal of Radiation Research and Applied Sciences, 17(1), 100782. https://doi.org/10.1016/j.jrras.2023.100782
Dafre, R., & Wasnik, P. (2023). Current diagnostic and treatment methods of Alzheimer’s disease: A narrative review. Cureus. https://doi.org/10.7759/cureus.45649
Divon, G., & Tal, A. (2018). Viewpoint Estimation---Insights & Model. European Conference on Computer Vision, 252–268. Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M.,
Minderer, M., Heigold, G., Gelly, S., Uszkoreit, J., & Houlsby, N. (2021). An image is worth 16X16 words: Transformers for image recognition at scale. ICLR 2021-9th International Conference on Learning Representations.
Fan, H., Xiong, B., Mangalam, K., Li, Y., Yan, Z., Malik, J., & Feichtenhofer, C. (2021). Multiscale vision transformers. Proceedings of the IEEE International Conference on Computer Vision, 6804–6815. https://doi.org/10.1109/ICCV48922.2021.00675
Gao, Y., & Liu, X. (2021). Secular trends in the incidence of and mortality due to Alzheimer’s disease and other forms of Dementia in China from 1990 to 2019: An age-period-cohort study and joinpoint analysis. Frontiers in Aging Neuroscience, 13. https://doi.org/10.3389/fnagi.2021. 709156
Gheflati, B., & Rivaz, H. (2022). Vision transformers for classification of breast ultrasound images. Proceedings of the Annual International Conference of the IEEE Engineering in Medicine and Biology Society, EMBS, 2022-July, 480–483. https://doi.org/10.1109/EMBC48229.2022. 9871809 Han, K., Wang, Y., Chen, H., Chen, X., Guo, J., Liu, Z., Tang, Y., Xiao, A., Xu, C., Xu, Y., Yang, Z.,
Zhang, Y., & Tao, D. (2023). A Survey on vision transformer. IEEE Transactions on Pattern
Analysis and Machine Intelligence, 45(1), 87–110. https://doi.org/10.1109/TPAMI.2022. 3152247
Hasnain, M., Pasha, M. F., Ghani, I., Imran, M., Alzahrani, M. Y., & Budiarto, R. (2020). Evaluating trust prediction and confusion matrix measures for web services ranking. IEEE Access, 8, 90847–90861. https://doi.org/10.1109/ACCESS.2020.2994222
Helaly, H. A., Badawy, M., & Haikal, A. Y. (2022). Deep learning approach for early detection of Alzheimer’s disease. Cognitive Computation, 14(5), 1711–1727. https://doi.org/10.1007/ s12559-021-09946-2
Irankhah, E. (2020). Evaluation of early detection methods for Alzheimer’s disease. Bioprocess Engineering, 4(1), 17. https://doi.org/10.11648/j.be.20200401.13
Joseph, V. R., & Vakayil, A. (2021). SPLIT: An optimal method for data splitting. Technometrics, 64(2), 166–176. https://doi.org/10.1080/00401706.2021.1921037
Juganavar, A., Joshi, A., & Shegekar, T. (2023). Navigating early Alzheimer’s diagnosis: A comprehensive review of diagnostic innovations. Cureus. https://doi.org/10.7759/cureus.44937
Juneja, M., Thakur, N., Thakur, S., Uniyal, A., Wani, A., & Jindal, P. (2020). GC-NET for classification of glaucoma in the retinal fundus image. Machine Vision and Applications, 31(5). https://doi.org/10.1007/s00138-020-01091-4
Karimijafarbigloo, S., Azad, R., Kazerouni, A., & Merhof, D. (2023). MS-Former: Multi-scale selfguided transformer for medical image segmentation. Proceedings of Machine Learning Research, 227, 680–694.
Karimijafarbigloo, S., Azad, R., Kazerouni, A., Ebadollahi, S., & Merhof, D. (2023). MMCFormer: Missing modality compensation transformer for brain tumor segmentation. Proceedings of Machine Learning Research, 227, 1144–1162.
Khan, S., Naseer, M., Hayat, M., Zamir, S. W., Khan, F. S., & Shah, M. (2022). Transformers in vision: A survey. ACM Computing Surveys, 54(10) 1-41. https://doi.org/10.1145/3505244
Kim, W. S., Lee, D. H., Kim, Y. J., Kim, T., Hwang, R. Y., & Lee, H. J. (2020). Path detection for autonomous traveling in orchards using patch-based CNN. Computers and Electronics in Agriculture, 175. https://doi.org/10.1016/j.compag.2020.105620
Krstinic, D., Seric, L., & Slapnicar, I. (2023). Comments on “MLCM: Multi-label confusion matrix.” In IEEE Access (Vol. 11, pp. 40692–40697). https://doi.org/10.1109/ACCESS.2023.3267672
Li, X., Feng, X., Sun, X., Hou, N., Han, F., & Liu, Y. (2022). Global, regional, and national burden of Alzheimer’s disease and other dementias, 1990–2019. Frontiers in Aging Neuroscience, 14. https://doi.org/10.3389/fnagi.2022.937486
Liu, H., Wang, C., & Peng, Y. (2021). Data augmentation with illumination correction in sematic segmentation. Journal of Physics Conference Series, 2025(1), 012009. https://doi.org/10.1088/ 1742-6596/2025/1/012009
Lyu, Y., Yu, X., Zhu, D., & Zhang, L. (2022). Classification of Alzheimer’s disease via vision transformer. ACM International Conference Proceeding Series, 463–468. https://doi.org/10.1145/3529190.3534754
Malik, P., & Singh, S. (2024). Alzheimer’s disease classification using neuroimaging modalities and deep learning. Journal of Harbin Engineering University, 45(6), 433-448.
Mi, J., Liu, C., Chen, H., Qian, Y., Zhu, J., Zhang, Y., Liang, Y., Wang, L., & Ta, D. (2024). Light on Alzheimer’s disease: From basic insights to preclinical studies. In Frontiers in Aging Neuroscience (Vol. 16). Frontiers Media SA. https://doi.org/10.3389/fnagi.2024.1363458
Muraina, I. O. (2022). Ideal dataset splitting ratios in Machine learning algorithms: General concerns for data scientists and data analysts. In 7th International Mardin Artuklu Scientific Researches Conference. Nichols, E., Steinmetz, J. D., Vollset, S. E., Fukutaki, K., Chalek, J., Abd-Allah, F., Abdoli, A., Abualhasan, A., Abu-Gharbieh, E., Akram, T. T., Al Hamad, H., Alahdab, F., Alanezi, F. M.,
Alipour, V., Almustanyir, S., Amu, H., Ansari, I., Arabloo, J., Ashraf, T., … Vos, T. (2022). Estimation of the global prevalence of dementia in 2019 and forecasted prevalence in 2050: An analysis for the Global Burden of Disease Study 2019. The Lancet Public Health, 7(2), e105– e125. https://doi.org/10.1016/S2468-2667(21)00249-8 Nichols, E., Szoeke, C. E. I., Vollset, S. E., Abbasi, N., Abd-Allah, F., Abdela, J., Aichour, M. T. E., Akinyemi, R. O., Alahdab, F., Asgedom, S. W., Awasthi, A., Barker-Collo, S. L., Baune, B. T., Béjot, Y., Belachew, A. B., Bennett, D. A., Biadgo, B., Bijani, A., Bin Sayeed, M. S., …
Murray, C. J. L. (2019). Global, regional, and national burden of Alzheimer’s disease and other dementias, 1990–2016: A systematic analysis for the Global Burden of Disease Study 2016. The Lancet Neurology, 18(1), 88–106. https://doi.org/10.1016/S1474-4422(18)30403-4 Pulido, M. L. B., Hernández, J. B. A., Ballester, M. Á. F., González, C. M. T., Mekyska, J., & Smékal,
Z. (2020). Alzheimer’s disease and automatic speech analysis: A review. In Expert Systems with Applications (Vol. 150). https://doi.org/10.1016/j.eswa.2020.113213
Ramachandran, P., Bello, I., Parmar, N., Levskaya, A., Vaswani, A., & Shlens, J. (2019). Stand-alone self-attention in vision models. Advances in Neural Information Processing Systems, 32.
Shaffi, N., Viswan, V., & Mahmud, M. (2024). Ensemble of vision transformer architectures for efficient Alzheimer’s Disease classification. Brain Informatics, 11(1). https://doi.org/10.1186/ s40708-024-00238-7
Shin, H., Jeon, S., Seol, Y., Kim, S., & Kang, D. (2023). Vision transformer approach for classification of Alzheimer’s disease using 18F-Florbetaben brain images. Applied Sciences (Switzerland), 13(6). https://doi.org/10.3390/app13063453
Suganthe, R. C., Geetha, M., Sreekanth, G. R., Gowtham, K., Deepakkumar, S., & Elango, R. (2021). Multi-class Classification of Alzheimer’s Disease Using Hybrid Deep Convolutional Neural Network. Volatiles & Essent. Oils, 8(5), 145–153.
Tang, Y., Xiong, X., Tong, G., Yang, Y., & Zhang, H. (2024). Multi-modal diagnosis model of Alzheimer’s disease based on improved Transformer. BioMedical Engineering Online, 23(1). https://doi.org/10.1186/s12938-024-01204-4
Vaswani, A., Ramachandran, P., Srinivas, A., Parmar, N., Hechtman, B., & Shlens, J. (2021). Scaling local self-attention for parameter efficient visual backbones. Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, 12889–12899. https://doi.org/10.1109/CVPR46437.2021.01270 Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., & Polosukhin,
I. (2017). Attention is all you need. Advances in Neural Information Processing Systems, 2017Decem, 5999–6009.
Vimala, B. B., Srinivasan, S., Mathivanan, S. K., Mahalakshmi, Jayagopal, P., & Dalu, G. T. (2023). Detection and classification of brain tumor using hybrid deep learning models. Scientific Reports, 13(1). https://doi.org/10.1038/s41598-023-50505-6
Wong, K. K. L., Xu, J., Chen, C., Ghista, D., & Zhao, H. (2023). Functional magnetic resonance imaging providing the brain effect mechanism of acupuncture and moxibustion treatment for depression. Frontiers in Neurology, 14. https://doi.org/10.3389/fneur.2023.1151421 Wong, K. K. L., Xu, W., Ayoub, M., Fu, Y. L., Xu, H., Shi, R., Zhang, M., Su, F., Huang, Z., & Chen,
W. (2023). Brain image segmentation of the corpus callosum by combining Bi-Directional Convolutional LSTM and U-Net using multi-slice CT and MRI. Computer Methods and Programs in Biomedicine, 238. https://doi.org/10.1016/j.cmpb.2023.107602
World Alzheimer Report. (2019). World Alzheimer report 2019, attitudes to dementia. Alzheimer’s Disease International: London.
Xhumari, E., & Haloci, S. (2023). A comparative study of credit scoring and risk management Techniques in Fintech: Machine Learning vs. Regression Analysis. CEUR Workshop Proceedings, 3402, 13–20.
Zhong, Z., Zheng, L., Kang, G., Li, S., & Yang, Y. (2020). Random erasing data augmentation. Proceedings of the AAAI Conference on Artificial Intelligence, 34(07), 13001–13008. https://doi.org/10.1609/aaai.v34i07.7000
Zhou, T., Fu, H., Chen, G., Shen, J., & Shao, L. (2020). Hi-Net: Hybrid-fusion network for multi-modal MR image synthesis. IEEE Transactions on Medical Imaging, 39(9), 2772–2781. https://doi.org/10.1109/TMI.2020.2975344
Zhu, X., Su, W., Lu, L., Li, B., Wang, X., & Dai, J. (2021). Deformable Detr: Deformable transformers for end-to-end object detection. ICLR 2021-9th International Conference on Learning Representations.
Published
Issue
Section
License
Copyright (c) 2025 Journal of Information and Communication Technology

This work is licensed under a Creative Commons Attribution 4.0 International License.
How to Cite
Research impact
Harvested 2026-09-06Counts differ between services because each indexes a different body of literature. None of them is the whole picture.
2002 - 2020






















