Two-stage GAN with Attention Gates for Brain MRI Inpainting: A Hybrid Framework for Preserving Diagnostic Features in Alzheimer's Disease Classification

PDF (1600KB), PP.171-189

Views: 0 Downloads: 0

Author(s)

Chhaya Yadav 1 Sunita Yadav 2 Arvind Panwar 3,*

1. School of Computer Science and Engineering, Galgotias University, Greater Noida, Uttar Pradesh, India

2. Inderprastha Engineering College, Ghaziabad, Uttar Pradesh, India

3. School of Computer Science and Engineering, Galgotias University, Greater Noida 201308, Uttar Pradesh, India

* Corresponding author.

DOI: https://doi.org/10.5815/ijitcs.2026.05.11

Received: 26 Mar. 2026 / Revised: 8 May 2026 / Accepted: 23 Jul. 2026 / Published: 8 Oct. 2026

Index Terms

Medical Image Inpainting, Generative Adversarial Networks, Attention Mechanisms, Brain MRI, Alzheimer's Disease Classification, Hybrid Inpainting, Artifact Removal

Abstract

Clinical brain MRI scans for Alzheimer's disease diagnosis often suffer from motion artifacts, signal dropout, and incomplete acquisitions. While deep learning methods like LaMa produce visually coherent inpainting, they may alter critical anatomical structures, and traditional methods such as OpenCV Telea maintain local continuity but introduce excessive smoothing. This study presents a dual-stage generative adversarial network with attention gates for medical image restoration, uniquely complemented by a novel gradient-weighted hybrid blending strategy that adaptively combines GAN outputs with classical inpainting based on local image gradients. The architecture employs an attention-enhanced U-Net generator and refinement network, supervised by patch and global discriminators through combined adversarial, reconstruction, perceptual, structural similarity, and gradient losses. Evaluation on 12,491 balanced Alzheimer's MRI scans across four severity stages with synthetic 10–30% occlusions shows that the proposed GAN reduces mean squared error by 48.7% versus LaMa and 41.3% versus OpenCV, achieving 17.44 dB PSNR and 0.9796 SSIM (computed over the brain region). The gradient-guided Hybrid-GAN maintains reconstruction quality with a total inference time of approximately 8.97 ms per image. Downstream VGG16 classification reveals Hybrid-GAN preserves diagnostic information most effectively, attaining 96.05% accuracy, only 0.96 points below the 97.01% baseline and surpassing all alternative methods. These results demonstrate that attention-driven generative models combined with structure-aware blending provide a practical and novel solution for artifact mitigation in neuroimaging diagnostics.

Cite This Paper

Chhaya Yadav, Sunita Yadav, Arvind Panwar, "Two-stage GAN with Attention Gates for Brain MRI Inpainting: A Hybrid Framework for Preserving Diagnostic Features in Alzheimer's Disease Classification", International Journal of Information Technology and Computer Science(IJITCS), Vol.18, No.5, pp.171-189, 2026. DOI:10.5815/ijitcs.2026.05.11

Reference

[1]E. Twiss, C. McPherson, and D. F. Weaver, “Global diseases deserve global solutions: Alzheimer’s disease,” Neurology International, vol. 17, no. 6, p. 92, 2025.
[2]W.-S. Tae, B.-J. Ham, S.-B. Pyun, and B.-J. Kim, “Current clinical applications of structural mri in neurological disorders,” Journal of Clinical Neurology (Seoul, Korea), vol. 21, no. 4, p. 277, 2025.
[3]R. Obuchowicz, J. Lasek, M. Wodzi´nski, A. Pi´orkowski, M. Strzelecki, and K. Nurzynska, “Artificial intelligence empowered radiology—current status and critical review,” Diagnostics, vol. 15, no. 3, p. 282, 2025.
[4]I. N. Sari, E. Horikawa, and W. Du, “Interactive image inpainting of large-scale missing region,” IEEE Access, vol. 9, pp. 56 430–56 442, 2021.
[5]Q. Wang, S. He, M. Su, and F. Zhao, “Image inpainting methods: A review of deep learning approaches,” Symmetry, vol. 18, no. 1, p. 94, 2026.
[6]R. G. Alexander, F. Yazdanie, S. Waite, Z. A. Chaudhry, S. Kolla, S. L. Macknik, and S. Martinez-Conde, “Visual illusions in radiology: untrue perceptions in medical images and their implications for diagnostic accuracy,” Frontiers in Neuroscience, vol. 15, p. 629469, 2021.
[7]J. C. Santos, H. Tom´as Pereira Alexandre, M. Seoane Santos, and P. Henriques Abreu, “The role of deep learning in medical image inpainting: A systematic review,” ACM Transactions on Computing for Healthcare, vol. 6, no. 3, pp. 1–24, 2025.
[8]A. Telea, “An image inpainting technique based on the fast marching method,” Journal of graphics tools, vol. 9, no. 1, pp. 23–34, 2004.
[9]M. Bertalmio, A. L. Bertozzi, and G. Sapiro, “Navier-stokes, fluid dynamics, and image and video inpainting,” in Proceedings of the 2001 IEEE Computer Society Conference on Computer Vision and Pattern Recognition. CVPR 2001, vol. 1. IEEE, 2001, pp. I–I.
[10]C. Barnes, E. Shechtman, A. Finkelstein, and D. B. Goldman, “Patchmatch: A randomized correspondence algorithm for structural image editing,” ACM Trans. Graph., vol. 28, no. 3, p. 24, 2009.
[11]A. Criminisi, P. P´erez, and K. Toyama, “Region filling and object removal by exemplar-based image inpainting,” IEEE Transactions on image processing, vol. 13, no. 9, pp. 1200–1212, 2004.
[12]D. Pathak, P. Krahenbuhl, J. Donahue, T. Darrell, and A. A. Efros, “Context encoders: Feature learning by inpainting,” in Proceedings of the IEEE conference on computer vision and pattern recognition, 2016, pp. 2536–2544.
[13]J. Yu, Z. Lin, J. Yang, X. Shen, X. Lu, and T. S. Huang, “Generative image inpainting with contextual attention,” in Proceedings of the IEEE conference on computer vision and pattern recognition, 2018, pp. 5505–5514.
[14]G. Liu, F. A. Reda, K. J. Shih, T.-C. Wang, A. Tao, and B. Catanzaro, “Image inpainting for irregular holes using partial convolutions,” in Proceedings of the European conference on computer vision (ECCV), 2018, pp. 85–100.
[15]J. Yu, Z. Lin, J. Yang, X. Shen, X. Lu, and T. S. Huang, “Free-form image inpainting with gated convolution,” in Proceedings of the IEEE/CVF international conference on computer vision, 2019, pp. 4471–4480.
[16]R. Suvorov, E. Logacheva, A. Mashikhin, A. Remizova, A. Ashukha, A. Silvestrov, N. Kong, H. Goka, K. Park, and V. Lempitsky, “Resolution-robust large mask inpainting with fourier convolutions,” in Proceedings of the IEEE/CVF winter conference on applications of computer vision, 2022, pp. 2149–2159.
[17]A. Lugmayr, M. Danelljan, A. Romero, F. Yu, R. Timofte, and L. Van Gool, “Repaint: Inpainting using denoising diffusion probabilistic models,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, 2022, pp. 11 461–11 471.
[18]Z. Wan, J. Zhang, D. Chen, and J. Liao, “High-fidelity pluralistic image completion with transformers,” in Proceedings of the IEEE/CVF international conference on computer vision, 2021, pp. 4692–4701. Volume (), Issue 17
[19]S. Iizuka, E. Simo-Serra, and H. Ishikawa, “Globally and locally consistent image completion,” vol. 36, no. 4. ACM New York, NY, USA, 2017, pp. 1–14.
[20]K. Nazeri, E. Ng, T. Joseph, F. Qureshi, and M. Ebrahimi, “Edgeconnect: Structure guided image inpainting using edge prediction,” in Proceedings of the IEEE/CVF international conference on computer vision workshops, 2019, pp. 0–0.
[21]J. Schlemper, O. Oktay, M. Schaap, M. Heinrich, B. Kainz, B. Glocker, and D. Rueckert, “Attention gated networks: Learning to leverage salient regions in medical images,” Medical image analysis, vol. 53, pp. 197–207, 2019.
[22]O. Oktay, J. Schlemper, L. L. Folgoc, M. Lee, M. Heinrich, K. Misawa, K. Mori, S. McDonagh, N. Y. Hammerla, B. Kainz et al., “Attention u-net: Learning where to look for the pancreas,” arXiv preprint arXiv:1804.03999, 2018.
[23]J. Johnson, A. Alahi, and L. Fei-Fei, “Perceptual losses for real-time style transfer and super-resolution,” in European conference on computer vision. Springer, 2016, pp. 694–711.
[24]L. A. Gatys, A. S. Ecker, and M. Bethge, “Image style transfer using convolutional neural networks,” in Proceedings of the IEEE conference on computer vision and pattern recognition, 2016, pp. 2414–2423.
[25]K. Armanious, C. Jiang, M. Fischer, T. K¨ustner, T. Hepp, K. Nikolaou, S. Gatidis, and B. Yang, “Medgan: Medical image translation using gans,” Computerized Medical Imaging and Graphics, vol. 79, p. 101684, 2020.
[26]W. H. Pinaya, P.-D. Tudosiu, J. Dafflon, P. F. Da Costa, V. Fernandez, P. Nachev, S. Ourselin, and M. J. Cardoso, “Brain imaging generation with latent diffusion models,” arXiv preprint arXiv:2209.07162, 2022.
[27]M. ¨Ozbey, O. Dalmaz, S. U. Dar, H. A. Bedel, S¸ . ¨Ozturk, A. G¨ung¨or, and T. C¸ ukur, “Unsupervised medical image translation with adversarial diffusion models,” IEEE Transactions on Medical Imaging, vol. 42, no. 12, pp. 3524–3539, 2023.
[28]D. Wang, Y. Kang, Y. Chen, Y. Gao, and S. Xu, “Sain: structure-aware image inpainting for large missing areas,” Journal of King Saud University Computer and Information Sciences, 2026.
[29]O. Elharrouss, N. Almaadeed, S. Al-Maadeed, and Y. Akbari, “Image inpainting: A review,” Neural Processing Letters, vol. 51, pp. 2007–2028, 2020.
[30]A. Kazerouni, E. K. Aghdam, M. Heidari, R. Azad, M. Fayyaz, I. Hacihaliloglu, and D. Merhof, “Diffusion models for medical image analysis: A comprehensive survey,” arXiv preprint arXiv:2211.07804, 2022.
[31]K. Velu and N. Jaisankar, “Design of a cnn–swin transformer model for alzheimer’s disease prediction using mri images,” IEEE Access, 2025.
[32]Kaggle, “Alzheimer’s mri preprocessed dataset,” https://www.kaggle.com/datasets/drsaeedmohsen/alzheimer dataset, accessed: 2024-12-28.
[33]N. V. Chawla, K. W. Bowyer, L. O. Hall, and W. P. Kegelmeyer, “Smote: synthetic minority over-sampling technique,” Journal of artificial intelligence research, vol. 16, pp. 321–357, 2002.
[34]L. Zhou, A. Jain, A. K. Dubey, S. K. Singh, N. Gupta, A. Panwar, S. Kumar, T. A. Althaqafi, V. Arya, W. Alhalabi et al., “Fpa-based weighted average ensemble of deep learning models for classification of lung cancer using ct scan images,” Scientific Reports, vol. 15, no. 1, p. 19369, 2025.