COMPARATIVE ANALYSIS OF DEEP LEARNING ARCHITECTURES FOR MONOCHROMATIC IMAGE COLORIZATION
DOI:
https://doi.org/10.68302/std2026.vol3.217Keywords:
computer vision, image colorization, deep learning, GANs, U-Net, CNNs, LPIPS, SSIM, PSNR, COCO datasetAbstract
Colorizing monochromatic images is crucial for enhancing the informativeness of visual data in modern information technology and computer vision systems. However, automatic colorization poses an inherently ill-posed mathematical challenge due to the multimodal nature of color distributions, where multiple valid color mappings exist for any given grayscale input. This article conducts a comparative analysis of classical algorithms (such as the Welch and Levin methods) against advanced deep learning pipelines. The evaluated architectures range from baseline convolutional neural networks (CNNs) and standalone U-Net models to generative adversarial networks (GANs) and a novel hybrid Fusion GAN that integrates global semantic priors extracted via a ResNet-18 backbone. All models were trained and rigorously evaluated in the CIE Lab* color space, which effectively separates luminance from chrominance. To ensure a robust evaluation, experiments were conducted using a diverse benchmark of 2,000 images from the COCO dataset. Quality assessment combined traditional metrics like Peak Signal-to-Noise Ratio (PSNR) and Structural Similarity Index (SSIM) with the Learned Perceptual Image Patch Similarity (LPIPS) metric, prioritizing human-like visual fidelity. Experimental results yielded counterintuitive insights. While the proposed Fusion GAN successfully surpassed classical methods, baseline CNNs, and standard GANs across most benchmarks, the standalone U-Net architecture secured the highest overall ranking. Specifically, U-Net achieved the top SSIM score of 0.945 (indicating superior structural preservation), the lowest LPIPS of 0.180 (best perceptual quality), and the fastest inference speed, enabling real-time applications.
Downloads
References
[1] S. Anwar, M. Tahir, C. Li, A. Mian, F. S. Khan, and A. W. Muzaffar, "Deep Learning for Image Colorization: A Survey," arXiv preprint arXiv:2008.10774, 2024. [Online]. Available: https://arxiv.org/abs/2008.10774. DOI: https://doi.org/10.48550/arXiv.2008.10774
[2] S. Bielievtsov, I. Ruban, K. Smelyakov, and D. Sumtsov, "Network technology for transmission of visual information," in Selected Papers of the XVIII Int. Sci.-Practical Conf. "Information Technologies and Security" (ITS 2018), Kyiv, Ukraine, Nov. 27, 2018, CEUR Workshop Proceedings, vol. 2318, 2018, pp. 160–175. [Online]. Available: https://ceur-ws.org/Vol-2318/
[3] T. Welsh, M. Ashikhmin, and K. Mueller, "Transferring Color to Greyscale Images," ACM Trans. Graph. (Proc. SIGGRAPH), vol. 21, no. 3, pp. 277–280, 2002. DOI: https://doi.org/10.1145/566654.566576
[4] A. Levin, D. Lischinski, and Y. Weiss, "Colorization Using Optimization," ACM Trans. Graph. (Proc. SIGGRAPH), vol. 23, no. 3, pp. 689–694, 2004. DOI: https://doi.org/10.1145/1186562.1015780
[5] R. Zhang, P. Isola, and A. A. Efros, "Colorful Image Colorization," in Proc. European Conf. Computer Vision (ECCV), Amsterdam, Netherlands, 2016, pp. 649–666. DOI: https://doi.org/10.48550/arXiv.1603.08511
[6] O. Ronneberger, P. Fischer, and T. Brox, "U-Net: Convolutional Networks for Biomedical Image Segmentation," in Proc. Int. Conf. Medical Image Computing and Computer-Assisted Intervention (MICCAI), Munich, Germany, 2015, pp. 234–241. DOI: https://doi.org/10.48550/arXiv.1505.04597
[7] P. Isola, J.-Y. Zhu, T. Zhou, and A. A. Efros, "Image-to-Image Translation with Conditional Adversarial Networks," in Proc. IEEE Conf. Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA, 2017, pp. 1125–1134. DOI: https://doi.org/10.48550/arXiv.1611.07004
[8] I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, "Generative Adversarial Nets," in Advances in Neural Information Processing Systems (NeurIPS), vol. 27, 2014. DOI: https://doi.org/10.48550/arXiv.1406.2661
[9] X. Kang, T. Yang, W. Ouyang, P. Ren, L. Li, and X. Xie, "DDColor: Towards Photo-Realistic Image Colorization via Dual Decoders," in Proc. IEEE/CVF Int. Conf. Computer Vision (ICCV), Paris, France, 2023, pp. 328–338. DOI: https://doi.org/10.48550/arXiv.2212.11613
[10] K. He, X. Zhang, S. Ren, and J. Sun, "Deep Residual Learning for Image Recognition," in Proc. IEEE Conf. Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA, 2016, pp. 770–778. DOI: https://doi.org/10.1109/CVPR.2016.90
[11] S. Iizuka, E. Simo-Serra, and H. Ishikawa, "Let there be Color!: Joint End-to-end Learning of Global and Local Image Priors for Automatic Image Colorization with Simultaneous Classification," ACM Trans. Graph. (Proc. SIGGRAPH), vol. 35, no. 4, pp. 110:1–110:11, 2016. DOI: https://doi.org/10.1145/2897824.2925974
[12] O. Zolotukhin, Y. Bodyanskiy, V. Filatov, M. Kudryavtseva, and Y. Yeriemieiev, "Stochastic Initialization for Neural Networks Based on the Analysis of Biological Systems," in Proc. 5th Int. Workshop of IT-Professionals on Artificial Intelligence (ProfIT AI 2025), Liverpool, United Kingdom, Oct. 15–17, 2025, pp. 415–426. [Online]. Available: https://ceur-ws.org/Vol-4164/
[13] R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang, "The Unreasonable Effectiveness of Deep Features as a Perceptual Metric," in Proc. IEEE Conf. Computer Vision and Pattern Recognition (CVPR), Salt Lake City, UT, USA, 2018, pp. 586–595. DOI: https://doi.org/10.1109/CVPR.2018.00068
[14] J. Johnson, A. Alahi, and L. Fei-Fei, "Perceptual Losses for Real-Time Style Transfer and Super-Resolution," in Proc. European Conf. Computer Vision (ECCV), Amsterdam, Netherlands, 2016, pp. 694–711, DOI: https://doi.org/10.1007/978-3-319-46475-6_43
[15] O. Byzkrovnyi, L. Savulioniene, K. Smelyakov, P. Sakalys, and A. Chupryna, "Comparison of Potential Road Accident Detection Algorithms for Modern Machine Vision System," Vide. Tehnologija. Resursi – Environment, Technology, Resources, vol. 3, pp. 50–55, 2023. DOI: https://doi.org/10.17770/etr2023vol3.7299
Downloads
Published
Issue
Section
License
Copyright (c) 2026 Ihor Panasenko, Kirill Smelyakov, Anastasiya Chupryna, Paulius Sakalys, Loreta Savulioniene, Dainius Savulionis

This work is licensed under a Creative Commons Attribution 4.0 International License.