NeXtSRGAN: enhancing super-resolution GAN with ConvNeXt discriminator for superior realism

N/ACitations
Citations of this article
14Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

Deep learning technologies have significantly advanced the field of single-image super-resolution (SISR), yet existing methods often prioritize peak signal-to-noise ratio (PSNR) over visual quality and realism. In this study, we propose NeXtSRGAN, which integrates a ConvNeXt-based discriminator to overcome these limitations and achieve more realistic and high-quality super-resolution (SR) images. NeXtSRGAN enhances image realism through its novel discriminator structure and generator residual block design, leveraging a relativistic discriminator and residual scaling. Experimental results on benchmark datasets demonstrate that NeXtSRGAN surpasses existing methods, including enhanced SRGAN (ESRGAN), with an average PSNR improvement of over 1.58 dB and an SSIM enhancement of over 0.05. Notably, NeXtSRGAN exhibits exceptional performance in facial image SR, as confirmed by the KID-F metric. By focusing on perceptual quality rather than solely PSNR, NeXtSRGAN sets a new standard for image restoration and holds promise for applications in various domains, such as video surveillance, medical imaging, and satellite photos. This code is available at https://github.com/PomKlementieff/NeXtSRGAN.

Cite

CITATION STYLE

APA

Park, S. W., Jung, S. H., & Sim, C. B. (2025). NeXtSRGAN: enhancing super-resolution GAN with ConvNeXt discriminator for superior realism. Visual Computer, 41(10), 7141–7167. https://doi.org/10.1007/s00371-024-03797-2

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free