Comparative Analysis the Super-Resolution Image Generation Performance Based on BigGAN and VQ-VAE-2
DOI:
https://doi.org/10.54097/hset.v41i.6812Keywords:
Super-resolution image reconstruction, deep learning, BigGAN, VQ-VAE-2.Abstract
Super-resolution image reconstruction has always been a popular research direction in the field of computer vision, which aims to recover high-resolution clear images from low-resolution images. Traditional super-resolution reconstruction algorithms mainly rely on the construction of constraints and the accuracy of registration between images to achieve the reconstruction effect, but their accuracy cannot meet the needs of practical applications with large multiples. Thanks to the rapid development of deep learning field, super-resolution image reconstruction based on deep learning has become the mainstream, and has achieved great success in reconstruction accuracy and speed. According to the different generative models used, the existing super-resolution image reconstruction methods mainly include two categories: GAN-based and VAEs-based. To quantitatively compare the limits of the two approaches' performance, this study selects two representative algorithms, BigGAN and VQ-VAE-2, and introduces the theoretical details and training process of these two methods, respectively. Furthermore, the reconstruction results of BigGAN and VQ-VAE-2 are further compared. Finally, this paper discusses the future development trend of super-resolution picture reconstruction with the current potential problems of BigGAN and VQ-VAE-2.
Downloads
References
Umehara, K., Ota, J. and Ishida, T. (2018) “Application of Super-Resolution Convolutional Neural Network for Enhancing Image Resolution in Chest Ct,” Journal of digital imaging, 31(4), pp. 441–450. doi: 10.1007/s10278-017-0033-z.
Mandanici, E. et al. (2019) “A Multi-Image Super-Resolution Algorithm Applied to Thermal Imagery,” Applied Geomatics, 11(3), pp. 215–228. doi: 10.1007/s12518-019-00253-y.
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. Generative adversarial nets. In Advances in neural information processing systems, pages 2672–2680, 2014.
Iederik P. Kingma and Max Welling. Auto-encoding variational bayes. CoRR, abs/1312.6114, 2013.
Hugo Larochelle and Iain Murray. The neural autoregressive distribution estimator. In Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics, pages 29–37, 2011.
Joyce, James M. "Kullback-leibler divergence." International encyclopedia of statistical science. Springer, Berlin, Heidelberg, 2011. 720-722.
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, Vicki Cheung, Alec Radford, and Xi Chen. Improved techniques for training gans. In Advances in neural information processing systems, pages 2234–2242, 2016.
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter. Gans trained by a two time-scale update rule converge to a local nash equilibrium. In I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett, editors, Advances in Neural Information Processing Systems 30, pages 6626–6637. Curran Associates, Inc., 2017.
A. Creswell, T. White, V. Dumoulin, K. Arulkumaran, B. Sengupta and A. A. Bharath, "Generative Adversarial Networks: An Overview," in IEEE Signal Processing Magazine, vol. 35, no. 1, pp. 53-65, Jan. 2018, doi: 10.1109/MSP.2017.2765202.
Brock, A., Donahue, J. and Simonyan, K., 2019. Large Scale GAN Training for High Fidelity Natural Image Synthesis. [online] arXiv.org. Available at: <https://arxiv.org/abs/1809.11096>.
Andrew Brock, Theodore Lim, J.M. Ritchie, and Nick Weston. Neural photo editing with introspective adversarial networks. In ICLR, 2017.
Oord, A., Vinyals, O. and Kavukcuoglu, K., 2017. Neural Discrete Representation Learning. [online] arXiv.org. Available at: <https://arxiv.org/abs/1711.00937v2>.
Razavi, A., Oord, A. and Vinyals, O., 2019. Generating Diverse High-Fidelity Images with VQ-VAE-2. [online] arXiv.org. Available at: <https://arxiv.org/abs/1906.00446v1>.
Jiang, Y., Li, X., Luo, H. et al. Quo vadis artificial intelligence? Discov Artif Intell 2, 4 (2022). https://doi.org/10.1007/s44163-022-00022-8.
Downloads
Published
Issue
Section
License

This work is licensed under a Creative Commons Attribution-NonCommercial 4.0 International License.







