Researches Advanced in Image Generation based on Deep Learning

Authors

  • Jianing Duan

DOI:

https://doi.org/10.54097/hset.v39i.6561

Keywords:

Image Generation; VAE and GAN; Deep Learning.

Abstract

Image generation has always been a study hotspot in machine learning, which aims to build models to learn specific semantic distributions from massive image data to generate realistic simulated images. Thanks to the deep learning technology’s quick development, generative models are constantly being developed and huge success has been achieved in image generation tasks. According to difference between generative models, the existing image generation methods based on deep learning can mainly be separated into three models: image generation based on Variational Autoencoder (VAE), image generation based on Generative Adversarial Network (GAN) and image generation combined the VAE and GAN. Focusing on the three frameworks, in this paper, the development process and related principles of each type of generation model are described respectively. After that, the different generation results of different generation models for the agreed training set are compared intuitively, the advantages and problems of various models are proposed, and reasonable improvement measures are proposed for some problems. Finally, the development prospects of various models are prospected.

Downloads

Download data is not yet available.

References

Kingma D P, Welling M. (2013) “Auto-encoding variational bayes”. arXiv preprint arXiv:1312.6114, 2013.

Kingma D P, Mohamed S, Jimenez Rezende D, et al. Semi-supervised learning with deep generative models[C]//Advances in neural information processing systems, 2014, 27.

Vincent P, Larochelle H, Bengio Y, et al. Extracting and composing robust features with denoising autoencoders[C]//Proceedings of the 25th international conference on Machine learning. 2008: 1096-1103.

Im Im D, Ahn S, Memisevic R, et al. Denoising criterion for variational auto-encoding framework [C]// Proceedings of the AAAI Conference on Artificial Intelligence. 2017, 31(1).

Higgins I, Matthey L, Pal A, et al. (2016) “beta-vae: Learning basic visual concepts with a constrained variational framework”.

Goodfellow I, Pouget-Abadie J, Mirza M, et al. (2020) “Generative adversarial networks”. Communications of the ACM, 63(11): 139-144.

Mirza M, Osindero S. (2014) “Conditional generative adversarial nets”. arXiv preprint arXiv:1411.1784, 2014.

Chen X, Duan Y, Houthooft R, et al. Infogan: Interpretable representation learning by information maximizing generative adversarial nets[J]. Advances in neural information processing systems, 2016, 29.

Zhao J, Mathieu M, LeCun Y. (2016) “Energy-based generative adversarial network”. arXiv preprint arXiv:1609.03126, 2016.

Mao X, Li Q, Xie H, et al. Least squares generative adversarial networks[C]//Proceedings of the IEEE international conference on computer vision. 2017: 2794-2802.

Odena A, Olah C, Shlens J. Conditional image synthesis with auxiliary classifier gans[C]//International conference on machine learning. PMLR, 2017: 2642-2651.

Arjovsky M, Chintala S, Bottou L. Wasserstein generative adversarial networks[C]//International conference on machine learning. PMLR, 2017: 214-223.

Gulrajani I, Ahmed F, Arjovsky M, et al. Improved training of wasserstein gans[C]//Advances in neural information processing systems, 2017, 30.

Kodali N, Abernethy J, Hays J, et al. (2017) “On convergence and stability of gans”. arXiv preprint arXiv:1705.07215, 2017.

Hjelm R D, Jacob A P, Che T, et al. (2017) “Boundary-seeking generative adversarial networks”. arXiv preprint arXiv:1702.08431, 2017.

Larsen A B L, Sønderby S K, Larochelle H, et al. Autoencoding beyond pixels using a learned similarity metric [C]//International conference on machine learning. PMLR, 2016: 1558-1566.

Bao J, Chen D, Wen F, et al. CVAE-GAN: fine-grained image generation through asymmetric training [C]//Proceedings of the IEEE international conference on computer vision. 2017: 2745-2754.

Donahue J, Krähenbühl P, Darrell T. (2016) “Adversarial feature learning”. arXiv preprint arXiv: 1605. 09782, 2016.

Ye F, Bors A G. InfoVAEGAN: learning joint interpretable representations by information maximization and maximum likelihood [C]//2021 IEEE International Conference on Image Processing (ICIP). IEEE, 2021: 749-753.

Downloads

Published

01-04-2023