Journals & Magazines >IEEE Transactions on Pattern ... >Volume: 41 Issue: 12

On the Effectiveness of Least Squares Generative Adversarial Networks

Download PDF
Download References
Request Permissions
Save to
Alerts

Abstract:

Unsupervised learning with generative adversarial networks (GANs) has proven to be hugely successful. Regular GANs hypothesize the discriminator as a classifier with the ...Show More

Metadata

Abstract:

Unsupervised learning with generative adversarial networks (GANs) has proven to be hugely successful. Regular GANs hypothesize the discriminator as a classifier with the sigmoid cross entropy loss function. However, we found that this loss function may lead to the vanishing gradients problem during the learning process. To overcome such a problem, we propose in this paper the Least Squares Generative Adversarial Networks (LSGANs) which adopt the least squares loss for both the discriminator and the generator. We show that minimizing the objective function of LSGAN yields minimizing the Pearson

$\chi ^2$ divergence. We also show that the derived objective function that yields minimizing the Pearson

$\chi ^2$ divergence performs better than the classical one of using least squares for classification. There are two benefits of LSGANs over regular GANs. First, LSGANs are able to generate higher quality images than regular GANs. Second, LSGANs perform more stably during the learning process. For evaluating the image quality, we conduct both qualitative and quantitative experiments, and the experimental results show that LSGANs can generate higher quality images than regular GANs. Furthermore, we evaluate the stability of LSGANs in two groups. One is to compare between LSGANs and regular GANs without gradient penalty. We conduct three experiments, including Gaussian mixture distribution, difficult architectures, and a newly proposed method — datasets with small variability, to illustrate the stability of LSGANs. The other one is to compare between LSGANs with gradient penalty (LSGANs-GP) and WGANs with gradient penalty (WGANs-GP). The experimental results show that LSGANs-GP succeed in training for all the difficult architectures used in WGANs-GP, including 101-layer ResNet.

Published in: IEEE Transactions on Pattern Analysis and Machine Intelligence ( Volume: 41, Issue: 12, 01 December 2019)

Page(s): 2947 - 2960

Date of Publication: 24 September 2018

ISSN Information:

PubMed ID: 30273144

DOI: 10.1109/TPAMI.2018.2872043

Funding Agency:

References is not available for this document.

Contents

1 Introduction

Deep learning has launched a profound reformation and even been applied to many real-world tasks, such as image classification [1], object detection [2], and segmentation [3]. These tasks fall into the scope of supervised learning, which means that a lot of labeled data is provided for the learning processes. Compared with supervised learning, however, unsupervised learning (such as generative models) obtains limited impact from deep learning. Although some deep generative models, e.g., RBM [4], DBM [5], and VAE [6], have been proposed, these models all face the difficulties of intractable functions (e.g., intractable partition function) or intractable inference, which in turn restricts the effectiveness of these models.

Select All

K. He, X. Zhang, S. Ren and J. Sun, "Deep residual learning for image recognition", Proc. IEEE Conf. Comput. Vis. Pattern Recognit., pp. 770-778, 2016.

On the Effectiveness of Least Squares Generative Adversarial Networks

Alerts

Abstract:

Metadata

Abstract:

ISSN Information:

Funding Agency:

1 Introduction

Authors

Figures

References

Citations

Keywords

Metrics

Footnotes

References

IEEE Account

Purchase Details

Profile Information

Need Help?