Skip to main navigation Skip to search Skip to main content

Understanding Estimation and Generalization Error of Generative Adversarial Networks

  • Ohio State University

Research output: Contribution to journalArticlepeer-review

21 Scopus citations

Abstract

This article investigates the estimation and generalization errors of the generative adversarial network (GAN) training. On the statistical side, we develop an upper bound as well as a minimax lower bound on the estimation error for training GANs. The upper bound incorporates the roles of both the discriminator and the generator of GANs, and matches the minimax lower bound in terms of the sample size and the norm of the parameter matrices of neural networks under ReLU activation. On the algorithmic side, we develop a generalization error bound for the stochastic gradient method (SGM) in training GANs. Such a bound justifies the generalization ability of the GAN training via SGM after multiple passes over the data and reflects the interplay between the discriminator and the generator. Our results imply that the training of the generator requires more samples than the training of the discriminator. This is consistent with the empirical observation that the training of the discriminator typically converges faster than that of the generator. The experiments validate our theoretical results.

Original languageEnglish
Article number9330788
Pages (from-to)3114-3129
Number of pages16
JournalIEEE Transactions on Information Theory
Volume67
Issue number5
DOIs
StatePublished - May 2021

Keywords

  • estimation error
  • GAN training
  • generalization error
  • neural networks
  • stochastic gradient method

Fingerprint

Dive into the research topics of 'Understanding Estimation and Generalization Error of Generative Adversarial Networks'. Together they form a unique fingerprint.

Cite this