Back to library

AI / Technology

Generative Adversarial Networks

Three key questions about this paper

What problem does Generative Adversarial Networks address?

Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014 · Primary source · PDF

What evidence supports the main claim in Generative Adversarial Networks?

The authors trained models on MNIST, the Toronto Face Database (TFD), and CIFAR-10. Their Parzen-window log-likelihood estimate on MNIST was 225 ± 2, compared with 214 ± 1.1 for Deep GSN, 138 ± 2 for DBN, and 121 ± 1.6 for stacked CAE. On TFD, the adversarial model scored 2057 ± 26: above Deep GSN (1890 ± 29) and DBN (1909 ± 66), but below stacked CAE (2110 ± 50). Higher is better for these reported log-likelihood estimates. [Paper, Table 1, p. 6]

What limitation should readers know about Generative Adversarial Networks?

There is also a coordination problem. If the generator advances too far without refreshing the discriminator, it can map many noise inputs to the same output and lose diversity—the failure the paper informally calls the “Helvetica scenario.” The method does not provide an explicit value for (pg(x)), making likelihood evaluation indirect. [Paper, §6, p. 7]

2 new free reports left todaySubscribe to Pro for unlimited reading and 10 new paper explanations each month.Upgrade to Pro