Generative Adversarial Networks or GANs are very powerful tools to generate data. However, training a GAN is not easy. More specifically, GANs suffer of three major issues such as instability of the training procedure, mode collapse and vanishing gradients.
In this episode I not only explain the most challenging issues one would encounter while designing and training Generative Adversarial Networks. But also some methods and architectures to mitigate them. In addition I elucidate the three specific strategies that researchers are considering to improve the accuracy and the reliability of GANs.
The most tedious issues of GANs
Convergence to equilibrium
A typical GAN is formed by at least two networks: a generator G and a discriminator D. The generator's task is to generate samples from random noise. In turn, the discriminator has to learn to distinguish fake samples from real ones. While it is theoretically possible that generators and discriminators converge to a Nash Equilibrium (at which both networks are in their optimal state), reaching such equilibrium is not easy.
Vanishing gradients
Moreover, a very accurate discriminator would push the loss function towards lower and lower values. This in turn, might cause the gradient to vanish and the entire network to stop learning completely.
Mode collapse
Another phenomenon that is easy to observe when dealing with GANs is mode collapse. That is the incapability of the model to generate diverse samples. This in turn, leads to generated data that are more and more similar to the previous ones. Hence, the entire generated dataset would be just concentrated around a particular statistical value.
The solution
Researchers have taken into consideration several approaches to overcome such issues. They have been playing with architectural changes, different loss functions and game theory.
Listen to the full episode to know more about the most effective strategies to build GANs that are reliable and robust.
Don't forget to join the conversation on our new Discord channel. See you there!
[RB] Replicating GPT-2, the most dangerous NLP model (with Aaron Gokaslan) (Ep. 83)
What is wrong with reinforcement learning? (Ep. 82)
Have you met Shannon? Conversation with Jimmy Soni and Rob Goodman about one of the greatest minds in history (Ep. 81)
Attacking machine learning for fun and profit (with the authors of SecML Ep. 80)
[RB] How to scale AI in your organisation (Ep. 79)
Replicating GPT-2, the most dangerous NLP model (with Aaron Gokaslan) (Ep. 78)
Training neural networks faster without GPU [RB] (Ep. 77)
How to generate very large images with GANs (Ep. 76)
[RB] Complex video analysis made easy with Videoflow (Ep. 75)
[RB] Validate neural networks without data with Dr. Charles Martin (Ep. 74)
How to cluster tabular data with Markov Clustering (Ep. 73)
Waterfall or Agile? The best methodology for AI and machine learning (Ep. 72)
Training neural networks faster without GPU (Ep. 71)
Validate neural networks without data with Dr. Charles Martin (Ep. 70)
Complex video analysis made easy with Videoflow (Ep. 69)
Episode 68: AI and the future of banking with Chris Skinner [RB]
Episode 67: Classic Computer Science Problems in Python
Episode 66: More intelligent machines with self-supervised learning
Episode 65: AI knows biology. Or does it?
Episode 64: Get the best shot at NLP sentiment analysis
Create your
podcast in
minutes
It is Free
Insight Story: Tech Trends Unpacked
Zero-Shot
Fast Forward by Tomorrow Unlocked: Tech past, tech future
Lex Fridman Podcast
The Unbelivable Truth - Series 1 - 26 including specials and pilot