
Sign up to save your podcasts
Or


This episode discusses the vanishing gradient - a problem that arises when training deep neural networks in which nearly all the gradients are very close to zero by the time back-propagation has reached the first hidden layer. This makes learning virtually impossible without some clever trick or improved methodology to help earlier layers begin to learn.
By Kyle Polich4.4
475475 ratings
This episode discusses the vanishing gradient - a problem that arises when training deep neural networks in which nearly all the gradients are very close to zero by the time back-propagation has reached the first hidden layer. This makes learning virtually impossible without some clever trick or improved methodology to help earlier layers begin to learn.

32,103 Listeners

30,680 Listeners

288 Listeners

1,094 Listeners

624 Listeners

583 Listeners

299 Listeners

344 Listeners

209 Listeners

201 Listeners

318 Listeners

98 Listeners

576 Listeners

100 Listeners

228 Listeners