flâneur — a map of the web's best reading

A Recipe for Training Neural Networks

karpathy.github.io · 3,882 words · saved by 13 readers

Musings of a Computer Scientist.

Some few weeks ago I posted a tweet on “the most common neural net mistakes”, listing a few common gotchas related to training neural nets. The tweet got quite a bit more engagement than I anticipated (including a webinar :)). Clearly, a lot of people have personally encountered the large gap between “here is how a convolutional layer works” and “our convnet achieves state of the art results”. So I thought it could be fun to brush off my dusty blog to expand my tweet to the long form that this topic deserves. However, instead of going into an enumeration of more common errors or fleshing them

Explore this link on the map →

saved by

related reading