flâneur — a map of the web's best reading

The Unreasonable Effectiveness of Recurrent Neural Networks

karpathy.github.io · 8,155 words · saved by 14 readers

Musings of a Computer Scientist.

There’s something magical about Recurrent Neural Networks (RNNs). I still remember when I trained my first recurrent network for Image Captioning . Within a few dozen minutes of training my first baby model (with rather arbitrarily-chosen hyperparameters) started to generate very nice looking descriptions of images that were on the edge of making sense. Sometimes the ratio of how simple your model is to the quality of the results you get out of it blows past your expectations, and this was one of those times. What made this result so shocking at the time was that the common wisdom was that RNN

Explore this link on the map →

saved by

related reading