Why you need to improve your training data, and how to do it « Pete Warden's blog
Photo by Lisha Li Andrej Karpathy showed this slide as part of his talk at Train AI and I loved it! It captures the difference between deep learning research and production perfectly. Academic pape…
May 28, 2018 By Pete Warden in Uncategorized 21 Comments Photo by Lisha Li Andrej Karpathy showed this slide as part of his talk at Train AI and I loved it! It captures the difference between deep learning research and production perfectly. Academic papers are almost entirely focused on new and improved models, with datasets usually chosen from a small set of public archives. Everyone I know who uses deep learning as part of an actual application spends most of their time worrying about the training data instead. There are lots of good reasons why researchers are so fixated on model architectu
Explore this link on the map →saved by
related reading
- A Recipe for Training Neural Networkskarpathy.github.io
- Thinking about High-Quality Human Data | Lil'Loglilianweng.github.io
- The Only Important Technology Is The Internet - Kevin Lukevinlu.ai
- A Recipe for Training Neural Networkskarpathy.github.io
- There Are No New Ideas in AI… Only New Datasetsblog.jxmo.io
- Training Data: What Is It? All About Machine Learning Training Dataappen.com
- The Little Book of Deep Learningfleuret.org
- DataRater: Meta-Learned Dataset Curationarxiv.org
- Jupyter Notebook Viewernbviewer.org
- MAI-Thinking-1: Building a Hill-Climbing Machinemicrosoft.ai
- Training, validation, and test data sets - Wikipediaen.wikipedia.org
- The Unreasonable Effectiveness of Easy Training Data for Hard Tasksalexandrabarr.beehiiv.com