flâneur — a map of the web's best reading

Mosaic ResNet Deep Dive

mosaicml.com · 4,345 words · saved by 1 readers

TL;DR: We recently released a set of recipes which can accelerate training of a ResNet-50 on ImageNet by up to 7x over standard baselines. In this report we take a deep dive into the technical details of our work and share the insights we gained about optimizing the efficiency of model training over a broad range of compute budgets.

Mosaic ResNet Deep Dive | Databricks Blog Skip to main content TL;DR: We recently released a set of recipes which can accelerate training of a ResNet-50 on ImageNet by up to 7x over standard baselines. In this report we take a deep dive into the technical details of our work and share the insights we gained about optimizing the efficiency of model training over a broad range of compute budgets. Introduction ResNets have established themselves as the go-to baseline and testbed for computer vision research (see ResNet Strikes Back or the PyTorch blog ). More efficient training recipes can save m

Explore this link on the map →

saved by

related reading