flâneur — a map of the web's best reading

Training LLMs with AMD MI250 GPUs and MosaicML

mosaicml.com · 2,385 words · saved by 2 readers

With the release of PyTorch 2.0 and ROCm 5.4, we are excited to announce that LLM training works out of the box on AMD datacenter GPUs, with zero code changes, and at high performance (144 TFLOP/s/GPU)! We are thrilled to see promising alternative options for AI hardware, and look forward to evaluating future devices and larger clusters soon.

Training LLMs with AMD MI250 GPUs and MosaicML | Databricks Blog Skip to main content With the release of PyTorch 2.0 and ROCm 5.4, we are excited to announce that LLM training works out of the box on AMD MI250 accelerators with zero code changes and at high performance! With MosaicML, the AI community has additional hardware + software options to choose from. At MosaicML, we've searched high and low for new ML training hardware on behalf of our customers. We do this to increase compute availability (as the world is in an NVIDIA supply crunch!), expand and educate the market, and ultimately re

Explore this link on the map →

saved by

related reading