flâneur — a map of the web's best reading

CLIMB: CLustering-based Iterative Data Mixture Bootstrapping for Language Model Pre-training

research.nvidia.com · 1,349 words · saved by 1 readers

CLIMB is a clustering-based iterative data mixture bootstrapping method for language model pre-training.

CLIMB: CLustering-based Iterative Data Mixture Bootstrapping for Language Model Pre-training 🧗🏻 CLIMB: CLustering-based Iterative Data Mixture Bootstrapping for Language Model Pre-training Shizhe Diao , Yu Yang , Yonggan Fu , Xin Dong , Dan Su , Markus Kliegl , Zijia Chen , Peter Belcak , Yoshi Suhara , Hongxu Yin , Mostofa Patwary , Yingyan (Celine) Lin , Jan Kautz , Pavlo Molchanov NVIDIA, Georgia Institute of Technology * Indicates Equal Contribution --> Paper 🤗 ClimbLab 🤗 ClimbMix Supplementary --> Code --> " target="_blank" class="external-link button is-normal is-rounded is-dark"> ar

Explore this link on the map →

related reading