2112.10769
arxiv.org · 7,531 words · saved by 1 readers
N/A
Published as a conference paper at ICLR 2023 ACCURATE N EURAL T RAINING WITH 4- BIT M ATRIX M ULTIPLICATIONS AT S TANDARD F ORMATS Brian Chmiel †◦ Ron Banner † Elad Hoffer † Hilla Ben Yaacov † Daniel Soudry ◦ † Habana Labs – An Intel company, Caesarea, Israel, ◦ Department of…
saved by
related reading
- lec06.pdfdropbox.com
- lec05.pdfdropbox.com
- The 4-bitter Lesson | humans&humansand.ai
- [2106.08295] A White Paper on Neural Network Quantizationarxiv.org
- A Visual Guide to Quantization - by Maarten Grootendorstnewsletter.maartengrootendorst.com
- Integer Quantization for Deep Learning Inference: Principles and Empirical Evaluationarxiv.org
- A Guide to Quantization in LLMs | Symbl.aisymbl.ai
- 2210.17323arxiv.org
- On neural scaling and the quanta hypothesisericjmichaud.com
- 2305.14314arxiv.org
- LoRA Without Regret - Thinking Machines Labthinkingmachines.ai
- Quantization from the ground upngrok.com