TheBloke/Mistral-7B-Instruct-v0.1-AWQ · Hugging Face
AWQ is an efficient, accurate and blazing-fast low-bit weight quantization method, currently supporting 4-bit quantization. Compared to GPTQ, it offers faster Transformers-based inference.
Mistral 7B Instruct v0.1 - AWQ Model creator: Mistral AI Original model: Mistral 7B Instruct v0.1 Description This repo contains AWQ model files for Mistral AI's Mistral 7B Instruct v0.1. About AWQ AWQ is an efficient, accurate and blazing-fast low-bit weight quantization method, currently supporting 4-bit quantization. Compared to GPTQ, it offers faster Transformers-based inference. Mistral AWQs These are experimental first AWQs for the brand-new model format, Mistral. As of September 29th 2023, they are only supported by AutoAWQ (version 0.1.1+) Repositories available AWQ…
related reading
- AI Model & API Providers Analysis | Artificial Analysisartificialanalysis.ai
- Hugging Face – The AI community building the future.huggingface.co
- Together AI | The AI Native Cloudtogether.ai
- Parsed | Custom, interpretable AI systems that continuously learnparsed.com
- Quantization · Hugging Facehuggingface.co
- Replicate - Run AI with an APIreplicate.com
- Compare AI Models: Pricing, Context & Benchmarks | OpenRouteropenrouter.ai
- There's An AI For That® — The front page of AItheresanaiforthat.com
- mistralai/Mistral-7B-Instruct-v0.2 · Hugging Facehuggingface.co
- Goodfire AIgoodfire.ai
- Melius | The AI-native operating system for creativesmelius.com
- Branches · HazyResearch/intelligence-per-watt · GitHubgithub.com