Mistral 7B | Mistral AI | Open source models
mistral.ai · 844 words · saved by 1 readers
The best 7B model to date, Apache 2.0
Research Mistral 7B September 27, 2023 By Mistral AI team Back to Blog 5 min read Share this post Copy url to clipboard Copied Mistral AI team is proud to release Mistral 7B, the most powerful language model for its size to date. Mistral 7B in short Mistral 7B is a 7.3B parameter model that: Outperforms Llama 2 13B on all benchmarks Outperforms Llama 1 34B on many benchmarks Approaches CodeLlama 7B performance on code, while remaining good at English tasks Uses Grouped-query attention (GQA) for faster inference Uses Sliding Window Attention (SWA) to handle longer sequences at smaller cost We'r
related reading
- Papers Explained 64: Mistralmedium.com
- mistralai/Mistral-7B-Instruct-v0.2 · Hugging Facehuggingface.co
- Mistral Mastery: Fine-Tuning & Fast Inference Guidemedium.com
- AI Model & API Providers Analysis | Artificial Analysisartificialanalysis.ai
- Mediumadithyask.medium.com
- TheBloke/Mistral-7B-Instruct-v0.1-AWQ · Hugging Facehuggingface.co
- GitHub - openlm-research/open_llama: OpenLLaMA, a permissively licensed open source reproduction of Meta AI’s LLaMA 7B trained on the RedPajama datasetgithub.com
- mistralai/Mistral-Large-Instruct-2411 · Hugging Facehuggingface.co
- Unsloth update: Mistral support + moreunsloth.ai
- GitHub - eugeneyan/open-llms: 📋 A list of open LLMs available for commercial use.github.com
- Welcome Mixtral - a SOTA Mixture of Experts on Hugging Facehuggingface.co
- AI Timeline — Complete History of 194+ Large Language Modelsllm-timeline.com