✳flâneur — a map of the web's best reading
Mistral 7B | Mistral AI | Open source models
mistral.ai · 844 words · saved by 1 readers
The best 7B model to date, Apache 2.0
Research Mistral 7B September 27, 2023 By Mistral AI team Back to Blog 5 min read Share this post Copy url to clipboard Copied Mistral AI team is proud to release Mistral 7B, the most powerful language model for its size to date. Mistral 7B in short Mistral 7B is a 7.3B parameter model that: Outperforms Llama 2 13B on all benchmarks Outperforms Llama 1 34B on many benchmarks Approaches CodeLlama 7B performance on code, while remaining good at English tasks Uses Grouped-query attention (GQA) for faster inference Uses Sliding Window Attention (SWA) to handle longer sequences at smaller cost We'r
Explore this link on the map →related reading
- AI Model & API Providers Analysis | Artificial Analysisartificialanalysis.ai
- Mediumadithyask.medium.com
- mistralai/Mistral-7B-Instruct-v0.2 · Hugging Facehuggingface.co
- Unsloth update: Mistral support + moreunsloth.ai
- Introducing Llama2-70B-Chat with MosaicML Inference | Databricks Blogmosaicml.com
- Fine-Tuning Llama-2: Tailoring Models to Unique Applicationsanyscale.com
- GitHub - karpathy/nanochat: The best ChatGPT that $100 can buy. · GitHubgithub.com
- Pathways Language Model (PaLM): Scaling to 540 Billion Parameters for Breakthrouai.googleblog.com
- GitHub - openai/parameter-golf: Train the smallest LM you can that fits in 16MB. Best model wins! · GitHubgithub.com
- 2506.17298arxiv.org
- Llama 2 · Hugging Facehuggingface.co
- Model optimization | OpenAI APIplatform.openai.com