OLMo from Ai2
OLMo 2 is a family of fully-open language models, developed start-to-finish with open and accessible training data, open-source training code, reproducible training recipes, transparent evaluations, intermediate checkpoints, and more. Details are in our technical report. OLMo 2 32B is the most capable and largest model in the OLMo 2 family, scaling up the OLMo 2 training recipe used for our 7B and 13B models. It is trained up to 6T tokens and post-trained using Tulu 3.1. OLMo 2 32B is the first fully-open model to outperform GPT3.5-Turbo and GPT-4o mini on a suite of popular, multi-skill academic benchmarks. OLMo 2 7B and 13B models are trained on up to 5T tokens. These models are on par with or better than equivalently-sized fully-open models, and competitive with open-weight models from Meta and Mistral on English academic benchmarks. OLMo 2 1B is the smallest member of the OLMo model family, surpassing peer models like Gemma 3 1B or Llama 3.2 1B. The 1B model size enables rapid iter
Olmo from Ai2 Olmo Our fully open language model and complete model flow. Chat with Olmo Build with Olmo The Olmo 3 model family Pick a variant to explore weights, code and reports. Every card includes instant links to artifacts. Read the technical report 32B-Base Achieves strong results in programming, reading comprehension, and math problem solving, maintains performance at extended context lengths, and works well with RL setups. 32B-Think Capable of reasoning through complex problems step by step. A strong platform for RL research and other advanced experiments that need serious horsepower.
Explore this link on the map →related reading
- Olmo 3: Charting a path through the model flow to lead open-source AI | Ai2allenai.org
- [2512.13961] Olmo 3arxiv.org
- Google "We Have No Moat, And Neither Does OpenAI"semianalysis.com
- Inkling: Our Open-Weights Model - Thinking Machines Labthinkingmachines.ai
- Composer2.pdfcursor.com
- State of AI 2025: 100T Token LLM Usage Study | OpenRouteropenrouter.ai
- Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Modelsarxiv.org
- o1 and Reasoning | AndoLogsblog.ando.ai
- Pathways Language Model (PaLM): Scaling to 540 Billion Parameters for Breakthrouai.googleblog.com
- Gemma 2: Improving Open Language Models at a Practical Sizearxiv.org
- The Llama Hitchiking Guide to Local LLMs – hackerllamaosanseviero.github.io
- RedPajama, a project to create leading open-source models, starts by reproducing LLaMA training dataset of over 1.2 trillion tokenstogether.xyz