Olmo is Ai2’s fully open language model family and model flow for researchers and developers. It includes base, thinking, and instruct variants and exposes artifacts such as weights, code, and reports.

Olmo preview

Overview

Olmo is Ai2’s fully open language model and complete model flow. The product page frames it as more than a single model release: it presents the model family, the artifacts behind each variant, and the training stages that make up the full lifecycle.

The page highlights Olmo 3 variants for different needs, including base, thinking, and instruct models at 32B and 7B scales. It also states that users can chat with Olmo or build with Olmo, making it relevant to researchers and developers who want to inspect, adapt, or use open language models in practice.

Features

Fully open model flow

Olmo is presented as a fully open language model, with the page emphasizing access to the complete model flow rather than only a finished endpoint.

Multiple model variants

The Olmo 3 family is shown with multiple variants, including 32B-Base, 32B-Think, 32B-Instruct, 7B-Base, 7B-Think, and 7B-Instruct.

Separated by training and use style

Variant descriptions distinguish between base models, reasoning-focused thinking models, and instruction-tuned chat models for different workloads.

Direct access to artifacts

The page says each card includes instant links to artifacts such as weights, code, and reports.

End-to-end training flow

The model flow graphic covers stages such as pretraining, mid-training, long context, instruct SFT, instruct DPO, instruct RL, thinking SFT, thinking DPO, thinking RL, and RL zero.

Documented training data mix

The source text states that the pretraining data is a fully open mixture curated from web, code, books, and scientific text, with deduplication and quality filtering.

Use Cases

  • Open-model research

    Use Olmo when you want an open model family with enough structure to study training stages, variant differences, and artifact links instead of only consuming an API output.

  • Model development and adaptation

    Use the base variants when your work involves experimentation, evaluation, or further training on a model that starts from a documented pretraining mixture.

  • Assistant-style applications

    Use the instruct variants for chat-style applications, tool use, and multi-turn dialogue where instruction tuning is the main requirement.

  • Reasoning and RL experiments

    Use the thinking variants when the task benefits from step-by-step reasoning or advanced reinforcement learning experiments.

Pros and Cons

Pros

  • Presents both the model family and the underlying model flow, giving more transparency than a single-endpoint release.
  • Offers multiple variants for different tasks, including base, thinking, and instruct models.
  • States that artifact links are available for weights, code, and reports.
  • Describes the pretraining data mix at a high level, including web, code, books, and scientific text.

Cons

  • The source does not provide pricing, licensing terms, or a clear access model beyond calling the system fully open.
  • Feature details come from the product page summary rather than deeper documentation, so some workflow specifics are not exposed in the provided evidence.

FAQ

What is Olmo?

Ai2 positions Olmo as its fully open language model and complete model flow. The page highlights that the model family includes base, thinking, and instruct variants, and that each card links to weights, code, and reports.

How is Olmo organized?

The source page says users can chat with Olmo and build with Olmo. It also presents the Olmo 3 model family and a model flow that includes pretraining, mid-training, long-context work, and fine-tuning stages.

Who is Olmo for?

The page explicitly frames Olmo as a fully open language model. It is aimed at researchers and developers who want access to the model flow as well as the model artifacts.

How much does Olmo cost?

The available source does not show pricing on the Olmo page. Ai2’s separate pricing URL returns a 404 page, so the access model is not clarified in the provided evidence.

Quick Facts

Category
Open Language Model
Source domain
allenai.org
Primary users
Researchers and developers
Model family
Olmo 3
Variants shown
32B and 7B base, think, and instruct models
Notable workflow
Complete model flow with pretraining, mid-training, long context, and fine-tuning stages