USASI
Model release

Olmo 3 7B (base)

Release in the OLMo family · version Olmo-3-1025-7B

Maintained by Ai2 (Allen Institute for AI)1

The pretrained base model at the 7B scale in Ai2's Olmo 3 release, trained in three stages (pretraining, mid-training, and long-context extension) on Dolma 3 data. The model card lists 5.93 trillion training tokens, 32 layers, and a 65,536-token context length.110

Last reviewedEntry updated Documented release Nov 20, 2025

Availability and license

Overall availability

Public

Downloadable from Hugging Face without gating. The card says the model is intended for research and educational use in line with Ai2's Responsible Use Guidelines.1

Availability is separate from permission: read the license before using or redistributing.

The model card states that the code and model are released under Apache 2.0 and that the model is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines, which list categories of prohibited use.19

Model-disclosure tier

Computed from the checklist below using USASI rubric v0.1. An editorial category, not a certification.
Model-disclosure tier (USASI rubric v0.1): Open-stack

Open-weight, plus published inference code, training code, and training recipe, and at least documented training-data composition.

How tiers are computed

Public materials checklist

Items for a model under USASI rubric v0.1. Unknown means unassessed or insufficient evidence.
Public materials checklist for Olmo 3 7B (base)
ItemStatusNotes and evidence
WeightsCan the general public download the model parameters for this release?PublicFinal weights and staged intermediate checkpoints are published on Hugging Face.1
Inference codeIs code for running the model published?PublicThe model card documents inference with Hugging Face Transformers and lists OLMo-core as the core repository for training, inference, and fine-tuning.12
Training codeIs the code used to train the model published?PublicOLMo-core publishes the official Olmo 3 7B pretraining, mid-training, and long-context scripts.23
Training-data informationPublic = the training data itself can be obtained. Partial = composition or sources are documented without full access.PartialThe Dolma 3 Mix used for pretraining is downloadable under ODC-BY, but its dataset card says some olmOCR science PDFs were redacted after Olmo 3 7B was trained (marked [REMOVED]) and that this affects reproducibility. Mid-training and long-context mixes are linked from the model card.41
Training recipeAre the training configuration and procedure documented in enough detail to follow?PublicThe technical report documents stages, schedules, and configurations, and the OLMo-core scripts encode the official runs.52
Evaluation materialsPublic = evaluation code or prompts that let others re-run the evaluations are published. Partial = results only.PublicThe model card publishes base-model evaluation results; OLMES documents commands for running the task suites from the Olmo 3 technical report.16

What it is useful for

Ai2 documents base models as starting points for custom fine-tuning and for research on pretrained models; the model card publishes intermediate checkpoints from each training stage and fine-tuning instructions using OLMo-core.17

Run and use notes

Documented facts only. No hardware or performance claims are made without a cited source and stated assumptions.
  • The model card documents inference with Hugging Face Transformers version 4.57.0 or later, and optional 8-bit loading through bitsandbytes.1
  • Intermediate checkpoints are available as Hugging Face repository revisions named stage1-stepXXX (pretraining), stage2-stepXXX (mid-training), and stage3-stepXXX (long context).1

Organization context

Provenance and derivatives

Trained from scratch by Ai2. Per the model card, stage 1 used the Dolma 3 Mix (5.93T tokens), stage 2 used the Dolma 3 Dolmino Mix (100B tokens), and stage 3 used the Dolma 3 Longmino Mix (50B tokens).14

Catalog records that name this entry in their provenance:

Other releases in the OLMo family

OLMo family overview

U.S. eligibility

Project eligibility rests on documented governing or maintaining entities, not on contributors.

Eligible · basis: U.S. nonprofit or lab

The model card states the model was developed by the Allen Institute for AI (Ai2), which describes itself as a Seattle-based non-profit AI research institute.18

Assessed Sep 29, 2026

Sources

  1. 1.
  2. 2.
  3. 3.
  4. 4.
  5. 5.
    Olmo 3 (arXiv 2512.13961v2) (external site: arxiv.org)

    Team Olmo (Ai2) · Paper · published Apr 14, 2026 · accessed Sep 29, 2026

  6. 6.
    allenai/olmes (external site: github.com)

    Ai2 · Repository · accessed Sep 29, 2026

  7. 7.
    Olmo | Ai2 platform documentation (external site: docs.allenai.org)

    Ai2 · Documentation · accessed Sep 29, 2026

  8. 8.
    About us | Ai2 (external site: allenai.org)

    Ai2 · Official page · accessed Sep 29, 2026

  9. 9.
    Responsible use guidelines | Ai2 (external site: allenai.org)

    Ai2 · Official page · accessed Sep 29, 2026

  10. 10.
    Olmo 3: Charting a path through the model flow to lead open-source AI (external site: allenai.org)

    Ai2 · Announcement · published Nov 20, 2025 · accessed Sep 29, 2026

This listing is not an endorsement, a safety assessment, or a federal approval.

Support Us

Help keep USASI useful.

Optional. No USASI account required. Payment takes place on the linked provider’s website (Buy Me a Coffee).

About supporting this project