Olmo 3 7B Instruct
Release in the OLMo family · version Olmo-3-7B-Instruct
Maintained by Ai2 (Allen Institute for AI)1
A 7B chat model in Ai2's Olmo 3 release, post-trained from the Olmo 3 7B base model with supervised fine-tuning, direct preference optimization, and reinforcement learning with verifiable rewards on the Dolci datasets. It responds without a separate reasoning trace, unlike the Olmo 3 Think models.1811
- Model hub: Model card (Hugging Face) (external site: huggingface.co)
- Repository: open-instruct Olmo 3 post-training scripts (external site: github.com)
- Paper: Olmo 3 technical report (arXiv 2512.13961) (external site: arxiv.org)
- Documentation: Olmo developer documentation (external site: docs.allenai.org)
Availability and license
Overall availability
Downloadable from Hugging Face without gating. The card says the model is intended for research and educational use in line with Ai2's Responsible Use Guidelines.1
Availability is separate from permission: read the license before using or redistributing.
Apache License 2.0 (open-instruct) (external site: raw.githubusercontent.com)3
Apache License 2.0 (OLMo-core) (external site: raw.githubusercontent.com)5
The model card states the model is licensed under Apache 2.0 and intended for research and educational use in accordance with Ai2's Responsible Use Guidelines, which list categories of prohibited use.113
Model-disclosure tier
Open-weight, plus published inference code, training code, and training recipe, and at least documented training-data composition.
Public materials checklist
| Item | Status | Notes and evidence |
|---|---|---|
| WeightsCan the general public download the model parameters for this release? | Public | Final weights and post-training checkpoints (step revisions) are published on Hugging Face.1 |
| Inference codeIs code for running the model published? | Public | The model card documents inference with Hugging Face Transformers and vLLM; Ai2's documentation adds a vLLM function-calling example.110 |
| Training codeIs the code used to train the model published? | Public | open-instruct publishes the Olmo 3 7B Instruct SFT, DPO, and RL scripts; its README states that SFT runs on OLMo-core and DPO and RL on open-instruct. Base-model training scripts are in OLMo-core.24 |
| Training-data informationPublic = the training data itself can be obtained. Partial = composition or sources are documented without full access. | Partial | The Dolci post-training datasets are downloadable under ODC-BY, but the Dolma 3 Mix used to pretrain the underlying base model has some olmOCR science PDFs redacted after training, which its dataset card says affects reproducibility.671 |
| Training recipeAre the training configuration and procedure documented in enough detail to follow? | Public | The technical report describes the Instruct post-training pipeline, and the open-instruct README lists the scripts, commits, and experiment logs used for each stage.82 |
| Evaluation materialsPublic = evaluation code or prompts that let others re-run the evaluations are published. Partial = results only. | Public | The model card publishes post-training evaluation results; OLMES documents commands for running the Olmo 3 evaluation suites.19 |
What it is useful for
Run and use notes
- The model card documents inference with Hugging Face Transformers 4.57.0 or later and with vLLM, and recommends temperature 0.6 and top_p 0.95.1
- Ai2's documentation gives a vLLM (0.11.1 or later) serving command with automatic tool choice and the olmo3 tool-call parser, which exposes an OpenAI-compatible endpoint.10
Organization context
Provenance and derivatives
Fine-tuned by Ai2 from its own Olmo 3 7B base model (Olmo-3-1025-7B) through the Olmo-3-7B-Instruct-SFT and Olmo-3-7B-Instruct-DPO checkpoints, followed by RLVR on the Dolci Instruct RL data.16
- Derived from: Olmo 3 7B (base) — Base model.
- Derived from: Dolci Instruct SFT (external site: huggingface.co) — Supervised fine-tuning mixture.
Catalog records that name this entry in their provenance:
Other releases in the OLMo family
- Olmo 3 7B (base)Model-disclosure tier (USASI rubric v0.1): Open-stack
- Olmo 3.1 32B ThinkModel-disclosure tier (USASI rubric v0.1): Fully open
U.S. eligibility
Sources
This listing is not an endorsement, a safety assessment, or a federal approval.