Independent project. Not a U.S. government website.

USASI
Model family

gpt-oss-safeguard

Family overview — summarizes releases; licenses and availability belong to each release.

Maintained by OpenAI12

gpt-oss-safeguard is a pair of open-weight, text-only reasoning models from OpenAI, gpt-oss-safeguard-120b and gpt-oss-safeguard-20b, post-trained from the corresponding gpt-oss models to classify content against a policy that the user supplies. OpenAI's technical report is dated October 29, 2025. OpenAI lists the models as a model partner of the Robust Open Online Safety Tools (ROOST) Model Community, a group of safety practitioners working with open models.1234Fact reviewed Oct 1, 2026

Last reviewedEntry updated Documented release Oct 2025

Releases assessed

Each release is assessed on its own. This family has no family-wide openness label.
Releases in the gpt-oss-safeguard family
ReleaseWeightsTier (USASI rubric v0.2)LicenseReleased
gpt-oss-safeguard-120bWeights: PublicModel-disclosure tier (USASI rubric v0.2): Open-weightApache-2.0, Apache-2.0Oct 2025
gpt-oss-safeguard-20bWeights: PublicModel-disclosure tier (USASI rubric v0.2): Open-weightApache-2.0, Apache-2.0Oct 2025

What it is useful for

The model cards list safety uses such as filtering LLM inputs and outputs, online content labeling, and offline labeling for trust-and-safety work. OpenAI recommends the base gpt-oss models for other applications, including end-user chat.31Fact reviewed Oct 1, 2026

Organization context

Provenance and derivatives

Both models are fine-tunes of OpenAI's gpt-oss models. The technical report says they were trained without any additional biological or cybersecurity data.134

  • Derived from: gpt-oss — Base model family (gpt-oss-120b and gpt-oss-20b).

U.S. eligibility

Project eligibility rests on documented governing or maintaining entities, not on contributors.

Eligible · basis: U.S. headquarters

Developed and published by OpenAI. Its operating company, OpenAI Group PBC, is described as a Delaware public benefit corporation at 1455 3rd Street, San Francisco, California, in a February 2026 agreement filed with the SEC.17

Assessed Oct 1, 2026

Sources

  1. 1.
    Technical Report: Performance and baseline evaluations of gpt-oss-safeguard-120b and gpt-oss-safeguard-20b (external site: cdn.openai.com)

    OpenAI · Paper · published Oct 29, 2025 · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  2. 2.
    openai/gpt-oss-safeguard README (external site: raw.githubusercontent.com)

    OpenAI · Repository · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  3. 3.
    openai/gpt-oss-safeguard-120b model card (external site: huggingface.co)

    OpenAI (Hugging Face) · Model card · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  4. 4.
    openai/gpt-oss-safeguard-20b model card (external site: huggingface.co)

    OpenAI (Hugging Face) · Model card · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  5. 5.
    Hugging Face model metadata for openai/gpt-oss-safeguard-120b (external site: huggingface.co)

    Hugging Face · Repository · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  6. 6.
    Hugging Face model metadata for openai/gpt-oss-safeguard-20b (external site: huggingface.co)

    Hugging Face · Repository · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  7. 7.
    Exhibit 10.1: Equity commitment letter agreement between OpenAI Group PBC and Amazon (external site: sec.gov)

    U.S. Securities and Exchange Commission (Amazon.com, Inc. filing) · Filing · published Feb 27, 2026 · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

Support Us

Help keep USASI useful.

Optional. No USASI account required. Payment takes place on the linked provider’s website (Buy Me a Coffee).

About supporting this project