Independent project. Not a U.S. government website.

USASI
Model family

Microsoft Florence-2

Family overview — summarizes releases; licenses and availability belong to each release.

Maintained by Microsoft (Azure AI)13

Florence-2 is a family of vision foundation models from Microsoft's Azure AI group that handle captioning, object detection, visual grounding, segmentation, and OCR through text prompts, using a sequence-to-sequence design with a DaViT vision encoder. The models were trained on FLD-5B, a set of 5.4 billion visual annotations on 126 million images that the team built with an iterative automated annotation process. Microsoft published four checkpoints on Hugging Face in June 2024: Florence-2-base (0.23B parameters) and Florence-2-large (0.77B), plus versions of each fine-tuned on a collection of downstream tasks (base-ft and large-ft).1325Fact reviewed Oct 1, 2026

Last reviewedEntry updated Documented release Jun 2024

Releases assessed

Each release is assessed on its own. This family has no family-wide openness label.
Releases in the Microsoft Florence-2 family
ReleaseWeightsTier (USASI rubric v0.2)LicenseReleased
Florence-2-largeWeights: PublicModel-disclosure tier (USASI rubric v0.2): Open-weightMITJun 2024

What it is useful for

The model card shows prompt-selected tasks including captioning at several levels of detail, object detection, dense region captioning, region proposals, caption-to-phrase grounding, OCR, and OCR with regions.3Fact reviewed Oct 1, 2026

Organization context

U.S. eligibility

Project eligibility rests on documented governing or maintaining entities, not on contributors.

Eligible · basis: U.S. headquarters

The technical report lists its authors under Azure AI, Microsoft, and the checkpoints are published in Microsoft's Hugging Face organization with an MIT license naming Microsoft Corporation. Microsoft Corporation lists One Microsoft Way, Redmond, Washington as its principal executive offices on the cover page of its Form 10-K for fiscal year 2026.1349

Assessed Oct 1, 2026

Sources

  1. 1.
    Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks (external site: arxiv.org)

    Microsoft (arXiv) · Paper · published Nov 10, 2023 · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  2. 2.
    Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks - Microsoft Research (external site: microsoft.com)

    Microsoft Research · Official page · published Jun 2024 · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  3. 3.
    microsoft/Florence-2-large model card (external site: huggingface.co)

    Microsoft (Hugging Face) · Model card · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  4. 4.
    microsoft/Florence-2-large LICENSE (MIT) (external site: huggingface.co)

    Microsoft (Hugging Face) · License · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  5. 5.
    Hugging Face model listing for microsoft Florence models (API) (external site: huggingface.co)

    Hugging Face · Other · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  6. 6.
    Hugging Face model metadata for microsoft/Florence-2-base (external site: huggingface.co)

    Hugging Face · Other · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  7. 7.
    Hugging Face model metadata for microsoft/Florence-2-base-ft (external site: huggingface.co)

    Hugging Face · Other · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  8. 8.
    Hugging Face model metadata for microsoft/Florence-2-large-ft (external site: huggingface.co)

    Hugging Face · Other · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

  9. 9.
    Microsoft Corporation Form 10-K for the fiscal year ended June 30, 2026 (external site: sec.gov)

    Microsoft Corporation (U.S. SEC filing) · Filing · accessed Oct 1, 2026 · evidence reviewed Oct 1, 2026

Support Us

Help keep USASI useful.

Optional. No USASI account required. Payment takes place on the linked provider’s website (Buy Me a Coffee).

About supporting this project