<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Speech, vision, and multimodal AI — USASI news</title>
    <link>https://unitedstatesofamericasuperintelligence.com/hubs/speech-vision-and-multimodal/</link>
    <atom:link href="https://unitedstatesofamericasuperintelligence.com/hubs/speech-vision-and-multimodal/feed.xml" rel="self" type="application/rss+xml"/>
    <description>Dated, sourced news items about the organizations and records featured in the USASI Speech, vision, and multimodal AI hub. Independent project. Not a United States government website.</description>
    <language>en-us</language>
    <lastBuildDate>Sun, 11 Oct 2026 12:00:00 GMT</lastBuildDate>
    <item>
      <title>White House science summit lists industry partners in the Genesis Mission Consortium</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/white-house-science-summit-genesis-consortium-partners/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/white-house-science-summit-genesis-consortium-partners/</guid>
      <pubDate>Sun, 11 Oct 2026 12:00:00 GMT</pubDate>
      <category>policy</category>
      <description>On October 8, 2026 the White House published a fact sheet on science initiatives announced at its &quot;Science: A New Golden Age&quot; summit. It says eleven industry partners committed SI-for-science tools and compute credits to the Genesis Mission Consortium, and names NVIDIA, AMD, OpenAI, Anthropic, Google, AMP, Emerald AI, AWS, Armada, Crusoe, and Micron. The same fact sheet says NIH, DOE, and Biohub launched a virtual biology initiative to build data and models for virtual cells. The fact sheet describes announcements and commitments rather than an executive order.</description>
    </item>
    <item>
      <title>OpenAI releases Codex CLI 0.162.0 with managed Git worktree tools</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/codex-cli-0-162-0-released/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/codex-cli-0-162-0-released/</guid>
      <pubDate>Fri, 09 Oct 2026 12:00:00 GMT</pubDate>
      <category>release</category>
      <description>On October 8, 2026 OpenAI published version 0.162.0 of its open-source Codex CLI on GitHub. The release notes add tools for creating and listing managed Git worktrees from trusted local projects when the worktrees feature is enabled, and add task pinning in the agent Command Center with the `p` key. They also say live web access and remote compaction can be configured for custom Responses-compatible model providers, and that a signed PowerShell installer is now published with Windows releases. USASI has not tested these changes.</description>
    </item>
    <item>
      <title>OpenAI releases Codex CLI 0.161.0 with GPT-6.1 Sol as the default model</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/codex-cli-0-161-0-released/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/codex-cli-0-161-0-released/</guid>
      <pubDate>Thu, 08 Oct 2026 12:00:00 GMT</pubDate>
      <category>release</category>
      <description>On October 7, 2026 OpenAI published version 0.161.0 of its open-source Codex CLI on GitHub. The release notes say GPT-6.1 Sol is now the default model in the bundled and Amazon Bedrock catalogs, and that Amazon Bedrock supports multi-agent V2 and Ultra reasoning on compatible models. The notes also add a `/mcp login &lt;name&gt;` command for signing in to MCP servers from an active terminal session, and make Daybreak opt-in through `--enable cli_daybreak`. USASI has not tested these changes.</description>
    </item>
    <item>
      <title>Diffusers 0.41.0 adds a Qwen-Image 2.1 pipeline and deprecates ONNX support</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/diffusers-0-41-0-released/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/diffusers-0-41-0-released/</guid>
      <pubDate>Thu, 08 Oct 2026 12:00:00 GMT</pubDate>
      <category>release</category>
      <description>On October 6, 2026 Hugging Face published Diffusers 0.41.0 on GitHub. The release notes say it adds Qwen-Image 2.1, with text-to-image generation, image editing, native transparency, and LoRA training, and introduces tensor-parallel checkpoint loading in which each rank reads its own slice of sharded weights. The notes deprecate ONNX support in favor of Optimum and remove previously deprecated APIs, including the `LuminaText2ImgPipeline` and `Lumina2Text2ImgPipeline` aliases. They also say minor releases will now be coordinated around new model integrations, as in Transformers. Qwen-Image 2.1 is a third-party model and is not a catalog entry; the notes do not state its license.</description>
    </item>
    <item>
      <title>Ai2 open-sources AstaBrief 8B, a model for writing cited scientific reports</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/ai2-open-sources-astabrief-8b/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/ai2-open-sources-astabrief-8b/</guid>
      <pubDate>Sun, 04 Oct 2026 12:00:00 GMT</pubDate>
      <category>release</category>
      <description>On October 2, 2026 Ai2 announced that it is open-sourcing AstaBrief 8B, a model that turns a research question and retrieved literature excerpts into a cited report, along with its training data. Ai2 says the model was built from Qwen3-8B using supervised fine-tuning and direct preference optimization, and that it is available as Fast mode in the &quot;Generate a report&quot; feature of Asta, its platform for scientific work. The Hugging Face model card lists the Apache 2.0 license. Ai2 notes that most of the training and evaluation was completed in 2025 and that it has not rerun the full evaluation against today's frontier models.</description>
    </item>
    <item>
      <title>Ai2 releases OLMo-core 3, its open training library, with mixture-of-experts support</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/ai2-releases-olmo-core-3/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/ai2-releases-olmo-core-3/</guid>
      <pubDate>Thu, 01 Oct 2026 12:00:00 GMT</pubDate>
      <category>release</category>
      <description>Ai2 released version 3 of OLMo-core, the PyTorch training library behind its Olmo models, on October 1, 2026. The release redesigns the library's mixture-of-experts training system, adding expert and pipeline parallelism, a distributed optimizer, grouped matrix multiplication for experts, and support for the MXFP8 number format; Ai2 says it is designed to scale mixture-of-experts training into the trillion-parameter range. The code is licensed under Apache 2.0, and the repository includes Ai2's official training scripts for Olmo 3.</description>
    </item>
    <item>
      <title>NVIDIA releases Kumo Tabular, an open-weight model for predictions on tables</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/nvidia-releases-kumo-tabular/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/nvidia-releases-kumo-tabular/</guid>
      <pubDate>Thu, 01 Oct 2026 12:00:00 GMT</pubDate>
      <category>release</category>
      <description>NVIDIA published Kumo Tabular on September 29, 2026: a pretrained model for classification and regression on tabular data that takes labeled example rows as context and predicts values for new rows without task-specific training. It comes in three sizes, from 28 million to 215 million parameters, and NVIDIA says it was pretrained entirely on artificial tables. The weights are on Hugging Face under the OpenMDW-1.1 license, and inference runs through NVIDIA's open-source structured-data-models library.</description>
    </item>
    <item>
      <title>OpenAI adds GPT-6.1 Sol to its API</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/openai-releases-gpt-6-1-sol-in-api/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/openai-releases-gpt-6-1-sol-in-api/</guid>
      <pubDate>Thu, 01 Oct 2026 12:00:00 GMT</pubDate>
      <category>release</category>
      <description>OpenAI's API changelog lists GPT-6.1 Sol (model ID gpt-6.1-sol) as released on September 29, 2026, describing it as a model for complex coding and professional work at a lower cost than GPT-6 Astra. The changelog says it is available through the Responses and Chat Completions endpoints. Separate changelog entries on the same date add computer use to the Agents API and an Ultrafast mode for GPT-6 Astra in the Responses API. GPT-6.1 Sol is a hosted model; OpenAI has not published its weights.</description>
    </item>
    <item>
      <title>NVIDIA releases Nemotron 3 Diarization, an open-weight speaker diarization model</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/nvidia-releases-nemotron-3-diarization/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/nvidia-releases-nemotron-3-diarization/</guid>
      <pubDate>Thu, 01 Oct 2026 12:00:00 GMT</pubDate>
      <category>release</category>
      <description>NVIDIA released Nemotron 3 Diarization on September 23, 2026. The model identifies which speaker is talking at each point in an audio recording, for up to eight speakers, and can run on streaming audio as well as complete recordings. It is a Transformer encoder with about 100 million parameters, built with NVIDIA's NeMo speech tools. The weights are on Hugging Face under the OpenMDW License Agreement, version 1.1, and the model card says they are ready for commercial and non-commercial use.</description>
    </item>
    <item>
      <title>NVIDIA agrees to acquire Hugging Face; the deal has not yet closed</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/nvidia-agrees-to-acquire-hugging-face/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/nvidia-agrees-to-acquire-hugging-face/</guid>
      <pubDate>Tue, 29 Sep 2026 12:00:00 GMT</pubDate>
      <category>acquisition</category>
      <description>NVIDIA entered into a definitive agreement to acquire Hugging Face, Inc. on September 2, 2026, and announced it the next day. NVIDIA's Form 8-K says the transaction is expected to close in the first half of 2027, subject to closing conditions including regulatory approvals, so Hugging Face is still listed as a privately held company with no parent organization. NVIDIA says the Hugging Face platform will stay open to models from other developers and will not require NVIDIA hardware.</description>
    </item>
    <item>
      <title>Meta releases Muse Glimmer as an open-weight model under Apache 2.0</title>
      <link>https://unitedstatesofamericasuperintelligence.com/news/meta-releases-muse-glimmer-open-weights/</link>
      <guid isPermaLink="true">https://unitedstatesofamericasuperintelligence.com/news/meta-releases-muse-glimmer-open-weights/</guid>
      <pubDate>Tue, 29 Sep 2026 12:00:00 GMT</pubDate>
      <category>release</category>
      <description>In August 2026 Meta Superintelligence Labs published Muse Glimmer, a model of about 30 billion parameters that takes text and images as input and was distilled from Muse Spark. Its weights can be downloaded from Hugging Face without an access request under the Apache 2.0 license, and the repository also includes a separate Muse Glimmer Usage Policy that lists prohibited uses. Muse Glimmer has its own release record under the Muse family.</description>
    </item>
  </channel>
</rss>
