AI news, roughly super. A daily briefing from what the AI YouTube world actually said.

Wednesday, September 9, 2026

Coverage: 93 videos reviewed (0 partial or status unknown); videos under 45 s were not reviewed. Dates are UTC upload dates. Vendor claims are labelled as such.

New today

OpenAI says agents on an unreleased model produced a Navier-Stokes blow-up proof
OpenAI said in a blog post that a group of agents running on an unreleased next-generation model, described as significantly more capable than GPT-6 Astra, produced a proof of finite-time singularity formation for the forced 3D incompressible Navier-Stokes equations, a Clay Millennium Prize problem. Channels relaying the post reported that the run lasted 88 hours from Sept. 1 to Sept. 5, 2026, with 4.9 million agent messages and 300 billion output tokens; Wes Roth said 10,000 coordinating agents were involved. The proof has not been independently verified in any of these videos. OpenAI's statement, as read by Wes Roth, said the model has been trained since Aug. 28 and gave no name, benchmarks or release date.

OpenAI released GPT-6 Astra on Sept. 3, 2026; channels relay API specs and rollout to Pro
OpenAI released GPT-6 Astra on Sept. 3, 2026, according to Julian Goldie, who read an API page listing a 1,050,000-token context window, 128,000-token maximum output and an April 30, 2026 knowledge cutoff. Goldie also said Astra adds mid-turn steering and asynchronous tool calling, which he credited with part of a claimed 47% time reduction on simulated tasks. Fireship said in a video dated Sept. 9 that Astra rolled out to Pro subscribers the previous day and that Nvidia's Jensen Huang posted on X that AGI had arrived, noting Astra was trained on more than 100,000 Grace Blackwell GPUs with 400,000 more coming. Rollout tier and the Huang figures are relayed and unverified; no pricing was given.

Mathematicians and OpenAI dispute credit and data use in Navier-Stokes result
Mathematician Tristan Buckmaster, who had worked for a year with Codex, and OpenAI disagreed publicly over whether OpenAI's model drew on his work, according to channel readings of posts on X. OpenAI said its team and agents saw none of their work and no specific user data was accessed, but said it could not rule out de-identified usage data helping improve its models. Buckmaster and Levent Alpoge, who is an Anthropic employee acting personally, reported finite-time blowup results for related equations using Claude, Codex, a GPT-5.6 model and Astra. Matthew Berman and Wes Roth reported only one side's public posts; both drew opinion conclusions (Roth that OpenAI did nothing wrong, Berman that builders should assume vendors may learn from their data).

OpenAI-published Astra launch benchmarks include ARC-AGI-3 near 99%, relayed by three channels
Julian Goldie relayed OpenAI's own launch figures for GPT-6 Astra against GPT-5.6 Sol: OSWorld 2.0 72.6% versus 65.7%, Terminal Bench 4.0 57.9% versus 37.3%, Deep SWE 1.1 74.1% versus 72.7%, Automation Bench 41.4% versus 18.1% and ARC-AGI-3 99.9% versus 7.8%. A Mastra host read a chart showing ARC-AGI-3 at 98.6% for Astra, 7.8% for GPT-5.6 Sol and 30% for the prior best, Claude Opus 5. Fireship said a Berkeley team had reached 99% on ARC-AGI with Opus 4.8 and Fable 5 through a better harness, without naming the source or version. None of the channels reproduced the figures.

AI Advantage blind test of 50 one-shot sites: Astra preferred 35 to 15 over Fable 5.1
In a test of 50 one-shot website builds via API, one reviewer at The AI Advantage preferred GPT-6 Astra to Claude Fable 5.1 in 35 cases to 15, and Fable 5.1 to Fable 5 in 30 cases to 17 with 3 ties. AI judges on visuals picked Astra 47 times (3 ties) with an OpenAI-model judge and 48 times with Fable 5.1 as judge. AI-judged functionality passed 48 of 50 sites for Astra and 47 of 50 each for Fable 5.1 and Fable 5. API cost for the 50 sites was $20.64 for Astra, $29.15 for Fable 5.1 and $21.24 for Fable 5, with no caching or batch. The test is one rater, one run, and websites only; per-category samples were about five sites.

Anthropic released Claude Fable 5.1 on Sept. 1, 2026, per channels relaying its materials
Anthropic released Claude Fable 5.1 on Sept. 1, 2026, according to Julian Goldie and Mastra hosts, who described a coding and knowledge-work model with an always-on thinking mode, effort levels from low to max, and a 1 million-token context. Goldie said the API name is Claude-Fable-51 and that Mythos 5.1 is the same model with fewer guardrails. The AI Advantage said Anthropic released it to get ahead of Astra; Mastra hosts said it trails Astra on most benchmarks shown. Riley Brown called it the best coding model as of Sept. 3, without benchmarks. All are relayed accounts.

Continuing stories

Also notable

Models & learning