AI news, roughly super. A daily briefing from what the AI YouTube world actually said.

Wednesday, September 2, 2026

Coverage: 75 videos reviewed (1 partial or status unknown); videos under 45 s were not reviewed. Dates are UTC upload dates. Vendor claims are labelled as such.

New today

Alibaba releases Qwen 3.8 Max 0902 snapshot; weights reportedly coming open
Qwen3.8-Max-0902 is an upgraded snapshot with claimed gains in coding and long agentic runs. Julian Goldie reports 2.4T MoE (~95B active), 1M context, open weights plus a 27B on Hugging Face; Fahd Mirza says weights are only "coming soon". Alibaba numbers put it behind Fable 5 and GPT-5.6 on several benchmarks (HLE 43.6, SWE-Bench Pro 67.6). Mirza saw it fix a planted bug via Hermes.

Google releases Gemini 3.8 Flash at $0.75 input with strong claimed benchmarks
Third Flash release in six weeks. Google charts claim leadership on financial analysis, Harvey legal and expert-reasoning benchmarks; Prompt Engineering reports roughly Opus 5 parity on one benchmark but Opus over 2.5x better on Terminal Bench, up to 300 tok/s and up to 30% more output tokens per task (Artificial Analysis). Gemini 3.8 Flash Cyber limited to trusted partners.

OpenAI Astra persistent agents previewed to executives; cyber-capability and looped-transformer reports
Reportedly a few dozen executives saw Astra in August (16 agents on a math problem). The Information reportedly says it uses looped/recurrent-depth transformers, unconfirmed by OpenAI; Wes Roth says an OpenAI post suggests Astra may reach critical cyber capability with safeguards first. Goldie relays leaders claiming near-AGI and an automated research intern benchmark.

Continuing stories

Also notable

Models & learning