Superintelligence, give or take. The daily AI briefing: what the AI world actually said, sorted by how much it matters.

Wednesday, September 2, 2026

Coverage: 75 videos reviewed (1 partial or status unknown); videos under 45 s were not reviewed. Dates are UTC upload dates. Vendor claims are labelled as such.

New today

Alibaba updated Qwen 3.8 Max to the 0902 snapshot; open-weights status is disputed
Alibaba released a Qwen3.8-Max-0902 snapshot with claimed gains in coding and long agentic runs, per Fahd Mirza and Julian Goldie. Alibaba's own numbers, relayed by Goldie, put it behind Fable 5 and GPT-5.6 on several benchmarks (Humanity's Last Exam 43.6 vs 53.3 and 47.2; SWE-Bench Pro 67.6 vs 80 for Fable 5). Mirza fixed a planted sort-order bug with it in a Docker app and rated its multilingual answers well. The two channels disagree on weights availability.

Google released Gemini 3.8 Flash, priced at $0.75 input; benchmark claims mixed
Google released Gemini 3.8 Flash, its third Flash release in six weeks, per Fahd Mirza and Prompt Engineering. Mirza cited vendor charts showing a $0.75 input price and leading results on financial-analysis and Harvey legal benchmarks. Prompt Engineering said it is roughly level with Opus 5 on Deep Sweep 1.1 but Opus is more than 2.5 times better on the new Terminal Bench. Artificial Analysis, as cited, measured up to 300 tokens per second and up to 30% more output tokens per task than the prior Flash.

OpenAI revealed Jalapeno inference chip, claiming wins over Nvidia GB200 and GB300 per kilowatt
Nate B Jones said OpenAI reported Jalapeno beat GB200 and GB300 systems on latency and throughput per kilowatt across three open-weight model tests, and that AI-written design code ran 1.5 to 1.8 times faster than human-expert versions. It is an inference chip only, and OpenAI still has about 12 GW of Nvidia systems. The results are OpenAI claims.

Continuing stories

Also notable

Models & learning