AI digest
The useful part of the release cycle, when it matters
A short editorial update for people who need to know what changed without refreshing every model page. New launches, open-weight notes, benchmark context, and the decision that follows.
Signal as of 12 August 2026
New does not mean ready to switch.
The digest tracks the release, then points you to the test that would prove whether it belongs in your workflow.
Subscribe to release RSSNewsletter
Save your email for digest updates
Optional updates when a release looks worth your time. Automated email delivery is still being rolled out — we store your preference now and you can use RSS in the meantime.
Updates only — checklists stay free from the resource library, with or without joining. Automated email delivery is still rolling out.
Current window
8 verified releases in the last 28 days
Showing 6 here; the live tracker is the source of truth for dates, prices, context, and open-weight status.
Z.ai
GLM 5.3
Z.ai's 2026-08-14 Coding Plan flagship — same base as GLM 5.2, with the documented gains from post-training only. Direct token API and open weights are not published yet.
Qwen
Qwen3.8-27B
Apache 2.0 Qwen3.8 dense VLM for local and self-hosted work — native 262K context, image and video input, thinking on by default but can be turned off. Distinct from hosted Qwen3.8-Max.
Gemini 3.7 Flash
Google's 2026-08-13 Flash workhorse for coding and agents, three weeks after 3.6 Flash, at an introductory $0.75 / $3.75 per million tokens.
SpaceXAI
Grok 4.6
SpaceXAI's current code and chat default — same $2/$6 list as Grok 4.5, with a 500K context window and image input.
Qwen
Qwen3.8-2.4T-A95B
Official Hugging Face checkpoint for Qwen3.8 open weights — a text base model, not the hosted Qwen3.8-Max API.
Meta
Muse Glimmer 30B
Meta's open-weight multimodal agentic model distilled from Muse Spark for local consumer hardware (~24–32GB class with 4-bit).
Read next
Launch context, not launch noise
Use these guides to decide what deserves a controlled test.
Qwen
Qwen3.8-27B: Apache 2.0 dense VLM for local work
Qwen3.8-27B is a 27B dense vision-language checkpoint under Apache 2.0 — a practical local path that is not the hosted Qwen3.8-Max API.
Z.ai
GLM 5.3: Coding Plan flagship, API and weights still pending
GLM 5.3 is the current GLM Coding Plan default — same base as 5.2, with documented post-training gains. Direct token pricing and open weights are not published yet.
Gemini 3.7 Flash: the new coding and agent workhorse
Gemini 3.7 Flash arrives three weeks after 3.6 Flash as Google’s current Flash workhorse for coding and agents, with introductory API rates through the end of 2026.
Meta
Muse Spark 1.2: a model co-trained with its coding harness
Muse Spark 1.2 is Meta's coding-focused model update, trained with Muse Code to improve long-horizon repository work and tool use.