Skip to main content
AI Choice EngineAI Choice Engine
Back to blog

Framework

AI API pricing snapshot — September 2026: verified list prices across the frontier

A buyer-ready snapshot of verified AI API list prices as of 2026-09-05, now including GPT-6 Astra, Gemini 3.8 Flash and Muse Spark 1.3. Every figure is sourced; none is a speculated cut or cliff.

Published August 28, 2026Updated September 5, 20265 min readAI ToolsBy AI Choice Engine Editorial

Verification note (accessed 2026-09-05): Every price, context-window and release-date figure below is verified against two independent sources listed in Sources and against the site's model catalog, where each row carries its own check date. The August edition of this snapshot has been superseded by the late-August/September release wave: GPT-6 Astra, Gemini 3.8 Flash, Muse Spark 1.3 and Claude Fable 5.1 all shipped inside one week.

If you only read one paragraph: as of 2026-09-05 the frontier API tiers sit at GPT-6 Astra $10/$50 (OpenAI's new computer-use flagship), GPT-5.6 Sol $4/$20, Claude Fable 5.1 / Claude Opus 5 $10/$50 and $5/$25, Gemini 3.8 Flash $0.75/$3.75 (intro), Muse Spark 1.3 $1.25/$4.25, Grok 4.6 $2/$6, with strong open-weight budget options from DeepSeek V4 Pro ($1.32/$3.96 peak), GLM 5.3 Flash ($0.15/$0.50 third-party) and Mistral Large 3 ($0.50/$1.50). None of this is a forecast — it is the live list, and every catalog row shows the date it was last checked.

Verified list prices (per 1M tokens, standard tier)

ProviderModelInput $/MTokOutput $/MTokContextNotes
OpenAIGPT-6 Astra10.0050.001.05MNew computer-use flagship (2026-09-03); >263K bills $20/$100; cached $1
OpenAIGPT-5.6 Sol4.0020.001.05MCut from $5/$30 launch list 2026-08-21, promo through 2026-11-21
OpenAIGPT-5.6 Terra2.0012.001.05MBalanced tier; cache $0.20/MTok
OpenAIGPT-5.6 Luna0.201.201.05MCheapest GPT-5.6 tier; cache $0.02/MTok
AnthropicClaude Fable 5.1 / Opus 510.00 / 5.0050.00 / 25.001MFable 5.1 (2026-09-01) priced with Fable 5; cache reads cut to $0.25/MTok
AnthropicClaude Sonnet 52.0010.001MBalanced tier
AnthropicClaude Haiku 4.51.005.00200KEarliest-possible retirement 2026-10-15 per Anthropic's deprecation floor
GoogleGemini 3.8 Flash0.753.751MNew workhorse Flash (2026-09-02); intro price through 2026-12-31, then $1.50/$7.50
GoogleGemini 3.7 Flash0.753.751MSame intro pricing as 3.8; Google keeps both available
GoogleGemini 3.1 Pro2.0012.001.05M>200K bills $4/$18; still Preview
MetaMuse Spark 1.31.254.251MNew (2026-09-02); cached $0.15; contributor tier $0.10/$0.20 trains on your data
xAIGrok 4.62.006.00500K<200K prompt; >200K bills $4/$12
DeepSeekV4 Pro1.323.961MMIT open weights; official peak rates, off-peak half price
DeepSeekV4 Flash0.441.321MMIT open weights; official peak rates
MistralLarge 30.501.50262KApache 2.0 open weights

Why this matters for buyers

  • The frontier repriced twice in five weeks. GPT-5.6 Sol's 2026-08-21 cut was followed by four new flagships in the 2026-09-01 → 09-03 window. Any forecast built on the August snapshot is already one generation behind — check the model table, where every row carries its check date.
  • One model no longer has one price. Surface, region, cache, and provider all change the number; the cheapest tier almost always carries conditions (Gemini 3.7/3.8 Flash intro pricing doubles on 2027-01-01 — verify the route before locking an annual commitment).
  • Route by workload. With GPT-5.6 Luna at $0.20 and Gemini 3.8 Flash at $0.75 intro, per-task model routing beats single-model standardization for high-volume work. The model picker applies this logic to your answers.
  • Open weights change the math. Mistral Large 3, DeepSeek V4 and GLM 5.3 weights are free to self-host, removing per-token API lock-in for capable tiers — the trade is serving ops.

Where the new releases land on price

  • GPT-6 Astra ($10/$50) is the most expensive OpenAI tier — a computer-use flagship, not a Luna replacement. Long-context work above 263K tokens bills $20/$100.
  • Gemini 3.8 Flash ($0.75/$3.75 intro) undercuts nothing in the table on capability per dollar if Google's Terminal-Bench numbers hold; the price doubles on 2027-01-01.
  • Muse Spark 1.3 ($1.25/$4.25) undercuts every closed frontier row — with two caveats worth reading: the headline benchmark used a reasoning mode that is not broadly available yet, and the cheaper "contributor" tier trains on your data.

Sources (accessed 2026-09-05 unless noted)

This post is research-derived and editorial. It does not replace a vendor quote or a committed-use contract. Verify live prices before purchasing or publishing a forecast.

Editorial note

AI Choice Engine publishes editorial guides to help readers understand fit, trade-offs, and next steps before choosing a tool or provider.

Newsletter

Get the buyer checklist that goes with this guide

Subscribe to download the matching checklist. Automated email delivery is still rolling out — this is not an inbox confirmation.

A practical scorecard for comparing fit, cost, rollout risk, support, and lock-in.

Updates only — checklists stay free from the resource library, with or without joining. Automated email delivery is still rolling out.

Buying guides

Guide pages connected to this article

These guides go one level deeper for readers who want a longer-form buying view before choosing a provider.

Next steps

Next step across the network

Continue with a focused hub page instead of restarting your research from scratch.