📍 Selection Hub · Updated 2026-09

How to Choose an AI Coding Model
in H2 2026: The Full Picture

Not "which is strongest" but "which fits your case". Three vendors shipped in the first three days of September 2026: Claude Fable 5.1 (09-01), Gemini 3.8 Flash (09-02) and GPT-6 Astra (09-03), which reshuffled the top tier. Here is a decision framework across budget, task complexity and context, a snapshot of all three camps, and a single entry point to every selection guide on this site.

Updated 2026-09-20

#Decision Framework#Three Camps#Navigation#2026-09
Current state: all three camps have shipping flagships, one is still rumored

Claude (Opus 5 / Fable 5.1) and OpenAI (GPT-6 Astra plus the three GPT-5.6 tiers) are live on QCode; on the open-weight side DeepSeek has moved to deepseek-v4.1-flash (2026-09-10) and deepseek-v4-pro (GA 08-13, peak/off-peak pricing), with GLM-5.3 (08-14) and Kimi K3 (07-16, weights 07-27) on sale. Google's directory now lists Gemini 3.8 Flash and 3.8 Live, while Gemini 4 Pro is only confirmed as starting pre-training on 2026-07-21 — no date has ever been given.

Landscape as of 2026-09-18.

Decision Framework: Three Axes

Answer these three questions before comparing benchmark scores — it's more effective.

Axis Ask yourself Lean toward Typical scenario Note
Budget sensitivity Is this workload large enough that cost control matters? Tight budget → Luna / deepseek-v4.1-flash / Sonnet 5; room to spend → Sol / Fable 5 Bulk completion, CI scripts, light classification Luna lists at about 1/25 of Sol; DeepSeek's V4.1-Flash runs $0.15/$0.6 off-peak and doubles at peak, so volume gaps stay wide
Task complexity Is this a long-horizon refactor or high-difficulty reasoning task, or routine feature work? High difficulty → Sol / Fable 5; routine → Terra / Sonnet 5 Cross-repo refactors, complex debugging, architecture decisions On the hardest slice of tasks, the flagship success-rate gap far outweighs the price gap
Context needs How much context does a single task need? Astra / Sol, Terra, Luna (1.05M) · Fable 5.1 (1M) · Gemini 3.8 Flash Long-document Q&A, large cross-file refactors Current flagship context windows across all three camps are similar — not the main differentiator

Three-Camp Snapshot

Claude (Anthropic)

Fable 5.1: the current top tier (released 2026-09-01, $10/$50, 1M context, 128K output, reliable knowledge cutoff June 2026, Thinking is Adaptive and always on). Opus 5: $5/$25, 1M context - Anthropic's own documentation recommends starting with Opus 5 for most workloads and reserving Fable 5.1 for demanding reasoning and long-horizon agentic work. Sonnet 5: $2/$10, the balanced workhorse. Haiku 4.5: $1/$5, the fastest tier.

OpenAI

GPT-6 Astra: the top tier released 2026-09-03 (model id gpt-6-astra, 1,050,000 context, 128,000 output cap including reasoning tokens, list price $10/$50, reasoning tiers low through max with no none). The three GPT-5.6 tiers: Sol (flagship, currently $4/$20 on the official pricing page, marked as promotional pricing available at least through 2026-11-21), Terra (balanced $2/$12) and Luna (fast $0.20/$1.20). Sol and Astra have identical context windows and output caps; the split is the reasoning tier and the knowledge cutoff. There is also the gated GPT-5.6-Cyber (Daybreak Red tier, not available through normal channels).

Google

Gemini 3.8 Flash and 3.8 Flash Cyber shipped on 2026-09-02 with a published entry price of $0.75 per 1M input and $3.75 per 1M output, stated as an introductory rate expiring 2026-12-31, after which — from 2027-01-01 — it becomes $1.50/$7.50; the Cyber build is limited to trusted defenders through the Fairwind programme. The earlier 3.6 Flash (07-21) and 3.7 Flash (08-13 GA) remain. Gemini 3.5 Pro is still not broadly available - the only checkable official wording is that it is testing with partners and will ship when ready, with no date ever given; the previously circulated July and 12 August dates both passed, so do not treat either as settled.

Open-weight camp (DeepSeek / GLM / Kimi)

DeepSeek's two live tiers are deepseek-v4.1-flash (2026-09-10, called by the release note “the smallest model in our new architecture family” with native visual understanding; 1M / 384K, MIT, official off-peak $0.15/$0.6) and deepseek-v4-pro (GA 2026-08-13, $0.435/$0.87 at launch). From 2026-09-10 deepseek-v4-flash and vision-exp were retired, with requests on those names served by V4.1-Flash at the Flash price (pricing footnote (1)); the changelog's 2026-09-10 entry withdrew the V4-Pro redirect, keeping V4 Pro with unchanged billing. GLM-5.3 (2026-08-14) and Kimi K3 (2026-07-16) remain. Compare on /deepseek-v4-1-flash-vs-v4-pro, prices on /deepseek-v4-1-flash-pricing.

GPT-5.4 Retirement Migration Guide

How the three GPT-5.6 tiers divide the work, and what is worth moving up to GPT-6 Astra.

Claude Fable 5 vs GPT-5.6 Sol

Specs, pricing, and SWE-Bench Pro scores for both flagships, side by side.

What Is the Relationship Between Claude Mythos 5 and Fable 5?

What everyday users can actually access is Fable 5 — Mythos 5 has no public commercial access.

Recent Timeline

Key milestones across all three camps over the past two months.

2026-07-09

GPT-5.6's three tiers fully launch

Sol, Terra, and Luna open to everyone via API and Codex; QCode adds support the same day.

2026-07-30

GPT-5.6 price cuts

Luna cut 80% to $0.20/$1.20, Terra cut 20% to $2/$12, Sol unchanged.

2026-08-14

August's packed release week

GLM-5.3 released (same base as 5.2, post-training scaling, coding +50%); Qwen3.8-Max launched 08-03 and open-sourced Max-tier weights for the first time on 08-12; GPT-5.6-Cyber entered the vetted Daybreak Red tier on 08-10.

2026-09-01 to 09-03

Three vendors in three days, top tier reshuffled

On 09-01 Anthropic released Claude Fable 5.1 and Mythos 5.1; on 09-02 Google shipped Gemini 3.8 Flash and 3.8 Flash Cyber; on 09-03 OpenAI released GPT-6 Astra. Astra and Fable 5.1 carry identical list prices ($10/$50) with near-identical context and output caps - the first same-price head-to-head across camps in the second half of 2026.

2026-09-10 ~ 09-14

DeepSeek changes gear: V4.1-Flash ships, the 0731 build retires

On 2026-09-10 DeepSeek released V4.1-Flash, cut the Flash price and retired deepseek-v4-flash and deepseek-v4-flash-vision-exp; the changelog's 2026-09-10 entry withdrew the V4 Pro redirect in response to demand. Around the same time the official API catalogue also listed Gemini 3.8 Live (id gemini-3.8-live).

About this page

This page is a decision framework and navigation hub; full first-party fact checks live on the per-model pages. Every September 2026 release date and flagship price here was rechecked against vendor documentation (OpenAI model docs and pricing, the Anthropic model overview, the Google model directory, and DeepSeek's news260910 note, changelog and pricing footnote (1)), captured 2026-09-18. Older entries keep the sources already recorded on the topic pages.

FAQ

What's the first thing to check when picking an AI coding model today?

Don't jump straight to benchmark scores. Answer three questions first: budget sensitivity, task complexity, and per-task context needs. Once those are answered, the candidate list is usually down to 1-2 tiers, and benchmark scores can settle the final call.

Do all GPT-5.4 users need to migrate?

No. The official line is that the 2026-08-31 cutoff covers ChatGPT-signed-in sessions only (rechecked 2026-08-30, still “starting August 31”). API or Codex-key access is outside this announcement; you can keep using it or move to the matching GPT-5.6 tier.

Is Gemini 3.5 Pro available yet?

Not yet. As of 2026-08-22 it's still in limited preview; both previously rumored launch dates (July and August 12) have slipped, and Google has confirmed no new date — treat it as a rumor, not a basis for decisions.

How is this page different from the 'AI Model Radar'?

The radar page is a snapshot of which models are available/preview/restricted. This page is a decision framework for which model fits your situation, plus a hub for this site's new and updated deep-dive content. They're complementary — read both.

Create a QCode account

Plans from ¥60 / $8.57. One key across Claude, GPT, and Gemini models.

Related Reading

Prices, benchmark scores, and launch dates on this page have been cross-checked; the Gemini 3.5 Pro launch date is an unconfirmed rumor. Information may change — always defer to each vendor's official announcements.