Model decision table

Rankings with the tradeoffs left in.

Scan the full catalog, narrow it to your workload, and compare up to three models. Specifications are source-reviewed; fit scores are editorial and their weighting is explicit.

Reviewed 2026-09-2038 models11 providers
Full catalog

Compare models

Showing 38 models

Compare#ModelEditorial fitCodingReasoningTool useContextPrice / 1M tokensEvidence
01
GPT-6 AstraOpenAI · AvailableHard end-to-end agent work · Computer use
10.0editorial10.010.010.01.05M$10→$50Tier caveat
02
Claude Fable 5.1Anthropic · AvailableDemanding reasoning · Long-horizon agents
9.9editorial9.910.09.91M$10→$50Tier caveat
03
GPT-5.6 SolOpenAI · AvailableComplex production workflows · Coding
9.9editorial9.99.99.91.05M$4→$20Tier caveat
04
Claude Fable 5Anthropic · AvailableLong-running agents · Highest-capability tasks
9.9editorial9.99.99.81M$10→$50
05
Claude Opus 5Anthropic · AvailableComplex reasoning · Long-running agents
9.9editorial9.99.99.81M$5→$25
06
GPT-5.5OpenAI · AvailableComplex reasoning · Coding
9.8editorial9.89.89.71M$5→$30
07
Grok 4.6xAI · AvailableAgentic coding · Knowledge work
9.8editorial9.89.89.7500K$2→$6Tier caveat
08
Claude Opus 4.8Anthropic · AvailableComplex reasoning · Agentic coding
9.8editorial9.89.89.61M$5→$25
09
Grok 4.5xAI · AvailableAgentic coding · Knowledge work
9.7editorial9.79.79.6500K$2→$6Tier caveat
10
GPT-5.4OpenAI · AvailableCoding · Agents
9.7editorial9.89.59.71M$2.5→$15
11
Gemini 3.8 FlashGoogle · AvailableLong-horizon software engineering · Autonomous agents
9.7editorial9.79.69.71.05M$0.75→$3.75Tier caveat
12
GPT-5.6 TerraOpenAI · AvailableProduction agents · Coding
9.6editorial9.69.69.71.05M$2→$12Tier caveat
13
Gemini 3.7 FlashGoogle · AvailableAgentic coding · Multimodal reasoning
9.6editorial9.69.59.61.05M$0.75→$3.75Tier caveat
14
Qwen3.8-MaxAlibaba / Qwen · AvailableLong-horizon coding · Complex multimodal agents
9.5editorial9.59.59.51MN/ATier caveat
15
Claude Sonnet 5Anthropic · AvailableBalanced performance · Production workloads
9.5editorial9.69.59.31M$2→$10Tier caveat
16
GPT-5.2-CodexOpenAI · AvailableCoding-focused tasks · Type inference
9.5editorial9.79.39.4400K$1.75→$14
17
Gemini 3.6 FlashGoogle · AvailableAgentic coding · Multimodal tasks
9.5editorial9.59.49.51M$1.5→$7.5Tier caveat
18
Gemini 3.5 FlashGoogle · AvailableFast multimodal agents · Search grounding
9.4editorial9.49.49.41M$1.5→$9Tier caveat
19
Kimi K3Moonshot AI · AvailableLong-horizon coding · Knowledge work
9.3editorial9.49.39.21M$3→$15Tier caveat
20
GLM-5.3Z.ai · AvailableComplex software engineering · Long-context analysis
9.3editorial9.39.39.21M$1.4→$4.4Tier caveat
21
GPT-5.2OpenAI · AvailableGeneral-purpose · Balanced tasks
9.2editorial9.39.29.0400K$1.75→$14
22
GPT-OSS-120BOpenAI · AvailableSelf-hosted · Privacy
9.2editorial9.39.29.0128K$0→$0Tier caveat
23
GLM-5Zhipu AI · AvailableBilingual (CN/EN) · Value-focused
9.2editorial9.29.39.0200K$1→$3.2
24
GPT-5.6 LunaOpenAI · AvailableHigh-volume workflows · Subagents
9.1editorial9.19.19.31.05M$0.2→$1.2Tier caveat
25
GLM-5.3-FlashZ.ai · AvailableCost-sensitive multimodal coding · Long-context agents
9.1editorial9.29.19.11M$0.15→$0.5Tier caveat
26
DeepSeek V4.1 FlashDeepSeek · AvailableLow-cost multimodal agents · High-throughput coding
9.1editorial9.29.19.11M$0.15→$0.6Tier caveat
27
Kimi K2.7 CodeMoonshot AI · AvailableLong-horizon software engineering · Coding agents
9.1editorial9.29.09.2262K$0.95→$4Tier caveat
28
DeepSeek V4 ProDeepSeek · AvailableDeepSeek agent workloads · Long-context text workflows
9.1editorial9.19.28.91M$0.66→$1.98Tier caveat
29
Mistral Medium 3.5Mistral · AvailableEuropean compliance · Agentic coding
9.0editorial9.19.18.9256K$1.5→$7.5
30
Qwen3.8-FlashAlibaba / Qwen · AvailableHigh-volume multimodal agents · Long-context coding
9.0editorial9.08.99.01M$0.15→$0.47Tier caveat
31
Qwen3.8-Flash-NextAlibaba / Qwen · AvailableSelf-hosted multimodal agents · Cost-efficient long-context inference
8.9editorial8.98.88.9256KN/ATier caveat
32
Claude Haiku 4.5Anthropic · AvailableFast responses · High-volume tasks
8.8editorial8.98.88.7200K$1→$5
33
Gemini 3.5 Flash-LiteGoogle · AvailableHigh-volume automation · Subagents
8.8editorial8.88.78.91M$0.3→$2.5Tier caveat
34
GPT-Live-1OpenAI · AvailableFull-duplex voice agents · Telephony
1.0editorial1.01.01.0Not statedN/ATier caveat
35
North Small TranslateCohere · Research/non-commercialMachine translation · Private translation deployments
1.0editorial1.01.01.016KN/ATier caveat
—
Gemini 3.8 LiveGoogle · AvailableLow-latency voice agents · Real-time dialogue
N/AN/AN/AN/A131K$0.75→$4.5Tier caveat
—
Gemini 3.8 Live Extended ThinkingGoogle · AvailableComplex voice agents · Multi-step voice workflows
N/AN/AN/AN/A131K$0.75→$4.5Tier caveat
—
Qwen3.8-Omni-FlashAlibaba / Qwen · AvailableAudio and video understanding · Multimodal content analysis
N/AN/AN/AN/A1MN/ATier caveat
Score methodology

Orientation, not manufactured certainty.

The editorial fit score weights coding at 40%, reasoning at 35%, and tool use at 25%. It is a compact decision aid built from the dimensions tracked in this catalog—not a claim that one model is universally best. Specialized models that do not map to those dimensions are listed but marked N/A and excluded from editorial ranking.

Prices show input then output cost per one million tokens. Provider tiers, caching, batches, and hosts can change the actual bill. Open the evidence links and caveat labels before committing to a production choice.

Inspect the source log and review policy →