ApertureAug 2026
FieldCompareLabsGuideMethod
  • Field
  • Compare
  • Labs
  • Guide

Field guide · 28 Aug 2026

Which model is actually useful — not just first on a leaderboard.

Aperture ranks 34 current and recent models from 12 labs on intelligence, coding, agents, price, and speed. Filter the field, sort by the job, compare up to four, and leave with a pick you can defend.

Tracked
34
Current
24
Open weight
12

If you only need one

Why these

Frontier

A

Claude Opus 5

Anthropic

Highest measured intelligence that you can actually buy.

63.1 · $5 / $25

Coding

A

Claude Fable 5

Anthropic

Repo work, SWE-bench Pro, Terminal-Bench.

62.1 · $10 / $50

Agents

Z

GLM-5.3

Z.ai (Zhipu)

Tool loops, computer use, long jobs.

59.5 · $1.40 / $4.40

Value

Z

GLM-5.3 Flash

Z.ai (Zhipu)

At least 55 intelligence, best points per dollar.

57.5 · $0.15 / $0.50

Open weight

K

Kimi K3

Moonshot AI

Highest open model with real evidence.

59.7 · $3 / $15

Speed

G

Gemini 3.7 Flash

Google DeepMind

Fastest model that still clears 50 intelligence.

56 · $0.75 / $3.75

Multimodal

G

Gemini 3.7 Flash

Google DeepMind

Vision plus video, ranked by intelligence.

56 · $0.75 / $3.75

Volume

Z

GLM-5.3 Flash

Z.ai (Zhipu)

Best model at or under $0.30 / 1M input.

57.5 · $0.15 / $0.50

The field

Intelligence Index is the Artificial Analysis composite (approx. 0–65). Cards are the default on a phone; switch to table or the price map on a larger screen.

25 models

01
A

Claude Mythos 5

Anthropic · Jul 2026 · Restricted

63.4
Index
Intel63.4
Code78
Agents58

Anthropic’s limited-access peak. Tops composite boards, but you cannot buy it on the public API. Treat Fable 5 as the available twin.

Price /1M
$10 / $50
Context
1M
Speed
—
Status
Restricted
Eval researchInternal Anthropic-class workloads
Details
02
A

Claude Opus 5

Anthropic · Jul 2026 · Proprietary

63.1
Index
Intel63.1
Code78
Agents59.2

Highest generally-available Intelligence Index. Half of Fable’s price, first-choice for agents and knowledge work.

Price /1M
$5 / $25
Context
1M
Speed
52 t/s
Status
Current
AgentsHard debugging
Details
03
A

Claude Fable 5

Anthropic · Jun 2026 · Proprietary

62.1
Index
Intel62.1
Code76.5
Agents56.6

The available Claude that wins the hardest coding evals. Expensive and slower than Opus 5. Default when the repo is the product.

Price /1M
$10 / $50
Context
1M
Speed
65 t/s
Status
Current
SWE-bench Pro class workLong refactors
Details
04
O

GPT-5.6 Sol

OpenAI · Jul 2026 · Proprietary

60.9
Index
Intel60.9
Code77.4
Agents57.8

OpenAI’s flagship. Leads ARC-AGI-2 and Terminal-Bench. Token-efficient per task even though list price looks high.

Price /1M
$5 / $30
Context
1.05M
Speed
74 t/s
Status
Current
Abstract reasoningCodex-style agents
Details
05
x

Grok 4.6

xAI · Aug 2026 · Proprietary

60.9
Index
Intel60.9
Code76.8
Agents58.7

Ties GPT-5.6 Sol on intelligence at $2 / $6. Strong agents and GPQA. No public SWE-bench; 500K context is the ceiling.

Price /1M
$2 / $6
Context
500K
Speed
58 t/s
Status
Current
Frontier quality at mid priceAgents
Details
06
K

Kimi K3

Moonshot AI · Jul 2026 · Open weight

59.7
Index
Intel59.7
Code76.2
Agents54.3

Largest open-weight model that is actually good. 97% of the lead score at a lower output price. Native multimodal.

Price /1M
$3 / $15
Context
1.05M
Speed
36 t/s
Status
Current
Open-weight frontierLong-horizon coding
Details
07
Z

GLM-5.3

Z.ai (Zhipu) · Aug 2026 · Open weight

59.5
Index
Intel59.5
Code74.8
Agents59.1

One point off Opus 5 on the Intelligence Index. Best agentic score in the open-weight set. Text-only; weights were staged late August.

Price /1M
$1.40 / $4.40
Context
1M
Speed
66 t/s
Status
Current
Agents on a budgetOpen-weight intelligence
Details
08
Q

Qwen3.8 Max

Alibaba (Qwen) · Aug 2026 · Open weight

58.1
Index
Intel58.1
Code71.8
Agents58.4

Best open-weight flagship on evidence-backed boards. Slow to sample. Strong multilingual and agentic for the license.

Price /1M
$2 / $6
Context
1M
Speed
21 t/s
Status
Current
Open-weight frontierSelf-host at scale
Details
09
Z

GLM-5.3 Flash

Z.ai (Zhipu) · Aug 2026 · Open weight

57.5
Index
Intel57.5
Code71.5
Agents58.2

Almost GLM-5.3 intelligence at Flash prices. If it holds under independent harnesses, it is the value model of August.

Price /1M
$0.15 / $0.50
Context
1M
Speed
110 t/s
Status
Current
Cheap agentsHigh volume reasoning
Details
10
∞

Muse Spark 1.2

Meta · Aug 2026 · Proprietary

56.8
Index
Intel56.8
Code71
Agents52

Meta’s closed frontier. Fast on 1.1, strong HLE, still thinner independent coding evidence than Claude or GPT.

Price /1M
$1.25 / $4.25
Context
1M
Speed
—
Status
Current
Meta stackKnowledge work
Details
11
O

GPT-5.6 Terra

OpenAI · Jul 2026 · Proprietary

56.6
Index
Intel56.6
Code76.7
Agents50.2

The middle GPT-5.6. Keeps most of Sol’s coding, drops some agentic depth, much faster and cheaper.

Price /1M
$2 / $12
Context
1.05M
Speed
149 t/s
Status
Current
Balanced productionCoding assistants
Details
12
G

Gemini 3.7 Flash

Google DeepMind · Aug 2026 · Proprietary

56
Index
Intel56
Code76.1
Agents45.1

Fastest serious reasoning model in the field. GPQA leader on several boards. Agentic depth still behind Claude and Grok.

Price /1M
$0.75 / $3.75
Context
1M
Speed
330 t/s
Status
Current
Low-latency appsMultimodal
Details
13
A

Claude Sonnet 5

Anthropic · Jun 2026 · Proprietary

55.3
Index
Intel55.3
Code71.5
Agents49.7

Anthropic’s daily driver. Not the crown, but the one you actually leave on in production if Opus is too rich.

Price /1M
$2 / $10
Context
1M
Speed
82 t/s
Status
Current
Product chatDocs
Details
14
D

DeepSeek V4 Pro

DeepSeek · Apr 2026 · Open weight

53.2
Index
Intel53.2
Code68.8
Agents49.6

The default cheap-and-good open model for coding. Peak/off-peak API pricing. Not a 60-index flagship.

Price /1M
$0.87 / $1.74
Context
1M
Speed
79 t/s
Status
Current
Open-weight codingCost control
Details
15
O

GPT-5.6 Luna

OpenAI · Jul 2026 · Proprietary

52.3
Index
Intel52.3
Code71.4
Agents46.9

The punchline of the 5.6 stack: near-Terra coding at Luna prices. Default OpenAI pick for volume.

Price /1M
$0.20 / $1.20
Context
1.05M
Speed
202 t/s
Status
Current
High volumeCustomer support
Details
16
Q

Qwen3.8 27B

Alibaba (Qwen) · Aug 2026 · Open weight

52
Index
Intel52
Code68.1
Agents50.9

Dense 27B that hangs with models 50× its size. The local / workstation Qwen, not the datacenter Max.

Price /1M
$2.55 / $2.55
Context
262K
Speed
47 t/s
Status
Current
Local 24–80GB boxesFine-tunes
Details
17
D

DeepSeek V4 Flash

DeepSeek · Jul 2026 · Open weight

51.8
Index
Intel51.8
Code69.1
Agents48.4

Flash sibling of V4 Pro. Absurdly cheap, still 50+ intelligence. The volume open-weight pick.

Price /1M
$0.078 / $0.16
Context
1M
Speed
141 t/s
Status
Current
BatchSelf-host efficiency
Details
18
Q

Qwen3.8 Flash

Alibaba (Qwen) · Aug 2026 · Open weight

50
Index
Intel50
Code64
Agents44

Qwen’s cheap Fast lane. Use it as the open-weight Luna, not as a Max replacement.

Price /1M
$0.16 / $0.47
Context
1M
Speed
90 t/s
Status
Current
Open-weight volumeBatch
Details
19
G

Gemini 3.1 Pro

Google DeepMind · May 2026 · Proprietary

47.7
Index
Intel47.7
Code68.8
Agents23

Google’s Pro that actually shipped. Knowledge and multimodal are real; agentic scores are the hole in the card.

Price /1M
$2 / $12
Context
1M
Speed
80 t/s
Status
Current
Multimodal analysisLong documents
Details
20
mi

MiMo-V2.5-Pro

Xiaomi · Jun 2026 · Proprietary

47
Index
Intel47
Code55
Agents34

Xiaomi’s first model that belongs on a frontier page. Arena is real; published eval coverage is still thin.

Price /1M
Unlisted
Context
1M
Speed
—
Status
Current
Watch listChina-stack diversity
Details
21
m

MiniMax M3

MiniMax · Jun 2026 · Open weight

45.4
Index
Intel45.4
Code58.6
Agents36.1

Vendor SWE-bench is high; independent intelligence is mid. A specialist, not a general flagship.

Price /1M
$0.60 / $2.40
Context
1M
Speed
108 t/s
Status
Current
Open-weight coding at mid price
Details
22
M

Mistral Medium 3.5

Mistral · May 2026 · Proprietary

42
Index
Intel42
Code58
Agents32

Best generally available European API in this set. Context is 128K — a generation behind the 1M club.

Price /1M
$0.40 / $2
Context
128K
Speed
90 t/s
Status
Current
EU residencyMid-tier chat and coding
Details
23
A

Claude Haiku 4.5

Anthropic · Jan 2026 · Proprietary

38
Index
Intel38
Code55
Agents28

Anthropic’s fast cheap tier. Useful for classification and routing, not for the frontier jobs on this board.

Price /1M
$1 / $5
Context
200K
Speed
140 t/s
Status
Current
ClassificationRouting
Details
24
O

gpt-oss-120B

OpenAI · Aug 2025 · Open weight

35
Index
Intel35
Code50
Agents22

OpenAI’s open-weight line. Fine to self-host; do not confuse it with the 5.6 API stack.

Price /1M
Unlisted
Context
128K
Speed
—
Status
Current
Self-hostFine-tunes
Details
25
∞

Llama 4 Maverick

Meta · Apr 2025 · Open weight

32
Index
Intel32
Code45
Agents18

The Llama you can actually fine-tune. Not on the 2026 intelligence frontier. Still the default open Meta line.

Price /1M
$0.20 / $0.60
Context
1M
Speed
120 t/s
Status
Current
Fine-tunesPermissive-ish Meta license
Details

Aperture is a field guide, not a vendor. Scores compiled 28 Aug 2026.

MethodologyHow to pick