Skip to models
MODELINDEX
Edition 07.15.269 tracked models
Methodology
Independent model field guide

Choose the model.
Not the hype.

A decision-first index of the frontier: benchmark evidence, real API cost, context, openness, and what each model is actually useful for.

7major labs
9current models
3compare at once
DATA SNAPSHOTJUL
15
2026 · 1ST PARTY
Optimize for
Value frontier

Capability, meet cost.

Higher is better. Further left costs less. Only models with a published Terminal-Bench result and a listed API rate are plotted.

908070
$0$25$50 / 1M output
Terminal-Bench · reported
Model directory

Find your working model.

9 of 9 models · ranked for all-rounder

Lab
CompareModelFitContextInput / 1MOutput / 1M
FIT94
Context1M
Input / 1M$2
Output / 1M$6
FIT93
Context1M
Input / 1M$5
Output / 1M$30
Best for

High-stakes engineering, tool-use workflows, and professional knowledge work.

Coding agentsTool useProfessional work
Terminal-Bench (reported)82.7
SWE-Bench Pro58.6
Coding Agent Index—
MCP Atlas75.3
Inputs

Text · Vision · Tools

Watch for

Premium API price; reported scores use xhigh reasoning effort in a research environment.

OpenAI launch & evaluations
FIT93
Context1M
Input / 1M$5
Output / 1M$25
FIT91
Context1M
Input / 1MNot listed
Output / 1MNot listed
FIT88
Context1M
Input / 1M$30
Output / 1M$180
FIT88
Context256K
Input / 1MSelf-host
Output / 1MSelf-host
FIT87
Context1M
Input / 1M$0.43
Output / 1M$0.87
FIT81
Context1M
Input / 1MSelf-host
Output / 1MSelf-host
FIT78
Context10M
Input / 1MSelf-host
Output / 1MSelf-host
Read the numbers correctly

Benchmarks are evidence.
They are not a verdict.

Benchmark scores are transcribed from first-party launch posts, model cards, and API documentation. A dash means no sufficiently comparable figure was found—not a zero.

Fit scores are an editorial decision aid based on reported capability, price, modalities, openness, and deployment tradeoffs. They are intentionally separate from benchmark values.

Prices are USD per million tokens at the standard API rate where publicly listed. “Self-host” means weights are available; your infrastructure still has a cost. “Not listed” is intentionally not estimated.

MODELINDEX

Built by GPT-5.6 Sol High · Codex
Source-backed. Updated July 15, 2026.

Back to top