ApertureAug 2026
FieldCompareLabsGuideMethod
  • Field
  • Compare
  • Labs
  • Guide
Anthropic
A

Claude Fable 5

Jun 2026 · Proprietary · Current · supported evidence

The available Claude that wins the hardest coding evals. Expensive and slower than Opus 5. Default when the repo is the product.

Open compareLab site

Intelligence

62.1

AA Index

API price

$10 / $50

Input / output per 1M

Context

1M

65 tok/s

Capability

Intelligence Index62.1
Coding index76.5
Agentic index56.6
Arena Elo (offset)207

Use it when

  • ▸SWE-bench Pro class work
  • ▸Long refactors
  • ▸Chat preference

Skip if

  • –Budget or latency is the constraint

Benchmarks

GPQA Diamond
Graduate-level science questions designed so Google search is not enough. Still one of the cleanest knowledge/reasoning splits.
92.6%
Humanity’s Last Exam
Expert-written questions across fields. Harder than MMLU; the current differentiator for “does this model actually know things.”
55.5%
MMLU-Pro
Harder, less-saturated successor to MMLU. Classic MMLU is above 90% for every flagship and no longer ranks the field.
—
SWE-bench Verified
500 human-validated GitHub issues. Score swings 5–15 points by harness — treat vendor numbers as an upper bound.
95%
SWE-bench Pro
Harder, contamination-resistant coding eval. Currently the best public split between “can code” and “can maintain a repo.”
80.3%
Terminal-Bench 2.1
End-to-end tasks in a real terminal. Better proxy for coding agents than HumanEval, which is fully saturated.
83.4%
ARC-AGI-2
Abstract visual puzzles. Rewards generalization over memorization. GPT-5.6 Sol currently leads the published set.
89.2%
AIME 2025
American Invitational Mathematics Examination. Contest math; reasoning-mode models dominate.
—
MMMU
College-level multimodal understanding across diagrams, charts, and exam figures.
—

Modalities

text · vision · tools

Reasoning mode

hybrid

Hybrid and reasoning models spend tokens thinking. That raises GPQA and agents, and also raises latency and bill.

Close on intelligence

  • AClaude Opus 563.1
  • OGPT-5.6 Sol60.9
  • xGrok 4.660.9

Same lab

  • Claude Mythos 5
  • Claude Opus 5
  • Claude Sonnet 5
  • Claude Opus 4.8

Aperture is a field guide, not a vendor. Scores compiled 28 Aug 2026.

MethodologyHow to pick