Claude Fable 5
Hard coding agents, long computer-use sessions, and high-stakes knowledge work.
Snapshot 2026-07-15
Filter by lab and use case, sort by Arena, Artificial Analysis, SWE-bench, or price, then compare up to four models side by side — including Cursor Grok 4.5, who built this site.
Snapshot dated 2026-07-15. Scores change weekly; treat this as a decision aid, not a live API.
19 models · preset Frontier overall
Hard coding agents, long computer-use sessions, and high-stakes knowledge work.
General frontier reasoning, mixed multimodal apps, and OpenAI ecosystem tooling.
Coding agents, computer use, and careful long-form reasoning in production.
Coding agents in Cursor / Grok Build, long-running engineering tasks, and near-frontier intelligence at a fraction of Opus/GPT Pro spend.
Scientific reasoning, huge document corpora, and Google Cloud / Workspace stacks.
Day-to-day coding copilots, customer agents, and high-volume Claude deployments.
Self-hosting, cost-sensitive production, and research fine-tunes.
High-throughput apps, classification, and cost-sensitive Google stack workloads.