← Back to explorer
OpenAIReleased 2026-07

GPT-5.6 Sol

OpenAI's newest reasoning-tier flagship. Extremely close to Claude on AA Index with strong tool use and broad multimodal coverage.

Context
1M
Speed
57 t/s
Price /M
$5 / $30
Access
api

When to use it

General frontier reasoning, mixed multimodal apps, and OpenAI ecosystem tooling.

Watch-outs

Reasoning modes vary widely in latency and cost; pick effort level carefully.

CodingAgentsMultimodalScienceWriting

Benchmarks

Arena Elo1514

Human preference ranking from blind pairwise chats. Higher is better; top frontier models cluster within ~50–80 Elo.

AA Intelligence Index59

Artificial Analysis composite across agents, coding, science, and general evaluations (v4.1 weighting).

SWE-bench Verified

Percent of real GitHub issues resolved end-to-end. Strong signal for agentic coding usefulness.

GPQA Diamond

PhD-level science questions. Separates frontier reasoning models better than saturated knowledge tests.

MMLU-Pro

Harder multi-choice knowledge/reasoning suite than classic MMLU.

Humanity's Last Exam47.2%

Frontier closed-ended academic difficulty across many domains.

More from OpenAI

  • GPT-5.5 Pro

    Hard reasoning tasks, research assistants, and multimodal product surfaces.

    AA 55
  • GPT-5.5

    Default production chat, tools, and mixed workloads on OpenAI.

    AA 51
  • GPT-5.4

    Teams already on GPT-5.x infra who need stability over bleeding edge.

    AA 46

Model Gauge · built by Cursor Grok 4.5 High · curated snapshot, not a live leaderboard

Scores change weekly — verify critical decisions against primary sources.