← Back to explorer
OpenAIReleased 2026-02

GPT-5.4

Still widely deployed prior GPT-5 generation — strong computer-use and coding history, often cheaper via existing contracts.

Context
1.1M
Speed
80 t/s
Price /M
$2.5 / $15
Access
api

When to use it

Teams already on GPT-5.x infra who need stability over bleeding edge.

Watch-outs

Superseded by 5.5/5.6 on preference and newest composites.

CodingAgentsMultimodal

Benchmarks

Arena Elo1463

Human preference ranking from blind pairwise chats. Higher is better; top frontier models cluster within ~50–80 Elo.

AA Intelligence Index46

Artificial Analysis composite across agents, coding, science, and general evaluations (v4.1 weighting).

SWE-bench Verified80%

Percent of real GitHub issues resolved end-to-end. Strong signal for agentic coding usefulness.

GPQA Diamond92%

PhD-level science questions. Separates frontier reasoning models better than saturated knowledge tests.

MMLU-Pro

Harder multi-choice knowledge/reasoning suite than classic MMLU.

Humanity's Last Exam36.6%

Frontier closed-ended academic difficulty across many domains.

More from OpenAI

  • GPT-5.6 Sol

    General frontier reasoning, mixed multimodal apps, and OpenAI ecosystem tooling.

    AA 59
  • GPT-5.5 Pro

    Hard reasoning tasks, research assistants, and multimodal product surfaces.

    AA 55
  • GPT-5.5

    Default production chat, tools, and mixed workloads on OpenAI.

    AA 51

Model Gauge · built by Cursor Grok 4.5 High · curated snapshot, not a live leaderboard

Scores change weekly — verify critical decisions against primary sources.