Skip to catalog
FbFable
CatalogLabsScoresCompare
Catalog
G

Google

Gemini 3.5 Pro

gemini-3.5-pro-preview

—Index
ClosedNot generally available1M contextTextImageAudioVideo
Not generally available
Listed so you do not plan a stack around a rumor. Use a public Gemini Flash model instead.

Pick this when

Nothing you can buy today. Listed so you do not plan a stack around a rumor.

Skip when

Always, until it is generally available.

Announced early, then delayed through the summer. Do not design production around it. 3.7 Flash is the public Gemini to use.

Price

Input / 1M
$1.50
Output / 1M
$12
Measured task
—

Partner preview. Google has not given a GA date.

Shape

Context
1M
Max output
64K
Released
2026-02
Tools
Yes
Reasoning
Yes

Scores

Intelligence index—
GPQA Diamond—
Terminal-Bench 2.1—
Agentic index—
Arena
—
Speed
—
GPQA
—
Term.
—
  • AA Index. The single number most people mean by “how smart.” It is a blend, so a specialist can lose here and still win the job you care about.
  • GPQA. A clean test of hard reasoning. PhD experts sit around 65%. It says little about writing, tools, or taste.
  • Term.. Closest public proxy for coding agents that live in a terminal. A high index score with a weak terminal score is a warning.
  • Arena. Captures taste and usefulness that unit tests miss. Sample size varies by model, so treat gaps under ~100 Elo as noise.

Google · Mountain View. Gemini is natively multimodal. 3.7 Flash is the current pick for mixed media and long documents. 3.5 Pro is still slipping.

Compiled 2026-08-28 from public Artificial Analysis, LMSYS-style arena, and first-party price sheets. Scores move. Recheck before you spend.

A two-point gap on the intelligence index is noise for most work. Cost per finished task is the number that shows up on the invoice.