Skip to catalog
FbFable
CatalogLabsScoresCompare
Catalog
M

Meta

Llama 4 Scout

llama-4-scout

41Index
Llama 4 Community · restrictedGenerally available10M context17BTextImage

Pick this when

Enormous documents on a budget. Local and hosted Llama that still has a real ecosystem.

Skip when

You need 2026 frontier quality. Llama is in maintenance mode.

The family that made local AI ordinary. Still downloadable and widely supported. No longer where the action is.

Price

Input / 1M
$0.11
Output / 1M
$0.34
Measured task
$0.09

Hosted on Bedrock and others. 10M context is the headline.

Shape

Context
10M
Max output
16K
Released
2026-04
Tools
Yes
Reasoning
No

Scores

Intelligence index41
GPQA Diamond58
Terminal-Bench 2.132
Agentic index14
Arena
—
Speed
190/s
GPQA
58.2%
Term.
32.4%
  • AA Index. The single number most people mean by “how smart.” It is a blend, so a specialist can lose here and still win the job you care about.
  • GPQA. A clean test of hard reasoning. PhD experts sit around 65%. It says little about writing, tools, or taste.
  • Term.. Closest public proxy for coding agents that live in a terminal. A high index score with a weak terminal score is a warning.
  • Arena. Captures taste and usefulness that unit tests miss. Sample size varies by model, so treat gaps under ~100 Elo as noise.

Meta · Menlo Park. Muse is closed and powers Meta AI across Facebook, Instagram, and WhatsApp. Llama 4 still ships for local work. Glimmer is the small Apache agent.

Compiled 2026-08-28 from public Artificial Analysis, LMSYS-style arena, and first-party price sheets. Scores move. Recheck before you spend.

A two-point gap on the intelligence index is noise for most work. Cost per finished task is the number that shows up on the invoice.