Skip to catalog
FbFable
CatalogLabsScoresCompare
Catalog
M

Meta

Muse Spark 1.2

muse-spark-1.2

57Index
ClosedGenerally available1M contextTextImageAudioVideo

Pick this when

Cheap closed multimodal. Powers free Meta AI across WhatsApp, Instagram, Facebook.

Skip when

You need weights, or the data cannot go to Meta.

Most widely deployed model in this list by consumer reach. Meta stopped leading with Llama and made Muse the product.

Price

Input / 1M
$1.25
Output / 1M
$4.25
Measured task
$0.40

Contributor tier at $0.10 / $0.20 if Meta may train on your traffic. Easy no for confidential work.

Shape

Context
1M
Max output
64K
Released
2026-07
Tools
Yes
Reasoning
Yes

Scores

Intelligence index57
GPQA Diamond87
Terminal-Bench 2.169
Agentic index37
Arena
1070
Speed
140/s
GPQA
87.2%
Term.
69.4%
  • AA Index. The single number most people mean by “how smart.” It is a blend, so a specialist can lose here and still win the job you care about.
  • GPQA. A clean test of hard reasoning. PhD experts sit around 65%. It says little about writing, tools, or taste.
  • Term.. Closest public proxy for coding agents that live in a terminal. A high index score with a weak terminal score is a warning.
  • Arena. Captures taste and usefulness that unit tests miss. Sample size varies by model, so treat gaps under ~100 Elo as noise.

Meta · Menlo Park. Muse is closed and powers Meta AI across Facebook, Instagram, and WhatsApp. Llama 4 still ships for local work. Glimmer is the small Apache agent.

Compiled 2026-08-28 from public Artificial Analysis, LMSYS-style arena, and first-party price sheets. Scores move. Recheck before you spend.

A two-point gap on the intelligence index is noise for most work. Cost per finished task is the number that shows up on the invoice.