One email each morning — yesterday's AI, sortedGet it in your inbox

METAL LAB

Meta's Muse Spark 1.2 ties for third with Intelligence Index score of 54

Third model release in four months, per Artificial Analysis tracking, with improved agentic work capabilities

이미지: X — 벤치마크·평가

Summary

  • Meta has released Muse Spark 1.2, which scored 54 on the Intelligence Index tracked by evaluator Artificial Analysis.
  • The model marks Meta's third release in four months and was noted for significant improvement in agentic knowledge work capabilities compared to the previous version.
  • The score of 54 reportedly ties the model for third place alongside SpaceXAI, though detailed category-by-category scores were not disclosed.
모델명
Muse Spark 1.2
개발사
메타(Meta)
평가 지표
Artificial Analysis Intelligence Index
점수
54점
순위
SpaceXAI와 공동 3위
릴리스 빈도
4개월 동안 세 번째 공개
개선 영역
에이전틱 지식 노동(agentic knowledge work) 능력
공개 확인 시점
2026년 8월 5일(평가 기관 게시 기준)

Meta releases Muse Spark 1.2

Meta has released a new model called Muse Spark 1.2. Evaluator Artificial Analysis reported that the model scored 54 on its Intelligence Index. This score matches that of SpaceXAI, putting the two in a tie for third place, according to the evaluator. The disclosure did not specify which models hold the first and second spots, nor the exact score gaps between models.

Image related to the Artificial Analysis Intelligence Index
Muse Spark 1.2 evaluation results · Source: Artificial Analysis

Third release in four months

One notable point is the pace of release. According to Artificial Analysis, Muse Spark 1.2 is Meta's third model release within a four-month span. The pattern of frequent version updates aimed at raising scores appears to reflect an effort to close the gap with top-tier competitors.

ItemConfirmed detail
ModelMuse Spark 1.2
Intelligence Index54
RankingTied for third (with SpaceXAI)
Release cadence3 releases in 4 months
Key improvementAgentic knowledge work

Improvement focus: "agentic knowledge work"

The evaluator noted a clear improvement in agentic task performance compared to the previous release. Artificial Analysis described it as "significantly improving agentic knowledge work capabilities." This suggests gains in multi-step tasks involving documents, search, and tool use, rather than single-turn question answering. However, no detailed figures were provided on which specific sub-benchmarks improved or by how much.

For ongoing coverage of Intelligence Index scores across models and shifts in agent performance evaluation methods, see METAL LAB's benchmark coverage.

What remains unconfirmed

The information currently available is based on a single post from the evaluator. Technical specifications — including whether model weights are open, context length, pricing, and training/inference infrastructure — remain unconfirmed. The specific model name for SpaceXAI, mentioned as tying at third place, was also not specified.

CategoryStatus
Intelligence Index scoreConfirmed
RankingConfirmed (tied for third)
Detailed benchmark breakdownUndisclosed
Model specs/release formatUndisclosed
Top 1st/2nd place modelsNot mentioned

Further benchmark results and an official announcement from Meta are expected to clarify the significance of the score and the competitive landscape.