llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model Directory

Scoring & Data

Scoring & Data
1224 models729 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll Benchmarks
llmboard.aiCopyright 2026 llmboard.ai

Meta model product

Muse Spark 1.2

2 is a proprietary multimodal reasoning model from Meta Superintelligence Labs for long-horizon coding and whole-repository agentic work.

Updated Sep 8, 2026. Default version: Muse Spark 1.2

LLMBoard Score79.0Muse Spark 1.2
Coverage0%2 benchmark families
Context window1MTokens
Official input price$1.25Meta API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Similar models
  • About
  • FAQ

Muse Spark 1.2 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Muse Spark 1.2 LLMBoard score breakdown

Muse Spark 1.2 Benchmark Results

Benchmark scores for Muse Spark 1.2.

15 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkMeta Internal Coding BenchScore70.60%Rank01Participants1Percentile100.00%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Text Style ControlScore1,498.99 ratingRank05Participants210Percentile98.09%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena Text FactualityScore1,485.79 ratingRank06Participants121Percentile95.83%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena Vision Style ControlScore1,291.76 ratingRank06Participants107Percentile95.28%EvidenceAEvaluatedAug 27, 2026
BenchmarkLM Arena TextScore1,488.82 ratingRank08Participants210Percentile96.65%EvidenceAEvaluatedSep 2, 2026
BenchmarkLM Arena VisionScore1,304.12 ratingRank09Participants107Percentile92.45%EvidenceAEvaluatedAug 27, 2026
BenchmarkDeepSWE 1.1Score59.30%Rank15Participants31Percentile53.33%EvidenceCEvaluatedSep 8, 2026
BenchmarkTerminal-Bench 2.1Score82.90%Rank16Participants35Percentile55.88%EvidenceCEvaluatedSep 8, 2026
BenchmarkLM Arena Agent Bash Recovery StepsScore7.46%Rank17Participants49Percentile66.67%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent Task Outcome ExplicitScore6.60%Rank18Participants49Percentile64.58%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent Tool HallucinationScore0.77%Rank18Participants49Percentile64.58%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena WebdevScore1,534.19 ratingRank27Participants97Percentile72.92%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent LeaderboardScore0.46%Rank28Participants49Percentile43.75%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent Praise ComplaintScore-6.33%Rank36Participants49Percentile27.08%EvidenceAEvaluatedSep 5, 2026
BenchmarkLM Arena Agent SteerabilityScore-6.19%Rank44Participants49Percentile10.42%EvidenceAEvaluatedSep 5, 2026

Muse Spark 1.2 Arena Results

Preference and agent-evaluation results for the default version.

30 of 72 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenatext style controlCategoryindustry business and management and financial operationsRank01Rating / score1,515.90Votes643ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry legal and governmentRank01Rating / score1,541.03Votes272ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry medicine and healthcareRank01Rating / score1,530.07Votes232ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorymulti turnRank01Rating / score1,518.39Votes514ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry legal and governmentRank02Rating / score1,531.56Votes272ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry business and management and financial operationsRank03Rating / score1,496.55Votes643ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry medicine and healthcareRank03Rating / score1,515.04Votes232ObservationsN/AResult dateSep 2, 2026
ArenatextCategorymulti turnRank03Rating / score1,506.65Votes514ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry life and physical and social scienceRank03Rating / score1,527.17Votes522ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorynon englishRank03Rating / score1,492.54Votes1,852ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryrussianRank03Rating / score1,514.48Votes330ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryenglishRank04Rating / score1,495.70Votes1,388ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry life and physical and social scienceRank05Rating / score1,514.43Votes522ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryrussianRank05Rating / score1,505.33Votes330ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryexclude tiesRank05Rating / score1,498.37Votes2,402ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryexclude tiesRank05Rating / score1,519.51Votes2,402ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryindustry software and it servicesRank05Rating / score1,528.55Votes1,326ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryoverallRank05Rating / score1,498.99Votes3,240ObservationsN/AResult dateSep 2, 2026
Arenatext factualityCategoryoverallRank06Rating / score1,485.79Votes3,240ObservationsN/AResult dateSep 2, 2026
Arenavision style controlCategoryocrRank06Rating / score1,305.77Votes1,265ObservationsN/AResult dateAug 27, 2026
Arenavision style controlCategoryoverallRank06Rating / score1,291.76Votes1,844ObservationsN/AResult dateAug 27, 2026
ArenatextCategoryenglishRank07Rating / score1,492.45Votes1,388ObservationsN/AResult dateSep 2, 2026
ArenatextCategorynon englishRank07Rating / score1,481.88Votes1,852ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryenglishRank07Rating / score1,501.27Votes1,388ObservationsN/AResult dateSep 2, 2026
ArenavisionCategoryocrRank07Rating / score1,313.40Votes1,265ObservationsN/AResult dateAug 27, 2026
ArenatextCategoryexclude tiesRank08Rating / score1,501.32Votes2,402ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryoverallRank08Rating / score1,488.82Votes3,240ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategorycodingRank08Rating / score1,533.36Votes935ObservationsN/AResult dateSep 2, 2026
ArenatextCategoryindustry software and it servicesRank09Rating / score1,502.03Votes1,326ObservationsN/AResult dateSep 2, 2026
Arenatext style controlCategoryhard promptsRank09Rating / score1,512.88Votes2,124ObservationsN/AResult dateSep 2, 2026

Muse Spark 1.2 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$1.25 input, $4.25 output per 1M
Official provider
Meta
Lowest third-party
From $1.25 input, $4.25 output per 1M via OrcaRouter
Tracked offerings
13
13 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderOrcaRouterProvider model IDmeta/muse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderNanoGPTProvider model IDmeta/muse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderMetaProvider model IDmuse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderOpenRouterProvider model IDmeta/muse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderOpperProvider model IDmeta/muse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderLLM GatewayProvider model IDmeta/muse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderMerge GatewayProvider model IDmeta/muse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderOpenCode ZenProvider model IDmuse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderEmpirioLabs AIProvider model IDmuse-spark-1-2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderVercel AI GatewayProvider model IDmeta/muse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderDevPass (LLM Gateway)Provider model IDmuse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderKilo GatewayProvider model IDmeta/muse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026
ProviderAbacusProvider model IDmuse-spark-1.2RegionglobalInput / 1M$1.25Output / 1M$4.25Context1MUpdatedSep 8, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Muse Spark 1.2 Runtime Performance

Provider-specific output speed and catalog latency for Muse Spark 1.2. Runtime does not affect the capability score.

1 row
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderMeta Model APIOutput Speed28.44 tok/sCatalog Latency4.65 sMax Input1MMax Output131.1KUpdatedSep 8, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Muse Spark 1.2 Specifications

Technical details for the model's default version.

Version
Muse Spark 1.2
Released
Aug 5, 2026
Knowledge cutoff
Unknown
Parameters
N/A
Context window
1M
Max output
131.1K
Inputs
image, text, video
Outputs
text
Open weights
No
License
Proprietary

Muse Spark 1.2 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 row
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionMuse Spark 1.2ReleasedAug 5, 2026LLMBoard78.98ParametersN/AContext1MMax output131.1KOpen weightsNoLicenseProprietary

Models similar to Muse Spark 1.2

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#14+6.19
ME

Muse Spark 1.1

Meta

85.17 LLMBoard

Details
#40-6.89
ME

Muse Spark

Meta

72.09 LLMBoard

Details
#6+11.84
ME

Muse Spark 1.3

Meta

90.82 LLMBoard

Details
#79-21.59
ME

Muse Glimmer 30B

Meta

57.39 LLMBoard

Details
#200-55.54
ME

Llama 4 Maverick

Meta

23.44 LLMBoard

Details
#203-56.10
ME

Llama 3.1 405B

Meta

22.88 LLMBoard

Details

What is Muse Spark 1.2?

Key information about Muse Spark 1.2 and its available data.

2 is Meta Superintelligence Labs' proprietary multimodal reasoning model, with a 1,048,576-token context window for long-horizon coding and whole-repository agentic work. It is available through the Meta Model API in standard and contributor tiers.

Data as of 2026-09-08.

FAQ

Common questions about Muse Spark 1.2.

When was Muse Spark 1.2 released?

Muse Spark 1.2's default version was released on Aug 5, 2026.

How much does Muse Spark 1.2 cost?

Muse Spark 1.2's official API price is $1.25 per million input tokens and $4.25 per million output tokens via Meta. The lowest tracked third-party offer starts at $1.25 input and $4.25 output via OrcaRouter.

Who created Muse Spark 1.2?

Muse Spark 1.2 was created by Meta.

What is the context window for Muse Spark 1.2?

The default version has a 1M token context window.

Is Muse Spark 1.2 open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Muse Spark 1.2?

13 provider offerings are linked to the default version.