
Muse Spark 1.2 ranks #5 on GDPval-AA v2, our benchmark of agentic real-world knowledge work, at 1631 Elo - behind e.g. Claude Opus 5 (max, 1852), GPT-5.6 Sol (max, 1730), and Kimi K3 (1685), and ahead of e.g. Claude Opus 4.8 (max, 1588). Muse Spark 1.1 scored 1371 at its launch last month

Direct Elo comparison against the models practitioners already run is the fastest way to judge whether a new coding is worth trialing.
postMeta has released Muse Spark 1.3, their fourth Muse Spark model release in five…Artificial Analysis
postThe 3-point gain over Muse Spark 1.1 on the Artificial Analysis Intelligence…Artificial Analysis
postMuse Spark 1.3 scores 64 on coding tasks at a fifth of Opus 5's cost per taskArtificial AnalysisChecking sign-in…
Loading comments…