OnlySOTA

Claude Haiku 5.5 vs Muse Spark 1.3

Both are SOTA right now. Each row below is one board’s call on the two, and the marked figure is the one that board places higher. No row is added to another, so there is no overall winner: the boards measure different things, and where they split, that is the finding.

Board Claude Haiku 5.5 Muse Spark 1.3
Artificial Analysis LLM Leaderboard Artificial Analysis Intelligence Index 43.4 No. 19 of 375 48.1 No. 9 of 375
Arena — Agent Overall — 0.040 No. 18 of 51
Artificial Analysis Coding Agents Artificial Analysis Coding Agent Index 0.414 No. 21 of 23 0.543 No. 13 of 23
Arena — Code Overall — 1656 No. 13 of 141
Arena — Text Overall (style control) — 1494 No. 9 of 413
Arena — Vision Overall (style control) — 1290 No. 12 of 159

Each rank is the board’s own: the rank it publishes, or the model’s place in the board’s own order where it publishes none. Scores from different boards cannot be compared.

At a glance

Developer Anthropic Meta
Weights closed closed
Released ——
Price Blended Price, USD per 1M tokens (3:1 input:output) $0.20$2.00
Value: where it sits on the kill line Executioner SOTA