OnlySOTA

DeepSeek V4.1 Flash vs Grok 4.7

Both are SOTA right now. Each row below is one board’s call on the two, and the marked figure is the one that board places higher. No row is added to another, so there is no overall winner: the boards measure different things, and where they split, that is the finding.

Board DeepSeek V4.1 Flash Grok 4.7
Artificial Analysis LLM Leaderboard Artificial Analysis Intelligence Index 39.5 No. 27 of 372 46.4 No. 12 of 372
Arena — Agent Overall 0.040 No. 17 of 49 0.040 No. 16 of 49
Artificial Analysis Coding Agents Artificial Analysis Coding Agent Index — 0.563 No. 11 of 22
Arena — Code Overall 1620 No. 21 of 109 1638 No. 15 of 109
Arena — Text Overall (style control) 1474 No. 38 of 306 1442 No. 91 of 306

Each rank is the board’s own: the rank it publishes, or its place in its own order where it publishes none. The bar is how much of that board the model is ahead of. Scores from different boards cannot be compared.

At a glance

Developer DeepSeek SpaceXAI
Weights open closed
Released ——
Price Blended Price, USD per 1M tokens (3:1 input:output) $0.52$3.00
Value: where it sits on the kill line SOTA SOTA

DeepSeek V4.1 Flash against the other SOTA models

Grok 4.7 against the other SOTA models