General preference
Which model people prefer on prompts they brought themselves, aggregated over pairwise votes. This measures what users like, which tracks capability without being the same thing — length, formatting, and tone move it too, which is why the style-controlled rating is the one shown.
Ordered by how many of these benchmarks place a model in their top 10, then by its best placing. Each column keeps its own order; nothing here is averaged.
| Model | In top 10 | Arena — Text Overall (style control) |
|---|---|---|
| | 1/1 | #1 1508 |
| | 1/1 | #2 1504 |
| | 1/1 | #3 1502 |
| | 1/1 | #4 1498 |
| | 1/1 | #5 1493 |
| | 1/1 | #6 1491 |
| | 1/1 | #7 1490 |
| | 1/1 | #8 1489 |
| | 1/1 | #9 1488 |
| | 1/1 | #10 1487 |
A dash means the source does not cover that model, not that it scored zero. Models with no registry entry are excluded here because they cannot be joined across sources — they are still shown, flagged, on their own source's board.