tokenstat
tokenstat

Model catalog

Benchmarks

Arena Elo from free open daily snapshots on GitHub (text, code, vision, agent), plus a sparse Artificial Analysis subset OpenRouter embeds. One score set is cross-linked onto provider clones of the same model. Arena may score the same base model at different reasoning efforts, so those rows stay separate and link to the base model only when it exists in the catalog. The open dumps publish 20 models per Arena board (10 on agent), so a full board is short by design. Empty on a single board is a gap, not a ranking of zero.

19 scored rows on this board · 417 linked into the catalog · as of · Models · AI plans

#ModelScore
1
Anthropic: Claude Fable 5Anthropic · claude-fable-5 · high effort
1310Arena vision
2
Qwen3.8 MaxAlibaba · qwen3.8 · max effort
1302Arena vision
3
Anthropic: Claude 4.7 Opus ThinkingAnthropic · claude-opus-4-7 · high effort
1301Arena vision
4
Anthropic: Claude 4.7 Opus ThinkingAnthropic · claude-opus-4-7
1300Arena vision
5
Anthropic: Claude 4.6 Opus ThinkingAnthropic · claude-opus-4-6 · high effort
1299Arena vision
6
Meta: Muse Spark 1.3Meta · muse-spark-1.3 · max effort
1294Arena vision
7
Anthropic: Claude 4.6 Opus ThinkingAnthropic · claude-opus-4.6
1293Arena vision
8
Meta: Muse Spark 1.2Meta · muse-spark-1.2 · xhigh effort
1292Arena vision
9
Anthropic: Claude Fable 5.1Anthropic · claude-fable-5.1 · max effort
1289Arena vision
10
Anthropic: Claude Opus 5Anthropic · claude-opus-5 · high effort
1289Arena vision
11
Google: Gemini 3 ProGoogle · gemini-3-pro
1289Arena vision
12
OpenAI: GPT-5.5OpenAI · gpt-5.5
1287Arena vision
13
OpenAI: GPT-5.6 SolOpenAI · gpt-5.6-sol · xhigh effort
1286Arena vision
14
Gpt 5.4 HighOpenAI · gpt-5.4 · high effort
1285Arena vision
15
Google: Gemini 3.5 FlashGoogle · gemini-3.5-flash · high effort
1284Arena vision
16
OpenAI: GPT-6 AstraOpenAI · gpt-6-astra · max effort
1284Arena vision
17
Anthropic: Claude Opus 4.8 ThinkingAnthropic · claude-opus-4-8 · high effort
1283Arena vision
18
Google: Gemini 3.5 FlashGoogle · gemini-3.5-flash · medium effort
1283Arena vision
19
OpenAI: GPT-5.5OpenAI · gpt-5.5 · high effort
1283Arena vision

Sources

Arena Elo from Arena AI, via open dumps by oolong-tea-2026. Artificial Analysis indices via OpenRouter (subset). Artificial Analysis for the indices themselves. Plus vendor-reported launch evals collected from announcements, analyst writeups, launch blogs, and community benchmarks. As of .

Turn benchmark scores into working sessions

Run agents in a local Git-enabled workspace, connect to another machine when needed, and track the sessions and tokens behind your work. See what your AI subscriptions covered.

CLI for macOS, Linux, and Windows

Data and list-price estimates may contain mistakes. Not investment advice.