Model catalog
Benchmarks
Arena Elo from free open daily snapshots on GitHub (text, code, vision, agent), plus a sparse Artificial Analysis subset OpenRouter embeds. One score set is cross-linked onto provider clones of the same model. Arena may score the same base model at different reasoning efforts, so those rows stay separate and link to the base model only when it exists in the catalog. The open dumps publish 20 models per Arena board (10 on agent), so a full board is short by design. Empty on a single board is a gap, not a ranking of zero.
| # | Model | Score |
|---|---|---|
| 1 | Anthropic: Claude Fable 5Anthropic · claude-fable-5 · high effort | 1310Arena vision |
| 2 | Qwen3.8 MaxAlibaba · qwen3.8 · max effort | 1302Arena vision |
| 3 | Anthropic: Claude 4.7 Opus ThinkingAnthropic · claude-opus-4-7 · high effort | 1301Arena vision |
| 4 | Anthropic: Claude 4.7 Opus ThinkingAnthropic · claude-opus-4-7 | 1300Arena vision |
| 5 | Anthropic: Claude 4.6 Opus ThinkingAnthropic · claude-opus-4-6 · high effort | 1299Arena vision |
| 6 | Meta: Muse Spark 1.3Meta · muse-spark-1.3 · max effort | 1294Arena vision |
| 7 | Anthropic: Claude 4.6 Opus ThinkingAnthropic · claude-opus-4.6 | 1293Arena vision |
| 8 | Meta: Muse Spark 1.2Meta · muse-spark-1.2 · xhigh effort | 1292Arena vision |
| 9 | Anthropic: Claude Fable 5.1Anthropic · claude-fable-5.1 · max effort | 1289Arena vision |
| 10 | Anthropic: Claude Opus 5Anthropic · claude-opus-5 · high effort | 1289Arena vision |
| 11 | Google: Gemini 3 ProGoogle · gemini-3-pro | 1289Arena vision |
| 12 | OpenAI: GPT-5.5OpenAI · gpt-5.5 | 1287Arena vision |
| 13 | OpenAI: GPT-5.6 SolOpenAI · gpt-5.6-sol · xhigh effort | 1286Arena vision |
| 14 | Gpt 5.4 HighOpenAI · gpt-5.4 · high effort | 1285Arena vision |
| 15 | Google: Gemini 3.5 FlashGoogle · gemini-3.5-flash · high effort | 1284Arena vision |
| 16 | OpenAI: GPT-6 AstraOpenAI · gpt-6-astra · max effort | 1284Arena vision |
| 17 | Anthropic: Claude Opus 4.8 ThinkingAnthropic · claude-opus-4-8 · high effort | 1283Arena vision |
| 18 | Google: Gemini 3.5 FlashGoogle · gemini-3.5-flash · medium effort | 1283Arena vision |
| 19 | OpenAI: GPT-5.5OpenAI · gpt-5.5 · high effort | 1283Arena vision |
Sources
Arena Elo from Arena AI, via open dumps by oolong-tea-2026. Artificial Analysis indices via OpenRouter (subset). Artificial Analysis for the indices themselves. Plus vendor-reported launch evals collected from announcements, analyst writeups, launch blogs, and community benchmarks. As of .
Turn benchmark scores into working sessions
Run agents in a local Git-enabled workspace, connect to another machine when needed, and track the sessions and tokens behind your work. See what your AI subscriptions covered.
Data and list-price estimates may contain mistakes. Not investment advice.