tokenstat
tokenstat

Model catalog

Models and list rates

Published API list rates (USD per million tokens), who lists each model, and whether it can do tools, vision, or reasoning. Recommended order is newer first; among similarly new models, larger context wins. Free/$0 rows stay hidden unless you ask. Not a bill: measure real sessions with the CLI.

3,231 models in the merged catalog · as of . Sources: LiteLLM, OpenRouter, models.dev. · Providers · Plans · Benchmarks

Showing 48 of 2,670 matches · page 28 of 56 · 561 free/$0 rows hidden (show them)

ModelInputOutputContextOffers
Aion-RP 1.0 (8B)aion-labs/aion-rp-llama-3.1-8b
llama
$0.2006models.dev$0.200633K3
DeepSeek R1 Distill Qwen 14Balibaba-cn/deepseek-r1-distill-qwen-14b
reasoningtoolsqwen
$0.07LiteLLM$0.0733K3
DeepSeek R1 Distill Qwen 7Balibaba-cn/deepseek-r1-distill-qwen-7b
reasoningtoolsqwen
$0.072models.dev$0.14433K2
Phi-4-reasoning-plusazure-cognitive-services/phi-4-reasoning-plus
reasoningopenphi
$0.125models.dev$0.532K2
Llama 3.3 70B Euryaledeepinfra/Sao10K/L3.3-70B-Euryale-v2.3
llama
$0.493models.dev$0.49320K4
Evayale 70bSteelskull/L3.3-MS-Evayale-70B
llama
$0.493models.dev$0.49316K1
Llama 3.3 70B Cu MaiSteelskull/L3.3-Cu-Mai-R1-70b
llama
$0.493models.dev$0.49316K1
Magnum v4 72Banthracite-org/magnum-v4-72b
openllama
$2.01models.dev$2.9916K3
MS Evalebis 70bSteelskull/L3.3-MS-Evalebis-70b
llama
$0.493models.dev$0.49316K1
Qwen-MT Plusalibaba-cn/qwen-mt-plus
qwen
$0.25LiteLLM$0.7516K3
Qwen-MT Turboalibaba-cn/qwen-mt-turbo
qwen
$0.101models.dev$0.2816K2
Steelskull Electra R1 70bSteelskull/L3.3-Electra-R1-70b
llama
$0.6999models.dev$0.699916K1
Steelskull Nevoria 70bSteelskull/L3.3-MS-Nevoria-70b
llama
$0.493models.dev$0.49316K1
Steelskull Nevoria R1 70bSteelskull/L3.3-Nevoria-R1-70b
llama
$0.493models.dev$0.49316K1
The Drummer Cydonia 24B v2TheDrummer/Cydonia-24B-v2
$0.1003models.dev$0.120716K1
Dolphin 72baccounts/fireworks/models/dolphin-2-9-2-qwen2-72b
qwen
$0.306models.dev$0.3068K4
DeepSeek-V3.2-fastdeepseek-ai/DeepSeek-V3.2-fast
reasoningtoolsopen
$0.4models.dev$28K1
Claude-Sonnet-3.5anthropic/claude-sonnet-3.5
toolsvisionclaude-sonnet
$2.60models.dev$13189K1
Qwen2.5 32B Instructalibaba-cn/qwen2-5-32b-instruct
toolsopenqwen
$0.06LiteLLM$0.2131K3
Qwen2.5 72B Instructalibaba-cn/qwen2-5-72b-instruct
toolsopenqwen
$0.12LiteLLM$0.39131K3
Qwen2.5-VL 72B Instructalibaba-cn/qwen2-5-vl-72b-instruct
toolsvisionopenqwen
$0.13LiteLLM$0.4131K7
Hermes 3 70B Instructdeepinfra/NousResearch/Hermes-3-Llama-3.1-70B
opennousresearch
$0.12LiteLLM$0.3131K7
Qwen2.5 14B Instructalibaba-cn/qwen2-5-14b-instruct
toolsopenqwen
$0.144models.dev$0.431131K2
Qwen2.5 7B Instructalibaba-cn/qwen2-5-7b-instruct
toolsopenqwen
$0.04LiteLLM$0.1131K4
Qwen2.5-Coder 7B Instructalibaba-cn/qwen2-5-coder-7b-instruct
toolsopenqwen
$0.01LiteLLM$0.03131K2
MiniMax-M2.5-fastMiniMaxAI/MiniMax-M2.5-fast
reasoningtoolsopen
$0.3models.dev$1.208K1
Qwen2.5-VL 7B Instructalibaba-cn/qwen2-5-vl-7b-instruct
toolsvisionopenqwen
$0.287models.dev$0.717131K2
Llama 3.2 1B Instructmeta-llama/llama-3.2-1b-instruct
reasoningopenllama
$0.027models.dev$0.201131K4
Llama 3.2 3B Instructmeta-llama/llama-3.2-3b-instruct
reasoningopenllama
$0.02LiteLLM$0.02131K8
Llama 3.1 8B Instructneon/meta-llama-3-1-8b-instruct
scorestoolsopenllama
$0.1LiteLLM$0.1131K4
Qwen 2.5 72B InstructQwen/Qwen2.5-72B-Instruct
toolsopenqwen
$0.11models.dev$0.38128K4
Qwen 2.5 Coder 32Babacus/qwen-2.5-coder-32b
toolsopenqwen
$0.79models.dev$0.79128K1
Llama 3.2 11B Vision Instruct@cf/meta/llama-3.2-11b-vision-instruct
visionopenllama
$0.0485models.dev$0.676128K2
Command R (08-2024)cohere/command-r-08-2024
toolsopencommand-r
$0.15models.dev$0.6128K5
Command R+ (08-2024)cohere/command-r-plus-08-2024
toolsopencommand-r
$2.50models.dev$10128K6
Llama 3.1 70B Instructamazon-bedrock/meta.llama3-1-70b-instruct-v1:0
toolsopenllama
$0.72models.dev$0.72128K4
Llama 3.1 8B Instructmeta-llama/Meta-Llama-3.1-8B-Instruct
scorestoolsopenllama
$0.02models.dev$0.05128K2
Llama 3.1 8B Instructamazon-bedrock/meta.llama3-1-8b-instruct-v1:0
scorestoolsopenllama
$0.22models.dev$0.22128K4
Llama 3.2 3B Instruct@cf/meta/llama-3.2-3b-instruct
openllama
$0.0509models.dev$0.33580K2
Llama 3.2 1B Instruct@cf/meta/llama-3.2-1b-instruct
openllama
$0.027models.dev$0.20160K2
QwQ 32BQwen/QwQ-32B
reasoningtoolsopenqwen
$0.15LiteLLM$0.433K2
R1 Distill Llama 70Bdeepseek/deepseek-r1-distill-llama-70b
reasoningopendeepseek
$0.75LiteLLM$0.998K4
Qwen-Plusqwen/qwen-plus
reasoningtoolsqwen
$0.115models.dev$0.2871M10
Qwen2.5 7B Instructqwen/qwen-2.5-7b-instruct
toolsopenqwen
$0.1models.dev$0.233K2
Osmosis Structure 0.6Bosmosis/osmosis-structure-0.6b
toolsopenosmosis
$0.1models.dev$0.54K1
Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct
toolsopenqwen
$0.062models.dev$0.231131K8
Llama 3.1 70B Instructmeta-llama/llama-3.1-70b-instruct
toolsopenllama
$0.4models.dev$0.4131K3
Qwen Coder Plusllmgateway/qwen-coder-plus
toolsqwen
$0.502models.dev$1131K1

Price this against your own usage

List rates are only half the story. The CLI reads the session logs your coding harnesses already write and shows tokens, cache share, and usage value on your machine.

$ curl -fsSL https://tokenstat.ai/install.sh | bash