tokenstat
tokenstat

Model catalog

Models and list rates

Published API list rates (USD per million tokens), who lists each model, and whether it can do tools, vision, or reasoning. Recommended order is newer first; among similarly new models, larger context wins. Free/$0 rows stay hidden unless you ask. Not a bill: measure real sessions with the CLI.

3,231 models in the merged catalog · as of . Sources: LiteLLM, OpenRouter, models.dev. · Providers · Plans · Benchmarks

Showing 30 of 2,670 matches · page 56 of 56 · 561 free/$0 rows hidden (show them)

ModelInputOutputContextOffers
Meta-Llama-3-8B-Instructmeta-llama/Meta-Llama-3-8B-Instruct
$0.15LiteLLM$0.15N/A1
Meta: Llama 2 7bmeta/llama-2-7b
$0.05LiteLLM$0.25N/A3
Meta: Llama 2 7b Chatmeta/llama-2-7b-chat
$0.05LiteLLM$0.25N/A3
Meta: Llama 2 7b Chat Hfmeta-llama/Llama-2-7b-chat-hf
$0.15LiteLLM$0.15N/A3
Meta: Llama 3 8bmeta/llama-3-8b
$0.05LiteLLM$0.08N/A4
DeepSeek R1 Distill Llama 8Bfireworks_ai/accounts/fireworks/models/deepseek-r1-distill-llama-8b
$0.2LiteLLM$0.2N/A1
DeepSeek R1 Distill Llama 8Baccounts/fireworks/models/deepseek-r1-distill-llama-8b
$0.2LiteLLM$0.2N/A1
DeepSeek R1 Distill Llama 8Bnscale/deepseek-ai/DeepSeek-R1-Distill-Llama-8B
$0.025LiteLLM$0.025N/A1
Llama 3 8B Instructllama-3-8b-instruct
$0.05LiteLLM$0.25N/A1
Llama 3 8B Instructnovita/meta-llama/llama-3-8b-instruct
$0.04LiteLLM$0.04N/A1
Llama 3 8B Instructreplicate/meta/llama-3-8b-instruct
$0.05LiteLLM$0.25N/A1
Llama 3.2 1B Instructllama-3-2-1b-instruct
$0.027LiteLLM$0.201N/A2
Llama 3.2 1B Instructwatsonx/meta-llama/llama-3-2-1b-instruct
$0.1LiteLLM$0.1N/A1
Llama 3.2 3B Instructwatsonx/meta-llama/llama-3-2-3b-instruct
$0.15LiteLLM$0.15N/A1
Meta-Llama-3-8B-Instructanyscale/meta-llama/Meta-Llama-3-8B-Instruct
$0.15LiteLLM$0.15N/A1
Meta-Llama-3-8B-Instructdeepinfra/meta-llama/Meta-Llama-3-8B-Instruct
$0.03LiteLLM$0.06N/A1
Meta-Llama-3-8B-InstructMeta-Llama-3-8B-Instruct
$0.15LiteLLM$0.15N/A1
Codellama 7b Instruct Awqcloudflare/@hf/thebloke/codellama-7b-instruct-awq
$1.92LiteLLM$1.92N/A3
DeepSeek R1 Distill Qwen 1.5Bdeepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
$0.09LiteLLM$0.09N/A1
DeepSeek R1 Distill Qwen 14Bdeepseek-ai/DeepSeek-R1-Distill-Qwen-14B
$0.07LiteLLM$0.07N/A1
DeepSeek R1 Distill Qwen 7Bdeepseek-ai/DeepSeek-R1-Distill-Qwen-7B
$0.2LiteLLM$0.2N/A1
DeepSeek R1 Distill Qwen 14Bfireworks_ai/accounts/fireworks/models/deepseek-r1-distill-qwen-14b
$0.2LiteLLM$0.2N/A1
DeepSeek R1 Distill Qwen 7Bfireworks_ai/accounts/fireworks/models/deepseek-r1-distill-qwen-7b
$0.2LiteLLM$0.2N/A1
DeepSeek R1 Distill Qwen 1.5Bnscale/deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
$0.09LiteLLM$0.09N/A1
DeepSeek R1 Distill Qwen 14Baccounts/fireworks/models/deepseek-r1-distill-qwen-14b
$0.2LiteLLM$0.2N/A1
DeepSeek R1 Distill Qwen 14Bnovita/deepseek/deepseek-r1-distill-qwen-14b
$0.15LiteLLM$0.15N/A1
DeepSeek R1 Distill Qwen 14Bnscale/deepseek-ai/DeepSeek-R1-Distill-Qwen-14B
$0.07LiteLLM$0.07N/A1
DeepSeek R1 Distill Qwen 7Baccounts/fireworks/models/deepseek-r1-distill-qwen-7b
$0.2LiteLLM$0.2N/A1
DeepSeek R1 Distill Qwen 7Bnscale/deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
$0.2LiteLLM$0.2N/A1
Llama 2 7b Chat Int8cloudflare/@cf/meta/llama-2-7b-chat-int8
$1.92LiteLLM$1.92N/A3

Price this against your own usage

List rates are only half the story. The CLI reads the session logs your coding harnesses already write and shows tokens, cache share, and usage value on your machine.

$ curl -fsSL https://tokenstat.ai/install.sh | bash