tokenstat
tokenstat

Model catalog

Models and list rates

Published API list rates (USD per million tokens), who lists each model, and if it can do tools, vision, or reasoning. Recommended order puts known labs first, then newer releases. Free/$0 rows stay hidden unless you ask. Measure real sessions with the app or the CLI.

2,615 models in the merged catalog · as of . Sources: LiteLLM, OpenRouter, models.dev. · Providers · AI plans · Benchmarks

Showing 48 of 1,040 matches · page 18 of 22 · 322 free/$0 rows hidden (show them) · 93 beta/latest variants hidden (show them)

ModelInputOutputContextOffers
Mistral: Saba 24bmistral/mistral-saba-24b
$0.79LiteLLM$0.79N/A3
Mistral: Tinymistral/mistral-tiny
$0.25LiteLLM$0.25N/A2
Mistral: Vibe Cli Fastmistral/mistral-vibe-cli-fast
$0.15LiteLLM$0.6N/A2
Mistral: Vibe Cli With Toolsmistral/mistral-vibe-cli-with-tools
$1.50LiteLLM$7.50N/A2
Mistral: Voxtral Mini Transcribemistralai/voxtral-mini-transcribe
$50OpenRouter$0N/A1
Mistral: Voxtral Small 24B 2507 STTmistralai/voxtral-small-24b-2507-stt
$50OpenRouter$0N/A1
Qwen-Audio-3.0-TTS Flashqwen/qwen-audio-3.0-tts-flash
$15OpenRouter$0N/A1
Qwen-Audio-3.0-TTS Plusqwen/qwen-audio-3.0-tts-plus
$20OpenRouter$0N/A1
Qwen2.5 Coder 3B InstructQwen/Qwen2.5-Coder-3B-Instruct
$0.01LiteLLM$0.03N/A6
Qwen2.5 Coder 7BQwen/Qwen2.5-Coder-7B
$0.01LiteLLM$0.03N/A8
MiniMax: Speech 2.8 HDminimax/speech-2.8-hd
$100OpenRouter$0N/A1
MiniMax: Speech 2.8 Turbominimax/speech-2.8-turbo
$60OpenRouter$0N/A1
Cohere: Command Rcohere/command-r
$0.15LiteLLM$0.6N/A3
Cohere: Command R Pluscohere/command-r-plus
$2.50LiteLLM$10N/A4
Cohere: Embed V4.0cohere/embed-v4.0
$0.12LiteLLM$0N/A9
Perplexity: Codellama 34b Instructperplexity/codellama-34b-instruct
$0.35LiteLLM$1.40N/A2
Perplexity: Codellama 70b Instructperplexity/codellama-70b-instruct
$0.7LiteLLM$2.80N/A2
Perplexity: Pplx 70b Chatperplexity/pplx-70b-chat
$0.7LiteLLM$2.80N/A2
Perplexity: Pplx 70b Onlineperplexity/pplx-70b-online
$0LiteLLM$2.80N/A2
Perplexity: Pplx Embed Context V1 0.6bperplexity/pplx-embed-context-v1-0.6b
$0.008LiteLLM$0N/A2
Perplexity: Pplx Embed Context V1 4bperplexity/pplx-embed-context-v1-4b
$0.05LiteLLM$0N/A2
Perplexity: Sonar Medium Chatperplexity/sonar-medium-chat
$0.6LiteLLM$1.80N/A2
Perplexity: Sonar Medium Onlineperplexity/sonar-medium-online
$0LiteLLM$1.80N/A2
Perplexity: Sonar Small Chatperplexity/sonar-small-chat
$0.07LiteLLM$0.28N/A2
Perplexity: Sonar Small Onlineperplexity/sonar-small-online
$0LiteLLM$0.28N/A2
NVIDIA: Cosmos3 Super Reasonernvidia/Cosmos3-Super-Reasoner
$0.1LiteLLM$0.3N/A3
NVIDIA: Nemotron 3.5 ASR Streaming Multilingual 0.6Bnvidia/nemotron-3.5-asr-streaming-multilingual-0.6b
$3.33OpenRouter$0N/A1
NVIDIA: Nemotron Content Safety 3.5nvidia/Nemotron-Content-Safety-3.5
$0.2LiteLLM$0.2N/A3
NVIDIA: Parakeet TDT 0.6B v3nvidia/parakeet-tdt-0.6b-v3
$25OpenRouter$0N/A1
Amazon: Titan Embed Text V2amazon/titan-embed-text-v2
$0.02LiteLLM$0N/A3
Chronos Hermes 13b V2fireworks_ai/accounts/fireworks/models/chronos-hermes-13b-v2
$0.2LiteLLM$0.2N/A3
Code Llama 13bfireworks_ai/accounts/fireworks/models/code-llama-13b
$0.2LiteLLM$0.2N/A3
Code Llama 13b Instructfireworks_ai/accounts/fireworks/models/code-llama-13b-instruct
$0.2LiteLLM$0.2N/A3
Code Llama 13b Pythonfireworks_ai/accounts/fireworks/models/code-llama-13b-python
$0.2LiteLLM$0.2N/A3
Code Llama 34bfireworks_ai/accounts/fireworks/models/code-llama-34b
$0.9LiteLLM$0.9N/A3
Code Llama 34b Instructfireworks_ai/accounts/fireworks/models/code-llama-34b-instruct
$0.9LiteLLM$0.9N/A3
Code Llama 34b Pythonfireworks_ai/accounts/fireworks/models/code-llama-34b-python
$0.9LiteLLM$0.9N/A3
Code Llama 70bfireworks_ai/accounts/fireworks/models/code-llama-70b
$0.9LiteLLM$0.9N/A3
Code Llama 70b Instructfireworks_ai/accounts/fireworks/models/code-llama-70b-instruct
$0.9LiteLLM$0.9N/A3
Code Llama 70b Pythonfireworks_ai/accounts/fireworks/models/code-llama-70b-python
$0.9LiteLLM$0.9N/A3
Codegemma 2bfireworks_ai/accounts/fireworks/models/codegemma-2b
$0.1LiteLLM$0.1N/A3
Cogito 671b V2 P1fireworks_ai/accounts/fireworks/models/cogito-671b-v2-p1
$1.20LiteLLM$1.20N/A3
Cogito V1 Preview Llama 70bfireworks_ai/accounts/fireworks/models/cogito-v1-preview-llama-70b
$0.9LiteLLM$0.9N/A3
Dbrx Instructfireworks_ai/accounts/fireworks/models/dbrx-instruct
$1.20LiteLLM$1.20N/A3
Deepseek Coder V2 Instructfireworks_ai/accounts/fireworks/models/deepseek-coder-v2-instruct
$1.20LiteLLM$1.20N/A4
Deepseek Coder V2 Lite Basefireworks_ai/accounts/fireworks/models/deepseek-coder-v2-lite-base
$0.5LiteLLM$0.5N/A4
Deepseek Coder V2 Lite Instructfireworks_ai/accounts/fireworks/models/deepseek-coder-v2-lite-instruct
$0.5LiteLLM$0.5N/A4
Deepseek Prover V2fireworks_ai/accounts/fireworks/models/deepseek-prover-v2
$1.20LiteLLM$1.20N/A3

Set up a workspace and measure your agents

Download tokenstat to set up local Git-enabled workspaces, bring remote terminals and desktops into one window, and track the tokens from your own sessions.

CLI for macOS, Linux, and Windows