Model catalog
Models and list rates
Published API prices, in US dollars per million tokens, so you can compare models. tokenstat does not charge these rates. Your subscription bill is separate. The table also shows who lists each model, and whether it can do tools, vision, or reasoning. Recommended order puts known labs first, then newer releases. Free and $0 rows stay hidden unless you ask. Measure real sessions with the app or the CLI.
Showing 48 of 2,421 matches · page 23 of 51 · 97 beta/latest variants hidden (show them) · including free and $0 prices
| Model | Input | Output | Context | Offers |
|---|---|---|---|---|
Meta: Llama 2 70b Chat Hfmeta-llama/Llama-2-70b-chat-hf | $1LiteLLM | $1 | N/A | 3 |
Meta: Llama 3 70bmeta/llama-3-70b | $0.59LiteLLM | $0.79 | N/A | 4 |
Meta: Llama 3 70B Instructmeta-llama/Meta-Llama-3-70B-Instruct | $0.12LiteLLM | $0.3 | N/A | 7 |
Meta: Llama 3 70B Instruct Turbometa-llama/Meta-Llama-3-70B-Instruct-Turbo | $0.88LiteLLM | $0.88 | N/A | 3 |
NVIDIA: Nemotron 3 Embed 1B (free)nvidia/nemotron-3-embed-1b-20260716 | $0OpenRouter | $0 | 33K | 1 |
OpenAI: Gpt 4 Turbo Previewopenai/gpt-4-turbo-preview | $10LiteLLM | $30 | N/A | 3 |
Deepseek R1 0528 Distill Qwen3 8bfireworks_ai/accounts/fireworks/models/deepseek-r1-0528-distill-qwen3-8b | $0.2LiteLLM | $0.2 | N/A | 3 |
Google: Flan T5 Xl 3bgoogle/flan-t5-xl-3b | $0.6LiteLLM | $0.6 | N/A | 3 |
Google: Gemma 7b Itgoogle/gemma-7b-it | $0.05LiteLLM | $0.08 | N/A | 6 |
Google: Gemma 7b It Loragoogle/gemma-7b-it-lora | $0LiteLLM | $0 | N/A | 3 |
Meta: Llama3 8b Instruct Maasmeta/llama3-8b-instruct-maas | $0LiteLLM | $0 | N/A | 3 |
Mistral: 7b Instructmistralai/mistral-7b-instruct | $0.07LiteLLM | $0.28 | N/A | 4 |
Mistral: 7b Instruct V0.1mistral/mistral-7b-instruct-v0.1 | $0.15LiteLLM | $0.15 | N/A | 9 |
Mistral: 7b Instruct V0.2 Loramistral/mistral-7b-instruct-v0.2-lora | $0LiteLLM | $0 | N/A | 3 |
Mistral: 7b V0.1mistralai/mistral-7b-v0.1 | $0.05LiteLLM | $0.25 | N/A | 3 |
Mistral: Ministral 3 14b 2512mistral/ministral-3-14b-2512 | $0.2LiteLLM | $0.2 | N/A | 2 |
Mistral: Ministral 3 3b 2512mistral/ministral-3-3b-2512 | $0.1LiteLLM | $0.1 | N/A | 2 |
Mistral: Ministral 3 8b 2512mistral/ministral-3-8b-2512 | $0.15LiteLLM | $0.15 | N/A | 2 |
Perplexity: Pplx 7b Chatperplexity/pplx-7b-chat | $0.07LiteLLM | $0.28 | N/A | 2 |
Perplexity: Pplx 7b Onlineperplexity/pplx-7b-online | $0LiteLLM | $0.28 | N/A | 2 |
Code Llama 7bfireworks_ai/accounts/fireworks/models/code-llama-7b | $0.2LiteLLM | $0.2 | N/A | 3 |
Code Llama 7b Instructfireworks_ai/accounts/fireworks/models/code-llama-7b-instruct | $0.2LiteLLM | $0.2 | N/A | 3 |
Code Llama 7b Pythonfireworks_ai/accounts/fireworks/models/code-llama-7b-python | $0.2LiteLLM | $0.2 | N/A | 3 |
Code Qwen 1p5 7bfireworks_ai/accounts/fireworks/models/code-qwen-1p5-7b | $0.2LiteLLM | $0.2 | N/A | 3 |
Codegemma 7bfireworks_ai/accounts/fireworks/models/codegemma-7b | $0.2LiteLLM | $0.2 | N/A | 3 |
Cogito V1 Preview Llama 3bfireworks_ai/accounts/fireworks/models/cogito-v1-preview-llama-3b | $0.1LiteLLM | $0.1 | N/A | 3 |
Cogito V1 Preview Llama 8bfireworks_ai/accounts/fireworks/models/cogito-v1-preview-llama-8b | $0.2LiteLLM | $0.2 | N/A | 3 |
Cogito V1 Preview Qwen 14bfireworks_ai/accounts/fireworks/models/cogito-v1-preview-qwen-14b | $0.2LiteLLM | $0.2 | N/A | 3 |
Deepseek Coder 1b Basefireworks_ai/accounts/fireworks/models/deepseek-coder-1b-base | $0.1LiteLLM | $0.1 | N/A | 3 |
Deepseek Coder 7b Basefireworks_ai/accounts/fireworks/models/deepseek-coder-7b-base | $0.2LiteLLM | $0.2 | N/A | 3 |
Deepseek Coder 7b Base V1p5fireworks_ai/accounts/fireworks/models/deepseek-coder-7b-base-v1p5 | $0.2LiteLLM | $0.2 | N/A | 3 |
Deepseek Coder 7b Instruct V1p5fireworks_ai/accounts/fireworks/models/deepseek-coder-7b-instruct-v1p5 | $0.2LiteLLM | $0.2 | N/A | 3 |
Gemma 7bfireworks_ai/accounts/fireworks/models/gemma-7b | $0.2LiteLLM | $0.2 | N/A | 3 |
Hermes 2 Pro Mistral 7bfireworks_ai/accounts/fireworks/models/hermes-2-pro-mistral-7b | $0.2LiteLLM | $0.2 | N/A | 3 |
Internvl3 8bfireworks_ai/accounts/fireworks/models/internvl3-8b | $0.2LiteLLM | $0.2 | N/A | 3 |
Llama Guard 2 8bfireworks_ai/accounts/fireworks/models/llama-guard-2-8b | $0.2LiteLLM | $0.2 | N/A | 3 |
Llama Guard 3 1bfireworks_ai/accounts/fireworks/models/llama-guard-3-1b | $0.1LiteLLM | $0.1 | N/A | 3 |
Llama V2 7bfireworks_ai/accounts/fireworks/models/llama-v2-7b | $0.2LiteLLM | $0.2 | N/A | 3 |
Llama V2 7b Chatfireworks_ai/accounts/fireworks/models/llama-v2-7b-chat | $0.2LiteLLM | $0.2 | N/A | 3 |
Llama V3 8bfireworks_ai/accounts/fireworks/models/llama-v3-8b | $0.2LiteLLM | $0.2 | N/A | 3 |
Llama V3 8b Instruct Hffireworks_ai/accounts/fireworks/models/llama-v3-8b-instruct-hf | $0.2LiteLLM | $0.2 | N/A | 3 |
Llama V3p1 70b Instruct 1bfireworks_ai/accounts/fireworks/models/llama-v3p1-70b-instruct-1b | $0.1LiteLLM | $0.1 | N/A | 3 |
Llama V3p1 8b Instructfireworks_ai/accounts/fireworks/models/llama-v3p1-8b-instruct | $0.1LiteLLM | $0.1 | N/A | 3 |
Llama V3p2 1bfireworks_ai/accounts/fireworks/models/llama-v3p2-1b | $0.1LiteLLM | $0.1 | N/A | 3 |
Llama V3p2 1b Instructfireworks_ai/accounts/fireworks/models/llama-v3p2-1b-instruct | $0.1LiteLLM | $0.1 | N/A | 3 |
Llama V3p2 3bfireworks_ai/accounts/fireworks/models/llama-v3p2-3b | $0.1LiteLLM | $0.1 | N/A | 3 |
Llama V3p2 3b Instructfireworks_ai/accounts/fireworks/models/llama-v3p2-3b-instruct | $0.1LiteLLM | $0.1 | N/A | 3 |
Llamaguard 7bfireworks_ai/accounts/fireworks/models/llamaguard-7b | $0.2LiteLLM | $0.2 | N/A | 3 |
See these prices on your own sessions
tokenstat reads the logs your coding tools already write and prices the tokens at the published API rate. Work your subscription already covered is not an extra bill.