tokenstat
tokenstat

Llama 3.2 3B Instruct

meta-llama/llama-3.2-3b-instruct · rates as of · Benchmarks · Plans · Calculator

Open Llama instruction model for multilingual chat, reasoning, and coding

Cheapest input
$0.02
$/MTok · LiteLLM
Cheapest output
$0.02
$/MTok
Context
131K
tokens
Offers
8
3 sources

Offers

Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.

SourceProviderInputOutputCache readContext
LiteLLMcheapestmeta-llama/Llama-3.2-3B-Instruct$0.02$0.02$0.002
models.devmeta-llama/llama-3.2-3b-instructNovita AI$0.03$0.05$033K
LiteLLMmeta-llama/llama-3.2-3b-instruct$0.03$0.05$0.003
models.devmeta-llama/llama-3.2-3b-instructNanoGPT$0.0306$0.0493$0131K
models.devmeta-llama/llama-3.2-3b-instructKilo Gateway$0.05$0.33$0131K
OpenRoutermeta-llama/llama-3.2-3b-instructOpenRouter$0.05$0.33$0131K
models.devmeta-llama/Llama-3.2-3B-InstructPioneer$0.1$0.335$0.1131K
LiteLLMmeta-llama/llama-3-2-3b-instruct$0.15$0.15$0.015

Capabilities

  • Reasoning
  • Structured output
  • Attachments
  • Open weights
  • In: text
  • In: pdf
  • Out: text
Knowledge cutoff: 2023-12-31Released: 2024-09-25Family: llama

See what you actually spent

tokenstat prices your local sessions at list rates like these, with the cache split kept honest. Free, open source, runs on your machine.

$ curl -fsSL https://tokenstat.ai/install.sh | bash