tokenstat
tokenstat

Meta: Llama 3.2 3B

meta/llama-3.2-3b · rates as of · Benchmarks · Plans · Calculator

Open Llama instruction model for multilingual chat, reasoning, and coding

Cheapest input
$0.04
$/MTok · LiteLLM
Cheapest output
$0.08
$/MTok
Context
128K
tokens
Offers
5
2 sources

Offers

Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.

SourceProviderInputOutputCache readContext
LiteLLMcheapestllamagate/llama-3.2-3b$0.04$0.08$0.004
models.devllama-3.2-3bVenice AI$0.15$0.6$0128K
LiteLLMmeta/llama-3.2-3b$0.15$0.15$0.015
LiteLLMllama-3.2-3b$0.15$0.15$0.015
LiteLLMvercel_ai_gateway/meta/llama-3.2-3b$0.15$0.15$0.015

Capabilities

  • Tool calling
  • Open weights
  • In: text
  • Out: text
Released: 2024-10-03Family: llama

See what you actually spent

tokenstat prices your local sessions at list rates like these, with the cache split kept honest. Free, open source, runs on your machine.

$ curl -fsSL https://tokenstat.ai/install.sh | bash