tokenstat
tokenstat

WizardLM-2 8x22B

deepinfra/microsoft/WizardLM-2-8x22B · rates as of · Benchmarks · Plans · Calculator

Open-weight instruction model for adaptable chat and self-hosted production workloads

Cheapest input
$0.48
$/MTok · LiteLLM
Cheapest output
$0.48
$/MTok
Context
66K
tokens
Offers
10
3 sources

Offers

Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.

SourceProviderInputOutputCache readContext
LiteLLMcheapestdeepinfra/microsoft/WizardLM-2-8x22B$0.48$0.48$0.048
LiteLLMcheapestmicrosoft/WizardLM-2-8x22B$0.48$0.48$0.048
LiteLLMcheapestWizardLM-2-8x22B$0.48$0.48$0.048
models.devmicrosoft/wizardlm-2-8x22bNanoGPT$0.493$0.493$066K
models.devmicrosoft/wizardlm-2-8x22bKilo Gateway$0.62$0.62$066K
models.devmicrosoft/wizardlm-2-8x22bNovita AI$0.62$0.62$066K
OpenRoutermicrosoft/wizardlm-2-8x22bOpenRouter$0.62$0.62$066K
LiteLLMmicrosoft/wizardlm-2-8x22b$0.62$0.62$0.062
LiteLLMwizardlm-2-8x22b$0.62$0.62$0.062
LiteLLMnovita/microsoft/wizardlm-2-8x22b$0.62$0.62$0.062

Capabilities

  • Attachments
  • Open weights
  • In: text
  • In: pdf
  • Out: text
Knowledge cutoff: 2024-04-30Released: 2024-04-16Family: gpt

See what you actually spent

tokenstat prices your local sessions at list rates like these, with the cache split kept honest. Free, open source, runs on your machine.

$ curl -fsSL https://tokenstat.ai/install.sh | bash