tokenstat
tokenstat

Qwen3 VL 32B Instruct

qwen/qwen3-vl-32b-instruct · rates as of · Benchmarks · Plans · Calculator

Qwen vision-language model for visual reasoning, documents, and agent tasks

Cheapest input
$0.104
$/MTok · models.dev
Cheapest output
$0.416
$/MTok
Context
262K
tokens
Offers
8
3 sources

Offers

Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.

SourceProviderInputOutputCache readContext
models.devcheapestqwen/qwen3-vl-32b-instructKilo Gateway$0.104$0.416$0131K
OpenRouterqwen/qwen3-vl-32b-instructOpenRouter$0.104$0.416$0131K
LiteLLMdashscope/qwen3-vl-32b-instruct$0.16$0.64$0.016
LiteLLMqwen3-vl-32b-instruct$0.16$0.64$0.016
models.devQwen/Qwen3-VL-32B-InstructSiliconFlow (China)$0.2$0.6$0262K
models.devQwen/Qwen3-VL-32B-InstructSiliconFlow$0.2$0.6$0262K
LiteLLMaccounts/fireworks/models/qwen3-vl-32b-instruct$0.9$0.9$0.09
LiteLLMfireworks_ai/accounts/fireworks/models/qwen3-vl-32b-instruct$0.9$0.9$0.09

Capabilities

  • Tool calling
  • Structured output
  • Attachments
  • Open weights
  • In: text
  • In: image
  • Out: text
Released: 2025-10-23Family: qwen

See what you actually spent

tokenstat prices your local sessions at list rates like these, with the cache split kept honest. Free, open source, runs on your machine.

$ curl -fsSL https://tokenstat.ai/install.sh | bash