tokenstat
tokenstat

Qwen3 VL 32B Thinking

Qwen/Qwen3-VL-32B-Thinking · rates as of · Benchmarks · Plans · Calculator

Qwen vision-language model for visual reasoning, documents, and agent tasks

Cheapest input
$0.16
$/MTok · LiteLLM
Cheapest output
$2.87
$/MTok
Context
262K
tokens
Offers
4
2 sources

Offers

Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.

SourceProviderInputOutputCache readContext
LiteLLMcheapestdashscope/qwen3-vl-32b-thinking$0.16$2.87$0.016
LiteLLMcheapestqwen3-vl-32b-thinking$0.16$2.87$0.016
models.devQwen/Qwen3-VL-32B-ThinkingSiliconFlow (China)$0.2$1.50$0262K
models.devQwen/Qwen3-VL-32B-ThinkingSiliconFlow$0.2$1.50$0262K

Capabilities

  • Reasoning
  • Tool calling
  • Structured output
  • Attachments
  • In: text
  • In: image
  • Out: text
Released: 2025-10-21Family: qwen

See what you actually spent

tokenstat prices your local sessions at list rates like these, with the cache split kept honest. Free, open source, runs on your machine.

$ curl -fsSL https://tokenstat.ai/install.sh | bash