Nano Gpt
Qwen3.5 122B A10B Thinking
Compact GPT model for low-latency assistance and high-volume workloads
Cheapest input
$0.25
$/MTok · LiteLLM
Cheapest output
$1.75
$/MTok
Context
260K
tokens
Offers
2
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| LiteLLMcheapestqwen3.5-122b-a10b-thinking | — | $0.25 | $1.75 | $0.025 | — |
| models.devqwen3.5-122b-a10b:thinking | NanoGPT | $0.36 | $2.88 | $0 | 260K |
Capabilities
- Reasoning
- Attachments
- In: text
- In: image
- In: video
- Out: text
Released: 2026-02-24