Requesty
inkling-256k
Inkling 256K is the extended context variant of Inkling, a large MoE hybrid reasoning model from Thinking Machines with audio and vision input support and a 256K context window.
Cheapest input
$1.87
$/MTok · models.dev
Cheapest output
$4.68
$/MTok
Context
262K
tokens
Sources
1
1 listing
Price changes
Cheapest list input/output only. Rates move rarely, so this is an event log. Down means cheaper. Showing the 8 latest changes.
- ↑ higherOutput $3.74 → $4.68
- ↑ higherInput $1.50 → $1.87
Sources
Same model, different sources. Our first-party rate is listed first, then router and aggregator rows stay separate so you can see the spread instead of a single blended number.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestinkling-256k | Requesty | $1.87 | $4.68 | $0.374 | 262K |
Capabilities
- Reasoning
- Tool calling
- Attachments
- In: text
- In: image
- Out: text
Released: 2026-07-16
Run this model in a real workspace
Launch agents from the project they work in, use remote management to reach another machine, and see this model's sessions and token usage in your own history. Free to use, runs on your device.
Data and list-price estimates may contain mistakes. Not investment advice.