Llama 4 Maverick 17B Instruct
Open multimodal Llama for strong reasoning with efficient everyday serving
Cheapest input
$0.14
$/MTok · models.dev
Cheapest output
$0.59
$/MTok
Context
1.0M
tokens
Offers
7
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestmeta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8 | Abacus | $0.14 | $0.59 | $0 | 1.0M |
| models.devmeta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8 | IO.NET | $0.15 | $0.6 | $0.075 | 430K |
| LiteLLMdeepinfra/meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8 | — | $0.15 | $0.6 | $0.015 | — |
| LiteLLMmeta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8 | — | $0.15 | $0.6 | $0.015 | — |
| models.devmeta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8 | DeepInfra | $0.2 | $0.8 | $0 | 1.0M |
| models.devmeta-llama/llama-4-maverick-17b-128e-instruct-fp8 | Novita AI | $0.27 | $0.85 | $0 | 1.0M |
| LiteLLMmeta-llama/llama-4-maverick-17b-128e-instruct-fp8 | — | $0.27 | $0.85 | $0.027 | — |
Capabilities
- Tool calling
- Structured output
- Attachments
- Open weights
- In: text
- In: image
- Out: text
Knowledge cutoff: 2024-08Released: 2025-04-05Family: llama