Meta: Llama 4 Maverick 17B 128E Instruct
Open multimodal Llama model for strong reasoning and fast responses
Cheapest input
$0.35
$/MTok · models.dev
Cheapest output
$1.15
$/MTok
Context
524K
tokens
Offers
4
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestmeta/llama-4-maverick-17b-128e-instruct-maas | Google Vertex AI | $0.35 | $1.15 | $0 | 524K |
| LiteLLMmeta/llama-4-maverick-17b-128e-instruct-maas | — | $0.35 | $1.15 | $0.035 | — |
| LiteLLMllama-4-maverick-17b-128e-instruct-maas | — | $0.35 | $1.15 | $0.035 | — |
| LiteLLMvertex_ai/meta/llama-4-maverick-17b-128e-instruct-maas | — | $0.35 | $1.15 | $0.035 | — |
Capabilities
- Tool calling
- Structured output
- Attachments
- Open weights
- In: text
- In: image
- Out: text
Knowledge cutoff: 2024-08Released: 2025-04-29Family: llama