Llama-4-Maverick-17B-128E-Instruct-FP8
Open multimodal Llama model for strong reasoning and fast responses
Cheapest input
$0.2
$/MTok · LiteLLM
Cheapest output
$0.6
$/MTok
Context
128K
tokens
Offers
1
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number. Free/$0 listings are hidden by default (1 omitted). Show free/$0 offers.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| LiteLLMcheapestmeta/llama-4-maverick | — | $0.2 | $0.6 | $0.02 | — |
Capabilities
- Tool calling
- Attachments
- Open weights
- In: text
- In: image
- Out: text
Knowledge cutoff: 2024-08Released: 2025-04-05Family: llama
Scores
Sparse coverage on purpose. Missing a score means we have not linked one yet, not that the model is unranked. Full boards on /benchmarks.
- AA intelligence14.3
- AA coding16.3
- AA agentic1.3
Sources: Arena AI, Artificial Analysis. See /benchmarks.