Llama 3.1 8B
Compact Llama instruction model for fast chat and local deployment
Cheapest input
$0.05
$/MTok · models.dev
Cheapest output
$0.08
$/MTok
Context
131K
tokens
Offers
3
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestllama-3.1-8b-instant | Groq | $0.05 | $0.08 | $0.005 | 131K |
| models.devllama-3.1-8b-instant | Helicone | $0.05 | $0.08 | $0 | 131K |
| LiteLLMllama-3.1-8b-instant | — | $0.05 | $0.08 | $0.005 | — |
Capabilities
- Tool calling
- Open weights
- In: text
- Out: text
Knowledge cutoff: 2023-12Released: 2024-07-23Family: llama
Scores
Sparse coverage on purpose. Missing a score means we have not linked one yet, not that the model is unranked. Full boards on /benchmarks.
- AA intelligence7.6
- AA coding5.4
- AA agentic0.5
Sources: Arena AI, Artificial Analysis. See /benchmarks.