Llama-3.2-11B-Vision-Instruct
Open Llama multimodal model for image understanding and text reasoning
Cheapest input
$0
$/MTok · models.dev
Cheapest output
$0
$/MTok
Context
128K
tokens
Offers
3
1 source
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number. Showing free/$0 offers. Hide them.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestmeta/llama-3.2-11b-vision-instruct | GitHub Models | $0 | $0 | $0 | 128K |
| models.devmeta/llama-3.2-11b-vision-instruct | NVIDIA | $0 | $0 | $0 | 128K |
| models.devmeta/llama-3.2-11b-vision-instruct | Inference | $0.055 | $0.055 | $0 | 16K |
Capabilities
- Reasoning
- Tool calling
- Structured output
- Attachments
- Open weights
- In: text
- In: image
- In: audio
- Out: text
Knowledge cutoff: 2023-12Released: 2024-09-25Family: llama