Llama 3.1 Nemotron 70B Instruct
Nemotron model for efficient reasoning, coding, and specialized AI agents
Cheapest input
$0
$/MTok · models.dev
Cheapest output
$0
$/MTok
Context
128K
tokens
Offers
4
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number. Showing free/$0 offers. Hide them.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestnvidia/llama-3.1-nemotron-70b-instruct | NVIDIA | $0 | $0 | $0 | 128K |
| LiteLLMnvidia/Llama-3.1-Nemotron-70B-Instruct | — | $0.6 | $0.6 | $0.06 | — |
| LiteLLMLlama-3.1-Nemotron-70B-Instruct | — | $0.6 | $0.6 | $0.06 | — |
| LiteLLMdeepinfra/nvidia/Llama-3.1-Nemotron-70B-Instruct | — | $0.6 | $0.6 | $0.06 | — |
Capabilities
- Tool calling
- Open weights
- In: text
- Out: text
Released: 2025-04-15Family: nemotron