Llama 3.1 Nemotron 70B Instruct
Nemotron model for efficient reasoning, coding, and specialized AI agents
Cheapest input
$0.6
$/MTok · LiteLLM
Cheapest output
$0.6
$/MTok
Context
128K
tokens
Offers
3
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number. Free/$0 listings are hidden by default (1 omitted). Show free/$0 offers.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| LiteLLMcheapestnvidia/Llama-3.1-Nemotron-70B-Instruct | — | $0.6 | $0.6 | $0.06 | — |
| LiteLLMcheapestLlama-3.1-Nemotron-70B-Instruct | — | $0.6 | $0.6 | $0.06 | — |
| LiteLLMcheapestdeepinfra/nvidia/Llama-3.1-Nemotron-70B-Instruct | — | $0.6 | $0.6 | $0.06 | — |
Capabilities
- Tool calling
- Open weights
- In: text
- Out: text
Released: 2025-04-15Family: nemotron