Llama 3.3 Nemotron Super 49B v1.5
Nemotron model for efficient reasoning, coding, and specialized AI agents
Cheapest input
$0.05
$/MTok · models.dev
Cheapest output
$0.25
$/MTok
Context
131K
tokens
Offers
5
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number. Free/$0 listings are hidden by default (1 omitted). Show free/$0 offers.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestnvidia/Llama-3_3-Nemotron-Super-49B-v1_5 | NanoGPT | $0.05 | $0.25 | $0 | 128K |
| LiteLLMdeepinfra/nvidia/Llama-3.3-Nemotron-Super-49B-v1.5 | — | $0.1 | $0.4 | $0.01 | — |
| LiteLLMnvidia/Llama-3.3-Nemotron-Super-49B-v1.5 | — | $0.1 | $0.4 | $0.01 | — |
| LiteLLMLlama-3.3-Nemotron-Super-49B-v1.5 | — | $0.1 | $0.4 | $0.01 | — |
| models.devnvidia/Llama-3.3-Nemotron-Super-49B-v1.5 | DeepInfra | $0.4 | $0.4 | $0 | 131K |
Capabilities
- Reasoning
- Tool calling
- Structured output
- Open weights
- In: text
- Out: text
Released: 2025-07-25Family: nemotron