NVIDIA Nemotron Nano 9B v2
Compact Nemotron model for efficient reasoning and deployable AI agents
Cheapest input
$0.04
$/MTok · LiteLLM
Cheapest output
$0.16
$/MTok
Context
131K
tokens
Offers
9
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number. Free/$0 listings are hidden by default (1 omitted). Show free/$0 offers.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| LiteLLMcheapestnvidia/NVIDIA-Nemotron-Nano-9B-v2 | — | $0.04 | $0.16 | $0.004 | — |
| LiteLLMcheapestNVIDIA-Nemotron-Nano-9B-v2 | — | $0.04 | $0.16 | $0.004 | — |
| LiteLLMcheapestdeepinfra/nvidia/NVIDIA-Nemotron-Nano-9B-v2 | — | $0.04 | $0.16 | $0.004 | — |
| models.devnvidia.nemotron-nano-9b-v2 | Amazon Bedrock | $0.06 | $0.23 | $0 | 128K |
| LiteLLMnvidia.nemotron-nano-9b-v2 | — | $0.06 | $0.23 | $0.006 | — |
| models.devnvidia/nvidia-nemotron-nano-9b-v2 | NanoGPT | $0.17 | $0.68 | $0 | 128K |
| LiteLLMnvidia-nemotron-nano-9b-v2 | — | $0.2 | $0.2 | $0.02 | — |
| LiteLLMaccounts/fireworks/models/nvidia-nemotron-nano-9b-v2 | — | $0.2 | $0.2 | $0.02 | — |
| LiteLLMfireworks_ai/accounts/fireworks/models/nvidia-nemotron-nano-9b-v2 | — | $0.2 | $0.2 | $0.02 | — |
Capabilities
- Reasoning
- Tool calling
- Structured output
- Open weights
- In: text
- Out: text
Knowledge cutoff: 2024-09Released: 2025-08-18Family: nemotron