llama-3_2-nemoretriever-300m-embed-v1
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
Cheapest input
$0
$/MTok · models.dev
Cheapest output
$0
$/MTok
Context
33K
tokens
Offers
1
1 source
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number. Showing free/$0 offers. Hide them.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestnvidia/llama-3_2-nemoretriever-300m-embed-v1 | NVIDIA | $0 | $0 | $0 | 33K |
Capabilities
- Open weights
- In: text
- Out: text
Released: 2025-07-24