Venice
Hermes 3 405B Instruct
Open Llama instruction model for multilingual chat, reasoning, and coding
Cheapest input
$1
$/MTok · models.dev
Cheapest output
$1
$/MTok
Context
131K
tokens
Offers
7
3 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestnousresearch/hermes-3-llama-3.1-405b | Kilo Gateway | $1 | $1 | $0 | 131K |
| OpenRouternousresearch/hermes-3-llama-3.1-405b | OpenRouter | $1 | $1 | $0 | 131K |
| LiteLLMHermes-3-Llama-3.1-405B | — | $1 | $1 | $0.1 | — |
| LiteLLMNousResearch/Hermes-3-Llama-3.1-405B | — | $1 | $1 | $0.1 | — |
| LiteLLMdeepinfra/NousResearch/Hermes-3-Llama-3.1-405B | — | $1 | $1 | $0.1 | — |
| LiteLLMnebius/NousResearch/Hermes-3-Llama-3.1-405B | — | $1 | $3 | $0.1 | — |
| models.devhermes-3-llama-3.1-405b | Venice AI | $1.10 | $3 | $0 | 128K |
Capabilities
- Structured output
- Open weights
- In: text
- Out: text
Knowledge cutoff: 2023-12-31Released: 2025-09-25Family: hermes