Nano Gpt
Yi Large
Compact GPT model for low-latency assistance and high-volume workloads
Cheapest input
$3
$/MTok · LiteLLM
Cheapest output
$3
$/MTok
Context
32K
tokens
Offers
4
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| LiteLLMcheapestyi-large | — | $3 | $3 | $0.3 | — |
| LiteLLMcheapestaccounts/fireworks/models/yi-large | — | $3 | $3 | $0.3 | — |
| LiteLLMcheapestfireworks_ai/accounts/fireworks/models/yi-large | — | $3 | $3 | $0.3 | — |
| models.devyi-large | NanoGPT | $3.20 | $3.20 | $0 | 32K |
Capabilities
- In: text
- Out: text
Released: 2024-05-13