Nano Gpt
Phi 4 Mini
Compact GPT model for low-latency assistance and high-volume workloads
Cheapest input
$0
$/MTok · models.dev
Cheapest output
$0
$/MTok
Context
131K
tokens
Offers
8
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number. Showing free/$0 offers. Hide them.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestmicrosoft/phi-4-mini-instruct | GitHub Models | $0 | $0 | $0 | 128K |
| models.devmicrosoft/phi-4-mini-instruct | NVIDIA | $0 | $0 | $0 | 131K |
| LiteLLMPhi-4-mini-instruct | — | $0.075 | $0.3 | $0.0075 | — |
| LiteLLMazure_ai/Phi-4-mini-instruct | — | $0.075 | $0.3 | $0.0075 | — |
| models.devmicrosoft/Phi-4-mini-instruct | Weights & Biases | $0.08 | $0.35 | $0.08 | 128K |
| models.devphi-4-mini-instruct | NanoGPT | $0.17 | $0.68 | $0 | 128K |
| LiteLLMmicrosoft/Phi-4-mini-instruct | — | $8000 | $35000 | $800 | — |
| LiteLLMwandb/microsoft/Phi-4-mini-instruct | — | $8000 | $35000 | $800 | — |
Capabilities
- Reasoning
- Tool calling
- Structured output
- Open weights
- In: text
- Out: text
Knowledge cutoff: 2023-10Released: 2025-07-26Family: phi