Google: Gemini 2.0 Flash Lite
Low-latency Gemini model for high-volume multimodal and agent workloads
Cheapest input
$0.052
$/MTok · models.dev
Cheapest output
$0.21
$/MTok
Context
2M
tokens
Offers
6
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestgoogle/gemini-2.0-flash-lite | Poe | $0.052 | $0.21 | $0 | 990K |
| models.devgemini-2.0-flash-lite | 302.AI | $0.075 | $0.3 | $0 | 2M |
| models.devgemini-2.0-flash-lite | $0.075 | $0.3 | $0.0075 | 1.0M | |
| LiteLLMgemini/gemini-2.0-flash-lite | — | $0.075 | $0.3 | $0.0187 | — |
| LiteLLMgemini-2.0-flash-lite | — | $0.075 | $0.3 | $0.0187 | — |
| LiteLLMvercel_ai_gateway/google/gemini-2.0-flash-lite | — | $0.075 | $0.3 | $0.0075 | — |
Capabilities
- Tool calling
- Structured output
- Attachments
- In: text
- In: image
- In: audio
- In: video
- In: pdf
- Out: text
Knowledge cutoff: 2024-11Released: 2025-06-16Family: gemini-flash-lite