Gemini 3.1 Flash Live Preview
High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications
Cheapest input
$0.75
$/MTok · models.dev
Cheapest output
$4.50
$/MTok
Context
131K
tokens
Offers
2
2 sources
Offers
Same model, different listings. First-party, router, and aggregator rows stay separate so you can see the spread instead of a single blended number.
| Source | Provider | Input | Output | Cache read | Context |
|---|---|---|---|---|---|
| models.devcheapestgemini-3.1-flash-live-preview | $0.75 | $4.50 | $0 | 131K | |
| LiteLLMgemini-3.1-flash-live-preview | — | $0.75 | $4.50 | $0.075 | — |
Capabilities
- Reasoning
- Tool calling
- Attachments
- In: text
- In: image
- In: video
- In: audio
- Out: text
- Out: audio
Knowledge cutoff: 2025-01Released: 2026-03-26Family: gemini-flash