Gemini 2.5 Flash Lite
GoogleGemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
In
$0.10
Out
$0.40
Blended
$0.18
Where to buy this model
2 access routes| Route | In | Out | Blended | vs |
|---|---|---|---|---|
| ● Google AI Studio cheapest | $0.05 | $0.20 | $0.09 | - |
| $0.10 | $0.40 | $0.18 | +100% |
● official API · ● cloud platform · ● inference host · USD per 1M tokens. Same model, different providers - price and quantization vary.
- Ctx
- 1.048576M
- max output
- 66K
- cache read /1M
- $0.01
- Family
- gemini
- Modalities
- text · image · file · audio · video
- Features
- json · reasoning · tools
Live pricing, updated 2026-09-15.