Google: Gemini 2.5 Flash Lite
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
- Context window
- 1,048,576 tokens
- Input cost
- —
- Output cost
- —
- Latency (p50)
- 18 ms
