Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Google: Gemini 2.5 Flash Lite — Google. Context window 1,048,576 tokens. API pricing (per 1M tokens): Input $0.1, Output $0.4. License: Proprietary.
| Provider | |
|---|---|
| API model ID | google/gemini-2.5-flash-lite |
| Tier | Frontier |
| Context window | 1,048,576 tokens |
| Max output | 65,535 tokens |
| License | Proprietary |
| API pricing (per 1M tokens) · Input | $0.1 |
| API pricing (per 1M tokens) · Output | $0.4 |