Google: Gemini 2.5 Flash Lite
google/gemini-2.5-flash-lite
Gemini 2.5 Flash Lite is Google's budget option at $0.1/$0.4 per million input/output tokens, with a 1.05M-token context. It sits mid-board on capability, at #157 of 235 on the TryAii Score — its best showing is #12 on MATH. It's among the fastest models we track, at roughly 110-183 tokens/sec.
Prompt $/1M
$0.100
Completion $/1M
$0.400
Context
1,048,576 tok
Best provider
220 tok/s · Google AI Studio
Specs
- Model id
- google/gemini-2.5-flash-lite
- Display name
- Google: Gemini 2.5 Flash Lite
- Modality
- text+image+file+audio+video->text
- Tokenizer
- Gemini
- Instruct type
- —
- Context length
- 1,048,576
- Max completion
- 65,535
- Prompt $/1M
- $0.100
- Completion $/1M
- $0.400
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Google AI Studiofastest | 220 |
| 73 | |
| googlefastest | — |
| Artificial Analysis | 0 |
Prompt caching by provider
| Provider | Min tokens to cache |
|---|---|
| Google AI Studio | 2,048 tokens |
| Google Vertex | 2,048 tokens |