Google: Gemini 2.5 Flash Lite (batch)
google/gemini-2.5-flash-lite:batch
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Prompt $/1M
$0.050
Completion $/1M
$0.200
Context
1,048,576 tok
Best provider
—
Specs
- Model id
- google/gemini-2.5-flash-lite:batch
- Display name
- Google: Gemini 2.5 Flash Lite (batch)
- Modality
- text+image+file+audio+video->text
- Tokenizer
- Gemini
- Instruct type
- —
- Context length
- 1,048,576
- Max completion
- 65,535
- Prompt $/1M
- $0.050
- Completion $/1M
- $0.200
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| googlefastest | — |