Google: Gemma 4 31B
google/gemma-4-31b-it
Gemma 4 31B is a 31B open-weight model, ranking #91 of 235 on the TryAii Score. Its best showing is #18 on IFBench. It's served by 20 independent hosts from $0.09/$0.34 per million input/output tokens — about a quarter of what models at its level usually charge.
Prompt $/1M
$0.090
Completion $/1M
$0.340
Context
262,144 tok
Best provider
511 tok/s · Cerebras
Specs
- Model id
- google/gemma-4-31b-it
- Display name
- Google: Gemma 4 31B
- Modality
- text+image+video->text
- Tokenizer
- Gemma
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 16,384
- Prompt $/1M
- $0.090
- Completion $/1M
- $0.340
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionyes
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Cerebrasfastest | 511 |
| ModelRun | 119 |
| Friendli | 117 |
| Artificial Analysis | 73 |
| CoreWeave | 50 |
| SambaNova | 42 |
| Crusoe | 39 |
| OpenInference | 38 |
| WandB | 35 |
| Venice | 34 |
| SiliconFlow | 30 |
| Parasail | 26 |
| DeepInfra | 24 |
| Novita | 21 |
| Morph | 16 |
| Phala | 14 |
| Chutes | 13 |
| Together | 10 |
| Ambient | 5 |
| AkashML | 4 |
| NextBit | 4 |
| googlefastest | — |