Google: Gemma 4 26B A4B
google/gemma-4-26b-a4b-it
Gemma 4 26B A4B is a 26B open-weight model (4B active per token), ranking #111 of 235 on the TryAii Score. Its best showing is #33 on IFBench. It's served by 13 independent hosts from $0.07/$0.34 per million input/output tokens — about a third of what models at its level usually charge.
Prompt $/1M
$0.070
Completion $/1M
$0.340
Context
262,144 tok
Best provider
76 tok/s · Wafer
Specs
- Model id
- google/gemma-4-26b-a4b-it
- Display name
- Google: Gemma 4 26B A4B
- Modality
- text+image+video->text
- Tokenizer
- Gemma
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 16,384
- Prompt $/1M
- $0.070
- Completion $/1M
- $0.340
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Waferfastest | 76 |
| Artificial Analysis | 70 |
| DekaLLM | 47 |
| Cloudflarefastest | 44 |
| SiliconFlow | 37 |
| Parasail | 32 |
| 32 | |
| NextBit | 30 |
| Ambient | 27 |
| Venice | 23 |
| Novita | 23 |
| Io Net | 20 |
| DeepInfra | 18 |
| Ionstream | 7 |
| Darkbloom | 6 |
| googlefastest | — |