Benchmarks

Google: Gemma 4 26B A4B

google/gemma-4-26b-a4b-it

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Prompt $/1M
$0.070
Completion $/1M
$0.340
Context
262,144 tok
Best provider
52 tok/s · Wafer

Specs

Model id
google/gemma-4-26b-a4b-it
Display name
Google: Gemma 4 26B A4B
Modality
text+image+video->text
Tokenizer
Gemma
Instruct type
Context length
262,144
Max completion
16,384
Prompt $/1M
$0.070
Completion $/1M
$0.340
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionno
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
Waferfastest52
Ionstreamfastest50
Artificial Analysis45
Google34
Parasail32
Cloudflare29
Venice28
Ambient27
Novita23
SiliconFlow21
DeepInfra21
Io Net20
DekaLLM18
NextBit11
Google: Gemma 4 26B A4B - Benchmarks, Pricing and Speed | TryAii