Benchmarks

Google: Gemma 4 31B

google/gemma-4-31b-it

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Prompt $/1M
$0.120
Completion $/1M
$0.370
Context
262,144 tok
Best provider
90 tok/s · SambaNova

Specs

Model id
google/gemma-4-31b-it
Display name
Google: Gemma 4 31B
Modality
text+image+video->text
Tokenizer
Gemma
Instruct type
Context length
262,144
Max completion
16,384
Prompt $/1M
$0.120
Completion $/1M
$0.370
Free tier
no

Capabilities

  • Function callingno
  • JSON modeno
  • System promptno
  • Visionyes
  • Moderatedno

Benchmark scores

Each axis is one benchmark, normalized to the best model in the field.

No benchmark data for this model.

Speed by provider

Providertok/s
SambaNovafastest90
Artificial Analysis67
ModelRun62
Friendli56
WandB36
Venice30
Cerebras28
SiliconFlow21
Chutes20
OpenInference19
Parasail17
Together14
Phala12
DeepInfra10
Novita10
Ambient5
AkashML4
Google: Gemma 4 31B - Benchmarks, Pricing and Speed | TryAii