Appearance
← Back to Benchmarks
Model Detail

Gemma 4 31B

The better local model when quality matters more than speed and concurrency.

92/100 Benchmark score
Premium local escalation
Rank
#6 of 13
Source
Enterprise Ollama
Run date
Avg latency
22.59s
01 · Benchmark score 92/100
02 · Average latency 22.59s
03 · Role Escalation local model
Per-suite breakdown

Scorecard

Operator board #6

92/100 operator score

Messaging Quick execution pack

Messaging Tool Planning v2

External canon Fast local quality leader in current Gemma run

Public benchmark references

Cost Enterprise local

Best local quality of the Gemma pair, but significantly slower.

Strengths
  • Best quality in the Gemma local pair
  • Better routing judgment on the benchmark pack
  • Cleaner concise answers overall
Weaknesses
  • Much slower average latency
  • Worse fit for high-concurrency local loops
  • Still wrapped strict JSON in code fences
Operator read

Stronger output quality and better routing judgment than 26B, but roughly 3x slower in this quick benchmark pack.

Comparison

Where Gemma 4 31B lands

OperatorIndex
Source artifacts

Raw machine-readable files for anyone who wants to dig deeper or run their own analysis.

  • internal artifact output/benchmarks/2026-04-12-gemma-enterprise-benchmark/summary.md
  • internal artifact output/benchmarks/2026-04-12-gemma-enterprise-benchmark/results-raw.json
  • internal artifact output/benchmarks/2026-04-12-gemma-enterprise-benchmark/results-scored.json