gemma4:31b

gemma4 · 30.7B · Q4_K_M

THINKING MODEL

System manufacturer (AMD Ryzen 9 5950X 16-Core Processor)

32 GB · Microsoft Windows 11 Pro 10.0.26300

Tested on October 2, 2026 · Submitted by Qwen3.8vsGemma4
Top 75% Compare
Global Score
64 /100
Not Rec.
Hardware Fit
41/100
Quality
74/100

Get this model

Hardware

Machine
System manufacturer
CPU
AMD Ryzen 9 5950X 16-Core Processor
Cores
32 total (32 perf)
Frequency
3.4 GHz
RAM
32 GB DDR4
GPU
NVIDIA GeForce RTX 3090
OS
Microsoft Windows 11 Pro 10.0.26300
Arch
x64
Power Mode
balanced

Performance

Tokens/sec
9.5
Standard deviation
±0.5
First chunk latency
1.8 s
Time to first token
30.0 s
Load time
33.9 s
Memory usage
1.4 GB (4%)
Total tokens
1460
Thinking tokens (est.)
~720

Score breakdown

Speed
11/50
Time to first token
0/20
Memory
30/30

Quality

Reasoning
11/20
Coding
17/20
Instruction following
10/20
Structured output
15/15
Math
12/15
Multilingual
9/10

Category levels

Reasoning: Adequate Coding: Strong Instruction Following: Adequate Structured Output: Strong Math: Strong Multilingual: Strong

Metadata

Spec version
0.2.1
Runtime
Ollama 0.35.0
Model format
GGUF
Hardware profile
HIGH-END
Result hash
98dde570c79fb7d8ce673ce71d7d0f5c423ecb180cd0a1c2aaa0da786a3d1ea6

Interpretation

Hardware fit: 41/100. Overall suitability: NOT RECOMMENDED (Global 64/100). Category profile: Reasoning: Adequate, Coding: Strong, Instruction Following: Adequate, Structured Output: Strong, Math: Strong, Multilingual: Strong.

Warnings

  • Significant swap activity during benchmark (+0.8 GB). Model may exceed available RAM — results are severely degraded.

Disqualifiers

  • Time to first token too high: 30000ms (maximum: 15463ms for HIGH-END profile)

Bench Environment

Swap delta: +0.8 GB CPU load: avg 56% (peak 56%)

Run yours now

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest