qwen3:30b-a3b-instruct-2507-q4_K_M

qwen3moe · 30.5B · Q4_K_M

Dell Inc. PowerEdge R820 (Intel Xeon® E5-4620 0)

256 GB · Microsoft Windows Server 2022 Datacenter 10.0.20348

Tested on July 31, 2026 · Submitted by Csaba_R820
Top 60% Compare
Global Score
71 /100
Not Rec.
Hardware Fit
42/100
Quality
84/100

Get this model

Hardware

Machine
Dell Inc. PowerEdge R820
CPU
Intel Xeon® E5-4620 0
Cores
64 total (64 perf)
Frequency
2.2 GHz
RAM
256 GB DDR3
GPU
Microsoft Remote Display Adapter, Microsoft Basic Display Adapter
OS
Microsoft Windows Server 2022 Datacenter 10.0.20348
Arch
x64
Power Mode
balanced

Performance

Tokens/sec
4.0
Standard deviation
±0.2
First chunk latency
2.1 s
Time to first token
2.1 s
Load time
44.5 s
Memory usage
17.7 GB (7%)
Total tokens
993

Score breakdown

Speed
3/50
Time to first token
9/20
Memory
30/30

Quality

Reasoning
18/20
Coding
18/20
Instruction following
11/20
Structured output
15/15
Math
12/15
Multilingual
10/10

Category levels

Reasoning: Strong Coding: Strong Instruction Following: Adequate Structured Output: Strong Math: Strong Multilingual: Strong

Metadata

Spec version
0.2.1
Runtime
Ollama 0.32.5
Model format
GGUF
Hardware profile
HIGH-END
Result hash
a11d9a9690712c88f98026c6d74dc6085688a278c002d940ab87a70c6d7f3cae

Interpretation

Hardware fit: 42/100. Overall suitability: NOT RECOMMENDED (Global 71/100). Category profile: Reasoning: Strong, Coding: Strong, Instruction Following: Adequate, Structured Output: Strong, Math: Strong, Multilingual: Strong.

Disqualifiers

  • Token speed too low: 4.0 tok/s (minimum: 8 tok/s for HIGH-END profile)

Bench Environment

CPU load: avg 40% (peak 50%)

Run yours now

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest