qwen3-30b-a3b-t16-c32k-instruct:latest

qwen3moe · 30.5B · Q4_K_M

Dell Inc. Precision Tower 7910 (Intel Xeon® E5-2637 v3)

96 GB · Microsoft Windows 10 Pro 10.0.19045

Tested on August 18, 2026 · Submitted by Csaba
Top 1% Compare
Global Score
94 /100
Excellent
Hardware Fit
100/100
Quality
91/100

Get this model

Hardware

Machine
Dell Inc. Precision Tower 7910
CPU
Intel Xeon® E5-2637 v3
Cores
16 total (16 perf)
Frequency
3.5 GHz
RAM
96 GB DDR4
GPU
NVIDIA GeForce RTX 3090, NVIDIA GeForce RTX 3060
OS
Microsoft Windows 10 Pro 10.0.19045
Arch
x64
Power Mode
balanced

Performance

Tokens/sec
116.0
Standard deviation
±23.7
First chunk latency
575 ms
Time to first token
575 ms
Load time
10.2 s
Memory usage
21.0 GB (22%)
Total tokens
965

Score breakdown

Speed
50/50
Time to first token
20/20
Memory
30/30

Quality

Reasoning
17/20
Coding
19/20
Instruction following
16/20
Structured output
15/15
Math
14/15
Multilingual
10/10

Category levels

Reasoning: Strong Coding: Strong Instruction Following: Strong Structured Output: Strong Math: Strong Multilingual: Strong

Metadata

Spec version
0.2.1
Runtime
Ollama 0.32.1
Model format
GGUF
Hardware profile
HIGH-END
Result hash
9b2918bf71cf94a89d299645464379b1e216eb9b62cce1b16bda5b2cff9764dd

Interpretation

Hardware fit: 100/100. Overall suitability: EXCELLENT (Global 94/100). Category profile: Reasoning: Strong, Coding: Strong, Instruction Following: Strong, Structured Output: Strong, Math: Strong, Multilingual: Strong.

Bench Environment

CPU load: avg 20% (peak 26%)

Run yours now

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest