qwen3-235b-a22b-t48-c8k:latest

qwen3moe · 235.1B · Q4_K_M

Dell Inc. PowerEdge R820 (Intel Xeon® E5-4650 v2)

256 GB · Microsoft Windows Server 2022 Datacenter 10.0.20348

Tested on September 7, 2026
Top 72% Compare
Global Score
65 /100
Not Rec.
Hardware Fit
29/100
Quality
81/100

Get this model

Hardware

Machine
Dell Inc. PowerEdge R820
CPU
Intel Xeon® E5-4650 v2
Cores
80 threads (40 cores)
Frequency
2.4 GHz
RAM
256 GB DDR3
GPU
Microsoft Remote Display Adapter, NVIDIA Tesla P4, NVIDIA Tesla P4, Microsoft Basic Display Adapter
OS
Microsoft Windows Server 2022 Datacenter 10.0.20348
Arch
x64
Power Mode
balanced

Performance

Tokens/sec
2.4
Standard deviation
±0.0
First chunk latency
5.8 s
Time to first token
5.8 s
Load time
0.0 s
Memory usage
135.1 GB (53%)
Total tokens
909

Score breakdown

Speed
2/50
Time to first token
3/20
Memory
24/30

Quality

Reasoning
19/20
Coding
15/20
Instruction following
10/20
Structured output
14/15
Math
13/15
Multilingual
10/10

Category levels

Reasoning: Strong Coding: Adequate Instruction Following: Adequate Structured Output: Strong Math: Strong Multilingual: Strong

Metadata

Spec version
0.2.1
Runtime
Ollama 0.32.15
Model format
GGUF
Hardware profile
HIGH-END
Result hash
61c67cd65c8251f24b3bd3cc6eb6397f3954e18b4a23e307a885095ba81301f2

Interpretation

Hardware fit: 29/100. Overall suitability: NOT RECOMMENDED (Global 65/100). Category profile: Reasoning: Strong, Coding: Adequate, Instruction Following: Adequate, Structured Output: Strong, Math: Strong, Multilingual: Strong.

Disqualifiers

  • Token speed too low: 2.4 tok/s (minimum: 8 tok/s for HIGH-END profile)

Bench Environment

CPU load: avg 33% (peak 89%)

Run yours now

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest