qwen3.5-35b-a3b-t48-c8k:latest

qwen35moe · 36.0B · Q4_K_M

Dell Inc. PowerEdge R820 (Intel Xeon® E5-4620 0)

256 GB · Microsoft Windows Server 2022 Datacenter 10.0.20348

Tested on August 1, 2026 · Submitted by Csaba_R820
Top 69% Compare
Global Score
65 /100
Not Rec.
Hardware Fit
38/100
Quality
76/100

Get this model

Hardware

Machine
Dell Inc. PowerEdge R820
CPU
Intel Xeon® E5-4620 0
Cores
64 total (64 perf)
Frequency
2.2 GHz
RAM
256 GB DDR3
GPU
Microsoft Remote Display Adapter, Microsoft Basic Display Adapter
OS
Microsoft Windows Server 2022 Datacenter 10.0.20348
Arch
x64
Power Mode
balanced

Performance

Tokens/sec
3.4
Standard deviation
±0.1
First chunk latency
4.2 s
Time to first token
4.2 s
Load time
45.1 s
Memory usage
22.0 GB (9%)
Total tokens
1085

Score breakdown

Speed
3/50
Time to first token
5/20
Memory
30/30

Quality

Reasoning
16/20
Coding
15/20
Instruction following
9/20
Structured output
15/15
Math
11/15
Multilingual
10/10

Category levels

Reasoning: Strong Coding: Strong Instruction Following: Weak Structured Output: Strong Math: Adequate Multilingual: Strong

Metadata

Spec version
0.2.1
Runtime
Ollama 0.32.5
Model format
GGUF
Hardware profile
HIGH-END
Result hash
4332b7edb089bb5ac938d962b860c7cf8d7cee2c1861bc445a7fb5809442cb70

Interpretation

Hardware fit: 38/100. Overall suitability: NOT RECOMMENDED (Global 65/100). Category profile: Reasoning: Strong, Coding: Strong, Instruction Following: Weak, Structured Output: Strong, Math: Adequate, Multilingual: Strong.

Disqualifiers

  • Token speed too low: 3.4 tok/s (minimum: 8 tok/s for HIGH-END profile)

Bench Environment

CPU load: avg 60% (peak 75%)

Run yours now

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest