qwen3.5-35b-a3b-t48-c8k:latest
OLLAMA GGUFqwen35moe · 36.0B · Q4_K_M
Sep 6, 2026 · Intel Xeon® E5-4650 v2
gpt-oss:20b
OLLAMA GGUFgptoss · 20.9B · MXFP4
Jun 25, 2026 · Cortex-X925
Global Score
81 vs 92
Hardware Fit
63 vs 86
Quality Score
88 vs 95
Hardware
qwen3.5-35b-a3b-t4… gpt-oss:20b
MachineDell Inc. PowerEdge R820NVIDIA NVIDIA_DGX_Spark
CPUIntel Xeon® E5-4650 v2Cortex-X925
Cores8020
RAM256 GB122 GB
GPUMicrosoft Remote Display Adapter, NVIDIA Tesla P4, NVIDIA Tesla P4, Microsoft Basic Display AdapterDevice 2e12
OSMicrosoft Windows Server 2022 Datacenter 10.0.20348Ubuntu 24.04.4 LTS
Archx64arm64
Power Modebalancedperformance
Performance
qwen3.5-35b-a3b-t4… gpt-oss:20b
Tokens/sec17.260.7
First chunk845 ms542 ms
TTFT845 ms3.2 s
Load time103.1 s12.0 s
Memory usage21.9 GB12.0 GB
Memory %9%10%
HW Fit Score Breakdown
qwen3.5-35b-a3b-t48-c8k…
Speed
16/50
TTFT
17/20
Memory
30/30
gpt-oss:20b
Speed
50/50
TTFT
6/20
Memory
30/30
Quality
qwen3.5-35b-a3b-t48-c8k…
Reasoning
17/20
Coding
18/20
Instruction
16/20
Structured
15/15
Math
12/15
Multilingual
10/10
Reasoning: Strong Coding: Strong Instruction Following: Strong Structured Output: Strong Math: Strong Multilingual: Strong
gpt-oss:20b
Reasoning
19/20
Coding
19/20
Instruction
17/20
Structured
15/15
Math
15/15
Multilingual
10/10
Reasoning: Strong Coding: Strong Instruction Following: Strong Structured Output: Strong Math: Strong Multilingual: Strong
Run yours and compare
$
npm install -g metrillm@latest$
metrillmRequires Node 20+ and Ollama or LM Studio running
Or run without installing: npx metrillm@latest