Back to leaderboard

qwen3-vl-4b-instruct-q4_k_m

OPENAI GGUF
Good

Apr 22, 2026 · Apple M2 Max

qwen3-235b-a22b-t48-c8k:latest

OLLAMA GGUF

qwen3moe · 235.1B · Q4_K_M

Not Rec.

Sep 7, 2026 · Intel Xeon® E5-4650 v2

Global Score
62 vs 65
Hardware Fit
100 vs 29
Quality Score
46 vs 81

Hardware

qwen3-vl-4b-instru… qwen3-235b-a22b-t4…
MachineMac StudioDell Inc. PowerEdge R820
CPUApple M2 MaxIntel Xeon® E5-4650 v2
Cores1280
RAM96 GB256 GB
GPUApple M2 MaxMicrosoft Remote Display Adapter, NVIDIA Tesla P4, NVIDIA Tesla P4, Microsoft Basic Display Adapter
OSmacOS 26.1Microsoft Windows Server 2022 Datacenter 10.0.20348
Archarm64x64
Power Modebalancedbalanced

Performance

qwen3-vl-4b-instru… qwen3-235b-a22b-t4…
Tokens/sec67.22.4
First chunk66 ms5833 ms
TTFT68 ms5.8 s
Load time0.0 s0.0 s
Memory usage0.2 GB135.1 GB
Memory %0%53%

HW Fit Score Breakdown

qwen3-vl-4b-instruct-q4…

Speed
50/50
TTFT
20/20
Memory
30/30

qwen3-235b-a22b-t48-c8k…

Speed
2/50
TTFT
3/20
Memory
24/30

Quality

qwen3-vl-4b-instruct-q4…

Reasoning
16/20
Coding
0/20
Instruction
12/20
Structured
1/15
Math
8/15
Multilingual
9/10
Reasoning: Strong Coding: Poor Instruction Following: Adequate Structured Output: Poor Math: Adequate Multilingual: Strong

qwen3-235b-a22b-t48-c8k…

Reasoning
19/20
Coding
15/20
Instruction
10/20
Structured
14/15
Math
13/15
Multilingual
10/10
Reasoning: Strong Coding: Adequate Instruction Following: Adequate Structured Output: Strong Math: Strong Multilingual: Strong

Run yours and compare

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest