Back to leaderboard

qwen2.5:3b

OLLAMA GGUF

qwen2 · 3.1B · Q4_K_M

Good

Mar 2, 2026 · Apple M4

gemma-4-e2b-it@q5_k_xl

LM-STUDIO GGUF

gemma4 · Q5_K_XL

Good

Jun 29, 2026 · Intel Gen Intel® Core™ i9-11900H

Global Score
77 vs 79
Hardware Fit
100 vs 81
Quality Score
67 vs 78

Hardware

qwen2.5:3b gemma-4-e2b-it@q5_…
MachineMacBook AirDocker Container
CPUApple M4Intel Gen Intel® Core™ i9-11900H
Cores1016
RAM32 GB15 GB
GPUApple M4TigerLake-H GT1 [UHD Graphics]
OSmacOS 26.3Ubuntu 26.04 LTS
Archarm64x64
Power Modebalancedperformance

Performance

qwen2.5:3b gemma-4-e2b-it@q5_…
Tokens/sec47.217.2
First chunkN/A7 ms
TTFT168 ms1.6 s
Load time0.5 s8.9 s
Memory usage4.0 GB0.0 GB
Memory %13%0%

HW Fit Score Breakdown

qwen2.5:3b

Speed
40/50
TTFT
30/20
Memory
30/30

gemma-4-e2b-it@q5_k_xl

Speed
32/50
TTFT
19/20
Memory
30/30

Quality

qwen2.5:3b

Reasoning
12/20
Coding
14/20
Instruction
12/20
Structured
15/15
Math
5/15
Multilingual
9/10
Reasoning: Adequate Coding: Adequate Instruction Following: Adequate Structured Output: Strong Math: Weak Multilingual: Strong

gemma-4-e2b-it@q5_k_xl

Reasoning
11/20
Coding
17/20
Instruction
17/20
Structured
15/15
Math
8/15
Multilingual
10/10
Reasoning: Adequate Coding: Strong Instruction Following: Strong Structured Output: Strong Math: Adequate Multilingual: Strong

Run yours and compare

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest