microllm-250m:latest

OLLAMA GGUF

llama · 256.74M · F16

Not Rec.

Jul 5, 2026 · Apple M4 Pro

zai-org/glm-4.7-flash

LM-STUDIO GGUF

deepseek2 · 30B · Q4_K_M

Excellent

Jul 4, 2026 · AMD Ryzen 7 7800X3D 8-Core Processor

Global Score

34 vs 83

Hardware Fit

100 vs 86

Quality Score

6 vs 81

Hardware

microllm-250m:latest zai-org/glm-4.7-fl…

MachineMac miniASUS

CPUApple M4 ProAMD Ryzen 7 7800X3D 8-Core Processor

Cores1216

RAM24 GB31 GB

GPUApple M4 ProNVIDIA GeForce RTX 4090, Raphael

OSmacOS 26.5.2cachyos rolling

Archarm64x64

Power Modebalancedlow-power

Performance

microllm-250m:latest zai-org/glm-4.7-fl…

Tokens/sec291.6102.9

First chunk50 ms5 ms

TTFT79 ms153 ms

Load time0.0 sN/A

Memory usage0.6 GB20.8 GB

Memory %2%68%

HW Fit Score Breakdown

microllm-250m:latest

Speed

50/50

TTFT

20/20

Memory

30/30

zai-org/glm-4.7-flash

Speed

50/50

TTFT

20/20

Memory

16/30

Quality

microllm-250m:latest

Reasoning

1/20

Coding

0/20

Instruction

5/20

Structured

0/15

Math

0/15

Multilingual

0/10

Reasoning: Poor Coding: Poor Instruction Following: Weak Structured Output: Poor Math: Poor Multilingual: Poor

zai-org/glm-4.7-flash

Reasoning

13/20

Coding

19/20

Instruction

14/20

Structured

15/15

Math

10/15

Multilingual

10/10

Reasoning: Adequate Coding: Strong Instruction Following: Adequate Structured Output: Strong Math: Adequate Multilingual: Strong

View microllm-250m:latest details | View zai-org/glm-4.7-fl… details

Run yours and compare

$ npm install -g metrillm@latest

$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest