qwen3.8-27b-mlx

qwen3_5 · 27B · 8bit

MacBook Pro (Apple M5 Max)

64 GB · macOS 27.0

Tested on August 17, 2026 · Submitted by itxjobe
Top 81% Compare
Global Score
59 /100
Marginal
Hardware Fit
63/100
Quality
58/100

Get this model

Hardware

Machine
MacBook Pro
CPU
Apple M5 Max
Cores
18 threads (6 cores)
Frequency
2.4 GHz
RAM
64 GB LPDDR5
GPU
Apple M5 Max
OS
macOS 27.0
Arch
arm64
Power Mode
balanced

Performance

Tokens/sec
18.3
Standard deviation
±0.5
First chunk latency
4 ms
Time to first token
530 ms
Load time
N/A
Memory usage
38.5 GB (60%)
Total tokens
1743

Score breakdown

Speed
23/50
Time to first token
20/20
Memory
20/30

Quality

Reasoning
19/20
Coding
14/20
Instruction following
6/20
Structured output
1/15
Math
13/15
Multilingual
5/10

Category levels

Reasoning: Strong Coding: Adequate Instruction Following: Weak Structured Output: Poor Math: Strong Multilingual: Weak

Metadata

Spec version
0.2.1
Runtime
LM Studio
Model format
MLX
Hardware profile
HIGH-END
Result hash
d31f353f19a7d894e0d1cc98e8eb21471aa28e036a261797aaf4a82cead0a4a1

Interpretation

Hardware fit: 63/100. Overall suitability: MARGINAL (Global 59/100). Category profile: Reasoning: Strong, Coding: Adequate, Instruction Following: Weak, Structured Output: Poor, Math: Strong, Multilingual: Weak.

Warnings

  • Model memory footprint is estimated via LM Studio CLI rather than measured from a fresh load.

Bench Environment

Power: AC CPU load: avg 11% (peak 14%)

Run yours now

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest