qwen3-next:80b-a3b-instruct-q4_K_M

qwen3next · 79.7B · Q4_K_M

Dell Inc. Precision Tower 7910 (Intel Xeon® E5-2637 v3)

96 GB · Microsoft Windows 10 Pro 10.0.19045

Tested on August 18, 2026 · Submitted by Csaba
Top 45% Compare
Global Score
78 /100
Good
Hardware Fit
51/100
Quality
89/100

Get this model

Hardware

Machine
Dell Inc. Precision Tower 7910
CPU
Intel Xeon® E5-2637 v3
Cores
16 total (16 perf)
Frequency
3.5 GHz
RAM
96 GB DDR4
GPU
NVIDIA GeForce RTX 3090, NVIDIA GeForce RTX 3060
OS
Microsoft Windows 10 Pro 10.0.19045
Arch
x64
Power Mode
balanced

Performance

Tokens/sec
8.7
Standard deviation
±3.3
First chunk latency
1.3 s
Time to first token
1.3 s
Load time
61.6 s
Memory usage
48.3 GB (50%)
Total tokens
926

Score breakdown

Speed
9/50
Time to first token
17/20
Memory
25/30

Quality

Reasoning
18/20
Coding
19/20
Instruction following
14/20
Structured output
15/15
Math
13/15
Multilingual
10/10

Category levels

Reasoning: Strong Coding: Strong Instruction Following: Adequate Structured Output: Strong Math: Strong Multilingual: Strong

Metadata

Spec version
0.2.1
Runtime
Ollama 0.32.1
Model format
GGUF
Hardware profile
HIGH-END
Result hash
0b221155142a80a25edeb965c481d89ea695cca7a54a27bb9aa277d19d5a35b1

Interpretation

Hardware fit: 51/100. Overall suitability: GOOD (Global 78/100). Category profile: Reasoning: Strong, Coding: Strong, Instruction Following: Adequate, Structured Output: Strong, Math: Strong, Multilingual: Strong.

Warnings

  • Token speed is unstable (stddev 3.3 tok/s, mean 8.7 tok/s) — may indicate thermal throttling or memory pressure.

Bench Environment

CPU load: avg 21% (peak 26%)

Run yours now

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest