qwen3-next:80b-a3b-instruct-q4_K_M

qwen3next · 79.7B · Q4_K_M

Dell Inc. PowerEdge R820 (Intel Xeon® E5-4650 v2)

256 GB · Microsoft Windows Server 2022 Datacenter 10.0.20348

Tested on August 23, 2026 · Submitted by Csaba_R820
Top 55% Compare
Global Score
74 /100
Not Rec.
Hardware Fit
42/100
Quality
87/100

Get this model

Hardware

Machine
Dell Inc. PowerEdge R820
CPU
Intel Xeon® E5-4650 v2
Cores
80 threads (40 cores)
Frequency
2.4 GHz
RAM
256 GB DDR3
GPU
Microsoft Basic Display Adapter
OS
Microsoft Windows Server 2022 Datacenter 10.0.20348
Arch
x64
Power Mode
balanced

Performance

Tokens/sec
3.9
Standard deviation
±0.0
First chunk latency
2.1 s
Time to first token
2.1 s
Load time
73.3 s
Memory usage
46.9 GB (18%)
Total tokens
922

Score breakdown

Speed
3/50
Time to first token
9/20
Memory
30/30

Quality

Reasoning
18/20
Coding
19/20
Instruction following
12/20
Structured output
15/15
Math
13/15
Multilingual
10/10

Category levels

Reasoning: Strong Coding: Strong Instruction Following: Adequate Structured Output: Strong Math: Strong Multilingual: Strong

Metadata

Spec version
0.2.1
Runtime
Ollama 0.32.15
Model format
GGUF
Hardware profile
HIGH-END
Result hash
306ca72bd010eee82bf7212b00e111fc770405f3bcdb4807c854f28b1bc59f28

Interpretation

Hardware fit: 42/100. Overall suitability: NOT RECOMMENDED (Global 74/100). Category profile: Reasoning: Strong, Coding: Strong, Instruction Following: Adequate, Structured Output: Strong, Math: Strong, Multilingual: Strong.

Disqualifiers

  • Token speed too low: 3.9 tok/s (minimum: 8 tok/s for HIGH-END profile)

Bench Environment

CPU load: avg 44% (peak 50%)

Run yours now

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest