nvidia/nemotron-3-nano-omni

nemotron_h_moe · 30B · Q4_K_M

Micro-Star International Co., Ltd. MS-7E51 (AMD Ryzen 7 9700X 8-Core Processor)

31 GB · Microsoft Windows 11 Pro 10.0.26200

Tested on August 19, 2026
Top 39% Compare
Global Score
80 /100
Excellent
Hardware Fit
81/100
Quality
80/100

Get this model

Hardware

Machine
Micro-Star International Co., Ltd. MS-7E51
CPU
AMD Ryzen 7 9700X 8-Core Processor
Cores
16 total (16 perf)
Frequency
3.8 GHz
RAM
31 GB DDR5
GPU
NVIDIA GeForce RTX 5080, AMD Radeon(TM) Graphics
OS
Microsoft Windows 11 Pro 10.0.26200
Arch
x64
Power Mode
balanced

Performance

Tokens/sec
55.4
Standard deviation
±0.3
First chunk latency
32 ms
Time to first token
336 ms
Load time
N/A
Memory usage
24.3 GB (78%)
Total tokens
1140

Score breakdown

Speed
50/50
Time to first token
20/20
Memory
11/30

Quality

Reasoning
15/20
Coding
17/20
Instruction following
17/20
Structured output
14/15
Math
8/15
Multilingual
9/10

Category levels

Reasoning: Adequate Coding: Strong Instruction Following: Strong Structured Output: Strong Math: Adequate Multilingual: Strong

Metadata

Spec version
0.2.1
Runtime
LM Studio 0.4.15+2
Model format
GGUF
Hardware profile
BALANCED
Result hash
6d1805917d7645d3789afa3774d1a6ebcf8e716d7125838a786abd535970d330

Interpretation

Hardware fit: 81/100. Overall suitability: EXCELLENT (Global 80/100). Category profile: Reasoning: Adequate, Coding: Strong, Instruction Following: Strong, Structured Output: Strong, Math: Adequate, Multilingual: Strong.

Warnings

  • Model memory footprint is estimated via LM Studio CLI rather than measured from a fresh load.

Bench Environment

CPU load: avg 37% (peak 45%)

Run yours now

$ npm install -g metrillm@latest
$ metrillm

Requires Node 20+ and Ollama or LM Studio running

Or run without installing: npx metrillm@latest