NVIDIA
Nemotron 3 Nano 30B A3B
Nemotron 3 Nano 30B A3B is an open-weights 30B-parameter model by NVIDIA. It holds an Intelligence score of 42.4 on the local.ai leaderboard (2026-08 snapshot).
MoEopen weights
42.4
local.ai Intelligence
Parameters
30B
MoE · 3B active parameters
Benchmarks (local.ai snapshot)
HuggingFace metadata · updated 2026-07-23
Metadata refreshed from the HuggingFace API (snapshot 2026-08-15T15:53:42.405Z). Context length and long-form descriptions are the next pipeline stage.
Does it fit your hardware?
Detect or select your GPU first. Green = weights + context headroom, yellow = tight, grey = does not fit.
Q4_K_M 17.4GQ6_K 27GQ8_0 33GFP16 60G| Quantization | ≈ File size | Note |
|---|---|---|
| IQ2_XXS | 8.1 GB | Extreme compression, largest quality loss |
| IQ3_XXS | 10.8 GB | Very small |
| Q3_K_M | 13.8 GB | Small |
| NVFP4 | 15 GB | NVIDIA 4-bit |
| Q4_K_S | 16.2 GB | Common |
| Q4_K_M | 17.4 GB | Most common sweet spot |
| Q5_K_M | 22.8 GB | Balanced |
| Q6_K | 27 GB | Near-lossless |
| Q8_0 | 33 GB | Near-lossless |
| FP16 | 60 GB | Original weights |
Actual files on HuggingFace
| File | Size |
|---|---|
| model-00006-of-00013.safetensors | 4.66 GB |
| model-00012-of-00013.safetensors | 4.65 GB |
| model-00004-of-00013.safetensors | 4.65 GB |
| model-00009-of-00013.safetensors | 4.65 GB |
| model-00011-of-00013.safetensors | 4.65 GB |
| model-00008-of-00013.safetensors | 4.65 GB |
GGUF: lmstudio-community/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-GGUF · 62,148 downloads
GGUF: unsloth/Nemotron-3-Nano-30B-A3B-GGUF · 15,789 downloads
| File | Size | |
|---|---|---|
| BF16/Nemotron-3-Nano-30B-A3B-BF16-00001-of-00002.gguf | 46.49 GB | download |
| BF16/Nemotron-3-Nano-30B-A3B-BF16-00002-of-00002.gguf | 12.35 GB | download |
| Nemotron-3-Nano-30B-A3B-IQ4_NL.gguf | 16.93 GB | download |
| Nemotron-3-Nano-30B-A3B-IQ4_XS.gguf | 16.92 GB | download |
| Nemotron-3-Nano-30B-A3B-Q2_K_L.gguf | 16.85 GB | download |
| Nemotron-3-Nano-30B-A3B-Q3_K_M.gguf | 18.63 GB | download |
| Nemotron-3-Nano-30B-A3B-Q3_K_S.gguf | 16.88 GB | download |
| Nemotron-3-Nano-30B-A3B-Q4_0.gguf | 16.96 GB | download |
| Nemotron-3-Nano-30B-A3B-Q4_1.gguf | 18.68 GB | download |
| Nemotron-3-Nano-30B-A3B-Q4_K_M.gguf | 22.89 GB | download |
| Nemotron-3-Nano-30B-A3B-Q4_K_S.gguf | 20.51 GB | download |
| Nemotron-3-Nano-30B-A3B-Q5_K_M.gguf | 24.35 GB | download |
What you save self-hosting this model
Uses the model's own API price when it exists; otherwise the closest commercial equivalent.
Waiting for hardware detection…
RunLocal presents third-party leaderboard data (local.ai / Exo Labs, 2026-08) with its own fit and cost estimates. Verify pricing, licenses and file sizes at the source before deploying a model.