NVIDIA
Nemotron 3.5 Lightning
Nemotron 3.5 Lightning is an open-weights model by NVIDIA. It holds an Intelligence score of 65.5 on the local.ai leaderboard (2026-08 snapshot).
open weights
65.5
local.ai Intelligence
Parameters
undisclosed
Benchmarks (local.ai snapshot)
HuggingFace metadata · updated 2026-08-13
Metadata refreshed from the HuggingFace API (snapshot 2026-08-15T15:53:42.405Z). Context length and long-form descriptions are the next pipeline stage.
Does it fit your hardware?
Detect or select your GPU first. Green = weights + context headroom, yellow = tight, grey = does not fit.
Params not disclosed — no size estimate.Actual files on HuggingFace
| File | Size |
|---|---|
| model-00012-of-00014.safetensors | 4.65 GB |
| model-00004-of-00014.safetensors | 4.65 GB |
| model-00009-of-00014.safetensors | 4.65 GB |
| model-00011-of-00014.safetensors | 4.65 GB |
| model-00008-of-00014.safetensors | 4.65 GB |
| model-00010-of-00014.safetensors | 4.65 GB |
GGUF: unsloth/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF · 79,108 downloads
| File | Size | |
|---|---|---|
| BF16/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16-00001-of-00002.gguf | 45.81 GB | download |
| BF16/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16-00002-of-00002.gguf | 15.52 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-MXFP4_MOE.gguf | 21.62 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Q8_0.gguf | 32.60 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-UD-IQ1_M.gguf | 18.09 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-UD-IQ2_M.gguf | 18.10 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-UD-IQ2_XXS.gguf | 18.09 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-UD-IQ3_S.gguf | 19.70 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-UD-IQ3_XXS.gguf | 18.40 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-UD-IQ4_NL.gguf | 19.78 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-UD-Q3_K_XL.gguf | 19.78 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-UD-Q4_K_M.gguf | 23.53 GB | download |
GGUF: ggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF · 63,847 downloads
| File | Size | |
|---|---|---|
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16.gguf | 58.84 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Q4_0.gguf | 17.60 GB | download |
| NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Q8_0.gguf | 31.28 GB | download |
| mtp-NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16.gguf | 3.81 GB | download |
| mtp-NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Q4_0.gguf | 1.08 GB | download |
| mtp-NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Q8_0.gguf | 2.03 GB | download |
What you save self-hosting this model
Uses the model's own API price when it exists; otherwise the closest commercial equivalent.
Waiting for hardware detection…
RunLocal presents third-party leaderboard data (local.ai / Exo Labs, 2026-08) with its own fit and cost estimates. Verify pricing, licenses and file sizes at the source before deploying a model.