Qwen
Qwen3.5 9B
Qwen3.5 9B is an open-weights 9B-parameter model by Qwen. It holds an Intelligence score of 55.3 on the local.ai leaderboard (2026-08 snapshot).
open weights
55.3
local.ai Intelligence
Parameters
9B
Benchmarks (local.ai snapshot)
HuggingFace metadata · updated 2026-03-02
Metadata refreshed from the HuggingFace API (snapshot 2026-08-15T15:53:42.405Z). Context length and long-form descriptions are the next pipeline stage.
Does it fit your hardware?
Detect or select your GPU first. Green = weights + context headroom, yellow = tight, grey = does not fit.
Q4_K_M 5.2GQ6_K 8.1GQ8_0 9.9GFP16 18G| Quantization | ≈ File size | Note |
|---|---|---|
| IQ2_XXS | 2.4 GB | Extreme compression, largest quality loss |
| IQ3_XXS | 3.2 GB | Very small |
| Q3_K_M | 4.1 GB | Small |
| NVFP4 | 4.5 GB | NVIDIA 4-bit |
| Q4_K_S | 4.9 GB | Common |
| Q4_K_M | 5.2 GB | Most common sweet spot |
| Q5_K_M | 6.8 GB | Balanced |
| Q6_K | 8.1 GB | Near-lossless |
| Q8_0 | 9.9 GB | Near-lossless |
| FP16 | 18 GB | Original weights |
Actual files on HuggingFace
| File | Size |
|---|---|
| model.safetensors-00003-of-00004.safetensors | 5.00 GB |
| model.safetensors-00002-of-00004.safetensors | 4.97 GB |
| model.safetensors-00001-of-00004.safetensors | 4.91 GB |
| model.safetensors-00004-of-00004.safetensors | 3.10 GB |
GGUF: unsloth/Qwen3.5-9B-GGUF · 1,180,204 downloads
| File | Size | |
|---|---|---|
| Qwen3.5-9B-BF16.gguf | 16.69 GB | download |
| Qwen3.5-9B-IQ4_NL.gguf | 5.00 GB | download |
| Qwen3.5-9B-IQ4_XS.gguf | 4.81 GB | download |
| Qwen3.5-9B-Q3_K_M.gguf | 4.35 GB | download |
| Qwen3.5-9B-Q3_K_S.gguf | 4.02 GB | download |
| Qwen3.5-9B-Q4_0.gguf | 5.01 GB | download |
| Qwen3.5-9B-Q4_1.gguf | 5.44 GB | download |
| Qwen3.5-9B-Q4_K_M.gguf | 5.29 GB | download |
| Qwen3.5-9B-Q4_K_S.gguf | 5.02 GB | download |
| Qwen3.5-9B-Q5_K_M.gguf | 6.13 GB | download |
| Qwen3.5-9B-Q5_K_S.gguf | 5.92 GB | download |
| Qwen3.5-9B-Q6_K.gguf | 6.95 GB | download |
GGUF: DavidAU/Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF · 639,475 downloads
| File | Size | |
|---|---|---|
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-IQ2_M.gguf | 4.60 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-IQ3_M.gguf | 5.23 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-IQ4_NL.gguf | 6.16 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-IQ4_XS.gguf | 5.96 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-MTP-IQ2_M.gguf | 4.60 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-MTP-IQ3_M.gguf | 5.33 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-MTP-IQ4_NL.gguf | 6.29 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-MTP-IQ4_XS.gguf | 6.08 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-MTP-Q4_K_M.gguf | 6.50 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-MTP-Q4_K_S.gguf | 6.23 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-MTP-Q5_K_M.gguf | 7.30 GB | download |
| Qwen3.5-9B-The-Defiant-Fable-Uncnr-Heretic-NEO-MAX-MTP-Q5_K_S.gguf | 7.15 GB | download |
What you save self-hosting this model
Uses the model's own API price when it exists; otherwise the closest commercial equivalent.
Waiting for hardware detection…
RunLocal presents third-party leaderboard data (local.ai / Exo Labs, 2026-08) with its own fit and cost estimates. Verify pricing, licenses and file sizes at the source before deploying a model.