Google
Gemma 4 26B A4B
Gemma 4 26B A4B is an open-weights 26B-parameter model by Google. It holds an Intelligence score of 62.8 on the local.ai leaderboard (2026-08 snapshot).
MoEopen weights
62.8
local.ai Intelligence
Parameters
26B
MoE · 4B active parameters
Benchmarks (local.ai snapshot)
HuggingFace metadata · updated 2026-07-20
Metadata refreshed from the HuggingFace API (snapshot 2026-08-15T15:53:42.405Z). Context length and long-form descriptions are the next pipeline stage.
Does it fit your hardware?
Detect or select your GPU first. Green = weights + context headroom, yellow = tight, grey = does not fit.
Q4_K_M 15.1GQ6_K 23.4GQ8_0 28.6GFP16 52G| Quantization | ≈ File size | Note |
|---|---|---|
| IQ2_XXS | 7 GB | Extreme compression, largest quality loss |
| IQ3_XXS | 9.4 GB | Very small |
| Q3_K_M | 12 GB | Small |
| NVFP4 | 13 GB | NVIDIA 4-bit |
| Q4_K_S | 14 GB | Common |
| Q4_K_M | 15.1 GB | Most common sweet spot |
| Q5_K_M | 19.8 GB | Balanced |
| Q6_K | 23.4 GB | Near-lossless |
| Q8_0 | 28.6 GB | Near-lossless |
| FP16 | 52 GB | Original weights |
Actual files on HuggingFace
| File | Size |
|---|---|
| model-00001-of-00002.safetensors | 46.48 GB |
| model-00002-of-00002.safetensors | 1.59 GB |
GGUF: unsloth/gemma-4-26B-A4B-it-GGUF · 1,393,778 downloads
| File | Size | |
|---|---|---|
| BF16/gemma-4-26B-A4B-it-BF16-00001-of-00002.gguf | 46.49 GB | download |
| BF16/gemma-4-26B-A4B-it-BF16-00002-of-00002.gguf | 0.54 GB | download |
| MTP/mtp-gemma-4-26B-A4B-it-BF16.gguf | 0.80 GB | download |
| MTP/mtp-gemma-4-26B-A4B-it-F16.gguf | 0.80 GB | download |
| MTP/mtp-gemma-4-26B-A4B-it-Q8_0.gguf | 0.43 GB | download |
| gemma-4-26B-A4B-it-MXFP4_MOE.gguf | 15.41 GB | download |
| gemma-4-26B-A4B-it-Q8_0.gguf | 25.02 GB | download |
| gemma-4-26B-A4B-it-UD-IQ2_M.gguf | 9.33 GB | download |
| gemma-4-26B-A4B-it-UD-IQ2_XXS.gguf | 9.24 GB | download |
| gemma-4-26B-A4B-it-UD-IQ3_S.gguf | 10.51 GB | download |
| gemma-4-26B-A4B-it-UD-IQ3_XXS.gguf | 10.63 GB | download |
| gemma-4-26B-A4B-it-UD-IQ4_NL.gguf | 12.68 GB | download |
GGUF: unsloth/gemma-4-26B-A4B-it-qat-GGUF · 469,680 downloads
| File | Size | |
|---|---|---|
| MTP/mtp-gemma-4-26B-A4B-it-BF16.gguf | 0.80 GB | download |
| MTP/mtp-gemma-4-26B-A4B-it-F16.gguf | 0.80 GB | download |
| MTP/mtp-gemma-4-26B-A4B-it-Q4_0.gguf | 0.23 GB | download |
| MTP/mtp-gemma-4-26B-A4B-it-Q8_0.gguf | 0.43 GB | download |
| gemma-4-26B-A4B-it-qat-UD-Q4_K_XL.gguf | 13.27 GB | download |
| mmproj-BF16.gguf | 1.11 GB | download |
| mmproj-F16.gguf | 1.11 GB | download |
| mmproj-F32.gguf | 2.13 GB | download |
| mtp-gemma-4-26B-A4B-it.gguf | 0.23 GB | download |
What you save self-hosting this model
Uses the model's own API price when it exists; otherwise the closest commercial equivalent.
Waiting for hardware detection…
RunLocal presents third-party leaderboard data (local.ai / Exo Labs, 2026-08) with its own fit and cost estimates. Verify pricing, licenses and file sizes at the source before deploying a model.