Qwen

Qwen2.5 Coder 32B

Qwen2.5 Coder 32B is an open-weights 32B-parameter model by Qwen. It is listed as in-progress on the local.ai leaderboard.

codingopen weights
Parameters
32B
Benchmarks (local.ai snapshot)
τ²-bench
GAIA
GDPval
HuggingFace metadata · updated 2025-01-12
repo
Qwen/Qwen2.5-Coder-32B-Instruct
license
apache-2.0
downloads
1,279,430
likes
2,104

Metadata refreshed from the HuggingFace API (snapshot 2026-08-15T15:53:42.405Z). Context length and long-form descriptions are the next pipeline stage.

Does it fit your hardware?

Detect or select your GPU first. Green = weights + context headroom, yellow = tight, grey = does not fit.

Q4_K_M 18.6GQ6_K 28.8GQ8_0 35.2GFP16 64G
Quantization≈ File sizeNote
IQ2_XXS8.6 GBExtreme compression, largest quality loss
IQ3_XXS11.5 GBVery small
Q3_K_M14.7 GBSmall
NVFP416 GBNVIDIA 4-bit
Q4_K_S17.3 GBCommon
Q4_K_M18.6 GBMost common sweet spot
Q5_K_M24.3 GBBalanced
Q6_K28.8 GBNear-lossless
Q8_035.2 GBNear-lossless
FP1664 GBOriginal weights

Actual files on HuggingFace

FileSize
model-00001-of-00014.safetensors4.56 GB
model-00004-of-00014.safetensors4.54 GB
model-00005-of-00014.safetensors4.54 GB
model-00006-of-00014.safetensors4.54 GB
model-00007-of-00014.safetensors4.54 GB
model-00008-of-00014.safetensors4.54 GB
GGUF: bartowski/Qwen2.5-Coder-32B-Instruct-GGUF · 116,966 downloads
FileSize
Qwen2.5-Coder-32B-Instruct-IQ2_M.gguf10.49 GBdownload
Qwen2.5-Coder-32B-Instruct-IQ2_S.gguf9.67 GBdownload
Qwen2.5-Coder-32B-Instruct-IQ2_XS.gguf9.27 GBdownload
Qwen2.5-Coder-32B-Instruct-IQ2_XXS.gguf8.41 GBdownload
Qwen2.5-Coder-32B-Instruct-IQ3_M.gguf13.79 GBdownload
Qwen2.5-Coder-32B-Instruct-IQ3_XS.gguf12.76 GBdownload
Qwen2.5-Coder-32B-Instruct-IQ3_XXS.gguf11.96 GBdownload
Qwen2.5-Coder-32B-Instruct-IQ4_NL.gguf17.40 GBdownload
Qwen2.5-Coder-32B-Instruct-IQ4_XS.gguf16.48 GBdownload
Qwen2.5-Coder-32B-Instruct-Q2_K.gguf11.47 GBdownload
Qwen2.5-Coder-32B-Instruct-Q2_K_L.gguf12.18 GBdownload
Qwen2.5-Coder-32B-Instruct-Q3_K_L.gguf16.06 GBdownload
GGUF: Qwen/Qwen2.5-Coder-32B-Instruct-GGUF · 73,095 downloads
FileSize
qwen2.5-coder-32b-instruct-fp16-00001-of-00009.gguf7.29 GBdownload
qwen2.5-coder-32b-instruct-fp16-00002-of-00009.gguf7.27 GBdownload
qwen2.5-coder-32b-instruct-fp16-00003-of-00009.gguf7.27 GBdownload
qwen2.5-coder-32b-instruct-fp16-00004-of-00009.gguf7.27 GBdownload
qwen2.5-coder-32b-instruct-fp16-00005-of-00009.gguf7.27 GBdownload
qwen2.5-coder-32b-instruct-fp16-00006-of-00009.gguf7.27 GBdownload
qwen2.5-coder-32b-instruct-fp16-00007-of-00009.gguf7.27 GBdownload
qwen2.5-coder-32b-instruct-fp16-00008-of-00009.gguf7.27 GBdownload
qwen2.5-coder-32b-instruct-fp16-00009-of-00009.gguf2.89 GBdownload
qwen2.5-coder-32b-instruct-q2_k-00001-of-00002.gguf7.44 GBdownload
qwen2.5-coder-32b-instruct-q2_k-00002-of-00002.gguf4.03 GBdownload
qwen2.5-coder-32b-instruct-q2_k.gguf11.47 GBdownload

What you save self-hosting this model

Uses the model's own API price when it exists; otherwise the closest commercial equivalent.

Waiting for hardware detection…

RunLocal presents third-party leaderboard data (local.ai / Exo Labs, 2026-08) with its own fit and cost estimates. Verify pricing, licenses and file sizes at the source before deploying a model.