DeepSeek

DeepSeek V4 Flash

DeepSeek V4 Flash is an open-weights 162B-parameter model by DeepSeek. It holds an Intelligence score of 77.3 on the local.ai leaderboard (2026-08 snapshot).

open weights
77.3
local.ai Intelligence
Parameters
162B
Benchmarks (local.ai snapshot)
τ²-bench
80%
GAIA
82%
GDPval
74%
HuggingFace metadata · updated 2026-06-22
repo
deepseek-ai/DeepSeek-V4-Flash
license
mit
downloads
2,102,039
likes
2,104

Metadata refreshed from the HuggingFace API (snapshot 2026-08-15T15:53:42.405Z). Context length and long-form descriptions are the next pipeline stage.

Does it fit your hardware?

Detect or select your GPU first. Green = weights + context headroom, yellow = tight, grey = does not fit.

Q4_K_M 94GQ6_K 145.8GQ8_0 178.2GFP16 324G
Quantization≈ File sizeNote
IQ2_XXS43.7 GBExtreme compression, largest quality loss
IQ3_XXS58.3 GBVery small
Q3_K_M74.5 GBSmall
NVFP481 GBNVIDIA 4-bit
Q4_K_S87.5 GBCommon
Q4_K_M94 GBMost common sweet spot
Q5_K_M123.1 GBBalanced
Q6_K145.8 GBNear-lossless
Q8_0178.2 GBNear-lossless
FP16324 GBOriginal weights

Actual files on HuggingFace

FileSize
model-00004-of-00046.safetensors3.35 GB
model-00046-of-00046.safetensors3.35 GB
model-00012-of-00046.safetensors3.34 GB
model-00014-of-00046.safetensors3.34 GB
model-00016-of-00046.safetensors3.34 GB
model-00018-of-00046.safetensors3.34 GB
GGUF: huihui-ai/Huihui-DeepSeek-V4-Flash-abliterated-ds4-GGUF · 332,460 downloads
FileSize
DeepSeek-V4-Flash-MTP-Q4K-Q8_0-F32.gguf3.55 GBdownload
Huihui-DeepSeek-V4-Flash-BF16-abliterated-ds4-Q2.gguf80.76 GBdownload
Huihui-DeepSeek-V4-Flash-BF16-abliterated-ds4-Q2_K.gguf92.86 GBdownload
Huihui-DeepSeek-V4-Flash-BF16-abliterated-ds4-Q4_K.gguf153.33 GBdownload
GGUF: unsloth/DeepSeek-V4-Flash-0731-GGUF · 306,443 downloads
FileSize
UD-IQ1_M/DeepSeek-V4-Flash-0731-UD-IQ1_M-00001-of-00003.gguf0.00 GBdownload
UD-IQ1_M/DeepSeek-V4-Flash-0731-UD-IQ1_M-00002-of-00003.gguf46.05 GBdownload
UD-IQ1_M/DeepSeek-V4-Flash-0731-UD-IQ1_M-00003-of-00003.gguf34.88 GBdownload
UD-IQ1_S/DeepSeek-V4-Flash-0731-UD-IQ1_S-00001-of-00003.gguf0.00 GBdownload
UD-IQ1_S/DeepSeek-V4-Flash-0731-UD-IQ1_S-00002-of-00003.gguf45.72 GBdownload
UD-IQ1_S/DeepSeek-V4-Flash-0731-UD-IQ1_S-00003-of-00003.gguf31.14 GBdownload
UD-IQ2_M/DeepSeek-V4-Flash-0731-UD-IQ2_M-00001-of-00003.gguf0.00 GBdownload
UD-IQ2_M/DeepSeek-V4-Flash-0731-UD-IQ2_M-00002-of-00003.gguf46.53 GBdownload
UD-IQ2_M/DeepSeek-V4-Flash-0731-UD-IQ2_M-00003-of-00003.gguf38.15 GBdownload
UD-IQ2_XXS/DeepSeek-V4-Flash-0731-UD-IQ2_XXS-00001-of-00003.gguf0.00 GBdownload
UD-IQ2_XXS/DeepSeek-V4-Flash-0731-UD-IQ2_XXS-00002-of-00003.gguf46.46 GBdownload
UD-IQ2_XXS/DeepSeek-V4-Flash-0731-UD-IQ2_XXS-00003-of-00003.gguf38.15 GBdownload

What you save self-hosting this model

Uses the model's own API price when it exists; otherwise the closest commercial equivalent.

Waiting for hardware detection…

RunLocal presents third-party leaderboard data (local.ai / Exo Labs, 2026-08) with its own fit and cost estimates. Verify pricing, licenses and file sizes at the source before deploying a model.