deepseek-v4-gguf

Try Try

deepseek-v4-gguf has 284B parameters (13.0B active per token) and takes 156 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

antirez/deepseek-v4-gguf
Repo createdApr 26, 2026updated Aug 31, 2026
Model typeMixture of expertsFull attention (MQA)
InputsText
Total parameters284B
Active per token13.0B4.6% of the model
Experts6 of 256 activeplus 1 shared, always on
Max context (from config)1M tokens
Layers42plus 1 extra prediction layer
On disk156 GB1 file
PrecisionMXFP4 (94%), Q8_0 (4.2%), FP16 (1.4%)
QuantizationMXFP4 (GGUF)mix of MXFP4, Q8_0 · blocks of 32 · 100% of parameters
Fewest GPUs1× NVIDIA B200a single 180 GB card · weights only
Licensemit
GitHubNot linked