maple-preview-GGUF

Try Try

maple-preview-GGUF has 20.2B parameters (1.18B active per token) and takes 5.43 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

deepgrove/maple-preview-GGUF
Repo createdAug 6, 2026updated Aug 15, 2026
Model typeMixture of expertsFull attention (GQA)
InputsText
Total parameters20.2B
Active per token1.18B5.8% of the model
Experts8 of 256 active
Max context (from config)128K tokens
Layers24
On disk5.43 GB1 file
PrecisionTQ1_0 (76%), FP16 (23%), other (0.94%)
QuantizationF16 (GGUF)TQ1_0 · blocks of 256 · 97% of parameters · embeddings & output head kept in FP16
Fewest GPUs1× RTX 4060a single 8 GB card · weights only
Licensemit
GitHubNot linked