gemma-4-26B-A4B-it-AWQ-4bit

Try Try

gemma-4-26B-A4B-it-AWQ-4bit has 25.8B parameters and takes 17.2 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

cyankiwi/gemma-4-26B-A4B-it-AWQ-4bit
Repo createdApr 3, 2026updated Sep 10, 2026
Model typeMixture of expertsFull attention (GQA)
InputsText + images573M vision encoder
Total parameters25.8B
Active per token0
ExpertsNone of 128 active
Max context (from config)256K tokens
Layers30
On disk17.2 GB4 files
PrecisionINT4 (70%), FP16 (22%), scales (8.7%)
QuantizationINT4 (compressed-tensors)groups of 32 · 93% of parameters · embeddings & output head, vision encoder and dense MLP kept in FP16
Fewest GPUs1× RTX 4090a single 24 GB card · weights only
Licenseapache-2.0
GitHubNot linked