gemma-3-27b-it-int4-awq

Try Try

gemma-3-27b-it-int4-awq has 27.4B parameters (27.0B active per token) and takes 18.5 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

gaunernst/gemma-3-27b-it-int4-awq
Repo createdMar 21, 2025updated Apr 6, 2025
Model typeDenseFull attention (GQA)
InputsText + images423M vision encoder
Total parameters27.4B
Active per token27.0B98% of the model
ExpertsNone (dense)
Max context (from config)
Layers62
On disk18.5 GB4 files
PrecisionINT4 (69%), BF16 (20%), scales (11%)
QuantizationINT4 (AWQ)groups of 32 · 93% of parameters · embeddings & output head and vision encoder kept in BF16
Fewest GPUs1× RTX 4090a single 24 GB card · weights only
Licensegemma
GitHubNot linked