Huihui-Qwen3.8-27B-abliterated-GGUF

Try Try

Huihui-Qwen3.8-27B-abliterated-GGUF has 27.3B parameters (25.6B active per token) and takes 16.5 GB on disk. See where its parameters live, layer by layer, and what hardware it needs.

At a glance

huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF
Repo createdAug 16, 2026updated Sep 18, 2026
Model typeDenseHybrid attention: 48 linear + 17 full (GQA)
InputsText + imagesvision encoder in a separate mmproj file, not counted here
Total parameters27.3B
Active per token25.6B94% of the model
ExpertsNone (dense)
Max context (from config)256K tokens
Layers6448 linear + 17 full attention · plus 1 extra prediction layer
On disk16.5 GB1 file
PrecisionQ4_K (32%), IQ4_XS (24%), Q5_K (23%), Q6_K (16%), IQ4_NL (2.0%), Q3_K (1.6%), other (1.5%)
QuantizationQ4_K_M (GGUF)mix of Q4_K, IQ4_XS, Q5_K +4 more · blocks of 256 / blocks of 32 · 100% of parameters
Fewest GPUs1× RTX 4090a single 24 GB card · weights only
Licenseapache-2.0
GitHubNot linked